AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,979 results
Model Releases

BoostTaxo: Zero-Shot Taxonomy Induction via Boosting-Style Agentic Reasoning and Constraint-Aware Calibration

DGX agent

arXiv:2605.12520v1 Announce Type: cross Abstract: Taxonomy induction is crucial for organizing concepts into explicit and interpretable semantic hierarchies. While existing methods have achieved promi

model-releasesarxiv-cs-ai
14 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

LangSmith Engine is a phase shift because traces are no longer just records to be manually inspected, they’re now the catalyst for recursive…

DGX agent

LangSmith Engine is a phase shift because traces are no longer just records to be manually inspected, they’re now the catalyst for recursive agent self-improvement Engine looks at your traces, finds w

agentsharrison-chase--x
14 May 2026
Safety

Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers

DGX agent

arXiv:2605.11556v1 Announce Type: cross Abstract: Solving complex long-horizon tasks requires strong planning and reasoning capabilities. Although datasets with explicit chain-of-thought (CoT) rationa

safetyarxiv-cs-lg
13 May 2026
Agents

Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

DGX agent

arXiv:2605.12460v1 Announce Type: cross Abstract: The continued improvements in language model capability have unlocked their widespread use as drivers of autonomous agents, for example in coding or c

agentsarxiv-cs-cl
13 May 2026
Safety

AgentReview: Exploring Peer Review Dynamics with LLM Agents

DGX agent

arXiv:2406.12708v3 Announce Type: replace Abstract: Peer review is fundamental to the integrity and advancement of scientific publication. Traditional methods of peer review analyses often rely on exp

safetyarxiv-cs-cl
12 May 2026
Safety

Containment Verification: AI Safety Guarantees Independent of Alignment

DGX agent

arXiv:2605.09045v1 Announce Type: new Abstract: Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain

safetyarxiv-cs-ai
12 May 2026
Agents

Grok Voice is #1!

DGX agent

Grok Voice is #1! Announcing agentic performance benchmarking for Speech to Speech models on Artificial Analysis. We use 𝜏-Voice to measure tool calling and customer interaction voice agent capabiliti

agentselon-musk--x
12 May 2026
Model Releases

Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare

DGX agent

arXiv:2605.08445v1 Announce Type: new Abstract: AI models are increasingly deployed in live clinical environments where they must perform reliably across complex, high-stakes workflows that standard t

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

NARRA-Gym for Evaluating Interactive Narrative Agents

DGX agent

arXiv:2605.08503v1 Announce Type: new Abstract: Interactive narrative tasks require LLMs to sustain a coherent, evolving story while adapting to a user over multiple turns. However, suitable benchmark

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

DGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement

DGX agent

arXiv:2605.09730v1 Announce Type: new Abstract: Iterative self-refinement is a popular inference-time reliability technique, but its effectiveness in code-mode tool use depends heavily on the structur

model-releasesarxiv-cs-lg
12 May 2026
Agents

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models

DGX agent

arXiv:2409.13107v3 Announce Type: replace Abstract: Large language model-based (LLM) agents are emerging as a powerful enabler of robust embodied intelligence due to their capability of planning compl

agentsarxiv-cs-ro
12 May 2026
Safety

VISTA: A Generative Egocentric Video Framework for Daily Assistance

DGX agent

arXiv:2605.10579v1 Announce Type: new Abstract: Training AI agents to proactively assist humans in daily activities, from routine household tasks to urgent safety situations, requires large-scale visu

safetyarxiv-cs-cl
12 May 2026
Model Releases

Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents

DGX agent

arXiv:2605.07138v1 Announce Type: new Abstract: Reinforcement learning from verifiable emotion rewards RLVER has produced language models with strong empathetic performance, evaluated on benchmarks th

model-releasesarxiv-cs-ai
11 May 2026
Agents

Interactive Trajectory Planning with Learning-based Distributionally Robust Model Predictive Control and Markov Systems

DGX agent

arXiv:2605.07768v1 Announce Type: cross Abstract: We investigate interactive trajectory planning subject to uncertainty in the decisions of surrounding agents. To control the ego-agent, we aim to firs

agentsarxiv-cs-lg
11 May 2026
Model Releases

Interpreting Reinforcement Learning Agents with Susceptibilities

DGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Randomness is sometimes necessary for coordination

DGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

model-releasesarxiv-cs-ai
11 May 2026
Safety

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

DGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

safetyarxiv-cs-ai
11 May 2026
Agents

Maybe one of the only moats in 2026 is the context layer. AI improvements mean: ✅ UI/UX might simplify and consolidate. Instead of a lot of …

DGX agent

Maybe one of the only moats in 2026 is the context layer. AI improvements mean: ✅ UI/UX might simplify and consolidate. Instead of a lot of fancy buttons/knobs, you need simple, clean interfaces where

agentsjerry-liu--x
9 May 2026
Model Releases

Gemini 3.1 Flash-Lite is now generally available on Gemini Enterprise

DGX agent

Today, we’re thrilled to announce that Gemini 3.1 Flash-Lite, our fastest and most cost-efficient Gemini 3 series model yet, is now generally available. Designed for ultra-low latency, high-volume tas

model-releasesgoogle-cloud-ai
7 May 2026
Agents

ProgramBench: Can Language Models Rebuild Programs From Scratch?

DGX agent

arXiv:2605.03546v1 Announce Type: cross Abstract: Turning ideas into full software projects from scratch has become a popular use case for language models. Agents are being deployed to seed, maintain,

agentsarxiv-cs-ai
7 May 2026
Agents

SPHERE: Mitigating the Loss of Spectral Plasticity in Mixture-of-Experts for Deep Reinforcement Learning

DGX agent

arXiv:2605.04712v1 Announce Type: new Abstract: In deep reinforcement learning (DRL), an agent is trained from a stream of experience. In a continual learning setting, such agents can suffer from plas

agentsarxiv-cs-lg
7 May 2026
Agents

Stories like this is why we built E2B. Watching @genspark_ai go from zero to $250M ARR in 12 months on our infrastructure is one of the best…

DGX agent

Stories like this is why we built E2B. Watching @genspark_ai go from zero to 250M ARR in 12 months on our infrastructure is one of the best feelings in this job. Their Super Agent does deep research,

agentsyohei-nakajima--x
7 May 2026
Agents

arXiv Papers → LLM Artifacts This is how I keep up with AI research now. It's like having access to the most personalized arXiv feed. Automa…

DGX agent

arXiv Papers → LLM Artifacts This is how I keep up with AI research now. It's like having access to the most personalized arXiv feed. Automations run everyday to curate papers based a set of rules and

agentsdair-ai--x
6 May 2026
Model Releases

DataEvolver: Let Your Data Build and Improve Itself via Goal-Driven Loop Agents

DGX agent

arXiv:2605.01789v1 Announce Type: new Abstract: Constructing controllable visual data is a major bottleneck for image editing and multimodal understanding. Useful supervision is rarely produced by a s

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

LLM-ADAM: A Generalizable LLM Agent Framework for Pre-Print Anomaly Detection in Additive Manufacturing

DGX agent

arXiv:2605.03328v1 Announce Type: new Abstract: Additive manufacturing (AM) continues to transform modern manufacturing by enabling flexible, on-demand production of complex geometries across diverse

model-releasesarxiv-cs-lg
6 May 2026
Agents

Separating Intelligence from Execution: A Workflow Engine for the Model Context Protocol

DGX agent

arXiv:2605.00827v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly interact with external systems through tool-calling protocols such as the Model Context Protocol (MCP).

agentsarxiv-cs-ai
6 May 2026
Agents

Stop token maxxing 🛑 Our CEO @ashashutosh sat down with @a16z General Partner Peter Levine to discuss why we’re moving beyond the vector da…

DGX agent

Stop token maxxing 🛑 Our CEO @ashashutosh sat down with @a16z General Partner Peter Levine to discuss why we’re moving beyond the vector database: '85% of an agent’s work isn't the model; it's the und

agentspinecone--x
6 May 2026
Safety

T^2PO: Uncertainty-Guided Exploration Control for Stable Multi-Turn Agentic Reinforcement Learning

DGX agent

arXiv:2605.02178v1 Announce Type: new Abstract: Recent progress in multi-turn reinforcement learning (RL) has significantly improved reasoning LLMs' performances on complex interactive tasks. Despite

safetyarxiv-cs-ai
6 May 2026
Agents

Networked Information Aggregation for Binary Classification

DGX agent

arXiv:2605.01082v1 Announce Type: new Abstract: We study networked binary classification on a directed acyclic graph (DAG) where each agent observes only a subset of the feature columns of a shared da

agentsarxiv-cs-lg
5 May 2026
Applications

Autoformalizing Memory Specifications with Agents

DGX agent

arXiv:2605.00058v1 Announce Type: cross Abstract: The primary goal of Design Verification (DV) is to ensure that a proposed chip design implementation (either in code, or physical form) exactly matche

applicationsarxiv-cs-lg
4 May 2026
Agents

cofounder 2 fan art

DGX agent

cofounder 2 fan art Media Announcing Cofounder 2: Run an entire company with agents. It's the infrastructure for the one person billion dollar company - orchestrating agents across engineering, sales,

agentsyohei-nakajima--x
4 May 2026
Agents

Use open models in Fleet!

DGX agent

Use open models in Fleet! Not every step in an agent workflow needs the same model. Fleet now lets you customize which model each sub-agent uses, so you can route simple tasks to fast/cheap models and

agentsharrison-chase--x
4 May 2026
Agents

Waitlist is OFF go try Cofounder RIGHT NOW 🌻

DGX agent

Waitlist is OFF go try Cofounder RIGHT NOW 🌻 Announcing Cofounder 2: Run an entire company with agents. It's the infrastructure for the one person billion dollar company - orchestrating agents across

agentsyohei-nakajima--x
4 May 2026
Local Ai

Enhancing Linux Privilege Escalation Attack Capabilities of Local LLM Agents

DGX agent

arXiv:2604.27143v1 Announce Type: cross Abstract: Recent research has demonstrated the potential of Large Language Models (LLMs) for autonomous penetration testing, particularly when using cloud-based

local-aiarxiv-cs-ai
1 May 2026
Research

Learning to Aggregate Zero-Shot LLM Agents for Corporate Disclosure Classification

DGX agent

arXiv:2603.20965v2 Announce Type: replace-cross Abstract: This paper studies whether a lightweight supervised aggregator can combine diverse zero-shot large language model outputs into a stronger down

researcharxiv-cs-ai
1 May 2026
Agents

@Shopify https://github.com/NousResearch/hermes-agent/blob/main/optional-skills/productivity/shopify/SKILL.md

DGX agent

The Shopify skill for Hermes Agent enables integration with Shopify e-commerce platforms, likely allowing automated interactions with store management, product listings, orders, and customer data thro

agentsnous-research--x
1 May 2026
Agents

Shopify is the all-in-one commerce platform powering millions of businesses worldwide Thank you to the @Shopify team for building their own …

DGX agent

Shopify is the all-in-one commerce platform powering millions of businesses worldwide Thank you to the @Shopify team for building their own official Hermes Agent skill enabling your agent to manage pr

agentsnous-research--x
1 May 2026
Agents

So good

DGX agent

So good Shopify is the all-in-one commerce platform powering millions of businesses worldwide Thank you to the @Shopify team for building their own official Hermes Agent skill enabling your agent to m

agentsnous-research--x
1 May 2026
Agents

Synthetic Computers at Scale for Long-Horizon Productivity Simulation

DGX agent

arXiv:2604.28181v1 Announce Type: new Abstract: Realistic long-horizon productivity work is strongly conditioned on user-specific computer environments, where much of the work context is stored and or

agentsarxiv-cs-ai
1 May 2026
Hardware

Automating GPU Kernel Translation with AI Agents: cuTile Python to cuTile.jl

DGX agent

The TileGym project developed an AI-driven skill-based workflow that encodes 17 critical translation rules, static validation scripts, and example kernels, enabling automated conversion of cuTile Pyth

hardwarenvidia-developer
30 Apr 2026
Model Releases

Learning to Ask: When LLM Agents Meet Unclear Instruction

DGX agent

arXiv:2409.00557v4 Announce Type: replace-cross Abstract: Equipped with the capability to call functions, modern large language models (LLMs) can leverage external tools for addressing a range of task

model-releasesarxiv-cs-ai
30 Apr 2026
Agents

Fragmented data is stalling enterprise AI deployments before they ever ship

DGX agent

Enterprises are pushing past AI experiments and demanding production-grade deployments, but fragmented data and poorly scoped agents are stalling progress. The question is no longer whether AI agents

agentssiliconangle
29 Apr 2026
Tutorials

Organizing Agents’ memory at scale: Namespace design patterns in AgentCore Memory

DGX agent

In this post, you will learn how to design namespace hierarchies, choose the right retrieval patterns, and implement AWS Identity and Access Management (IAM)-based access control for AgentCore Memory.

tutorialsaws-ml-blog
29 Apr 2026
Research

The Download: storing nuclear waste and orchestrating agents

DGX agent

This is today’s edition of The Download, our weekday newsletter that provides a daily dose of what’s going on in the world of technology. It’s time to make a plan for nuclear waste Today, nuclear ener

researchmit-tech-review
29 Apr 2026
Model Releases

ClimAgent: LLM as Agents for Autonomous Open-ended Climate Science Analysis

DGX agent

arXiv:2604.16922v2 Announce Type: replace Abstract: Climate research is pivotal for mitigating global environmental crises, yet the accelerating volume of multi-scale datasets and the complexity of an

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Day 1 of #AIDevSF is live. 🙌 📍 Find us at booth 121 🎤 Catch our Chief Product Officer Or Dagan's session on Day 2: Efficiently Navigating…

DGX agent

Day 1 of #AIDevSF is live. 🙌 📍 Find us at booth 121 🎤 Catch our Chief Product Officer Or Dagan's session on Day 2: Efficiently Navigating the Agentic Action Space: Meta-Model Orchestration using AI21

agentsai21-labs--x
28 Apr 2026
Agents

In 2024, a high schooler cold DM'd us asking if he could work at Cognition before college. We said yes. We ended up helping him with his col…

DGX agent

In 2024, a high schooler cold DM'd us asking if he could work at Cognition before college. We said yes. We ended up helping him with his college essays and his US visa. If you've used DeepWiki or Devi

agentscognition-ai--x
28 Apr 2026
← Previous
1…191192193194195…375
Next →