AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,973 results
28 May 2026

Exploratory Experience Shapes the Geometry of Predictive Representations

Model ReleasesDGX agent

arXiv:2605.27929v1 Announce Type: cross Abstract: Active sensing links behavior and learning through an action-perception loop: actions determine the observations used to update internal predictive mo

Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis

AgentsDGX agent

arXiv:2605.28037v1 Announce Type: new Abstract: Prompt-based personality control is a key technique for designing large language model (LLM) dialogue agents that behave consistently across social cont

The Decision to Verify: How Warmth and User Characteristics Shape Reliance on Conversational Agents for Information Search

ResearchDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2605.28498v1 Announce Type: cross Abstract: Conversational artificial intelligence (AI) provides an efficient and convenient gateway to information access. However, it can cause overreliance whe

27 May 2026

AI Factories: The New Infrastructure of Intelligence

AgentsDGX agent

AI factories are token factories, converting power into intelligence in real time. And as agentic AI scales and autonomous, always-on special agents are deployed in the enterprise, performance per wat

26 May 2026

Adaptive Punishment for Cooperation in Mixed-Motive Games

AgentsDGX agent

arXiv:2605.24516v1 Announce Type: cross Abstract: Mixed-motive scenarios are ubiquitous in real-world multi-agent interactions, where self-interested agents often defect for immediate rewards, overloo

BlitzRank: Principled Zero-shot Ranking Agents with Tournament Graphs

ApplicationsDGX agent

arXiv:2602.05448v4 Announce Type: replace Abstract: Selecting the top m from n items via expensive k-wise comparisons is central to settings ranging from LLM-based document reranking to crowdsourced e

Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention

Local AiDGX agent

arXiv:2605.24999v1 Announce Type: cross Abstract: An agent must act on the situation before it, learn what it cannot yet represent, and model other agents well enough to coordinate. These faculties ar

KYA: A Framework-Agnostic Trust Layer for Autonomous Systems with Verifiable Provenance and Hierarchical Policy Composition

SafetyDGX agent

arXiv:2605.25376v1 Announce Type: cross Abstract: Observability tells operators when an agent is slow. KYA tells operators when an agent is wrong, drifting, leaking, or quietly going rogue. We present

Mixture of Complementary Agents for Robust LLM Ensemble

ResearchDGX agent

arXiv:2605.24048v1 Announce Type: cross Abstract: Multi-AI collaboration, such as ensembling or debating large language models (LLMs), is a promising paradigm for aggregating information and boosting

Multi-Level Strategic Classification: Incentivizing Improvement through Promotion and Relegation Dynamics

AgentsDGX agent

arXiv:2602.11439v2 Announce Type: replace Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outc

Personalize-then-Store: Benchmarking and Learning Personalized Memory for Long-horizon Agents

Model ReleasesDGX agent

arXiv:2605.25535v1 Announce Type: new Abstract: Existing large language model (LLM) based memory systems apply universal, static policies that overlook a fundamental reality: the contexts that are wor

Structure-Aware RAG: Structured Retrieval Augmented Generation from Noisy Data for Conversational Agents

ApplicationsDGX agent

arXiv:2605.24366v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely adopted in conversational applications. However, their reliance on parametric knowledge limits reliability

25 May 2026

Automatic Construction of Clinical Scoring Systems with LLM Agents

ResearchDGX agent

arXiv:2601.22324v2 Announce Type: replace Abstract: Modern clinical practice relies on evidence-based guidelines implemented as compact scoring systems composed of a small number of interpretable deci

22 May 2026

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

AgentsDGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

Echo: Learning from Experience Data via User-Driven Refinement

AgentsDGX agent

arXiv:2605.21984v1 Announce Type: cross Abstract: Static 'human data' faces inherent limitations: it is expensive to scale and bounded by the knowledge of its creators. Continuous learning from 'exper

Grok progress

AgentsDGX agent

Grok progress Grok Imagine Agent Mode is now available on Grok iOS app With Agent Mode, you can generate: • Consistent characters across generations • Multiple scenes with the same character • Differe

Open-source AI is pulling Dell’s entire portfolio into play

AgentsDGX agent

Open-source AI agents are changing software development, empowering both developers and citizen builders alike in a domain once reserved for professional engineers. But the influx of agents now spread

SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents

ResearchDGX agent

arXiv:2603.08403v3 Announce Type: replace Abstract: Long-horizon action-conditioned video generation aims to synthesize temporally coherent videos that follow complex action instructions over extended

21 May 2026

He's right. And it has a direct implication for code review that nobody's talking about. Think about why AI cracked coding before almost eve…

AgentsDGX agent

He's right. And it has a direct implication for code review that nobody's talking about. Think about why AI cracked coding before almost everything else. Not because code is simple. Because code is ch

Intelligent radiology workflow optimization with AI agents

ApplicationsDGX agent

Many healthcare organizations report that traditional worklist systems rely on rigid rules that ignore critical context, radiologist specialization, current workload, fatigue levels, and case complexi

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

AgentsDGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

20 May 2026

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

Model ReleasesDGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents

Model ReleasesDGX agent

arXiv:2605.19328v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) facilitate a new class of embodied AI systems, where these models are integrated into physical platfo

Stop Drawing Scientific Claims from LLM Social Simulations Without Robustness Audits

AgentsDGX agent

arXiv:2605.18890v1 Announce Type: cross Abstract: The scientific claims drawn from LLM social simulations should be no stronger than the robustness audits that support them. Generative agents bring ne

19 May 2026

An Efficient Streaming Video Understanding Framework with Agentic Control

SafetyDGX agent

arXiv:2605.17921v1 Announce Type: new Abstract: Streaming video requires handling dynamic information density under strict latency budgets. Yet, existing methods typically employ static strategies, su

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations?

SafetyDGX agent

arXiv:2605.17694v1 Announce Type: new Abstract: Power differences shape human communication through well documented socio cognitive effects, including language coordination, pronoun usage, authority b

Join us as a Forward Deployed Engineer!

AgentsDGX agent

Join us as a Forward Deployed Engineer! looking for someone to help us bring Hermes agent to the world. if you're a hermes agent power user and want to help others become one as well, hit my dms or em

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization

SafetyDGX agent

arXiv:2605.17877v1 Announce Type: new Abstract: A significant hurdle for current LLMs is the execution of complex, multi-stage tasks. Group Relative Policy Optimization (GRPO) has been emerging as a l

Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents

Local AiDGX agent

arXiv:2605.16671v1 Announce Type: new Abstract: Rapid biodiversity loss underscore the urgency of effective monitoring, yet manual surveys remain resource-intensive. While on-device AI offers a scalab

Tongyi DeepResearch Technical Report

AgentsDGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

18 May 2026

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming

AgentsDGX agent

arXiv:2605.15400v1 Announce Type: new Abstract: While AI agents are rapidly advancing from isolated tools to interactive collaborators, data-driven human-machine teaming (HMT) methods remain costly in

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

AgentsDGX agent

arXiv:2605.15665v1 Announce Type: new Abstract: Deploying large language model (LLM)-driven conversational agents in enterprise settings requires prompts that are simultaneously correct at launch and

Vera Arrives: NVIDIA’s First CPU Built for Agents Lands at Top AI Labs

HardwareDGX agent

The first NVIDIA Vera CPUs arrived at three of the world's leading AI labs on Friday — Anthropic in San Francisco, OpenAI in Mission Bay, SpaceXAI in Palo Alto — followed by a delivery to Oracle Cloud

15 May 2026

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents

Model ReleasesDGX agent

arXiv:2605.14153v1 Announce Type: cross Abstract: Exploitation is not a binary event. It is a ladder of acquiring progressive capabilities, from executing a single buggy line of code to taking full co

Key ideas I picked up at the @LangChain conference this week: 'AI engineering is data science. Look at your data.' — @sh_reya & @HamelHusain…

AgentsDGX agent

Key ideas I picked up at the @LangChain conference this week: 'AI engineering is data science. Look at your data.' — @sh_reya & @HamelHusain 'Data strategy *before* agents.' — @AndrewYNg 'Coding agent

14 May 2026

Beyond Individual Mimicry: Constructing Human-Like Social network with Graph-Augmented LLM Agents

Local AiDGX agent

arXiv:2605.12512v1 Announce Type: cross Abstract: Driven by large language models (LLMs), social bot can autonomously engage in local interactions, whose human-like behaviors enable them to evade soci

BoostTaxo: Zero-Shot Taxonomy Induction via Boosting-Style Agentic Reasoning and Constraint-Aware Calibration

Model ReleasesDGX agent

arXiv:2605.12520v1 Announce Type: cross Abstract: Taxonomy induction is crucial for organizing concepts into explicit and interpretable semantic hierarchies. While existing methods have achieved promi

LangSmith Engine is a phase shift because traces are no longer just records to be manually inspected, they’re now the catalyst for recursive…

AgentsDGX agent

LangSmith Engine is a phase shift because traces are no longer just records to be manually inspected, they’re now the catalyst for recursive agent self-improvement Engine looks at your traces, finds w

13 May 2026

Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers

SafetyDGX agent

arXiv:2605.11556v1 Announce Type: cross Abstract: Solving complex long-horizon tasks requires strong planning and reasoning capabilities. Although datasets with explicit chain-of-thought (CoT) rationa

Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

AgentsDGX agent

arXiv:2605.12460v1 Announce Type: cross Abstract: The continued improvements in language model capability have unlocked their widespread use as drivers of autonomous agents, for example in coding or c

12 May 2026

AgentReview: Exploring Peer Review Dynamics with LLM Agents

SafetyDGX agent

arXiv:2406.12708v3 Announce Type: replace Abstract: Peer review is fundamental to the integrity and advancement of scientific publication. Traditional methods of peer review analyses often rely on exp

Containment Verification: AI Safety Guarantees Independent of Alignment

SafetyDGX agent

arXiv:2605.09045v1 Announce Type: new Abstract: Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain

Grok Voice is #1!

AgentsDGX agent

Grok Voice is #1! Announcing agentic performance benchmarking for Speech to Speech models on Artificial Analysis. We use 𝜏-Voice to measure tool calling and customer interaction voice agent capabiliti

Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare

Model ReleasesDGX agent

arXiv:2605.08445v1 Announce Type: new Abstract: AI models are increasingly deployed in live clinical environments where they must perform reliably across complex, high-stakes workflows that standard t

NARRA-Gym for Evaluating Interactive Narrative Agents

Model ReleasesDGX agent

arXiv:2605.08503v1 Announce Type: new Abstract: Interactive narrative tasks require LLMs to sustain a coherent, evolving story while adapting to a user over multiple turns. However, suitable benchmark

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

Model ReleasesDGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement

Model ReleasesDGX agent

arXiv:2605.09730v1 Announce Type: new Abstract: Iterative self-refinement is a popular inference-time reliability technique, but its effectiveness in code-mode tool use depends heavily on the structur

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models

AgentsDGX agent

arXiv:2409.13107v3 Announce Type: replace Abstract: Large language model-based (LLM) agents are emerging as a powerful enabler of robust embodied intelligence due to their capability of planning compl

VISTA: A Generative Egocentric Video Framework for Daily Assistance

SafetyDGX agent

arXiv:2605.10579v1 Announce Type: new Abstract: Training AI agents to proactively assist humans in daily activities, from routine household tasks to urgent safety situations, requires large-scale visu

11 May 2026

Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents

Model ReleasesDGX agent

arXiv:2605.07138v1 Announce Type: new Abstract: Reinforcement learning from verifiable emotion rewards RLVER has produced language models with strong empathetic performance, evaluated on benchmarks th

Interactive Trajectory Planning with Learning-based Distributionally Robust Model Predictive Control and Markov Systems

AgentsDGX agent

arXiv:2605.07768v1 Announce Type: cross Abstract: We investigate interactive trajectory planning subject to uncertainty in the decisions of surrounding agents. To control the ego-agent, we aim to firs

Interpreting Reinforcement Learning Agents with Susceptibilities

Model ReleasesDGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

Randomness is sometimes necessary for coordination

Model ReleasesDGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

The Endogeneity of Miscalibration: Impossibility and Escape in Scored Reporting

SafetyDGX agent

arXiv:2605.07671v1 Announce Type: cross Abstract: Eliciting truthful reports from autonomous agents is a core problem in scalable AI oversight: a principal scores the agent's report using a strictly p

9 May 2026

Maybe one of the only moats in 2026 is the context layer. AI improvements mean: ✅ UI/UX might simplify and consolidate. Instead of a lot of …

AgentsDGX agent

Maybe one of the only moats in 2026 is the context layer. AI improvements mean: ✅ UI/UX might simplify and consolidate. Instead of a lot of fancy buttons/knobs, you need simple, clean interfaces where

7 May 2026

Gemini 3.1 Flash-Lite is now generally available on Gemini Enterprise

Model ReleasesDGX agent

Today, we’re thrilled to announce that Gemini 3.1 Flash-Lite, our fastest and most cost-efficient Gemini 3 series model yet, is now generally available. Designed for ultra-low latency, high-volume tas

ProgramBench: Can Language Models Rebuild Programs From Scratch?

AgentsDGX agent

arXiv:2605.03546v1 Announce Type: cross Abstract: Turning ideas into full software projects from scratch has become a popular use case for language models. Agents are being deployed to seed, maintain,

SPHERE: Mitigating the Loss of Spectral Plasticity in Mixture-of-Experts for Deep Reinforcement Learning

AgentsDGX agent

arXiv:2605.04712v1 Announce Type: new Abstract: In deep reinforcement learning (DRL), an agent is trained from a stream of experience. In a continual learning setting, such agents can suffer from plas

Stories like this is why we built E2B. Watching @genspark_ai go from zero to $250M ARR in 12 months on our infrastructure is one of the best…

AgentsDGX agent

Stories like this is why we built E2B. Watching @genspark_ai go from zero to 250M ARR in 12 months on our infrastructure is one of the best feelings in this job. Their Super Agent does deep research,

6 May 2026

arXiv Papers → LLM Artifacts This is how I keep up with AI research now. It's like having access to the most personalized arXiv feed. Automa…

AgentsDGX agent

arXiv Papers → LLM Artifacts This is how I keep up with AI research now. It's like having access to the most personalized arXiv feed. Automations run everyday to curate papers based a set of rules and

← Previous
1…152153154155156…300
Next →