AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Choosing the Lens: Strategic Perspective Activation in Context-Dependent Argumentation

DGX agent

arXiv:2605.31581v1 Announce Type: new Abstract: The same arguments often need to be evaluated under different external regimes. An agent with influence over the regime has a strategic lever that stand

agentsarxiv-cs-ai
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

COLLEAGUE.SKILL: Automated AI Skill Generation via Expert Knowledge Distillation

DGX agent

arXiv:2605.31264v1 Announce Type: new Abstract: LLM agents are increasingly expected not only to complete isolated tasks, but also to carry bounded representations of human expertise, judgment, and in

agentsarxiv-cs-ai
1 Jun 2026
Agents

S^3LDBO: A Snapshot Single-Loop Algorithm for Decentralized Bilevel Optimization

DGX agent

arXiv:2605.31311v1 Announce Type: cross Abstract: Networked AI systems increasingly rely on multiple agents that collaboratively learn and adapt models over communication networks. In such systems, bi

agentsarxiv-cs-lg
1 Jun 2026
Model Releases

CodeEvolve: an open source evolutionary coding agent for algorithmic discovery and optimization

DGX agent

arXiv:2510.14150v5 Announce Type: replace Abstract: We introduce CodeEvolve, an open-source framework that couples large language models with island-based evolutionary search for end-to-end algorithmi

model-releasesarxiv-cs-ai
29 May 2026
Applications

Evaluation of Conversational Agents: Understanding Culture, Context and Environment in Emotion Detection

DGX agent

arXiv:2605.30099v1 Announce Type: new Abstract: Valuable decisions and highly prioritized analysis now depend on applications such as facial biometrics, social media photo tagging, and human robots in

applicationsarxiv-cs-cv
29 May 2026
Model Releases

Exploratory Experience Shapes the Geometry of Predictive Representations

DGX agent

arXiv:2605.27929v1 Announce Type: cross Abstract: Active sensing links behavior and learning through an action-perception loop: actions determine the observations used to update internal predictive mo

model-releasesarxiv-cs-lg
28 May 2026
Agents

Personality, Role, and Expressive Style in Large Language Models: An Interactionist Analysis

DGX agent

arXiv:2605.28037v1 Announce Type: new Abstract: Prompt-based personality control is a key technique for designing large language model (LLM) dialogue agents that behave consistently across social cont

agentsarxiv-cs-cl
28 May 2026
Research

The Decision to Verify: How Warmth and User Characteristics Shape Reliance on Conversational Agents for Information Search

DGX agent

arXiv:2605.28498v1 Announce Type: cross Abstract: Conversational artificial intelligence (AI) provides an efficient and convenient gateway to information access. However, it can cause overreliance whe

researcharxiv-cs-ai
28 May 2026
Agents

Adaptive Punishment for Cooperation in Mixed-Motive Games

DGX agent

arXiv:2605.24516v1 Announce Type: cross Abstract: Mixed-motive scenarios are ubiquitous in real-world multi-agent interactions, where self-interested agents often defect for immediate rewards, overloo

agentsarxiv-cs-ai
26 May 2026
Applications

BlitzRank: Principled Zero-shot Ranking Agents with Tournament Graphs

DGX agent

arXiv:2602.05448v4 Announce Type: replace Abstract: Selecting the top m from n items via expensive k-wise comparisons is central to settings ranging from LLM-based document reranking to crowdsourced e

applicationsarxiv-cs-lg
26 May 2026
Local Ai

Interpretation, Learning, and Empathy as One Constraint: A Residual-Adequacy Architecture with Accountable Abstention

DGX agent

arXiv:2605.24999v1 Announce Type: cross Abstract: An agent must act on the situation before it, learn what it cannot yet represent, and model other agents well enough to coordinate. These faculties ar

local-aiarxiv-cs-ai
26 May 2026
Safety

KYA: A Framework-Agnostic Trust Layer for Autonomous Systems with Verifiable Provenance and Hierarchical Policy Composition

DGX agent

arXiv:2605.25376v1 Announce Type: cross Abstract: Observability tells operators when an agent is slow. KYA tells operators when an agent is wrong, drifting, leaking, or quietly going rogue. We present

safetyarxiv-cs-ai
26 May 2026
Research

Mixture of Complementary Agents for Robust LLM Ensemble

DGX agent

arXiv:2605.24048v1 Announce Type: cross Abstract: Multi-AI collaboration, such as ensembling or debating large language models (LLMs), is a promising paradigm for aggregating information and boosting

researcharxiv-cs-ai
26 May 2026
Agents

Multi-Level Strategic Classification: Incentivizing Improvement through Promotion and Relegation Dynamics

DGX agent

arXiv:2602.11439v2 Announce Type: replace Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outc

agentsarxiv-cs-lg
26 May 2026
Model Releases

Personalize-then-Store: Benchmarking and Learning Personalized Memory for Long-horizon Agents

DGX agent

arXiv:2605.25535v1 Announce Type: new Abstract: Existing large language model (LLM) based memory systems apply universal, static policies that overlook a fundamental reality: the contexts that are wor

model-releasesarxiv-cs-ai
26 May 2026
Applications

Structure-Aware RAG: Structured Retrieval Augmented Generation from Noisy Data for Conversational Agents

DGX agent

arXiv:2605.24366v1 Announce Type: new Abstract: Large Language Models (LLMs) have been widely adopted in conversational applications. However, their reliance on parametric knowledge limits reliability

applicationsarxiv-cs-cl
26 May 2026
Research

Automatic Construction of Clinical Scoring Systems with LLM Agents

DGX agent

arXiv:2601.22324v2 Announce Type: replace Abstract: Modern clinical practice relies on evidence-based guidelines implemented as compact scoring systems composed of a small number of interpretable deci

researcharxiv-cs-lg
25 May 2026
Agents

AutoRPA: Efficient GUI Automation through LLM-Driven Code Synthesis from Interactions

DGX agent

arXiv:2605.21082v1 Announce Type: new Abstract: Large Language Model (LLM) based agents have demonstrated proficiency in multi-step interactions with graphical user interfaces (GUIs). While most resea

agentsarxiv-cs-ai
22 May 2026
Agents

Echo: Learning from Experience Data via User-Driven Refinement

DGX agent

arXiv:2605.21984v1 Announce Type: cross Abstract: Static 'human data' faces inherent limitations: it is expensive to scale and bounded by the knowledge of its creators. Continuous learning from 'exper

agentsarxiv-cs-cl
22 May 2026
Research

SPIRAL: Self-Evolving Action-Conditioned Video Generation via Reflective Planning Agents

DGX agent

arXiv:2603.08403v3 Announce Type: replace Abstract: Long-horizon action-conditioned video generation aims to synthesize temporally coherent videos that follow complex action instructions over extended

researcharxiv-cs-cv
22 May 2026
Agents

Smaller Abstract State Spaces Enable Cross-Scale Generalization in Reinforcement Learning

DGX agent

arXiv:2605.20272v1 Announce Type: new Abstract: While humans readily generalize abstract concepts to more complex or larger tasks, building Reinforcement Learning (RL) systems with this ability remain

agentsarxiv-cs-lg
21 May 2026
Model Releases

ReacTOD: Bounded Neuro-Symbolic Agentic NLU for Zero-Shot Dialogue State Tracking

DGX agent

arXiv:2605.19077v1 Announce Type: cross Abstract: Task-oriented dialogue systems -- handling transactions, reservations, and service requests -- require predictable behavior, yet the moderately-sized

model-releasesarxiv-cs-ai
20 May 2026
Model Releases

RoboJailBench: Benchmarking Adversarial Attacks and Defenses in Embodied Robotic Agents

DGX agent

arXiv:2605.19328v1 Announce Type: cross Abstract: Recent advances in Vision-Language Models (VLMs) facilitate a new class of embodied AI systems, where these models are integrated into physical platfo

model-releasesarxiv-cs-ro
20 May 2026
Agents

Stop Drawing Scientific Claims from LLM Social Simulations Without Robustness Audits

DGX agent

arXiv:2605.18890v1 Announce Type: cross Abstract: The scientific claims drawn from LLM social simulations should be no stronger than the robustness audits that support them. Generative agents bring ne

agentsarxiv-cs-ai
20 May 2026
Safety

An Efficient Streaming Video Understanding Framework with Agentic Control

DGX agent

arXiv:2605.17921v1 Announce Type: new Abstract: Streaming video requires handling dynamic information density under strict latency budgets. Yet, existing methods typically employ static strategies, su

safetyarxiv-cs-cv
19 May 2026
Safety

Do LLM Agents Mirror Socio-Cognitive Effects in Power-Asymmetric Conversations?

DGX agent

arXiv:2605.17694v1 Announce Type: new Abstract: Power differences shape human communication through well documented socio cognitive effects, including language coordination, pronoun usage, authority b

safetyarxiv-cs-cl
19 May 2026
Safety

PAIR: Prefix-Aware Internal Reward Model for Multi-Turn Agent Optimization

DGX agent

arXiv:2605.17877v1 Announce Type: new Abstract: A significant hurdle for current LLMs is the execution of complex, multi-stage tasks. Group Relative Policy Optimization (GRPO) has been emerging as a l

safetyarxiv-cs-ai
19 May 2026
Local Ai

Sustainable Intelligence for the Wild: Democratizing Ecological Monitoring via Knowledge-Adaptive Edge Expert Agents

DGX agent

arXiv:2605.16671v1 Announce Type: new Abstract: Rapid biodiversity loss underscore the urgency of effective monitoring, yet manual surveys remain resource-intensive. While on-device AI offers a scalab

local-aiarxiv-cs-ai
19 May 2026
Agents

Tongyi DeepResearch Technical Report

DGX agent

arXiv:2510.24701v3 Announce Type: replace-cross Abstract: We present Tongyi DeepResearch, an agentic large language model, which is specifically designed for long-horizon, deep information-seeking res

agentsarxiv-cs-ai
19 May 2026
Agents

Beyond Partner Diversity: An Influence-Based Team Steering Framework for Zero-Shot Human-Machine Teaming

DGX agent

arXiv:2605.15400v1 Announce Type: new Abstract: While AI agents are rapidly advancing from isolated tools to interactive collaborators, data-driven human-machine teaming (HMT) methods remain costly in

agentsarxiv-cs-ai
18 May 2026
Agents

PRISM: Prompt Reliability via Iterative Simulation and Monitoring for Enterprise Conversational AI

DGX agent

arXiv:2605.15665v1 Announce Type: new Abstract: Deploying large language model (LLM)-driven conversational agents in enterprise settings requires prompts that are simultaneously correct at launch and

agentsarxiv-cs-ai
18 May 2026
Model Releases

ExploitBench: A Capability Ladder Benchmark for LLM Cybersecurity Agents

DGX agent

arXiv:2605.14153v1 Announce Type: cross Abstract: Exploitation is not a binary event. It is a ladder of acquiring progressive capabilities, from executing a single buggy line of code to taking full co

model-releasesarxiv-cs-ai
15 May 2026
Local Ai

Beyond Individual Mimicry: Constructing Human-Like Social network with Graph-Augmented LLM Agents

DGX agent

arXiv:2605.12512v1 Announce Type: cross Abstract: Driven by large language models (LLMs), social bot can autonomously engage in local interactions, whose human-like behaviors enable them to evade soci

local-aiarxiv-cs-ai
14 May 2026
Model Releases

BoostTaxo: Zero-Shot Taxonomy Induction via Boosting-Style Agentic Reasoning and Constraint-Aware Calibration

DGX agent

arXiv:2605.12520v1 Announce Type: cross Abstract: Taxonomy induction is crucial for organizing concepts into explicit and interpretable semantic hierarchies. While existing methods have achieved promi

model-releasesarxiv-cs-ai
14 May 2026
Safety

Hindsight Hint Distillation: Scaffolded Reasoning for SWE Agents from CoT-free Answers

DGX agent

arXiv:2605.11556v1 Announce Type: cross Abstract: Solving complex long-horizon tasks requires strong planning and reasoning capabilities. Although datasets with explicit chain-of-thought (CoT) rationa

safetyarxiv-cs-lg
13 May 2026
Agents

Multi-Stream LLMs: Unblocking Language Models with Parallel Streams of Thoughts, Inputs and Outputs

DGX agent

arXiv:2605.12460v1 Announce Type: cross Abstract: The continued improvements in language model capability have unlocked their widespread use as drivers of autonomous agents, for example in coding or c

agentsarxiv-cs-cl
13 May 2026
Safety

AgentReview: Exploring Peer Review Dynamics with LLM Agents

DGX agent

arXiv:2406.12708v3 Announce Type: replace Abstract: Peer review is fundamental to the integrity and advancement of scientific publication. Traditional methods of peer review analyses often rely on exp

safetyarxiv-cs-cl
12 May 2026
Safety

Containment Verification: AI Safety Guarantees Independent of Alignment

DGX agent

arXiv:2605.09045v1 Announce Type: new Abstract: Agentic frameworks are the software layer through which AI agents act in the world. Existing safety methods intervene on the model and therefore remain

safetyarxiv-cs-ai
12 May 2026
Model Releases

Measuring What Matters: Benchmarking Generative, Multimodal, and Agentic AI in Healthcare

DGX agent

arXiv:2605.08445v1 Announce Type: new Abstract: AI models are increasingly deployed in live clinical environments where they must perform reliably across complex, high-stakes workflows that standard t

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

NARRA-Gym for Evaluating Interactive Narrative Agents

DGX agent

arXiv:2605.08503v1 Announce Type: new Abstract: Interactive narrative tasks require LLMs to sustain a coherent, evolving story while adapting to a user over multiple turns. However, suitable benchmark

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

REI-Bench: Can Embodied Agents Understand Vague Human Instructions in Task Planning?

DGX agent

arXiv:2505.10872v4 Announce Type: replace-cross Abstract: Robot task planning decomposes human instructions into executable action sequences that enable robots to complete a series of complex tasks. A

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RubricRefine: Improving Tool-Use Agent Reliability with Training-Free Pre-Execution Refinement

DGX agent

arXiv:2605.09730v1 Announce Type: new Abstract: Iterative self-refinement is a popular inference-time reliability technique, but its effectiveness in code-mode tool use depends heavily on the structur

model-releasesarxiv-cs-lg
12 May 2026
Agents

Towards Robust Surgical Automation via Digital Twin Representations from Foundation Models

DGX agent

arXiv:2409.13107v3 Announce Type: replace Abstract: Large language model-based (LLM) agents are emerging as a powerful enabler of robust embodied intelligence due to their capability of planning compl

agentsarxiv-cs-ro
12 May 2026
Safety

VISTA: A Generative Egocentric Video Framework for Daily Assistance

DGX agent

arXiv:2605.10579v1 Announce Type: new Abstract: Training AI agents to proactively assist humans in daily activities, from routine household tasks to urgent safety situations, requires large-scale visu

safetyarxiv-cs-cl
12 May 2026
Model Releases

Can You Break RLVER? Probing Adversarial Robustness of RL-Trained Empathetic Agents

DGX agent

arXiv:2605.07138v1 Announce Type: new Abstract: Reinforcement learning from verifiable emotion rewards RLVER has produced language models with strong empathetic performance, evaluated on benchmarks th

model-releasesarxiv-cs-ai
11 May 2026
Agents

Interactive Trajectory Planning with Learning-based Distributionally Robust Model Predictive Control and Markov Systems

DGX agent

arXiv:2605.07768v1 Announce Type: cross Abstract: We investigate interactive trajectory planning subject to uncertainty in the decisions of surrounding agents. To control the ego-agent, we aim to firs

agentsarxiv-cs-lg
11 May 2026
Model Releases

Interpreting Reinforcement Learning Agents with Susceptibilities

DGX agent

arXiv:2605.08007v1 Announce Type: new Abstract: Susceptibilities are a technique for neural network interpretability that studies the response of posterior expectation values of observables to perturb

model-releasesarxiv-cs-lg
11 May 2026
Model Releases

Randomness is sometimes necessary for coordination

DGX agent

arXiv:2605.06825v1 Announce Type: new Abstract: Full parameter sharing is standard in cooperative multi-agent reinforcement learning (MARL) for homogeneous agents. Under permutation-symmetric observat

model-releasesarxiv-cs-ai
11 May 2026
← Previous
1…118119120121122…236
Next →