AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Pushing the Limits of LLM Tool Calling via Experiential Knowledge Integration and Activation

DGX agent

arXiv:2606.10875v1 Announce Type: new Abstract: Large language models (LLMs) rely on tool use to act as autonomous agents, yet often fail in multi-step execution due to insufficient tool-related knowl

agentsarxiv-cs-cl
10 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

The Distributed Detectability Band Against Marginal-Preserving Attacks

DGX agent

arXiv:2606.10456v1 Announce Type: cross Abstract: AI-control monitors score individual agent actions to detect misbehavior, but real harm can be distributed across many benign-looking steps, each indi

agentsarxiv-cs-ai
10 Jun 2026
Agents

Advancing Mathematics Research with AI-Driven Formal Proof Search

DGX agent

arXiv:2605.22763v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly excel at mathematical reasoning, but their unreliability limits their utility in mathematics research. A m

agentsarxiv-cs-ai
9 Jun 2026
Agents

Engagement Process: Rethinking the Temporal Interface of Action and Observation

DGX agent

arXiv:2605.11484v2 Announce Type: replace Abstract: Task completion in digital and physical environments increasingly involves complex temporal interaction, where actions and observations unfold over

agentsarxiv-cs-ai
9 Jun 2026
Agents

GIFT: LLM-Guided State-Reward Interface for Financial Reinforcement Learning

DGX agent

arXiv:2606.08450v1 Announce Type: new Abstract: Financial portfolio trading is naturally formulated as a reinforcement learning problem, where an agent sequentially rebalances assets under changing ma

agentsarxiv-cs-ai
9 Jun 2026
Agents

Motion planning for hundreds of floating robots

DGX agent

arXiv:2606.09620v1 Announce Type: new Abstract: Planning collision-free motion for large robot fleets is difficult because collision avoidance induces strong inter-agent coupling that grows rapidly wi

agentsarxiv-cs-ro
9 Jun 2026
Model Releases

RunAgent SuperBrowser: A Theory of Autonomous Web Navigation Grounded in Human Browsing Behaviour

DGX agent

arXiv:2606.09399v1 Announce Type: new Abstract: We present SUPERBROWSER, an autonomous web-navigation agent designed against a single guiding hypothesis: a web agent should browse the way a person bro

model-releasesarxiv-cs-ai
9 Jun 2026
Agents

Self-Paced Curriculum Reinforcement Learning for Autonomous Superbike Racing in Simulation

DGX agent

arXiv:2606.09236v1 Announce Type: cross Abstract: Autonomous Racing has seen remarkable progress through deep Reinforcement Learning (RL), primarily for four-wheeled vehicles. However, motorbikes intr

agentsarxiv-cs-ai
9 Jun 2026
Safety

Semantic Quorum Assurance: Collective Certification for Non-Deterministic AI Infrastructure

DGX agent

arXiv:2606.08021v1 Announce Type: cross Abstract: As large language model (LLM) agents are integrated into autonomous cloud operations, distributed systems face a semantic reliability problem: propose

safetyarxiv-cs-ai
9 Jun 2026
Agents

OGA-AID: Clinician-in-the-loop AI Report Drafting Assistant for Multimodal Observational Gait Analysis in Post-Stroke Rehabilitation

DGX agent

arXiv:2604.05360v2 Announce Type: replace-cross Abstract: Gait analysis is essential in post-stroke rehabilitation but remains time-intensive and cognitively demanding, especially when clinicians must

agentsarxiv-cs-ai
8 Jun 2026
Agents

Should You Use Your Large Language Model to Explore or Exploit?

DGX agent

arXiv:2502.00225v4 Announce Type: replace-cross Abstract: We evaluate the ability of the current generation of large language models (LLMs) to help a decision-making agent facing an exploration-exploi

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

Enhancing Software Engineering Through Closed-Loop Memory Optimization

DGX agent

arXiv:2606.05646v1 Announce Type: cross Abstract: Large language models (LLMs) have enabled powerful software engineering (SE) agents capable of navigating complex codebases and resolving real-world i

model-releasesarxiv-cs-ai
6 Jun 2026
Agents

ABBEL: Learning Natural-Language Belief States for Memory-Efficient Interaction

DGX agent

arXiv:2512.20111v2 Announce Type: replace Abstract: As the time horizons of sequential decision-making tasks grow, keeping full interaction histories in model context becomes increasingly costly. Rece

agentsarxiv-cs-cl
5 Jun 2026
Agents

Emergent Language as an Approach to Conscious AI

DGX agent

arXiv:2606.06380v1 Announce Type: new Abstract: The question of whether artificial systems can be conscious remains open, in part because existing approaches either evaluate systems against theory-der

agentsarxiv-cs-cl
5 Jun 2026
Agents

CADENCE: Predicting Realized MAPF Execution Time Beyond Sum of Costs

DGX agent

arXiv:2606.04746v1 Announce Type: new Abstract: Multi-Agent Path Finding (MAPF) algorithms are increasingly used to plan motion for robot teams in industrial warehouses and robotic shared workspaces,

agentsarxiv-cs-ro
4 Jun 2026
Local Ai

DPDL: Towards Differential Privacy Preservation in Decentralized Stochastic Learning on Non-IID Data

DGX agent

arXiv:2606.04399v1 Announce Type: new Abstract: In the paradigm of decentralized learning, a group of agents collaborate to train a global model using distributed datasets without a central server. Al

local-aiarxiv-cs-lg
4 Jun 2026
Model Releases

Evaluating Large Language Models in Dynamic Clinical Decision-Making with Standardized Patient Cases

DGX agent

arXiv:2606.05112v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly proposed as clinical agents, yet static, single-turn benchmarks cannot capture how a model dynamically del

model-releasesarxiv-cs-cl
4 Jun 2026
Agents

Position: Deployed Reinforcement Learning should be Continual

DGX agent

arXiv:2606.04029v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix para

agentsarxiv-cs-ai
4 Jun 2026
Agents

Stateful Visual Encoders for Vision-Language Models

DGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

agentsarxiv-cs-cl
4 Jun 2026
Agents

Decentralized Stochastic Nonconvex Optimization under the (L_0,L_1)-Smoothness

DGX agent

arXiv:2509.08726v3 Announce Type: replace-cross Abstract: This paper focuses on the decentralized stochastic optimization problem f(mathbf{x})=frac{1}{m}sum_{i=1}^m f_i(mathbf{x}) over a connected net

agentsarxiv-cs-lg
3 Jun 2026
Model Releases

Dynamics of Cognitive Heterogeneity: Investigating Behavioral Biases in Multi-Stage Supply Chains with LLM-Based Simulation

DGX agent

arXiv:2604.17220v2 Announce Type: replace-cross Abstract: Modeling coordination among generative agents in complex multi-round decision-making presents a core challenge for AI and operations managemen

model-releasesarxiv-cs-ai
3 Jun 2026
Agents

Minimax Optimal Strategy for Delayed Observations in Online Reinforcement Learning

DGX agent

arXiv:2603.03480v2 Announce Type: replace Abstract: We study reinforcement learning with delayed state observation, where the agent observes the current state after some random number of time steps. W

agentsarxiv-cs-lg
3 Jun 2026
Model Releases

Psi-Bench: Evaluating Persona-Sensitive Influencing in Persuasive Dialogues

DGX agent

arXiv:2606.02754v1 Announce Type: new Abstract: Personalization is a crucial capability of modern language agents. However, current research primarily positions personalized agents as passive responde

model-releasesarxiv-cs-lg
3 Jun 2026
Agents

Reinforcement Learning from Cross-domain Videos with Video Prediction Model

DGX agent

arXiv:2606.03201v1 Announce Type: cross Abstract: Reinforcement learning from expert videos across visually distinct domains is challenging due to the absence of reward signals and the presence of dom

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

Auditing Asset-Specific Preferences in Financial Large Language Models: Evidence from Bitcoin Representations and Portfolio Allocation

DGX agent

arXiv:2606.02528v1 Announce Type: cross Abstract: Large language models now power robo-advisors and trading agents, yet whether they carry built-in biases toward specific assets is largely untested. W

model-releasesarxiv-cs-lg
2 Jun 2026
Agents

Automated Conjecture Resolution with Formal Verification

DGX agent

arXiv:2604.03789v2 Announce Type: replace-cross Abstract: Recent advances in large language models have significantly improved their ability to perform mathematical reasoning, extending from elementar

agentsarxiv-cs-ai
2 Jun 2026
Safety

Beyond Independent Manipulation: Individual Fairness-aware Strategic Classification with Peer Imitation

DGX agent

arXiv:2606.00827v1 Announce Type: cross Abstract: Strategic classification (SC) investigates scenarios where agents manipulate their features to obtain favorable decisions from predictive models. Exis

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

From Segments to Scenes: Temporal Understanding in Autonomous Driving via Vision-Language Model

DGX agent

arXiv:2512.05277v3 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) are increasingly deployed as the perception and reasoning backbone of autonomous agents acting in the wild, with

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

GenPT: Beyond Self-Report for Reliable LLM Psychometrics via Generative Projective Testing

DGX agent

arXiv:2606.00860v1 Announce Type: cross Abstract: Self-report questionnaires remain the prevailing tool for probing the psychological states of persona-conditioned agents (PC-Agents). However, classic

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

Model-Native Computing Architecture: Envisioning Future System Architecture Through the Lens of Computer Architecture

DGX agent

arXiv:2606.00288v1 Announce Type: new Abstract: Large language models are undergoing a transition from model technology to system technology. As developers use Codex, Claude Code, AutoGPT, and related

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Needles at Scale: LLM-Assisted Target Selection for Windows Vulnerability Research

DGX agent

arXiv:2606.01364v1 Announce Type: cross Abstract: The attack surface of a modern operating system is a haystack: thousands of signed binaries and millions of functions, almost none relevant to any giv

agentsarxiv-cs-ai
2 Jun 2026
Agents

PlatonicNav: Unveiling Semantic Correspondence in Navigation with Platonic Topological Maps

DGX agent

arXiv:2606.01788v1 Announce Type: new Abstract: Embodied visual navigation, where an agent perceives a complex environment and acts to reach a goal from raw sensory input, underpins a wide range of ap

agentsarxiv-cs-cv
2 Jun 2026
Agents

Symmetry-Aware 9D Pose Estimation with Sim(3)-Consistent Feature and Spherical Inception Convolution

DGX agent

arXiv:2606.02219v1 Announce Type: new Abstract: Object pose estimation is a fundamental problem for an agent system to perceive or manipulate objects in images or videos. However, current instance-lev

agentsarxiv-cs-cv
2 Jun 2026
Model Releases

Task diversity produces systematic transfer but inhibits continual reinforcement learning

DGX agent

arXiv:2606.00880v1 Announce Type: cross Abstract: Continual reinforcement learning aims to produce agents that learn not only to improve at their current tasks but also to adapt as task distributions

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

Towards Sparse Video Understanding and Reasoning

DGX agent

arXiv:2602.13602v2 Announce Type: replace Abstract: We present revise (nderline{Re}asoning with nderline{Vi}deo nderline{S}parsity), a multi-round agent for video question answering (VQA). Instead of

agentsarxiv-cs-cv
2 Jun 2026
Agents

A Tight Theory of Error Feedback Algorithms in Distributed Optimization

DGX agent

arXiv:2605.31594v1 Announce Type: new Abstract: Communication costs are a major bottleneck in distributed learning and first-order optimization. A common approach to alleviate this issue is to compres

agentsarxiv-cs-lg
1 Jun 2026
Agents

Answer-Set-Programming-based Abstractions for Reinforcement Learning

DGX agent

arXiv:2605.31444v1 Announce Type: new Abstract: Reinforcement Learning (RL) enables autonomous agents to learn policies from experience, but realistic problems often involve enormous state spaces, mak

agentsarxiv-cs-ai
1 Jun 2026
Agents

If LLMs Have Human-Like Attributes, Then So Does Age of Empires II

DGX agent

arXiv:2605.31514v1 Announce Type: cross Abstract: Much research has been carried out on large language models (LLMs) and LLM-powered agentic workflows. However, many works within the field state emerg

agentsarxiv-cs-ai
1 Jun 2026
Agents

Learning to Perceive the World Through Control: Empowerment-Based Representation Learning

DGX agent

arXiv:2605.30656v1 Announce Type: new Abstract: In many practical reinforcement learning environments, observations are far higher-dimensional than the variables that matter for control. In this work,

agentsarxiv-cs-lg
1 Jun 2026
Agents

SpecDB: LLM-Generated Customized Databases via Feature-Oriented Decomposition

DGX agent

arXiv:2605.31097v1 Announce Type: cross Abstract: Mainstream relational databases ship a uniform feature set across deployments, although individual workloads exercise only a fraction of the available

agentsarxiv-cs-ai
1 Jun 2026
Agents

BitTP: The Lightweight Trajectory Prediction Model with BitLLM for Edge-Devices

DGX agent

arXiv:2605.29705v1 Announce Type: new Abstract: Trajectory prediction is a fundamental task for autonomous systems, requiring complex reasoning about multi-agent interactions and intents. Large langua

agentsarxiv-cs-ai
29 May 2026
Model Releases

Gram: Assessing sabotage propensities via automated alignment auditing

DGX agent

arXiv:2605.30322v1 Announce Type: cross Abstract: We introduce Gram, an automated alignment auditing framework to assess the propensity of AI agents to engage in sabotage. We evaluate Gemini models ac

model-releasesarxiv-cs-ai
29 May 2026
Agents

Honeyval: A Comprehensive Evaluation Framework for LLM-powered HTTP Honeypots

DGX agent

arXiv:2605.29963v1 Announce Type: cross Abstract: Honeypots are decoy systems mimicking real system components designed to defend against cyber attacks. Recently, LLMs increasingly serve as simulation

agentsarxiv-cs-ai
29 May 2026
Model Releases

AsyncTool: Evaluating the Asynchronous Function Calling Capability under Multi-Task Scenarios

DGX agent

arXiv:2605.27995v1 Announce Type: new Abstract: Large language model (LLM)-based agents have shown strong capabilities in using external tools to solve complex tasks. However, existing evaluations oft

model-releasesarxiv-cs-ai
28 May 2026
Agents

AtomComposer: Discovering Chemical Space from First Principles with Reinforcement Learning

DGX agent

arXiv:2605.28287v1 Announce Type: new Abstract: Discovering novel stable molecules without training data remains a grand scientific challenge. Current molecular generative models are trained on large,

agentsarxiv-cs-lg
28 May 2026
Agents

DeltaMCP: Incremental Regeneration via Spec-Aware Transformation for MCP servers

DGX agent

arXiv:2605.28148v1 Announce Type: cross Abstract: The rapid development of LLMs coupled with the introduction of Model Context Protocol (MCP) has revolutionized how intelligent agents interact with AP

agentsarxiv-cs-ai
28 May 2026
Model Releases

Do LLMs Favor Their Providers? Measuring Vertical Integration Bias in Code Generation

DGX agent

arXiv:2605.28515v1 Announce Type: cross Abstract: Large Language Models (LLMs) have become an integral part of software development, especially with the advent of agentic capabilities. Yet, many front

model-releasesarxiv-cs-ai
28 May 2026
Safety

Grimlock: Guarding High-Agency Systems with eBPF and Attested Channels

DGX agent

arXiv:2605.27488v1 Announce Type: cross Abstract: Agentic systems increasingly run user-authored orchestration code that invokes tools, spawns subtasks, and delegates work across machines and clouds.

safetyarxiv-cs-ai
28 May 2026
← Previous
1…129130131132133…236
Next →