AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Beyond Majority Voting: Efficient Best-Of-N with Radial Consensus Score

DGX agent

arXiv:2604.12196v1 Announce Type: new Abstract: Large language models (LLMs) frequently generate multiple candidate responses for a given prompt, yet selecting the most reliable one remains challengin

agentsarxiv-cs-cl
15 Apr 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Defining and Evaluation Method for External Human-Machine Interfaces

DGX agent

arXiv:2604.12293v1 Announce Type: new Abstract: As the number of fatalities involving Autonomous Vehicles increase, the need for a universal method of communicating between vehicles and other agents o

agentsarxiv-cs-ro
15 Apr 2026
Agents

PrivacyReasoner: Can LLM Emulate a Human-like Privacy Mind?

DGX agent

arXiv:2601.09152v2 Announce Type: replace Abstract: Prior work on LLM-based privacy focuses on norm judgment over synthetic vignettes, rather than how people think about a specific data practice and f

agentsarxiv-cs-ai
15 Apr 2026
Safety

The Stackelberg Speaker: Optimizing Persuasive Communication in Social Deduction Games

DGX agent

arXiv:2510.09087v2 Announce Type: replace Abstract: Large language model (LLM) agents have shown remarkable progress in social deduction games (SDGs). However, existing approaches primarily focus on i

safetyarxiv-cs-ai
15 Apr 2026
Agents

Consensus-based Recursive Multi-Output Gaussian Process

DGX agent

arXiv:2604.10146v1 Announce Type: new Abstract: Multi-output Gaussian Processes provide principled uncertainty-aware learning of vector-valued fields but are difficult to deploy in large-scale, distri

agentsarxiv-cs-lg
14 Apr 2026
Local Ai

Deep-Reporter: Deep Research for Grounded Multimodal Long-Form Generation

DGX agent

arXiv:2604.10741v1 Announce Type: cross Abstract: Recent agentic search frameworks enable deep research via iterative planning and retrieval, reducing hallucinations and enhancing factual grounding. H

local-aiarxiv-cs-ai
14 Apr 2026
Agents

M^3KG-RAG: Multi-hop Multimodal Knowledge Graph-enhanced Retrieval-Augmented Generation

DGX agent

arXiv:2512.20136v3 Announce Type: replace-cross Abstract: Retrieval-Augmented Generation (RAG) has recently been extended to multimodal settings, connecting multimodal large language models (MLLMs) wi

agentsarxiv-cs-ai
14 Apr 2026
Agents

Structure-Grounded Knowledge Retrieval via Code Dependencies for Multi-Step Data Reasoning

DGX agent

arXiv:2604.10516v1 Announce Type: new Abstract: Selecting the right knowledge is critical when using large language models (LLMs) to solve domain-specific data analysis tasks. However, most retrieval-

agentsarxiv-cs-cl
14 Apr 2026
Agents

AI-Induced Human Responsibility (AIHR) in AI-Human teams

DGX agent

arXiv:2604.08866v1 Announce Type: cross Abstract: As organizations increasingly deploy AI as a teammate rather than a standalone tool, morally consequential mistakes often arise from joint human-AI wo

agentsarxiv-cs-ai
13 Apr 2026
Agents

Contribution of task-irrelevant stimuli to drift of neural representations

DGX agent

arXiv:2510.21588v2 Announce Type: replace-cross Abstract: Biological and artificial learners are inherently exposed to a stream of data and experience throughout their lifetimes and must constantly ad

agentsarxiv-cs-lg
13 Apr 2026
Agents

MT-OSC: Path for LLMs that Get Lost in Multi-Turn Conversation

DGX agent

arXiv:2604.08782v1 Announce Type: new Abstract: Large language models (LLMs) suffer significant performance degradation when user instructions and context are distributed over multiple conversational

agentsarxiv-cs-cl
13 Apr 2026
Agents

One Life to Learn: Inferring Symbolic World Models for Stochastic Environments from Unguided Exploration

DGX agent

arXiv:2510.12088v2 Announce Type: replace Abstract: Symbolic world modeling requires inferring and representing an environment's transitional dynamics as an executable program. Prior work has focused

agentsarxiv-cs-ai
10 Apr 2026
Safety

PriPG-RL: Privileged Planner-Guided Reinforcement Learning for Partially Observable Systems with Anytime-Feasible MPC

DGX agent

arXiv:2604.08036v1 Announce Type: cross Abstract: This paper addresses the problem of training a reinforcement learning (RL) policy under partial observability by exploiting a privileged, anytime-feas

safetyarxiv-cs-ro
10 Apr 2026
Agents

The Illusion of Stochasticity in LLMs

DGX agent

arXiv:2604.06543v1 Announce Type: cross Abstract: In this work, we demonstrate that reliable stochastic sampling is a fundamental yet unfulfilled requirement for Large Language Models (LLMs) operating

agentsarxiv-cs-lg
10 Apr 2026
Agents

Working Paper: Towards a Category-theoretic Comparative Framework for Artificial General Intelligence

DGX agent

arXiv:2603.28906v2 Announce Type: replace Abstract: AGI has become the Holly Grail of AI with the promise of level intelligence and the major Tech companies around the world are investing unprecedente

agentsarxiv-cs-ai
10 Apr 2026
Agents

Distributed Online Submodular Maximization under Communication Delays: A Simultaneous Decision-Making Approach

DGX agent

arXiv:2603.27803v2 Announce Type: replace Abstract: We provide a distributed online algorithm for multi-agent submodular maximization under communication delays. We are motivated by the future distrib

agentsarxiv-cs-lg
14 Aug 2026
Model Releases

ERSkill: Evolving for Skill-Guided Adaptive Memory Retrieval

DGX agent

arXiv:2608.12720v1 Announce Type: cross Abstract: While Large Language Model (LLM) agents increasingly rely on long-term memory for persistent interactions, the retrieval mechanisms governing this mem

model-releasesarxiv-cs-ai
14 Aug 2026
Agents

Position: Reasoning is a Learnable Rule-Based Process

DGX agent

arXiv:2608.12325v1 Announce Type: new Abstract: Autonomous reasoning is among the most scientifically and economically motivating topics in AI today. Historically the purview of symbolic AI, recent ad

agentsarxiv-cs-ai
14 Aug 2026
Agents

SAP-Nav: Spatial Semantic Representation Meets Active Perception for Hierarchical Open-Vocabulary Object Navigation

DGX agent

arXiv:2608.12707v1 Announce Type: new Abstract: Hierarchical open-vocabulary object navigation (OVON) requires agents to follow free-form instructions that may specify targets through scene-, room-, r

agentsarxiv-cs-ro
14 Aug 2026
Agents

Sovereign by necessity? Frontier AI export controls, cyber security, and the limits of national AI capability

DGX agent

arXiv:2608.13272v1 Announce Type: new Abstract: A small number of firms based in two states produce the most capable frontier AI models. The governments of those states have shown both the legal power

agentsarxiv-cs-ai
14 Aug 2026
Agents

Credo: Declarative Control of LLM Pipelines via Beliefs and Policies

DGX agent

arXiv:2604.14401v2 Announce Type: replace Abstract: Agentic AI systems are becoming commonplace in domains that require long-lived, stateful decision-making in continuously evolving conditions. As suc

agentsarxiv-cs-ai
13 Aug 2026
Agents

DCM Bandits: Multiplayer Information Asymmetric Cascading Bandits for Multiple Clicks

DGX agent

arXiv:2608.11873v1 Announce Type: new Abstract: In this work, we extend the Dependent Click Model (DCM) Bandits to a multiplayer information-asymmetric setting, where multiple agents interact with a s

agentsarxiv-cs-lg
13 Aug 2026
Agents

Principal Trait Analysis: Towards Deriving 'Skills' in Human-AI Collaboration

DGX agent

arXiv:2608.11460v1 Announce Type: new Abstract: Large Language Model-powered agents are increasingly used in the workplace via human-artificial intelligence (AI) collaboration. In this new era of work

agentsarxiv-cs-cl
13 Aug 2026
Research

Evaluating Rational Contracting in Natural Language

DGX agent

arXiv:2608.10475v1 Announce Type: new Abstract: The emergence of language-based AI agents promises to transform the scope of machine economic activity. Instead of just proposing bids or following hard

researcharxiv-cs-ai
12 Aug 2026
Model Releases

Test-Time Self-Evolving GUI Visual Grounding via Reflection-Guided On-Policy Self-Distillation

DGX agent

arXiv:2608.11191v1 Announce Type: cross Abstract: GUI Visual Grounding is a fundamental capability for GUI agents. Existing models typically freeze their parameters after deployment, limiting their ab

model-releasesarxiv-cs-ai
12 Aug 2026
Agents

Branch2Skill: Efficient Skill Evolution Through Reasoning Trees

DGX agent

arXiv:2608.08677v1 Announce Type: new Abstract: Skill evolution improves agent skills through feedback over time, with failed trajectories often providing informative signals by revealing incomplete o

agentsarxiv-cs-ai
11 Aug 2026
Safety

Contextual Value Alignment via Multilayer Combinatorial Fusion

DGX agent

arXiv:2608.07642v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Persistent Semantic Entities in Tool-Augmented LLM Systems

DGX agent

arXiv:2608.07952v1 Announce Type: cross Abstract: Tool-augmented LLM agents can harbor implicit state that persists across sessions, activates through events, and propagates across agent boundaries---

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

RangeFactory: Scalable Construction of Multi-Hop Cyber Ranges

DGX agent

arXiv:2608.09526v1 Announce Type: cross Abstract: Real-world cyberattacks often require sustained progress across multiple hosts and network segments, making multi-hop cyber ranges essential infrastru

agentsarxiv-cs-ai
11 Aug 2026
Agents

Regret, equilibrium, and learning in games: A guided tour

DGX agent

arXiv:2608.09389v1 Announce Type: cross Abstract: This note aims to serve as an entry point to the literature on learning in games, a topic with significant theoretical appeal and a wide range of appl

agentsarxiv-cs-lg
11 Aug 2026
Agents

SkillAligner: Treating Retrieved Skills as Adaptable Drafts at Execution Time

DGX agent

arXiv:2608.06880v1 Announce Type: new Abstract: General-purpose skills promise reusable procedural knowledge for language agents, yet semantic relevance does not guarantee execution utility: a retriev

agentsarxiv-cs-lg
10 Aug 2026
Agents

CalibForge: Adversarial Solver Calibration for Scaling Learnable Terminal Tasks

DGX agent

arXiv:2608.06352v1 Announce Type: cross Abstract: Training terminal agents requires executable and verifiable tasks that are not merely solvable, but appropriately challenging for learning. Executable

agentsarxiv-cs-cl
7 Aug 2026
Agents

Robust Context-Aware Detection of Malicious Instructions in Text

DGX agent

arXiv:2608.05430v1 Announce Type: cross Abstract: The remarkable instruction-following ability of modern LLMs has enabled their practical use as the minds of agents that can autonomously complete incr

agentsarxiv-cs-lg
7 Aug 2026
Agents

CoPlan: A Trustworthy Co-Intelligence Interface for Care Planning through Role-Based Contestable Argument Graphs

DGX agent

arXiv:2608.05107v1 Announce Type: new Abstract: AI-supported care planning can help clinicians, patients, caregivers, and care teams coordinate complex decisions across clinical, functional, psychosoc

agentsarxiv-cs-ai
6 Aug 2026
Agents

Spoken Function Calling: A New Perspective on Spoken Language Understanding for Large Audio Language Models

DGX agent

arXiv:2608.05126v1 Announce Type: new Abstract: Spoken Language Understanding (SLU) is the core component of task-oriented dialogue systems and a pivotal link in achieving seamless human-agent interac

agentsarxiv-cs-cl
6 Aug 2026
Agents

AgentPanel: Toward a New Paradigm for Human--AI Collaboration in Exploring Scientific Questions

DGX agent

arXiv:2608.03283v1 Announce Type: new Abstract: Identifying promising scientific ideas remains an important challenge in research practice. Researchers commonly rely on small-group discussions or one-

agentsarxiv-cs-ai
5 Aug 2026
Safety

Implementing Causal Perception: Competing SCMs and Situated Fairness

DGX agent

arXiv:2608.03917v1 Announce Type: new Abstract: Causal perception occurs when agents with competing Structural Causal Models (SCMs) of the same system infer different probability distributions, includ

safetyarxiv-cs-ai
5 Aug 2026
Agents

A drone that learns to efficiently find non-uniformly distributed objects in agricultural fields: from simulation to the real world

DGX agent

arXiv:2505.09278v2 Announce Type: replace Abstract: Drones are promising for data collection in precision agriculture but are limited by battery capacity. Drone paths are usually planned using full co

agentsarxiv-cs-ro
4 Aug 2026
Agents

ShiJianBench: From Dialogue to Decision for Long-Horizon Evaluation of Investment Advisors

DGX agent

arXiv:2608.01204v1 Announce Type: new Abstract: Conversational investment advisors influence not only what users know, but also how they make subsequent decisions as market conditions evolve. Existing

agentsarxiv-cs-cl
4 Aug 2026
Agents

VC-Tooler: Learning Compositional and Adaptive Visual Tool Use

DGX agent

arXiv:2608.02217v1 Announce Type: new Abstract: Agentic multimodal reasoning extends passive image understanding by allowing VLMs to actively acquire and refine visual evidence through visual tool int

agentsarxiv-cs-cv
4 Aug 2026
Research

RRM: Experience-Driven Reflective Retrieval Memory for Long-Horizon Multimodal Reasoning

DGX agent

arXiv:2607.28156v1 Announce Type: new Abstract: Existing multimodal long-term memory agents use external memory to overcome the limited context available for long videos. However, most methods emphasi

researcharxiv-cs-cl
31 Jul 2026
Agents

SciDataSailor: Deep Scientific Data Exploring

DGX agent

arXiv:2607.28098v1 Announce Type: cross Abstract: Scientific datasets are commonly organized as hierarchical repositories containing heterogeneous and interdependent files, making their inspection, in

agentsarxiv-cs-cl
31 Jul 2026
Agents

DREvo: Distilling Recalibrated Historical Experience for Harness Self-Evolution

DGX agent

arXiv:2607.26722v1 Announce Type: cross Abstract: Harness plays a critical role in large language model agent performance, and building a high-performing harness requires substantial expert effort. Th

agentsarxiv-cs-lg
30 Jul 2026
Agents

Nanobot Algorithms for Treatment of Diffuse Cancer

DGX agent

arXiv:2509.06893v2 Announce Type: replace-cross Abstract: Motile nanosized particles, or 'nanobots', promise more effective and less toxic targeted drug delivery because of their unique scale and prec

agentsarxiv-cs-ro
29 Jul 2026
Model Releases

RSIBench-Data: Benchmarking Data-Centric Research for Recursive Self-Improvement

DGX agent

arXiv:2607.25886v1 Announce Type: cross Abstract: Recursive self-improvement requires turning evidence of model failures into better models. Data-centric post-training research entails diagnosing capa

model-releasesarxiv-cs-cl
29 Jul 2026
Agents

CodeEvo: Interaction-Driven Synthesis of Code-centric Data through Hybrid and Iterative Feedback

DGX agent

arXiv:2507.22080v2 Announce Type: replace-cross Abstract: Acquiring high-quality instruction-code pairs is essential for training Large Language Models for code generation. While automated synthesis h

agentsarxiv-cs-ai
28 Jul 2026
Safety

Plato-Bio: verification-first biological novelty screening with temporal rediscovery and structural benchmarks

DGX agent

arXiv:2607.23975v1 Announce Type: new Abstract: Large language model research agents can connect literature retrieval, analysis code, and manuscript preparation, but coherent output does not establish

safetyarxiv-cs-ai
28 Jul 2026
Agents

SetGo: Metadata Readiness for Scientific AI Datasets

DGX agent

arXiv:2607.22677v1 Announce Type: cross Abstract: Scientific datasets intended for AI use require both computational readiness for model training and metadata readiness for discovery, sharing, and reu

agentsarxiv-cs-ai
28 Jul 2026
← Previous
1…123124125126127…236
Next →