AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Memorization Diagnostics for Code LLMs Should be Scale-Aware

DGX agent

arXiv:2608.12771v1 Announce Type: cross Abstract: The extent to which large language models for code rely on memorization over genuine understanding remains highly debated. While current literature fr

agentsarxiv-cs-ai
14 Aug 2026
Agents

OmniScientist: An Omni-Modal Omni-Discipline AI Scientist

AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2608.13558v1 Announce Type: new Abstract: Recent advances in foundation models have enabled AI scientists to automate increasingly complete research workflows, from hypothesis generation and cod

agentsarxiv-cs-ai
14 Aug 2026
Agents

QUIETT: Query-Independent Table Transformation for Robust Reasoning

DGX agent

arXiv:2602.20017v2 Announce Type: replace Abstract: Real-world tables often contain schema inconsistencies, heterogeneous value formats, and implicit relational structures that degrade table reasoning

agentsarxiv-cs-cl
14 Aug 2026
Model Releases

Training AI Scientists to Replicate Research

DGX agent

arXiv:2608.13331v1 Announce Type: cross Abstract: The replicability of papers is a cornerstone of scientific knowledge, ensuring the reliability of existing results and providing a base for further ex

model-releasesarxiv-cs-ai
14 Aug 2026
Agents

No One to Blame: A Framework of Constitutive AI Unaccountability

DGX agent

arXiv:2608.12104v1 Announce Type: cross Abstract: The increasing deployment of autonomous, agentic AI systems challenges traditional accountability mechanisms. Existing research predominantly frames A

agentsarxiv-cs-ai
13 Aug 2026
Agents

Top-down Traffic Scenario Generation via Joint Initial-Goal Diffusion and Trajectory Infilling

DGX agent

arXiv:2608.11407v1 Announce Type: new Abstract: Robust traffic simulators are crucial for developing and testing autonomous vehicles to reduce the costly, labor-intensive real-world data collection pr

agentsarxiv-cs-ro
13 Aug 2026
Agents

Coordinating the Unknown Lipschitz Constant in Multiplayer Bandits

DGX agent

arXiv:2608.10526v1 Announce Type: cross Abstract: Motivated by decentralized applications, we study cooperative multi-agent bandits in continuous (Lipschitz) action spaces when the Lipschitz constant

agentsarxiv-cs-ai
12 Aug 2026
Agents

Emotion2Skill: Model-Internal Emotion Signals for Adaptive Skill Selection and Evolution

DGX agent

arXiv:2608.09248v1 Announce Type: new Abstract: Skill-based LLM agents select reusable procedures from an external library to solve complex tasks, yet their routing decisions rely entirely on text-lev

agentsarxiv-cs-ai
11 Aug 2026
Model Releases

Toward Metacognitive One-Shot Indirect Prompt Injection: Strategy Abstraction Via Outcome-Conditioned Reflection

DGX agent

arXiv:2608.08795v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents are vulnerable to indirect prompt injection (IPI), in which malicious instructions embedded in external o

model-releasesarxiv-cs-cl
11 Aug 2026
Agents

Dueling World Models: Advantage-Style Action Channels for Common-Mode Distractor Rejection

DGX agent

arXiv:2608.06706v1 Announce Type: cross Abstract: Latent world models plan by predicting future states from an action, but when a scene contains motion the agent does not control, they quietly go acti

agentsarxiv-cs-ai
10 Aug 2026
Agents

Predicting Task Difficulty Without Rollouts

DGX agent

arXiv:2608.05797v1 Announce Type: cross Abstract: Task difficulty dictates an agent's likelihood of success, and estimating it without rollouts means forecasting this directly from a task description

agentsarxiv-cs-cl
7 Aug 2026
Agents

Search2Skill: Skill Distillation Beyond Knowledge Boundaries Via Rubric-Based Reinforcement Learning

DGX agent

arXiv:2608.05245v1 Announce Type: new Abstract: Reusable skills, which encapsulate the procedural knowledge required to solve real-world professional tasks, offer LLM-based agents a path toward self-e

agentsarxiv-cs-ai
7 Aug 2026
Agents

CheMLFlow: An Open-Source Platform for Cheminformatics and Materials Informatics Applications

DGX agent

arXiv:2608.04942v1 Announce Type: cross Abstract: CheMLFlow is an open-source platform for building and executing end-to-end, high-throughput, and agentic workflows for scientific and technological ap

agentsarxiv-cs-ai
6 Aug 2026
Model Releases

Evidence-Ledger Adjudication for Claim-Evidence Traceability

DGX agent

arXiv:2607.26512v1 Announce Type: new Abstract: AI agents can draft claims faster than authors can check whether the cited or retrieved evidence supports them. We study evidence-ledger adjudication: a

model-releasesarxiv-cs-ai
31 Jul 2026
Agents

GVR-Coder: A Visual-Feedback Framework for Structured SVG Generation in Complex Document and Meeting Scenarios

DGX agent

arXiv:2607.28073v1 Announce Type: cross Abstract: In demanding professional environments and meeting review scenarios, lengthy text often imposes a high cognitive load. To facilitate efficient informa

agentsarxiv-cs-cv
31 Jul 2026
Model Releases

KernelGenBench: A Multi-Source and Multi-Chip Benchmark for LLM-based Kernel Generation

DGX agent

arXiv:2607.27231v1 Announce Type: cross Abstract: Large language models (LLMs) have significantly increased the demand for efficient accelerator kernels, but kernel development remains a highly specia

model-releasesarxiv-cs-lg
31 Jul 2026
Agents

PUDA: An AI-Native Hardware Harness for Self-Driving Laboratories

DGX agent

arXiv:2607.26464v1 Announce Type: cross Abstract: Physical Unified Device Architecture (PUDA) is an AI-native hardware harness for self-driving laboratories (SDLs). Rather than building a human-center

agentsarxiv-cs-ai
31 Jul 2026
Agents

Representation and Invariance in Reinforcement Learning

DGX agent

arXiv:2112.07752v4 Announce Type: replace-cross Abstract: Researchers have formalized reinforcement learning (RL) in different ways. If an agent in one RL framework is to run within another RL framewo

agentsarxiv-cs-lg
31 Jul 2026
Agents

Training Skills Like Parameters via Self-Supervised Semantic Diffusion

DGX agent

arXiv:2607.27557v1 Announce Type: new Abstract: While Large Language Models (LLMs) demonstrate remarkable general instruction-following capabilities, they often fall short of human experts in highly s

agentsarxiv-cs-cl
31 Jul 2026
Agents

Mental World Modeling

DGX agent

arXiv:2607.27201v1 Announce Type: new Abstract: World models enable a predictive substrate for planning and action, yet existing formulations merely answer a physical question: what/where it is, and h

agentsarxiv-cs-cl
30 Jul 2026
Safety

ResearchArena: Evaluating Sabotage and Monitoring in Automated AI R&D

DGX agent

arXiv:2607.19321v2 Announce Type: replace-cross Abstract: As AI agents begin to automate AI R&D, we need ways to assess whether their outputs are safe to deploy, even when the agents themselves may be

safetyarxiv-cs-lg
30 Jul 2026
Agents

SAFAARI: Schema-Aware Framework for Accelerated Advertiser Response Intelligence

DGX agent

arXiv:2607.25042v1 Announce Type: new Abstract: The evolution of customer support systems is rapidly advancing with agentic chatbots, yet these systems face significant limitations when accessing ente

agentsarxiv-cs-ai
29 Jul 2026
Agents

Specula: Scaling formal specifications for autonomous model checking of system code

DGX agent

arXiv:2607.25333v1 Announce Type: cross Abstract: Specula is a push-button agentic system that generates high-quality formal specifications for large, complex system code and uses the specifications f

agentsarxiv-cs-ai
29 Jul 2026
Safety

CRAFT: Learn the Schema, Execute the Plan

DGX agent

arXiv:2607.22642v1 Announce Type: new Abstract: Enterprise coding agents translate natural-language analytical requests into executable code over proprietary APIs, schemas, and metric definitions. Yet

safetyarxiv-cs-ai
28 Jul 2026
Agents

HiMemVLN: Enhancing Reliability of Open-Source Zero-Shot Vision-and-Language Navigation with Hierarchical Memory System

DGX agent

arXiv:2603.14807v3 Announce Type: replace Abstract: LLM-based agents have demonstrated impressive zero-shot performance in vision-language navigation (VLN) tasks. However, most zero-shot methods prima

agentsarxiv-cs-cv
28 Jul 2026
Agents

KG2Code: Bridging Knowledge Graphs and Large Language Models via Executable Code for Question Answering

DGX agent

arXiv:2607.22652v1 Announce Type: new Abstract: Recent research has explored the integration of knowledge graphs (KGs) with large language models (LLMs) to enhance their performance on downstream know

agentsarxiv-cs-ai
28 Jul 2026
Agents

PoTRE: Test-Time Reasoning inspired by Cognitive Heterogeneity

DGX agent

arXiv:2607.20268v1 Announce Type: new Abstract: While Large Language Models (LLMs) excel at many tasks, they frequently struggle with complex reasoning that requires long-horizon planning and iterativ

agentsarxiv-cs-ai
23 Jul 2026
Model Releases

UESF-Bench: Benchmarking and Probing for Unified Embodied Seeking and Following

DGX agent

arXiv:2607.13621v1 Announce Type: new Abstract: Language-guided human following is an important capability for embodied agents, but existing benchmarks typically assume that the target person is visib

model-releasesarxiv-cs-ai
16 Jul 2026
Agents

Differentiable Clone-Structured Causal Graphs for End-to-End Cognitive Map Learning from Image Sequences

DGX agent

arXiv:2607.12382v1 Announce Type: new Abstract: How can an agent build a structured map of its world from nothing but an ongoing sequence of raw sensory input and its own movements, especially when na

agentsarxiv-cs-lg
15 Jul 2026
Model Releases

In-Context Reinforcement Learning under Non-Stationarity: A Survey

DGX agent

arXiv:2607.11906v1 Announce Type: new Abstract: The development of decision-pretrained transformers, algorithm distillation, long-context meta-RL, and retrieval-augmented agents has renewed interest i

model-releasesarxiv-cs-ai
15 Jul 2026
Model Releases

KnowAct-GUIClaw: Know Deeply, Act Perfectly, Personal GUI Assistant with Self-Evolving Memory and Skill

DGX agent

arXiv:2607.12625v1 Announce Type: new Abstract: OpenClaw has emerged as a leading agent framework for complex task automation, yet it faces insufficient cross-platform GUI interaction support and a we

model-releasesarxiv-cs-cl
15 Jul 2026
Model Releases

Token Reduction Is Not Cost Reduction

DGX agent

arXiv:2607.12161v1 Announce Type: new Abstract: Context-reduction layers for API-based coding agents, including command-output compressors, retrieval rankers, and payload-optimizing proxies, are usual

model-releasesarxiv-cs-cl
15 Jul 2026
Agents

3100 Opinions on Code Review in an AI World: Building Causal Theory from Practitioner Discourse

DGX agent

arXiv:2607.07980v1 Announce Type: cross Abstract: Coding agents now author entire pull requests, and practitioners sharply disagree about what this does to code review: whether it becomes the bottlene

agentsarxiv-cs-ai
10 Jul 2026
Agents

Computation, Condensation, and the Incompleteness Between Them: A Coupled Foundation of Intelligence

DGX agent

arXiv:2303.04203v4 Announce Type: replace-cross Abstract: The theory of computation was built to answer Turing's question: what is effectively calculable by an unbounded, immortal, disembodied agent f

agentsarxiv-cs-cv
10 Jul 2026
Agents

Distributed Dynamic Associative Memory via Online Convex Optimization

DGX agent

arXiv:2511.23347v2 Announce Type: replace Abstract: An associative memory (AM) enables cue-response recall, and it has recently been recognized as a key mechanism underlying modern neural architecture

agentsarxiv-cs-lg
9 Jul 2026
Agents

Grounded autonomous research: a fault-tolerant LLM pipeline from corpus to manuscript in frontier computational physics

DGX agent

arXiv:2607.02329v1 Announce Type: new Abstract: Autonomous-research agents have demonstrated end-to-end LLM automation in machine-learning sandboxes where execution provides calibration. Frontier phys

agentsarxiv-cs-ai
3 Jul 2026
Agents

A Single Rewrite Suffices: Empirical Lessons from Production Skill Description Optimization

DGX agent

arXiv:2606.30775v1 Announce Type: cross Abstract: Enterprise AI agents route user queries to specialized skills by matching queries against natural language skill descriptions. When two skills share o

agentsarxiv-cs-ai
1 Jul 2026
Agents

Better Understanding, Understanding Better

DGX agent

arXiv:2606.31892v1 Announce Type: cross Abstract: 'Any fool can know; the point is to understand.' A well-known remark often attributed to Einstein captures a widely shared intuition: understanding is

agentsarxiv-cs-ai
1 Jul 2026
Agents

Emergent Culture in Minimal LLM Systems

DGX agent

arXiv:2606.30668v1 Announce Type: cross Abstract: What happens when LLM agents operate with no context outside a turn, minimal prompting, and simple tools? Inspired by swarm engineering, we give colle

agentsarxiv-cs-ai
1 Jul 2026
Agents

Nazrin: An Atomic Neural Proof Automation Tactic in Lean 4

DGX agent

arXiv:2602.18767v3 Announce Type: replace-cross Abstract: In Machine-Assisted Theorem Proving, a theorem proving agent searches for a sequence of expressions and tactics that can prove a statement in

agentsarxiv-cs-lg
1 Jul 2026
Agents

Developmental Trajectories of Situation Modeling and Mentalizing in Transformer Language Models

DGX agent

arXiv:2606.28524v1 Announce Type: new Abstract: Recent work suggests that Large Language Models (LLMs) are sensitive to the belief states of agents described by text, as measured by the false belief t

agentsarxiv-cs-cl
30 Jun 2026
Agents

UnfoldArt: Zero-Shot Recovery of Full Articulated 3D Objects from Text or Image

DGX agent

arXiv:2606.30608v1 Announce Type: new Abstract: Articulated 3D objects are essential for interactive environments in embodied AI, robotics, and virtual reality, but reconstructing their structure and

agentsarxiv-cs-cv
30 Jun 2026
Agents

Decoupling Reconnaissance and Exploitation: Measuring the Capability Boundaries of LLM-Based Web Penetration Testing

DGX agent

arXiv:2606.25332v1 Announce Type: cross Abstract: Large Language Models (LLMs) have shown promise for automated penetration testing, yet existing end-to-end black-box evaluations are highly susceptibl

agentsarxiv-cs-ai
25 Jun 2026
Agents

Tinker Tales: A Tangible Dialogue System for Child-AI Co-Creative Storytelling

DGX agent

arXiv:2602.04109v2 Announce Type: replace-cross Abstract: Conversational AI agents are increasingly explored as creative partners, yet how conversation design shapes child-AI dialogue in co-creative s

agentsarxiv-cs-ai
25 Jun 2026
Agents

From Task-Guided Conversational Graphs to Goal-Oriented Dialogue Runtimes

DGX agent

arXiv:2606.23797v1 Announce Type: cross Abstract: Graph and multi-agent orchestration frameworks make production large language model (LLM) workflows practical, but they do not by themselves solve con

agentsarxiv-cs-ai
24 Jun 2026
Agents

Reinforcement Learning to Disentangle Multiqubit Quantum States from Partial Observations

DGX agent

arXiv:2406.07884v2 Announce Type: replace-cross Abstract: Using partial knowledge of a quantum state to control multiqubit entanglement is a largely unexplored paradigm in the emerging field of quantu

agentsarxiv-cs-lg
23 Jun 2026
Agents

Sakana Fugu Technical Report

DGX agent

arXiv:2606.21228v1 Announce Type: new Abstract: The capabilities of frontier Large Language Models (LLMs) continue to advance, with different providers increasingly specializing in distinct domains. T

agentsarxiv-cs-lg
23 Jun 2026
Agents

UECP: Uncertainty-Enhanced Collaborative Perception

DGX agent

arXiv:2606.23046v1 Announce Type: new Abstract: Collaborative perception serves as a pivotal solution to enhance the perception capability of individual agents in autonomous driving, where a core chal

agentsarxiv-cs-cv
23 Jun 2026
← Previous
1…120121122123124…236
Next →