AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

EmbodiedGovBench: A Benchmark for Governance, Recovery, and Upgrade Safety in Embodied Agent Systems

DGX agent

arXiv:2604.11174v1 Announce Type: cross Abstract: Recent progress in embodied AI has produced a growing ecosystem of robot policies, foundation models, and modular runtimes. However, current evaluatio

model-releasesarxiv-cs-ai
14 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Escaping the Context Bottleneck: Active Context Curation for LLM Agents via Reinforcement Learning

DGX agent

arXiv:2604.11462v1 Announce Type: new Abstract: Large Language Models (LLMs) struggle with long-horizon tasks due to the 'context bottleneck' and the 'lost-in-the-middle' phenomenon, where accumulated

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

From Understanding to Creation: A Prerequisite-Free AI Literacy Course with Technical Depth Across Majors

DGX agent

arXiv:2604.09634v1 Announce Type: cross Abstract: Most AI literacy courses for non-technical undergraduates emphasize conceptual breadth over technical depth. This paper describes UNIV 182, a prerequi

agentsarxiv-cs-ai
14 Apr 2026
Agents

Learning to Focus and Precise Cropping: A Reinforcement Learning Framework with Information Gaps and Grounding Loss for MLLMs

DGX agent

arXiv:2603.27494v2 Announce Type: replace-cross Abstract: To enhance the perception and reasoning capabilities of multimodal large language models in complex visual scenes, recent research has introdu

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

MAVEN-T: Multi-Agent enVironment-aware Enhanced Neural Trajectory predictor with Reinforcement Learning

DGX agent

arXiv:2604.10169v1 Announce Type: new Abstract: Trajectory prediction remains a critical yet challenging component in autonomous driving systems, requiring sophisticated reasoning capabilities while m

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

On Feedback Speed Control for a Planar Tracking

DGX agent

arXiv:2604.09795v1 Announce Type: cross Abstract: This paper investigates a planar tracking problem between a leader and follower agent. We propose a novel feedback speed control law, paired with a co

agentsarxiv-cs-ro
14 Apr 2026
Model Releases

PaperScope: A Multi-Modal Multi-Document Benchmark for Agentic Deep Research Across Massive Scientific Papers

DGX agent

arXiv:2604.11307v1 Announce Type: new Abstract: Leveraging Multi-modal Large Language Models (MLLMs) to accelerate frontier scientific research is promising, yet how to rigorously evaluate such system

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

RTMC: Step-Level Credit Assignment via Rollout Trees

DGX agent

arXiv:2604.11037v1 Announce Type: cross Abstract: Multi-step agentic reinforcement learning benefits from fine-grained credit assignment, yet existing approaches offer limited options: critic-free met

agentsarxiv-cs-ai
14 Apr 2026
Agents

Towards Autonomous Mechanistic Reasoning in Virtual Cells

DGX agent

arXiv:2604.11661v1 Announce Type: cross Abstract: Large language models (LLMs) have recently gained significant attention as a promising approach to accelerate scientific discovery. However, their app

agentsarxiv-cs-ai
14 Apr 2026
Model Releases

Tracing the Roots: A Multi-Agent Framework for Uncovering Data Lineage in Post-Training LLMs

DGX agent

arXiv:2604.10480v1 Announce Type: new Abstract: Post-training data plays a pivotal role in shaping the capabilities of Large Language Models (LLMs), yet datasets are often treated as isolated artifact

model-releasesarxiv-cs-ai
14 Apr 2026
Agents

RIRF: Reasoning Image Restoration Framework

DGX agent

arXiv:2604.09511v1 Announce Type: new Abstract: Universal image restoration (UIR) aims to recover clean images from diverse and unknown degradations using a unified model. Existing UIR methods primari

agentsarxiv-cs-cv
13 Apr 2026
Agents

Blockchain and AI: Securing Intelligent Networks for the Future

DGX agent

arXiv:2604.06323v2 Announce Type: cross Abstract: Blockchain and artificial intelligence (AI) are increasingly proposed together for securing intelligent networks, but the literature remains fragmente

agentsarxiv-cs-ai
10 Apr 2026
Model Releases

Fighting AI with AI: AI-Agent Augmented DNS Blocking of LLM Services during Student Evaluations

DGX agent

arXiv:2604.02360v1 Announce Type: cross Abstract: The transformative potential of large language models (LLMs) in education, such as improving accessibility and personalized learning, is being eclipse

model-releasesarxiv-cs-ai
10 Apr 2026
Safety

MedVR: Annotation-Free Medical Visual Reasoning via Agentic Reinforcement Learning

DGX agent

arXiv:2604.08203v1 Announce Type: new Abstract: Medical Vision-Language Models (VLMs) hold immense promise for complex clinical tasks, but their reasoning capabilities are often constrained by text-on

safetyarxiv-cs-cv
10 Apr 2026
Model Releases

Mitigating Visual Context Degradation in Large Multimodal Models: A Training-Free Decoupled Agentic Framework

DGX agent

arXiv:2509.23322v2 Announce Type: replace Abstract: With the continuous expansion of Large Language Models (LLMs) and advances in reinforcement learning, LLMs have demonstrated exceptional reasoning c

model-releasesarxiv-cs-cv
10 Apr 2026
Agents

The Planetary Cost of AI Acceleration, Part II: The 10th Planetary Boundary and the 6.5-Year Countdown

DGX agent

arXiv:2604.04956v2 Announce Type: replace-cross Abstract: The recent, super-exponential scaling of autonomous Large Language Model (LLM) agents signals a broader, fundamental paradigm shift from machi

agentsarxiv-cs-ai
10 Apr 2026
Agents

AaLLM: An End-to-End Analog Circuit Design Framework from Topology Generation to Sizing Using Large Language Models

DGX agent

arXiv:2608.13472v1 Announce Type: cross Abstract: Analog circuit design is a time-consuming, iterative process in a nonlinear and high-dimensional design space that relies heavily on expert intuition.

agentsarxiv-cs-ai
14 Aug 2026
Model Releases

QuoteBench: How Matched Scores Can Hide Command-Path Failures

DGX agent

arXiv:2608.13547v1 Announce Type: new Abstract: LLM coding agents issue Bash commands through interfaces that may serialize, wrap, and reparse model output. Matched execution scores alone cannot disti

model-releasesarxiv-cs-ai
14 Aug 2026
Agents

MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games

DGX agent

arXiv:2602.24188v2 Announce Type: replace Abstract: We present a scalable and verifiable methodology for evaluating language models in multi-turn interactions, using a suite of collaborative games tha

agentsarxiv-cs-cl
12 Aug 2026
Agents

What Iterated Self-Feeding Probes of Language Models Measure, and a test that separates the construction from the model

DGX agent

arXiv:2608.10986v1 Announce Type: new Abstract: A growing class of methods probes a language model by feeding it its own output: self-consistency, iterated refinement, agentic loops. We ask what such

agentsarxiv-cs-cl
12 Aug 2026
Safety

Agentic Harnesses: LLM-Driven Verification Layers for Robot Autonomy

DGX agent

arXiv:2608.09857v1 Announce Type: cross Abstract: Advances in advanced artificial intelligence tools have sparked research in robot autonomy, but the development of such systems has largely focused on

safetyarxiv-cs-ai
11 Aug 2026
Safety

Agentic Visual Reasoning in Whole-Slide Pathology Images via Active Perception

DGX agent

arXiv:2608.08648v1 Announce Type: new Abstract: Whole-slide visual reasoning requires identifying sparse diagnostic evidence in gigapixel pathology slides and integrating observations across spatial s

safetyarxiv-cs-cv
11 Aug 2026
Research

Capability Is Not Propensity: Measuring Pressure-Robust Cooperative Behavior in Civic LLM Agents

DGX agent

arXiv:2608.09485v1 Announce Type: new Abstract: Cooperative capabilities in language models are dual-use. The same social reasoning that supports civic deliberation can also enable strategic omission,

researcharxiv-cs-ai
11 Aug 2026
Model Releases

PolicyKG: An Agentic LLM Pipeline for Translating Institutional Policies into SHACL Knowledge Graphs

DGX agent

arXiv:2608.09028v1 Announce Type: new Abstract: Institutional policies stay in natural language while the systems that check compliance demand machine-readable constraints. Bridging that gap is still

model-releasesarxiv-cs-ai
11 Aug 2026
Agents

Autonomous discovery of accelerator commissioning algorithms

DGX agent

arXiv:2608.07138v1 Announce Type: cross Abstract: Simulated commissioning has become essential for de-risking modern light-source design and commissioning, but the procedures being simulated are still

agentsarxiv-cs-ai
10 Aug 2026
Applications

KNOWPLAN: Knowledge-Driven AI Agents for Smart Degree Pathway Planning

DGX agent

arXiv:2608.06530v1 Announce Type: new Abstract: Planning a degree from official university sources requires solving two problems in order. The institution's curriculum must first be reconstructed from

applicationsarxiv-cs-ai
10 Aug 2026
Agents

Towards Assurance Closure in AI-Native Large-Scale Agile Software Development

DGX agent

arXiv:2608.07317v1 Announce Type: cross Abstract: The AI-Native Manifesto envisions large-scale agile software development in which humans increasingly govern intent, risk, and exceptions while agents

agentsarxiv-cs-ai
10 Aug 2026
Model Releases

TRIBE: Predicting Team Performance via Communication Behavior Ensembles

DGX agent

arXiv:2608.06926v1 Announce Type: new Abstract: Designing autonomous agents that effectively assist human teams hinges on understanding team dynamics, often without task specific knowledge. We present

model-releasesarxiv-cs-ai
10 Aug 2026
Research

Beyond Flat Policies: Hierarchical Post-Training for Embodied Agents in Robotic Manipulation

DGX agent

arXiv:2608.05999v1 Announce Type: new Abstract: Vision-language-action (VLA) models have demonstrated remarkable capabilities in robotic manipulation by leveraging pretrained vision-language models. H

researcharxiv-cs-ro
7 Aug 2026
Model Releases

VideoArgus: Agentic Rubric-Grounded Unified Evaluation for Video Generation and Editing

DGX agent

arXiv:2608.05485v1 Announce Type: new Abstract: Evaluating generated videos remains challenging because existing benchmarks rely on fixed evaluation content, cover only a subset of generation and edit

model-releasesarxiv-cs-cv
7 Aug 2026
Model Releases

Formal Analysis and Supply Chain Security for Agentic AI Skills

DGX agent

arXiv:2603.00195v2 Announce Type: replace-cross Abstract: 32 pages, 5 theorems with full proofs, 68 references, open-source tool: https://github.com/qualixar/skillfortify. v2: corrects the bibliograph

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

Scrouting: Cost-Aware Routing of Coding Agents by Scouting the Repository First

DGX agent

arXiv:2608.04804v1 Announce Type: cross Abstract: Frontier language models can resolve repository-level software issues, but each attempt is expensive, and existing routers select a model from the iss

model-releasesarxiv-cs-ai
6 Aug 2026
Model Releases

ANCHOR-RE: An Agentic Neuro-Symbolic Framework for Grounded Biomedical Relation Extraction

DGX agent

arXiv:2608.03154v1 Announce Type: new Abstract: Biomedical relation extraction (BioRE) extracts structured knowledge from biomedical literature for applications such as knowledge base construction and

model-releasesarxiv-cs-cl
5 Aug 2026
Research

Should We Type or Talk to LLM Agents? A Comprehensive Study of Voice and Keyboard Input Perturbations

DGX agent

arXiv:2608.03970v1 Announce Type: new Abstract: Human input reaches language models by typing or speaking, and each channel leaves a distinct signature: orthographic noise for keyboards; for voice, di

researcharxiv-cs-ai
5 Aug 2026
Agents

Abstention as an Action Can Kill Both the Reward Gradient and the KL Anchor: Collapse Law and Repair for Error-Penalized Reinforcement Learning

DGX agent

arXiv:2608.00301v1 Announce Type: cross Abstract: Error-penalized scoring rules (+1 for a correct answer, -lambda for a wrong one, 0 for abstaining) are increasingly prescribed against hallucination:

agentsarxiv-cs-cl
4 Aug 2026
Model Releases

LiveMem: Maintaining Memory State Continuity in Long-Running LLM Inference

DGX agent

arXiv:2608.02515v1 Announce Type: new Abstract: Long-running assistants and agents consume interaction streams that eventually outgrow the context. Existing context retention, summarization, and retri

model-releasesarxiv-cs-cl
4 Aug 2026
Model Releases

PackingGPT: 3D Packing Agent for Real Furniture in Last-Mile Delivery

DGX agent

arXiv:2608.01427v1 Announce Type: new Abstract: 3D bin packing rectangular items into standardised containers to maximise space utilisation under geometric shipping automation. Loading a furniture pur

model-releasesarxiv-cs-cv
4 Aug 2026
Model Releases

SIPTraj: Map-Free End-to-End Trajectory Prediction via Physics-Guided Scene Interaction

DGX agent

arXiv:2608.00779v1 Announce Type: new Abstract: Trajectory prediction of surrounding agents is a prerequisite for safe planning and decision making in autonomous driving. Without high-definition (HD)

model-releasesarxiv-cs-ro
4 Aug 2026
Agents

Auto-JEPA: A Latent World Model of Continuous Intent for End-to-End Autonomous Driving

DGX agent

arXiv:2607.29031v1 Announce Type: cross Abstract: Existing autonomous-driving world models typically perform dense prediction of future videos, occupancy states, BEV representations, or agent motion.

agentsarxiv-cs-ai
3 Aug 2026
Model Releases

MirrorCraft: Paired Evaluation under Hidden Rule Changes in Minecraft

DGX agent

arXiv:2607.29218v1 Announce Type: new Abstract: With the prosperity of the large language models (LLMs), it has become an interesting topic: how do LLM-based agents work in Minecraft? Unfortunately, m

model-releasesarxiv-cs-ai
3 Aug 2026
Agents

AI as Friction for Reflection Support in Ideation

DGX agent

arXiv:2607.26827v1 Announce Type: cross Abstract: Generative AI tools for creative work tend to be designed around the goal of removing friction, on the assumption that smoother iteration and faster o

agentsarxiv-cs-ai
31 Jul 2026
Local Ai

Auto Research for Materials: Auditable AI-Scientist Workflows with Held-Out Transfer

DGX agent

arXiv:2607.17100v2 Announce Type: replace-cross Abstract: Auto Research uses language-model agents to propose, implement, and evaluate machine-learning changes in a closed loop, but is usually judged

local-aiarxiv-cs-ai
31 Jul 2026
Agents

One Run Is Not an Idea: The Implementation Lottery in Automated Research

DGX agent

arXiv:2607.26587v1 Announce Type: cross Abstract: Automated research systems use experimental scores both to deliver artifacts and to decide which ideas to retain, transfer, and pursue. Yet one run sc

agentsarxiv-cs-ai
31 Jul 2026
Model Releases

OSReward: Instituting Standardized Evaluation for Cross-Platform Computer-Use Reward Models

DGX agent

arXiv:2607.28609v1 Announce Type: cross Abstract: Computer-using agents (CUAs) are advancing rapidly across the digital world. A CUA trajectory records the agent's actions, states, and reasoning. Veri

model-releasesarxiv-cs-cl
31 Jul 2026
Research

Pushing the Frontier on Approximate EFX Allocations

DGX agent

arXiv:2406.12413v3 Announce Type: replace-cross Abstract: We study the problem of allocating a set of indivisible goods to a set of agents with additive valuation functions, aiming to achieve approxim

researcharxiv-cs-ai
31 Jul 2026
Agents

HeteroPROPMT: A Real-time and Privacy-Preserving Heterogeneous Collaborative Perception Framework

DGX agent

arXiv:2607.26283v1 Announce Type: new Abstract: Collaborative Perception (CP) improves autonomous systems' awareness of their surroundings by sharing sensor data, intermediate features, and detection

agentsarxiv-cs-cv
30 Jul 2026
Safety

CAST: Game Solvers as Turn-Level Teachers for LLM Agents

DGX agent

arXiv:2607.25308v1 Announce Type: cross Abstract: Training large language models (LLMs) to act in long-horizon games is a promising step toward generalist decision-making, yet reinforcement learning w

safetyarxiv-cs-ai
29 Jul 2026
Model Releases

Towards Robust Reinforcement Learning for Small-Scale Language Model Agents

DGX agent

arXiv:2607.25091v1 Announce Type: new Abstract: The alignment of Small Language Models (SLMs) in the 70--500M parameter range using reinforcement learning is often considered unstable, though the unde

model-releasesarxiv-cs-ai
29 Jul 2026
← Previous
1…116117118119120…236
Next →