AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Hardware

AutoLLMResearch: Training Research Agents for Automating LLM Experiment Configuration -- Learning from Cheap, Optimizing Expensive

DGX agent

arXiv:2605.11518v1 Announce Type: cross Abstract: Effectively configuring scalable large language model (LLM) experiments, spanning architecture design, hyperparameter tuning, and beyond, is crucial f

hardwarearxiv-cs-cl
13 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Local Ai

Distributed Quantum Gaussian Processes for Multi-Agent Systems

DGX agent

arXiv:2602.15006v2 Announce Type: replace-cross Abstract: Gaussian Processes (GPs) are a powerful tool for probabilistic modeling, but their performance is often constrained in complex, large-scale re

local-aiarxiv-cs-lg
13 May 2026
Local Ai

An Uncertainty-Aware Resilience Micro-Agent for Causal Observability in the Computing Continuum

DGX agent

arXiv:2605.10718v1 Announce Type: cross Abstract: Grey failures in the computing continuum produce ambiguous overlapping symptoms that existing approaches fail to diagnose reliably, either due to a la

local-aiarxiv-cs-ai
12 May 2026
Safety

Breaking the Impasse: Dual-Scale Evolutionary Policy Training for Social Language Agents

DGX agent

arXiv:2605.08721v1 Announce Type: new Abstract: While Reinforcement Learning with Verifiable Rewards (RLVR) has proven effective for closed-ended tasks, extending it to open-ended social language game

safetyarxiv-cs-cl
12 May 2026
Model Releases

Flame3D: Zero-shot Compositional Reasoning of 3D Scenes with Agentic Language Models

DGX agent

arXiv:2605.09218v1 Announce Type: cross Abstract: 3D scene understanding spans reasoning about free space, object grounding, hypothetical object insertions, complex geometric relationships, and integr

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

Mirror, Mirror on the Wall: Can VLM Agents Tell Who They Are at All?

DGX agent

arXiv:2605.08816v1 Announce Type: new Abstract: In the animal kingdom, mirror self-recognition is a canonical probe of higher-order cognition, emerging only in some species. We ask whether an analogou

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

OPT-BENCH: Evaluating the Iterative Self-Optimization of LLM Agents in Large-Scale Search Spaces

DGX agent

arXiv:2605.08904v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and tool use. However, the fundamental cognitive faculties essential

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

RewardHarness: Self-Evolving Agentic Post-Training

DGX agent

arXiv:2605.08703v1 Announce Type: new Abstract: Evaluating instruction-guided image edits requires rewards that reflect subtle human preferences, yet current reward models typically depend on large-sc

model-releasesarxiv-cs-ai
12 May 2026
Model Releases

TRACER: Verifiable Generative Provenance for Multimodal Tool-Using Agents

DGX agent

arXiv:2605.09934v1 Announce Type: new Abstract: Multimodal large language models increasingly solve vision-centric tasks by calling external tools for visual inspection, OCR, retrieval, calculation, a

model-releasesarxiv-cs-cl
12 May 2026
Model Releases

UTS at PsyDefDetect: Multi-Agent Councils and Absence-Based Reasoning for Defense Mechanism Classification

DGX agent

arXiv:2605.09769v1 Announce Type: new Abstract: This paper describes our system for classifying psychological defense mechanisms in emotional support dialogues using the Defense Mechanism Rating Scale

model-releasesarxiv-cs-ai
12 May 2026
Safety

Cognitive Agent Compilation for Explicit Problem Solver Modeling

DGX agent

arXiv:2605.07040v1 Announce Type: cross Abstract: Large language models (LLMs) are widely used for tutoring, feedback generation, and content creation, but their broad pretraining makes them hard to c

safetyarxiv-cs-ai
11 May 2026
Safety

Many-to-Many Multi-Agent Pickup and Delivery

DGX agent

arXiv:2605.07835v1 Announce Type: new Abstract: Multi-robot systems in automated warehouses must manage continuous streams of pickup-and-delivery tasks while ensuring efficiency and safety. Prior work

safetyarxiv-cs-ro
11 May 2026
Safety

SOD: Step-wise On-policy Distillation for Small Language Model Agents

DGX agent

arXiv:2605.07725v1 Announce Type: cross Abstract: Tool-integrated reasoning (TIR) is difficult to scale to small language models due to instability in long-horizon tool interactions and limited model

safetyarxiv-cs-ai
11 May 2026
Model Releases

TEA-Bench: A Systematic Benchmarking of Tool-enhanced Emotional Support Dialogue Agent

DGX agent

arXiv:2601.18700v2 Announce Type: replace Abstract: Emotional Support Conversation requires not only affective expression but also grounded instrumental support to provide trustworthy guidance. Howeve

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Tools as Continuous Flow for Evolving Agentic Reasoning

DGX agent

arXiv:2605.07339v1 Announce Type: new Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in orchestrating tools for reasoning tasks. However, existing methods rely on a s

model-releasesarxiv-cs-ai
11 May 2026
Agents

cotomi Act: Learning to Automate Work by Watching You

DGX agent

arXiv:2605.03231v1 Announce Type: new Abstract: What if a browser agent could learn your work simply by watching you do it? We present cotomi Act, a browser-based computer-using agent that combines re

agentsarxiv-cs-ai
7 May 2026
Model Releases

From SWE-ZERO to SWE-HERO: Execution-free to Execution-based Fine-tuning for Software Engineering Agents

DGX agent

arXiv:2604.01496v2 Announce Type: replace-cross Abstract: We introduce SWE-ZERO to SWE-HERO, a two-stage SFT recipe that achieves state-of-the-art results on SWE-bench by distilling open-weight fronti

model-releasesarxiv-cs-cl
7 May 2026
Model Releases

Perceive, Verify and Understand Long Video: Multi-Granular Perception and Active Verification via Interactive Agents

DGX agent

arXiv:2509.24943v2 Announce Type: replace Abstract: Long videos, characterized by temporal complexity and sparse task-relevant information, pose significant reasoning challenges for AI systems. Althou

model-releasesarxiv-cs-cv
7 May 2026
Model Releases

AcademiClaw: When Students Set Challenges for AI Agents

DGX agent

arXiv:2605.02661v1 Announce Type: new Abstract: Benchmarks within the OpenClaw ecosystem have thus far evaluated exclusively assistant-level tasks, leaving the academic-level capabilities of OpenClaw

model-releasesarxiv-cs-ai
6 May 2026
Model Releases

AI Agents for Inventory Control: Human-LLM-OR Complementarity

DGX agent

arXiv:2602.12631v2 Announce Type: replace-cross Abstract: Inventory control is a fundamental operations problem in which ordering decisions are traditionally guided by theoretically grounded operation

model-releasesarxiv-cs-lg
6 May 2026
Model Releases

EO-Gym: A Multimodal, Interactive Environment for Earth Observation Agents

DGX agent

arXiv:2605.01250v1 Announce Type: new Abstract: Earth Observation (EO) analysis is inherently interactive: resolving uncertainty often requires expanding the region of interest, retrieving historical

model-releasesarxiv-cs-ai
6 May 2026
Safety

Governing What the EU AI Act Excludes: Accountability for Autonomous AI Agents in Smart City Critical Infrastructure

DGX agent

arXiv:2605.01091v1 Announce Type: cross Abstract: When a traffic signal controller adjusts green phases and a grid manager curtails power on the same corridor, each system may comply with its own obli

safetyarxiv-cs-ai
6 May 2026
Model Releases

Neuro-Symbolic Agents for Hallucination-Free Requirements Reuse

DGX agent

arXiv:2605.01562v1 Announce Type: cross Abstract: The Object-Oriented Method for Requirements Authoring and Management (OOMRAM) is a requirements reuse framework that relies on exact identifier matchi

model-releasesarxiv-cs-ai
6 May 2026
Safety

When Safety Geometry Collapses: Fine-Tuning Vulnerabilities in Agentic Guard Models

DGX agent

arXiv:2605.02914v1 Announce Type: new Abstract: A guard model fine-tuned on entirely benign data can lose all safety alignment -- not through adversarial manipulation, but through standard domain spec

safetyarxiv-cs-lg
6 May 2026
Safety

Agentic Learner with Grow-and-Refine Multimodal Semantic Memory

DGX agent

arXiv:2511.21678v2 Announce Type: replace-cross Abstract: MLLMs exhibit strong reasoning on isolated queries, yet they operate de novo -- solving each problem independently and often repeating the sam

safetyarxiv-cs-lg
5 May 2026
Safety

Causal Foundations of Collective Agency

DGX agent

arXiv:2605.00248v1 Announce Type: new Abstract: A key challenge for the safety of advanced AI systems is the possibility that multiple simpler agents might inadvertently form a collective agent with c

safetyarxiv-cs-ai
5 May 2026
Model Releases

Enhancing Judgment Document Generation via Agentic Legal Information Collection and Rubric-Guided Optimization

DGX agent

arXiv:2605.02011v1 Announce Type: new Abstract: Automating the drafting of judgment documents is pivotal to judicial efficiency, yet it remains challenging due to the dual requirements of comprehensiv

model-releasesarxiv-cs-cl
5 May 2026
Model Releases

Training-Free Time Series Classification via In-Context Reasoning with LLM Agents

DGX agent

arXiv:2510.05950v2 Announce Type: replace Abstract: Time series classification (TSC) spans diverse application scenarios, yet labeled data are often scarce, making task-specific training costly and in

model-releasesarxiv-cs-ai
5 May 2026
Model Releases

Scaling data and AI with Managed Service for Apache Airflow

DGX agent

Orchestration is no longer just about moving data; it is about governing enterprise intelligence. To reflect our deep commitment to and embrace of open-source software, we shared earlier that Cloud Co

model-releasesgoogle-cloud-ai
4 May 2026
Safety

Agent-Agnostic Evaluation of SQL Accuracy in Production Text-to-SQL Systems

DGX agent

arXiv:2604.28049v1 Announce Type: new Abstract: Text-to-SQL (T2SQL) evaluation in production environments poses fundamental challenges that existing benchmarks do not address. Current evaluation metho

safetyarxiv-cs-ai
1 May 2026
Model Releases

Agentic Education: Using Claude Code to Teach Claude Code

DGX agent

arXiv:2604.17460v2 Announce Type: replace-cross Abstract: AI coding assistants have proliferated rapidly, yet structured pedagogical frameworks for learning these tools remain scarce. Developers face

model-releasesarxiv-cs-ai
1 May 2026
Model Releases

Can Large Language Models Implement Agent-Based Models? An ODD-based Replication Study

DGX agent

arXiv:2602.10140v2 Announce Type: replace-cross Abstract: Large language models (LLMs) can now synthesize non-trivial executable code from textual descriptions, raising an important question: can LLMs

model-releasesarxiv-cs-ai
1 May 2026
Local Ai

Echo-{alpha}: Large Agentic Multimodal Reasoning Model for Ultrasound Interpretation

DGX agent

arXiv:2604.28011v1 Announce Type: new Abstract: Ultrasound interpretation requires both precise lesion localization and holistic clinical reasoning, yet existing methods typically excel at only one of

local-aiarxiv-cs-cv
1 May 2026
Safety

METASYMBO: Multi-Agent Language-Guided Metamaterial Discovery via Symbolic Latent Evolution

DGX agent

arXiv:2604.27300v1 Announce Type: new Abstract: Metamaterial discovery seeks microstructured materials whose geometry induces targeted mechanical behavior. Existing inverse-design methods can efficien

safetyarxiv-cs-ai
1 May 2026
Model Releases

Visual Generation in the New Era: An Evolution from Atomic Mapping to Agentic World Modeling

DGX agent

arXiv:2604.28185v1 Announce Type: new Abstract: Recent visual generation models have made major progress in photorealism, typography, instruction following, and interactive editing, yet they still str

model-releasesarxiv-cs-cv
1 May 2026
Agents

Run custom MCP proxies serverless on Amazon Bedrock AgentCore Runtime

DGX agent

This post shows you how to deploy a serverless MCP proxy on Amazon Bedrock AgentCore Runtime that gives you a programmable layer to implement proper governance, controls, and observability aligned wit

agentsaws-ml-blog
29 Apr 2026
Model Releases

Can Current Agents Close the Discovery-to-Application Gap? A Case Study in Minecraft

DGX agent

arXiv:2604.24697v1 Announce Type: new Abstract: Discovering causal regularities and applying them to build functional systems--the discovery-to-application loop--is a hallmark of general intelligence,

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

IntrAgent: An LLM Agent for Content-Grounded Information Retrieval through Literature Review

DGX agent

arXiv:2604.22861v1 Announce Type: cross Abstract: Scientific research relies on accurate information retrieval from literature to support analytical decisions. In this work, we introduce a new task, I

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Learning in Blocks: A Multi Agent Debate Assisted Personalized Adaptive Learning Framework for Language Learning

DGX agent

arXiv:2604.22770v1 Announce Type: cross Abstract: Most digital language learning curricula rely on discrete-item quizzes that test recall rather than applied conversational proficiency. When progressi

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Measuring Successful Cooperation in Human-AI Teamwork: Development and Validation of the Perceived Cooperativity and Teaming Perception Scales

DGX agent

arXiv:2604.24461v1 Announce Type: cross Abstract: As human-AI cooperation becomes increasingly prevalent, reliable instruments for assessing the subjective quality of cooperative human-AI interaction

agentsarxiv-cs-ai
28 Apr 2026
Agents

Process orchestration has become the critical mandate for enterprise AI

DGX agent

Enterprise AI orchestration has become the defining challenge of the agentic era — not because the models aren’t ready, but because most enterprises aren’t. The whole software stack is being reimagine

agentssiliconangle
28 Apr 2026
Model Releases

The Price of Agreement: Measuring LLM Sycophancy in Agentic Financial Applications

DGX agent

arXiv:2604.24668v1 Announce Type: new Abstract: Given the increased use of LLMs in financial systems today, it becomes important to evaluate the safety and robustness of such systems. One failure mode

model-releasesarxiv-cs-ai
28 Apr 2026
Model Releases

Zero-to-CAD: Agentic Synthesis of Interpretable CAD Programs at Million-Scale Without Real Data

DGX agent

arXiv:2604.24479v1 Announce Type: new Abstract: Computer-Aided Design (CAD) models are defined by their construction history: a parametric recipe that encodes design intent. However, existing large-sc

model-releasesarxiv-cs-cv
28 Apr 2026
Model Releases

An Artifact-based Agent Framework for Adaptive and Reproducible Medical Image Processing

DGX agent

arXiv:2604.21936v1 Announce Type: new Abstract: Medical imaging research is increasingly shifting from controlled benchmark evaluation toward real-world clinical deployment. In such settings, applying

model-releasesarxiv-cs-ai
27 Apr 2026
Agents

‘You can’t make great decisions without good data’: How fragmented data blocks enterprise AI success

DGX agent

Without trusted data as the foundation, even the most sophisticated models will produce outcomes that enterprises can’t act on with confidence. As a matter of fact, enterprises are paying a steep pric

agentssiliconangle
27 Apr 2026
Agents

You can’t prompt your way out of complexity: Why enterprise AI is turning to process intelligence

DGX agent

As enterprises navigate the complexities of scaling AI initiatives, the agentic AI blueprint for success lies in combining deep process intelligence with powerful cloud platforms, enabling organizatio

agentssiliconangle
27 Apr 2026
Model Releases

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

DGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Beyond Pixels: Introspective and Interactive Grounding for Visualization Agents

DGX agent

arXiv:2604.21134v1 Announce Type: new Abstract: Vision-Language Models (VLMs) frequently misread values, hallucinate details, and confuse overlapping elements in charts. Current approaches rely solely

model-releasesarxiv-cs-cl
24 Apr 2026
← Previous
1…176177178179180…375
Next →