AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence

DGX agent

arXiv:2605.26494v1 Announce Type: new Abstract: We introduce the MiniMax-M2 series, a family of Mixture-of-Experts language models built around the principle that mini activations can unleash maximum

agentsarxiv-cs-ai
27 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

For How Long Should We Be Punching? Learning Action Duration in Fighting Games

DGX agent

arXiv:2605.20911v1 Announce Type: cross Abstract: Fighting games such as Street Fighter II present unique challenges to reinforcement learning (RL) agents due to their fast-paced, real-time nature. In

agentsarxiv-cs-lg
21 May 2026
Local Ai

Learning to Hand Off: Provably Convergent Workflow Learning under Interface Constraints

DGX agent

arXiv:2605.19140v1 Announce Type: new Abstract: We study workflow learning in a setting where specialized agents hand off control through a shared artifact, each agent observes only a local function o

local-aiarxiv-cs-ai
20 May 2026
Agents

Sequential Resource Trading Using Comparison-Based Gradient Estimation

DGX agent

arXiv:2408.11186v4 Announce Type: replace-cross Abstract: We study sequential multi-issue trading between two greedily rational agents who exchange resources from a finite set of categories. Each agen

agentsarxiv-cs-ai
15 May 2026
Safety

Can a Single Message Paralyze the AI Infrastructure? The Rise of AbO-DDoS Attacks through Targeted Mobius Injection

DGX agent

arXiv:2605.11442v1 Announce Type: cross Abstract: Large Language Model (LLM) agents have emerged as key intermediaries, orchestrating complex interactions between human users and a wide range of digit

safetyarxiv-cs-cl
13 May 2026
Agents

PC3D: Zero-Shot Cooperation Across Variable Rosters via Personalized Context Distillation

DGX agent

arXiv:2605.10377v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning often assumes a fixed execution team, yet many decentralized systems must operate with varying numbers of

agentsarxiv-cs-lg
12 May 2026
Agents

How^{2}: How to learn from procedural How-to questions

DGX agent

arXiv:2510.11144v2 Announce Type: replace-cross Abstract: An agent facing a planning problem can use answers to how-to questions to reduce uncertainty and fill knowledge gaps, helping it solve both cu

agentsarxiv-cs-cl
5 May 2026
Agents

A Systematic Approach for Large Language Models Debugging

DGX agent

arXiv:2604.23027v1 Announce Type: new Abstract: Large language models (LLMs) have become central to modern AI workflows, powering applications from open-ended text generation to complex agent-based re

agentsarxiv-cs-ai
28 Apr 2026
Agents

An Analysis of the Coordination Gap between Joint and Modular Learning for Job Shop Scheduling with Transportation Resources

DGX agent

arXiv:2604.24117v1 Announce Type: new Abstract: Efficient job-shop scheduling with transportation resources is critical for high-performance manufacturing. With the rise of 'decentralized factories',

agentsarxiv-cs-ai
28 Apr 2026
Agents

Case-Specific Rubrics for Clinical AI Evaluation: Methodology, Validation, and LLM-Clinician Agreement Across 823 Encounters

DGX agent

arXiv:2604.24710v1 Announce Type: new Abstract: Objective. Clinical AI documentation systems require evaluation methodologies that are clinically valid, economically viable, and sensitive to iterative

agentsarxiv-cs-ai
28 Apr 2026
Agents

From Static to Interactive: Authoring Interactive Visualizations via Natural Language

DGX agent

arXiv:2601.17736v2 Announce Type: replace-cross Abstract: Interactivity is crucial for effective data visualizations. However, it is often challenging to implement interactions for existing static vis

agentsarxiv-cs-ai
28 Apr 2026
Agents

SCRIBE: Structured Mid-Level Supervision for Tool-Using Language Models

DGX agent

arXiv:2601.03555v2 Announce Type: replace Abstract: Training reliable tool-augmented agents remains a significant challenge, largely due to the difficulty of credit assignment in multi-step reasoning.

agentsarxiv-cs-ai
28 Apr 2026
Agents

SphUnc: Hyperspherical Uncertainty Decomposition and Causal Identification via Information Geometry

DGX agent

arXiv:2603.01168v2 Announce Type: replace-cross Abstract: Reliable decision-making in complex multi-agent systems requires calibrated predictions and interpretable uncertainty. We introduce SphUnc, a

agentsarxiv-cs-ai
23 Apr 2026
Safety

Shepherding UAV Swarm with Action Prediction Based on Movement Constraints

DGX agent

arXiv:2604.17189v1 Announce Type: new Abstract: In this study, we propose a new sheepdog-inspired control method for a swarm of small unmanned aerial vehicles (UAVs), which predicts the swarm behavior

safetyarxiv-cs-ro
21 Apr 2026
Agents

MCPThreatHive: Automated Threat Intelligence for Model Context Protocol Ecosystems

DGX agent

arXiv:2604.13849v1 Announce Type: cross Abstract: The rapid proliferation of Model Context Protocol (MCP)-based agentic systems has introduced a new category of security threats that existing framewor

agentsarxiv-cs-ai
17 Apr 2026
Agents

Young people's perceptions and recommendations for conversational generative artificial intelligence in youth mental health

DGX agent

arXiv:2604.13381v1 Announce Type: cross Abstract: Conversational generative artificial intelligence agents (or genAI chatbots) could benefit youth mental health, yet young people's perspectives remain

agentsarxiv-cs-ai
17 Apr 2026
Model Releases

From Imitation to Discrimination: Progressive Curriculum Learning for Robust Web Navigation

DGX agent

arXiv:2604.12666v1 Announce Type: cross Abstract: Text-based web agents offer computational efficiency for autonomous web navigation, yet developing robust agents remains challenging due to the noisy

model-releasesarxiv-cs-cl
15 Apr 2026
Agents

Automating Structural Analysis Across Multiple Software Platforms Using Large Language Models

DGX agent

arXiv:2604.09866v1 Announce Type: cross Abstract: Recent advances in large language models (LLMs) have shown the promise to significantly accelerate the workflow by automating structural modeling and

agentsarxiv-cs-ai
14 Apr 2026
Agents

Boosted Distributional Reinforcement Learning: Analysis and Healthcare Applications

DGX agent

arXiv:2604.04334v2 Announce Type: replace-cross Abstract: Researchers and practitioners are increasingly considering reinforcement learning to optimize decisions in complex domains like robotics and h

agentsarxiv-cs-ai
13 Apr 2026
Model Releases

Kill-Chain Canaries: Stage-Level Tracking of Prompt Injection Across Attack Surfaces and Model Safety Tiers

DGX agent

arXiv:2603.28013v3 Announce Type: replace-cross Abstract: Multi-agent LLM systems are entering production -- processing documents, managing workflows, acting on behalf of users -- yet their resilience

model-releasesarxiv-cs-ai
13 Apr 2026
Agents

Beyond Functional Correctness: Design Issues in AI IDE-Generated Large-Scale Projects

DGX agent

arXiv:2604.06373v1 Announce Type: cross Abstract: New generation of AI coding tools, including AI-powered IDEs equipped with agentic capabilities, can generate code within the context of the project.

agentsarxiv-cs-ai
10 Apr 2026
Agents

Predictive Representations for Skill Transfer in Reinforcement Learning

DGX agent

arXiv:2604.07016v1 Announce Type: new Abstract: A key challenge in scaling up Reinforcement Learning is generalizing learned behaviour. Without the ability to carry forward acquired knowledge an agent

agentsarxiv-cs-lg
10 Apr 2026
Agents

TOOLCAD: Exploring Tool-Using Large Language Models in Text-to-CAD Generation with Reinforcement Learning

DGX agent

arXiv:2604.07960v1 Announce Type: cross Abstract: Computer-Aided Design (CAD) is an expert-level task that relies on long-horizon reasoning and coherent modeling actions. Large Language Models (LLMs)

agentsarxiv-cs-cl
10 Apr 2026
Safety

Multi-AUV Ad-hoc network-based Target Tracking: A Value Gradient Guidance Multi-Agent Diffusion Reinforcement Learning Approach

DGX agent

arXiv:2608.12436v1 Announce Type: new Abstract: Multi-AUV ad-hoc network-based target tracking requires networked autonomous underwater vehicles (AUVs) to cooperatively track maneuvering targets under

safetyarxiv-cs-lg
14 Aug 2026
Model Releases

Predictive Allostatic Organization in Recurrent and Spiking Agents Under Partial Observability

DGX agent

arXiv:2608.11506v1 Announce Type: cross Abstract: Adaptive behavior under partial observability depends on internal organization that carries information beyond the current observation. Drawing on Bar

model-releasesarxiv-cs-lg
14 Aug 2026
Safety

AI Guardrail Survival under Single-Cycle Agentic Self-Summarization

DGX agent

arXiv:2608.11392v1 Announce Type: cross Abstract: Long-running agents periodically compact their context, replacing the transcript with a model-generated summary.Recent work shows that dropping a stan

safetyarxiv-cs-ai
13 Aug 2026
Model Releases

Convergent Detour Hijacking: Task-Preserving Resource Amplification in Skill-Based LLM Agents

DGX agent

arXiv:2608.12273v1 Announce Type: cross Abstract: LLM agents increasingly rely on third-party skills, using natural-language descriptions for selection and instruction bodies for planning. This progre

model-releasesarxiv-cs-ai
13 Aug 2026
Model Releases

TRACE Bench: Task-driven Roleplay Agentic Checklist Evaluation

DGX agent

arXiv:2608.11236v1 Announce Type: cross Abstract: Roleplay evaluation should do more than assign a single score: it should reveal which role requirements were tested, which failed, and which dialogue

model-releasesarxiv-cs-ai
13 Aug 2026
Research

Post-Hoc Sparse Coding of Latent Communication Between Vision-Language Model Agents

DGX agent

arXiv:2608.10198v1 Announce Type: new Abstract: Latent-space communication allows heterogeneous vision-language model agents to exchange continuous representations without serializing visual and reaso

researcharxiv-cs-ai
12 Aug 2026
Local Ai

Agentic Stage-One Stellarator Optimization: Autonomous Multi-Objective Search for Finite-Beta Equilibria

DGX agent

arXiv:2608.01344v2 Announce Type: replace Abstract: Stage-one stellarator design searches a high-dimensional family of three-dimensional plasma boundaries and fixed-boundary MHD equilibria for configu

local-aiarxiv-cs-ai
11 Aug 2026
Model Releases

AutoRefine: Compiling Trajectories into Validated Typed Agent Artifacts

DGX agent

arXiv:2601.22758v2 Announce Type: replace Abstract: Large language model agents repeatedly encounter related tasks, yet systems that learn from trajectories commit every lesson to one predefined artif

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

CMU-Drive and V2V-VLA: Cooperative Multi-agent Unified Driving with Reasoning Benchmark and Vehicle-to-Vehicle Vision-Language-Action Models

DGX agent

arXiv:2608.07621v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have recently achieved impressive performance for end-to-end autonomous driving, yet existing approaches are primari

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Context Is Not Authority: Structured Runtime Governance for Financial Market Agents

DGX agent

arXiv:2608.09025v1 Announce Type: new Abstract: Financial agents can turn correct context into an unauthorized effect: a customer-facing commitment, trade, or deployed policy. We present SAGE-Fin, a f

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Findings of the First Teaching Monster Challenge: A Benchmark of Pedagogical Content Knowledge in AI Agents

DGX agent

arXiv:2608.08852v1 Announce Type: new Abstract: AI agents can now solve problems, answer like subject experts, and generate long-form multimodal content. However, whether they can adapt a lesson to fi

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

ForestBench: A Unified Graph Framework for Evaluating Multi-Agent Collaboration

DGX agent

arXiv:2608.08605v1 Announce Type: new Abstract: Multi-agent systems (MAS) built on Large Language Models (LLMs) are proliferating rapidly, but their heterogeneous execution traces provide no common ba

model-releasesarxiv-cs-ai
11 Aug 2026
Tutorials

From Trajectories to Evidence: Auditable Experimental Records for Industrial Research Agents

DGX agent

arXiv:2608.05235v1 Announce Type: cross Abstract: Research agents increasingly conduct multi-round machine-learning experiments in industrial recommendation settings and retain the resulting trajector

tutorialsarxiv-cs-ai
11 Aug 2026
Model Releases

Large Multimodal Agents for Intelligent Transportation Systems: Architectures, Evidence, and Deployment Challenges

DGX agent

arXiv:2608.08184v1 Announce Type: new Abstract: Large multimodal agents (LMAs) are increasingly proposed for intelligent transportation systems (ITS), but existing studies often conflate multimodality

model-releasesarxiv-cs-ai
11 Aug 2026
Safety

Learning to Coordinate Symbolic Tools: LLM Agents for Verified Sum-of-Squares Certificates

DGX agent

arXiv:2608.00326v2 Announce Type: replace Abstract: Tool calling allows large language models (LLMs) to invoke external computation during problem solving, a useful capability in various fields includ

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

Optimal Multi-Agent Path Finding in Continuous Time

DGX agent

arXiv:2508.16410v3 Announce Type: replace-cross Abstract: Continuous-time Conflict Based Search (CCBS) has been widely used as an exact baseline for Continuous-time Multi-Agent Path Finding (MAPFR), a

model-releasesarxiv-cs-ro
11 Aug 2026
Model Releases

Rethinking Self-Evolving Agents: Do We Still Need Prescribed Optimization Pipelines?

DGX agent

arXiv:2608.09629v1 Announce Type: new Abstract: Self-evolving agents are usually built around prescribed optimization pipelines: the framework decides how to gather evidence, revise a persistent artif

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

RouteGuard: Certifying Routing Gain in LLM Multi-Agent Systems When Complementarity Is Not Enough

DGX agent

arXiv:2608.07583v1 Announce Type: cross Abstract: Multi-agent LLM systems route among model-backed advisors, yet a deployer rarely knows before shipping whether routing will help at all. Prevailing ro

model-releasesarxiv-cs-lg
11 Aug 2026
Safety

SCOUT: Self-Checking and Recovery-Aware Tool-Thought Agents for Ultra-Long Egocentric Video Reasoning

DGX agent

arXiv:2608.07959v1 Announce Type: new Abstract: Ultra-long egocentric video understanding requires reasoning over temporally sparse evidence distributed across hours or days, challenging current multi

safetyarxiv-cs-ai
11 Aug 2026
Model Releases

SkillReason: Reasoning-Enhanced Agent Skill Retrieval for Implicit User Requests

DGX agent

arXiv:2608.08640v1 Announce Type: new Abstract: Large language model agents increasingly rely on reusable skills to extend their capabilities beyond parametric knowl- edge. However, retrieving the app

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

SodaMem: Evidence-Grounded Temporal Graph Memory for LLM Agents

DGX agent

arXiv:2608.08055v1 Announce Type: new Abstract: Large language model (LLM) agents that assist users over weeks of conversation must remember what is currently true, not merely what was once said. Flat

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Stateful CARS: Exact Cross-History Reuse for Policy-Constrained LLM Agents

DGX agent

arXiv:2608.08282v1 Announce Type: new Abstract: Tool-using language-model agents face constraints whose meaning changes with observations and prior actions. We study exact sampling from the model dist

model-releasesarxiv-cs-lg
11 Aug 2026
Model Releases

TelemetrySuffBench: Is Agent Telemetry Sufficient for Failure-Origin Diagnosis?

DGX agent

arXiv:2608.07899v1 Announce Type: new Abstract: Agent systems increasingly expose execution traces, yet telemetry that reveals a failure may still be inadequate for identifying where that failure orig

model-releasesarxiv-cs-ai
11 Aug 2026
Model Releases

Verication-driven closed-loop multi-agent large language modelframework for code-compliant structural design

DGX agent

arXiv:2608.07978v1 Announce Type: cross Abstract: Multi-agent large language model(LLM)systems are applied to structural design,yet most use one-shot generation and cannot verify their output,leaving

model-releasesarxiv-cs-ai
11 Aug 2026
Local Ai

Weather- and Location-Aware Agentic Dining Recommendation: Leveraging LLM World Knowledge for Region-Sensitive Contextual Reasoning

DGX agent

arXiv:2608.07593v1 Announce Type: cross Abstract: Context-aware recommender systems have long recognized that factors such as location, time, and weather shape where and what people choose to eat. Exi

local-aiarxiv-cs-ai
11 Aug 2026
← Previous
1…9596979899…236
Next →