AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,215 results
27 May 2026

Learning GUI Grounding with Spatial Reasoning from Visual Feedback

AgentsDGX agent

arXiv:2509.21552v2 Announce Type: replace-cross Abstract: Graphical User Interface (GUI) grounding is commonly framed as a coordinate prediction task -- given a natural language instruction, generate

Learning to Act under Noise: Enhancing Agent Robustness via Noisy Environments

AgentsDGX agent

arXiv:2605.27209v1 Announce Type: new Abstract: Recent advances in large language models (LLMs) have facilitated the widespread deployment of LLMs as interactive agents capable of reasoning, planning,

Lessons from Penetration Tests on Large-Scale Agent Systems

AgentsDGX agent

arXiv:2605.27042v1 Announce Type: cross Abstract: As AI systems gain increasing autonomy and execution capability, the number of discovered security vulnerabilities continues to rise. However, many of


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

LiPUP-MA: A Residential Experience-centric Multi-Agent Framework for Living-in-the-loop Participatory Urban Planning

AgentsDGX agent

arXiv:2412.20505v2 Announce Type: replace Abstract: Participatory Urban Planning (PUP) is increasingly supported by LLM-based agents, yet existing methods largely rely on static preference elicitation

Managed Deep Agents is built for agents that need to work over long time horizons, use tools, preserve context, and produce artifacts. A few…

AgentsDGX agent

Managed Deep Agents is built for agents that need to work over long time horizons, use tools, preserve context, and produce artifacts. A few examples of what teams are building: ✅ Support + triage age

MedCollab: IBIS-Guided Multi-Agent Collaboration with Hierarchical Disease Relation Chains for Clinical Diagnosis

AgentsDGX agent

arXiv:2603.01131v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have shown promise in clinical diagnosis but remain limited by unreliable report generation, weak evidence ground

Mind the Tool Failures: Achieving Synergistic Tool Gains for Medical Agents

AgentsDGX agent

arXiv:2605.26691v1 Announce Type: new Abstract: Medical AI agents increasingly use external tools for diagnosis, treatment recommendation, and evidence retrieval, yet most existing approaches assume t

Modeling Agentic Technical Debt and Stochastic Tax: A Standalone Framework for Measurement, Simulation, and Dashboarding

AgentsDGX agent

arXiv:2605.27320v1 Announce Type: new Abstract: Agentic AI systems combine probabilistic reasoning with delegated action through tools, context, memory, orchestration, and external workflow integratio

MUSE-Autoskill: Self-Evolving Agents via Skill Creation, Memory, Management, and Evaluation

AgentsDGX agent

arXiv:2605.27366v1 Announce Type: new Abstract: Large language model (LLM) agents rely on reusable skills to solve complex tasks. However, existing skill creation approaches treat skills as isolated a

Neural Bayesian Sequential Routing

AgentsDGX agent

arXiv:2605.26147v1 Announce Type: new Abstract: Human decision-making is sequential and uncertainty-aware, yet standard neural networks often rely on static, dense forward computation with limited vis

oh yeah first* automations - instead of dumb crons kicking of agents, devin has made them smart. try it - it is the first non annoying proac…

AgentsDGX agent

oh yeah first* automations - instead of dumb crons kicking of agents, devin has made them smart. try it - it is the first non annoying proactive agent impl ive seen https://x.com/russelljkaplan/status

Optimal Rates for Feasible Payoff Set Estimation in Games

AgentsDGX agent

arXiv:2602.04397v2 Announce Type: replace-cross Abstract: We study a setting in which two players play a (possibly approximate) Nash equilibrium of a bimatrix game, while a learner observes only their

Persona Generators: Generating Diverse Synthetic Personas for Arbitrary Contexts

AgentsDGX agent

arXiv:2602.03545v2 Announce Type: replace Abstract: Evaluating AI systems that interact with humans requires understanding their behavior across diverse user populations, but collecting representative

Personalizing Embodied Multimodal Large Language Model Agents over Long-term User Interactions

AgentsDGX agent

arXiv:2605.26256v1 Announce Type: new Abstract: Multimodal large language model (MLLM)-based embodied agents have shown strong potential for solving complex tasks in physical environments. However, pe

PolyFusionAgent: A Multimodal Foundation Model and Autonomous AI Assistant for Polymer Property Prediction and Inverse Design

AgentsDGX agent

arXiv:2605.26543v1 Announce Type: new Abstract: Polymer discovery is central to fields ranging from energy storage to biomedicine, but it is hindered by an astronomically large chemical design space a

Powering agentic AI sales strategy with Amazon Bedrock AgentCore

AgentsDGX agent

As agent adoption scaled, we saw a common pattern emerge across enterprises, including our own sales organization: specialized agents deliver value, but without orchestration, users carry the cognitiv

Probing the Knowledge Boundary: An Interactive Agentic Framework for Deep Knowledge Extraction

AgentsDGX agent

arXiv:2602.00959v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) can be seen as compressed knowledge bases, but it remains unclear what knowledge they truly contain and how far t

QUACK: Questioning, Understanding, and Auditing Communicated Knowledge in Multimodal Social Deduction Agents

AgentsDGX agent

arXiv:2605.27068v1 Announce Type: cross Abstract: Social deduction games have become a popular testbed for probing reasoning, deception, coordination, and belief modeling in Large Language Model (LLM)

Real-Time Progress Prediction in Reasoning Language Models

AgentsDGX agent

arXiv:2506.23274v4 Announce Type: replace-cross Abstract: Recent reasoning language models, particularly those that employ long latent chains of thought, achieve strong performance on complex agentic

Robinhood opens its platform to AI agents for trading and credit card spending

AgentsDGX agent

Robinhood Markets Inc. today opened its trading and banking platforms to artificial intelligence agents, launching a beta Agentic Trading product and a virtual Agentic Credit Card that can place stock

Robinhood will let your AI agent trade stocks and make (or lose) lots of money

AgentsDGX agent

Robinhood is opening its trading platform to AI agents. In an announcement on Wednesday, Robinhood says traders can now create a separate account for an AI agent and add a specific amount of money, al

self improving agents are here http://agnost.ai

AgentsDGX agent

Yohei Nakajima discusses the emergence of self-improving AI agents, likely exploring how autonomous agents can iteratively enhance their capabilities and performance through feedback loops and learnin

Self-signals Driven Multi-LLM Debate for Efficient and Accurate Reasoning

AgentsDGX agent

arXiv:2510.06843v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) have exhibited impressive capabilities across diverse application domains. Recent work has explored Multi-LLM Age

Semigroup Consistency as a Diagnostic for Learned Physics Simulators

AgentsDGX agent

arXiv:2605.26324v1 Announce Type: cross Abstract: Learned physics simulators are often evaluated by one-step or short-horizon prediction error, but these metrics can miss failures in temporal composit

SetupX: Can LLM Agents Learn from Past Failures in Functionality-Correct Code Repository Setup?

AgentsDGX agent

arXiv:2605.26186v1 Announce Type: cross Abstract: Functionality-correct repository setup aims to configure execution environments (e.g., dependencies, build scripts) to successfully execute a reposito

SPEAR: Code-Augmented Agentic Prompt Optimization

AgentsDGX agent

arXiv:2605.26275v1 Announce Type: new Abstract: Automatic prompt engineering (APE) rewrites prompts to improve downstream task performance, but existing APE loops treat the optimizer itself as a fixed

sqlite AGENTS.md

AgentsDGX agent

sqlite AGENTS.md SQLite gained an AGENTS.md file five days ago - but it's not intended for their own development, it's presumably aimed at people who are pointing agents at the SQLite codebase. It inc

Stateful Inference for Low-Latency Multi-Agent Tool Calling

AgentsDGX agent

arXiv:2605.26289v1 Announce Type: new Abstract: Multi-agent tool calling is becoming the dominant interaction pattern for LLM-based systems, yet existing inference frameworks treat each tool call as a

Synergetic Empowerment: Wireless Communications Meets Embodied Intelligence

AgentsDGX agent

arXiv:2509.10481v2 Announce Type: replace-cross Abstract: Wireless communication is evolving into an agent era, where large-scale agents with inherent embodied intelligence are not just users but acti

Telenor Nordics Customer Service self-help corpus

AgentsDGX agent

arXiv:2605.26891v1 Announce Type: new Abstract: This paper presents a multilingual customer service self-help corpus comprising 1,122 manually validated documents in Finnish, Danish, Norwegian, and Sw

the future of Continual Learning will rely on building to systems to ingest, understand, & apply knowledge from Agent Traces at scale had a …

AgentsDGX agent

the future of Continual Learning will rely on building to systems to ingest, understand, & apply knowledge from Agent Traces at scale had a blast presenting LangSmith Engine with @bentannyhill to show

The MiniMax-M2 Series: Mini Activations Unleashing Max Real-World Intelligence

AgentsDGX agent

arXiv:2605.26494v1 Announce Type: new Abstract: We introduce the MiniMax-M2 series, a family of Mixture-of-Experts language models built around the principle that mini activations can unleash maximum

The Necessity of a Unified Framework for LLM-Based Agent Evaluation

AgentsDGX agent

arXiv:2602.03238v2 Announce Type: replace Abstract: With the advent of Large Language Models (LLMs), general-purpose agents have seen fundamental advancements. However, evaluating these agents present

The Need for an External Observer Formalizing the Sufficiency Gap: A Mathematical Extension of Mixture Identifiability and Contextual Grounding in Sequence Models

AgentsDGX agent

arXiv:2605.26711v1 Announce Type: new Abstract: We construct a binary mixed-regime process with one deterministic textual regime and one random regime governed by an unobserved latent state. Even an i

The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now,…

AgentsDGX agent

The Redpoint InfraRed 100 is now live. These are the companies building the infrastructure that powers everything happening in AI right now, from world models and agent runtimes to the sandboxes, data

The Sensation Modulating Network:Haltability as the architectural ground for object-directed phenomenology

AgentsDGX agent

arXiv:2605.26856v1 Announce Type: cross Abstract: Cognitive science remains split between cognitivism - which accounts for recursion and language but cannot ground formal symbols in meaning - and 4E a

The Strongest Teacher Is Not Always the Best Teacher: Student-Centric Answer Selection

AgentsDGX agent

arXiv:2605.26872v1 Announce Type: cross Abstract: LLM training increasingly relies on teacher-generated supervision, from synthetic responses to reasoning traces and tool-use demonstrations. Current p

this is a fantastic example of a robust agentic memory system built on langgraph! there's 4 key parts 1. retrieval (rag) 2. storage, fragmen…

AgentsDGX agent

this is a fantastic example of a robust agentic memory system built on langgraph! there's 4 key parts 1. retrieval (rag) 2. storage, fragmented by memory type 3. reasoning* 4. learning (often called d

this is cool. one paid api to wrap all paid apis

AgentsDGX agent

This post likely discusses a unified API service that aggregates access to multiple paid APIs, allowing developers to integrate various third-party services through a single payment and interface rath

This is pretty cool!

AgentsDGX agent

This is pretty cool! We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf, pypdf, markitdown, pdftotext, opendata

Today's Hermes Agent Jam starts in 2 hours, see you soon!

AgentsDGX agent

Nous Research announced an upcoming 'Hermes Agent Jam' event starting in 2 hours from the time of posting. The event likely involves developers and AI researchers collaborating on or demonstrating app

Uncertainty-Aware Gaussian Map for Vision-Language Navigation

AgentsDGX agent

arXiv:2605.26503v1 Announce Type: new Abstract: Vision-Language Navigation (VLN) requires an agent to navigate 3D environments following natural language instructions. During navigation, existing agen

Unveiling the Fragility of Vision-Language Models: Multi-Modal Adversarial Synergy via Texture-Constrained Perturbations and Cross-Modal Optimization

AgentsDGX agent

arXiv:2605.26501v1 Announce Type: cross Abstract: Large Vision-Language Models (LVLMs) have transformed multi-modal understanding, excelling in tasks like image captioning and visual question answerin

Vital Trace: Protocol-Constrained Patient-State Reasoning for Longitudinal Clinical Trajectories

AgentsDGX agent

arXiv:2602.12833v2 Announce Type: replace-cross Abstract: Longitudinal clinical reasoning over electronic health records requires tracking evolving physiological measurements, laboratory results, and

we are a car

AgentsDGX agent

we are a car We've created the world's fastest PDF parser ⚡️ And it's more accurate than any other open-source, model-free PDF parser out there (pymupdf, pypdf, markitdown, pdftotext, opendataloader,

We launched Context Hub as a way to manage skills, AGENTS.md files, and other context files an agent might need You can easily use it as a v…

AgentsDGX agent

We launched Context Hub as a way to manage skills, AGENTS.md files, and other context files an agent might need You can easily use it as a virtual filesystem in deepagents See this video for more info

We spent a ton of time making worktrees actually work well for agent swarms and large repos. When you’re running 10s of agents that ship, yo…

AgentsDGX agent

We spent a ton of time making worktrees actually work well for agent swarms and large repos. When you’re running 10s of agents that ship, you quickly realize you need worktree, but git’s defaults are

When Does Adaptive Guidance Help? Belief-Aware Privileged Distillation for Autonomous Driving Under Partial Observability

AgentsDGX agent

arXiv:2605.26155v1 Announce Type: cross Abstract: Guided Soft Actor-Critic (GSAC) distills knowledge from a privileged full-state teacher to a partial-observation student for autonomous driving, but u

WisdomAI brings plug-and-play conversational AI analytics capabilities to SaaS applications

AgentsDGX agent

Not content with its push into autonomous artificial intelligence workers, WisdomAI Inc. is bringing its comprehensive analytics capabilities to third-party software-as-a-service platforms with today’

Xage extends zero trust to autonomous AI agents across cloud, SaaS and edge

AgentsDGX agent

Zero-trust cybersecurity company Xage Security Inc. today unveiled new capabilities in its platform designed to give enterprises deterministic visibility into autonomous artificial intelligence agents

XGrammar-2: Efficient Dynamic Structured Generation Engine for Agentic LLMs

AgentsDGX agent

arXiv:2601.04426v3 Announce Type: replace Abstract: Modern LLM agents increasingly rely on dynamic structured generation, such as tool calling and response protocols. Unlike traditional structured gen

26 May 2026

A Multi-Agent LLM Framework for Rating the Quality of Surgical Feedback

AgentsDGX agent

arXiv:2605.25440v1 Announce Type: cross Abstract: Verbal feedback delivered by attending surgeons in the operating room plays a critical formative role in resident trainee skill acquisition. Yet, asse

A Token/KV-Cache Communication Media Selection and Resource Allocation Strategy for Multi-Agent Collaboration

AgentsDGX agent

arXiv:2605.25422v1 Announce Type: cross Abstract: The convergence of large language models (LLMs) with 6G networks is fostering a paradigm of autonomous multi-agent cooperation, which in turn is expec

Act or Clarify? Modeling Sensitivity to Uncertainty and Cost in Communication

AgentsDGX agent

arXiv:2602.02843v3 Announce Type: replace Abstract: When deciding how to act under uncertainty, agents may choose to act to reduce uncertainty or they may act despite that uncertainty. In communicativ

Activegraph's website, newsletter, and marketing are all run on Cofounder!

AgentsDGX agent

Activegraph's website, newsletter, and marketing are all run on Cofounder! [technical blog post] “Evidence Compilation Before Semantic Memory: ActiveGraph on LongMemEval-S” —— 🔍85.6% QA accuracy and 8

.@AdamRLucek on how we use traces to build evals for production agents.

AgentsDGX agent

.@AdamRLucek on how we use traces to build evals for production agents. Trace data is literally worth its weight in gold these days, if you know what to do with it! As has been established, creating e

Adaptive Punishment for Cooperation in Mixed-Motive Games

AgentsDGX agent

arXiv:2605.24516v1 Announce Type: cross Abstract: Mixed-motive scenarios are ubiquitous in real-world multi-agent interactions, where self-interested agents often defect for immediate rewards, overloo

Agent-as-Peer-Debriefer: A Multi-Agent Framework with Perspective-Based Refinement for Qualitative Analysis

AgentsDGX agent

arXiv:2605.24600v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used for qualitative data analysis (QDA), yet their outputs often miss the depth and nuance of human analy

“agent debt” is a new term, but it was inevitable, an instantiation of the AI-driven technical debt I keep warning about. what happens when …

AgentsDGX agent

“agent debt” is a new term, but it was inevitable, an instantiation of the AI-driven technical debt I keep warning about. what happens when you build it fast, but you don’t really know how to fix it.

Agent Manufacturing: Foundation-Model Agents as First-Class Industrial Entities

AgentsDGX agent

arXiv:2605.24823v1 Announce Type: new Abstract: Manufacturing has passed through four widely recognized paradigms - mechanization, electrification, programmable automation, and Smart Manufacturing - e

← Previous
1…5960616263…121
Next →