AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,156 results
12 Aug 2026

4D-WAM: 4D Consistent World Modeling for Autonomous Driving

AgentsDGX agent

arXiv:2608.10107v1 Announce Type: new Abstract: Emerging World-Action Models (WAMs) have demonstrated promising performance in autonomous driving by jointly modeling future driving scene evolution and

A Gateway Architecture for Enterprise MCP Authentication: Unifying Heterogeneous Auth, Identity Delegation, and the User / Non-User Persona Problem

AgentsDGX agent

arXiv:2608.10760v1 Announce Type: cross Abstract: The Model Context Protocol (MCP) has become the de-facto interface for connecting LLM agents to enterprise tools, and adoption has been explosive: wit

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data producti…

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data production. @mercor_ai’s @BrendanFoody joined us at our Sovereign AI

Agentic AI infrastructure shifts enterprise focus from model choice to platform control

AgentsDGX agent

As agentic AI infrastructure moves from experimentation into production, enterprises are confronting a more complex question than which model to use: how to control the cost, data exposure and infrast

Agentic Instruction Data Selection: Let DataMaster Interpret Your Intent

AgentsDGX agent

arXiv:2608.10579v1 Announce Type: new Abstract: Although existing instruction data selection methods have introduced various metrics, the inherent complexity of real-world datasets makes it impractica

Apexon targets stalled AI pilots with three AgentRise additions

AgentsDGX agent

Santa Clara-based technology services firm Apexon Inc. today expanded AgentRise, its agentic artificial intelligence platform, with three new components. The additions are named AgentRise Polaris, Age

Automating and Scaling Behavioral Scientific Research on AI Agents

AgentsDGX agent

arXiv:2608.10030v1 Announce Type: new Abstract: As AI agents are increasingly deployed in complex environments, understanding their behaviors becomes critical. Yet behavioral scientific research on AI

Autonomous Exploration-Based Precise Mapping for Mobile Robots through Stepwise and Consistent Motions

AgentsDGX agent

arXiv:2503.17005v3 Announce Type: replace Abstract: This paper presents an autonomous exploration framework. It is designed for indoor ground mobile robots that utilize laser Simultaneous Localization

Bandwidth-Efficient Multi-Agent Communication through Information Bottleneck and Vector Quantization

AgentsDGX agent

arXiv:2602.02035v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning systems deployed in real-world robotics applications face severe communication constraints that significant

Blacksmith raises $45M to aid AI code validation as agentic development grows

AgentsDGX agent

Blacksmith Software Inc. today announced it has raised 45 million in new funding for its continuous integration service, which combines code development with cloud-based testing instead of on the deve

Chain of Spatial Thoughts: Modality-Agnostic Spatial Grounding for Vision Language Models

AgentsDGX agent

arXiv:2608.10278v1 Announce Type: new Abstract: Spatial understanding is fundamental to embodied intelligence, underpinning applications such as robotic manipulation, embodied navigation, and autonomo

ChemWorld: Programmable Chemical Worlds for Controlled and Replayable Agent Experimentation

AgentsDGX agent

arXiv:2608.10792v1 Announce Type: new Abstract: Autonomous chemistry increasingly depends on environments in which agents can repeatedly act, observe, and adapt.Physical laboratories provide essential

Co-Evolution in Agentic Systems: Toward Self-Directed Evolution Beyond Human Design

AgentsDGX agent

arXiv:2608.10299v1 Announce Type: new Abstract: Agentic systems are increasingly expected to improve after deployment, yet single-entity self-evolution is often bounded by a static learning context, s

CodeRabbit bags $143M to help companies get a grip on the explosion of AI-generated code

AgentsDGX agent

CodeRabbit Inc., the creator of a popular tool that automatically reviews artificial intelligence-generated code, is becoming more ambitious after closing on its latest 143 million Series C round of f

Continuous Interaction Diffusion: A Diffusion-Native Runtime for Asynchronous Tool-Augmented Reasoning

AgentsDGX agent

arXiv:2608.10438v1 Announce Type: new Abstract: Large language models increasingly rely on external tools to access up-to-date information, perform computation, and interact with the outside world. Fo

Coordinating the Unknown Lipschitz Constant in Multiplayer Bandits

AgentsDGX agent

arXiv:2608.10526v1 Announce Type: cross Abstract: Motivated by decentralized applications, we study cooperative multi-agent bandits in continuous (Lipschitz) action spaces when the Lipschitz constant

Cross-View Sequential Visual Localization with Spatio-Temporal Context Modeling for Autonomous Driving

AgentsDGX agent

arXiv:2608.10660v1 Announce Type: cross Abstract: Continuous and reliable localization is essential for autonomous driving. Cross-view visual localization matches ground images with satellite maps, pr

DHS has lied about these incidents over and over again and been caught in the lies over and over again. They are lying now, and nobody shoul…

AgentsDGX agent

DHS has lied about these incidents over and over again and been caught in the lies over and over again. They are lying now, and nobody should believe them. NOW: DHS response to Virginia woman's video

Do Personalized Skills Help Coding Agents? An Empirical Study of Developer Interaction Histories

AgentsDGX agent

arXiv:2608.10319v1 Announce Type: cross Abstract: Large language model (LLM)-powered agents have rapidly evolved from code-completion tools into solvers of complex software engineering tasks. As devel

DOCSCHISEL: Adaptive Tool Documentation Optimization Framework for LLM Agents

AgentsDGX agent

arXiv:2608.10037v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly rely on external tools to accomplish complex real-world tasks, making tool documentation a critical groundin

DuplexWorld: Can voice agents help you get through the day?

AgentsDGX agent

arXiv:2608.10716v1 Announce Type: cross Abstract: Speech-to-speech (S2S) voice agents are increasingly being incorporated into enterprise for customer care and as daily companions for consumers owing

FlowScout: From Execution Feedback to Reliable Tool-Using Agent Workflows

AgentsDGX agent

arXiv:2608.10039v1 Announce Type: new Abstract: Agentic workflows have become an important abstraction for building reliable LLM-based automation systems by organizing large language models (LLMs), to

GAM-Agent: Game-Theoretic and Uncertainty-Aware Collaboration for Complex Visual Reasoning

AgentsDGX agent

arXiv:2505.23399v2 Announce Type: replace Abstract: We propose GAM-Agent, a game-theoretic multi-agent framework for enhancing vision-language reasoning. Unlike prior single-agent or monolithic models

GitSkills: A Dataset of Agent Skills on GitHub

AgentsDGX agent

arXiv:2608.10906v1 Announce Type: cross Abstract: An agent skill is a folder containing a SKILL.md file with instructions for a language-model agent, optionally accompanied by scripts and reference fi

Grok is now an AI ‘teammate’ you can assign work

AgentsDGX agent

SpaceXAI has introduced Grok Bot, an always-on AI agent service designed to behave like independent 'AI teammates' that can do your work for you. The bots share their own cloud-based computer environm

Hierarchical Compositionality for An Assistive AI Agent

AgentsDGX agent

arXiv:2608.10330v1 Announce Type: new Abstract: AI agents are increasingly being developed to assist humans in various applications, and Large Language Models and other deep network architectures are

How to Dogfood Your AI Chat Agent: A Three-Layer Evaluation Framework with Goal-Directed NPC Simulation

AgentsDGX agent

arXiv:2608.09939v1 Announce Type: cross Abstract: Production teams deploying LLM chat agents face a specific quality assurance gap: existing evaluation tools test individual responses or simulate soci

I believe the demand for compute is going to go up faster than the supply of compute, so the price of compute is going to increase significa…

AgentsDGX agent

I believe the demand for compute is going to go up faster than the supply of compute, so the price of compute is going to increase significantly in the future, perhaps as much as 10x in the next few y

If you’re interested in checking out LlamaParse for document extraction, sign up here: https://login.llamaindex.ai/sign-up

AgentsDGX agent

A 36‑page ArXiv whitepaper titled **ExtractBench** was released by Jerry Liu (jerryjliu0), describing a large‑scale, schema‑guided benchmark for real‑world document extraction from complex enterprise

Inferential Capability Does Not Determine Legal Scope

AgentsDGX agent

arXiv:2608.10601v1 Announce Type: cross Abstract: Two instruments of EU digital law place inference at their centre and mean different things by it. Article 3(1) of the AI Act uses the capability to i

LLM Agents Factory: Retrieval of Domain-Specific LLM Agents

AgentsDGX agent

arXiv:2608.09934v1 Announce Type: cross Abstract: Large language model (LLM) agents improve task performance by decomposing problems into role-specialized behaviors. However, their practical deploymen

MEGA: Self-Evolving Agent Optimization Infrastructure via Wisdom Graph

AgentsDGX agent

arXiv:2608.10504v1 Announce Type: new Abstract: As coding agents increasingly handle implementation, the central challenge shifts from building individual agents to building an infrastructure that sys

MESA:Task-Adaptive Multi-Structure Evidence Selection for Long-Horizon Agent Memory

AgentsDGX agent

arXiv:2608.10108v1 Announce Type: new Abstract: Long-horizon agents accumulate trajectories spanning hundreds of interleaved reasoning, action, and observation steps, where answering a query may depen

Mind Viruses: Self-Propagating Ideas in Multi-Agent LLM Systems

AgentsDGX agent

arXiv:2608.10218v1 Announce Type: new Abstract: AI agents are becoming more autonomous and increasingly interconnected, exposing them to new emergent risks arising from agent-to-agent interaction. One

MIRA: Medical Image Reflection for Agentic Diagnosis

AgentsDGX agent

arXiv:2608.10827v1 Announce Type: cross Abstract: Medical visual agents can use tools to inspect images and retrieve external knowledge, but indiscriminate tool use may introduce noisy or misleading e

Mitigating Context Interference for Reliable and Efficient Search Agents

AgentsDGX agent

arXiv:2608.10743v1 Announce Type: new Abstract: Recent research empowers Large Language Models (LLMs) as multi-turn search agents to iteratively retrieve and generate outputs until complex tasks are s

Most semantic search queries leave something unstated. The user knows what they mean. The system doesn't. Pinecone's text match filters scop…

AgentsDGX agent

Most semantic search queries leave something unstated. The user knows what they mean. The system doesn't. Pinecone's text match filters scope a vector search to a lexical condition (a machine number,

MT-PingEval: Evaluating Multi-Turn Collaboration with Private Information Games

AgentsDGX agent

arXiv:2602.24188v2 Announce Type: replace Abstract: We present a scalable and verifiable methodology for evaluating language models in multi-turn interactions, using a suite of collaborative games tha

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…

AgentsDGX agent

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco

Nebius shares jump 34% on continued AI infrastructure demand

AgentsDGX agent

Shares of Nebius Group NV closed 34% higher today after it reported second-quarter earnings that topped expectations across the board. The Netherlands-based company operates a cloud platform optimized

Nonlinear Model Predictive Control via Sequential Convex Programming for Drone-to-Drone Docking

AgentsDGX agent

arXiv:2608.10542v1 Announce Type: new Abstract: Autonomous mid-air docking of multi-rotor vehicles under disturbance-driven target motion poses a constrained non-linear trajectory optimization challen

On Understanding, Identifying, and Mitigating Vulnerabilities in Agentic Large Language Models

AgentsDGX agent

arXiv:2608.10530v1 Announce Type: cross Abstract: Large Language Models (LLMs) have undergone a shift from stateless conversational interfaces to autonomous agents capable of multi-step planning, tool

OpenPM: Auditable Point-in-Time Evaluation for LLM Portfolio-Management Agents

AgentsDGX agent

arXiv:2608.09988v1 Announce Type: cross Abstract: Large language models are increasingly used to read markets, assess risk, and allocate capital. However, reported results for LLM trading agents can b

Optimize Cheap, Deploy Strong: Cost-Aware Cross-Tier Transfer for Evolutionary Optimization

AgentsDGX agent

arXiv:2608.10694v1 Announce Type: cross Abstract: Evolutionary optimization of LLM prompts and agentic programs (e.g., GEPA) is dominated by fitness evaluation: scoring each candidate runs an answerin

Pay with confidence: How Solv Labs built verifiable, auditable agent payments on Amazon Bedrock AgentCore payments

AgentsDGX agent

Solv Labs built a governed agent-payments workflow on Amazon Bedrock AgentCore payments, where every transaction is authorized, attested in an AWS Nitro Enclave, priced for risk, and anchored to a pub

Protection Levels for Vision-Based Pose Estimation

AgentsDGX agent

arXiv:2608.10023v1 Announce Type: cross Abstract: Vision-based navigation complements Global Navigation Satellite Systems, but certification demands integrity guarantees that account for faulty measur

Recovering Wasted Compute in Autoresearch Agents

AgentsDGX agent

arXiv:2608.10424v1 Announce Type: new Abstract: A slew of recent works develop agents for solving research problems end-to-end, a paradigm increasingly referred to as autoresearch. Such agents have in

Risk-Aware Kinodynamic Motion Planning Under Uncertainty For Safe Navigation on Planetary Environments

AgentsDGX agent

arXiv:2608.11175v1 Announce Type: new Abstract: For autonomous space exploration, robotic agents need to perform motion planning in which environmental interactions may be unknown. Learning these inte

Robust Multi-Agent Bandits with Heavy-Tailed Rewards and Information Asymmetry

AgentsDGX agent

arXiv:2608.10529v1 Announce Type: cross Abstract: The multi-armed bandit problem is a central framework in sequential decision-making, extensively studied under sub-Gaussian reward assumptions. Howeve

Scaling AI agents with trustworthy data

AgentsDGX agent

Business and technology leaders need no convincing that the time of agentic AI is here. Organizations are rapidly adopting agents, and few executives doubt the technology’s potential to transform work

Seeing above the waves: A modular sensing framework for data acquisition at sea

AgentsDGX agent

arXiv:2608.10997v1 Announce Type: new Abstract: Advancing autonomy for surface vessels requires systematic evaluation of their sensing and perception subsystems. Yet, maritime environments impose uniq

Self-evolving Agentic Customer Support System at LinkedIn

AgentsDGX agent

arXiv:2608.10224v1 Announce Type: new Abstract: Enterprise support agents operate in rapidly changing environments where policies, product capabilities, and knowledge bases evolve continuously, making

SKILLER: Language-Level Reinforcement Learning for Reusable Skill Extraction in Small Language Models

AgentsDGX agent

arXiv:2608.10538v1 Announce Type: new Abstract: Agent skills represent a standardized format for packaging procedural knowledge and domain expertise, serving within agent harness systems as an essenti

SuperQuadricOcc: Real-Time Self-Supervised Semantic Occupancy Estimation with Superquadric Volume Rendering

AgentsDGX agent

arXiv:2511.17361v5 Announce Type: replace Abstract: Self-supervision for semantic occupancy estimation is appealing as it removes the labour-intensive manual annotation, thus allowing one to scale to

the @aiDotEngineer World Fair always one of the best events every year to talk to builders at the frontier of Research, Agents, Evals, Syste…

AgentsDGX agent

the @aiDotEngineer World Fair always one of the best events every year to talk to builders at the frontier of Research, Agents, Evals, Systems, etc A few weeks ago I gave a talk on - Continually Impro

The Deliberative Deficit: An Empirical Critique of LLMs in Democratic Discourse

AgentsDGX agent

arXiv:2608.10186v1 Announce Type: cross Abstract: LLMs are increasingly deployed in settings that require collective reasoning on complex, value-laden problems. Confidence in these deployments rests l

The Signal Rail: A Deterministic Motion Grammar for Communicating Conversational Agent State in Terminal Interfaces

AgentsDGX agent

arXiv:2608.10689v1 Announce Type: cross Abstract: Terminal interfaces to conversational agents report rich internal state (listening, thinking, executing tools, awaiting input, failing) almost entirel

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workl…

AgentsDGX agent

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workloads. Start building: https://www.together.ai/models/qwen3-8

VDC-Agent: When Video Detailed Captioners Evolve Themselves via Agentic Self-Reflection

AgentsDGX agent

arXiv:2511.19436v2 Announce Type: replace-cross Abstract: Existing Video Detailed Captioning (VDC) methods predominantly rely on costly human annotations or distillation from powerful proprietary mode

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting …

AgentsDGX agent

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting each host to pass them on, then measure what governs the spr

← Previous
123…120
Next →