AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,214 results
4 Jun 2026

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for in…

AgentsDGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 198B sparse MoE VLM designed by @StepFun_ai for inference from the start. 196B language backbone with a 1.8B v

MapAgent: An Industrial-Grade Agentic Framework for City-scale Lane-level Map Generation

AgentsDGX agent

arXiv:2606.04513v1 Announce Type: new Abstract: Lane-level maps are critical infrastructure for autonomous driving and lane-level navigation, yet constructing and maintaining standardized lane network

MetaPoint: Unlocking Precise Spatial Control in Agentic Visual Generation

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.05031v1 Announce Type: new Abstract: Generative visual models fundamentally struggle with precise spatial control. This arises from a core disconnect: models can process textual description

MIRAGE: Mobile Agents with Implicit Reasoning and Generative World Models

AgentsDGX agent

arXiv:2606.04627v1 Announce Type: new Abstract: Mobile agents are increasingly expected to operate everyday applications from screenshots and language goals, where reliable control requires reasoning

More about the guarantee and eligibility criteria: https://devin.ai/guarantee

AgentsDGX agent

Devin is an AI software engineer product that likely includes a performance guarantee and specific eligibility requirements for users, with details available on Cognition AI's official guarantee page.

New in the LangSmith Sandboxes GA Release: Sandbox CLI ✅Build snapshots from Dockerfiles ✅Manage sandboxes ✅Open interactive consoles ✅Tunne…

AgentsDGX agent

New in the LangSmith Sandboxes GA Release: Sandbox CLI ✅Build snapshots from Dockerfiles ✅Manage sandboxes ✅Open interactive consoles ✅Tunnel raw TCP ✅Use standard tools (ssh, scp, rsync, sftp) agains

nice post from @Harvey and @LangChain Labs, worth a read. improve agent feedback loop without setting $$ on fire

AgentsDGX agent

nice post from @Harvey and @LangChain Labs, worth a read. improve agent feedback loop without setting $$ on fire Can we design legal agent verifiers that are up to 1,000x cheaper? Verifiers are LLM ju

Notarized Agents: Receiver-Attested Confidential Receipts for AI Agent Actions

AgentsDGX agent

arXiv:2606.04193v1 Announce Type: cross Abstract: Current AI agent observability is structurally compromised: the entity producing the activity log is the same entity whose activity is being logged. A

okay but need to save some $$$ for tokens

AgentsDGX agent

This post likely discusses cost optimization strategies for managing API token usage and expenses, particularly relevant to users of language models or LLM-based services. The author emphasizes the pr

On the latest episode of Max Agency, @hwchase17 sat down with @nlarusstone, Head of AI at @benchling for a conversation on building agents f…

AgentsDGX agent

On the latest episode of Max Agency, @hwchase17 sat down with @nlarusstone, Head of AI at @benchling for a conversation on building agents for scientific work. ⏯️ YouTube: https://www.youtube.com/watc

Optimizing the Cost-Quality Tradeoff of Agentic Theorem Provers in Lean

AgentsDGX agent

arXiv:2606.04883v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used in workflows for generating formal proofs in Lean. These workflows often decompose problems into smal

Parthenon Law: A Self-Evolving Legal-Agent Framework

AgentsDGX agent

arXiv:2606.04602v1 Announce Type: new Abstract: As agents grow more capable, legal-domain LLM agents promise to turn document-heavy matters into reviewable work products -- yet reliable deployment fac

Poke, which lets users access AI agents via text message, becomes the first AI agent approved for Apple's Messages for Business platform (Sarah Perez/TechCrunch)

AgentsDGX agent

Sarah Perez / TechCrunch: Poke, which lets users access AI agents via text message, becomes the first AI agent approved for Apple's Messages for Business platform — Poke, a startup that turns using AI

Position: Deployed Reinforcement Learning should be Continual

AgentsDGX agent

arXiv:2606.04029v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has received increasing attention and adoption in real-world use cases. Most of these systems follow a train-then-fix para

Proud to announce that Cohere has been awarded first place in NATO’s Agentic AI for Cognitive Warfare Innovation Challenge. Congratulations …

AgentsDGX agent

Proud to announce that Cohere has been awarded first place in NATO’s Agentic AI for Cognitive Warfare Innovation Challenge. Congratulations to our fellow finalists: OpenMinds, which secured second pla

Provably Auditable and Safe LLM Agents from Human-Authored Ontologies

AgentsDGX agent

arXiv:2606.04903v1 Announce Type: cross Abstract: We introduce the LLM agent architecture Agentic Redux, intended for use with nontrivial problem domains that require linear auditability. Using the ty

Radiant Logic extends identity visibility platform to enterprise AI agents with real-time risk scoring

AgentsDGX agent

Radiant Logic Inc., a platform that provides identity visibility and intelligence, today announced it’s extending its services to agentic artificial intelligence to help companies control and govern t

Real-World Deployment of a 5G-Connected Edge-Controlled Aerial Robot in Industrial Subterranean Mines

AgentsDGX agent

arXiv:2606.04818v1 Announce Type: new Abstract: This article presents the first real-world autonomous flight of a 5G-connected aerial robot controlled by an edge-offloaded controller, and aims to brid

Recent Advances and Trends in Learning-based 3D Representations

AgentsDGX agent

arXiv:2606.04871v1 Announce Type: new Abstract: The selection of an appropriate 3D representation is a fundamental design decision that dictates the efficiency, quality, and capabilities of modern com

Scaling Datasets for Multi-Sensor, Multi-Agent, and Multi-Domain Learning in Autonomous Systems

AgentsDGX agent

arXiv:2606.04444v1 Announce Type: cross Abstract: Existing datasets cannot support large-scale learning in multi-agent, multi-sensor, or multi-domain autonomy, where diversity and coordination are ess

Self-Reflective APIs: Structure Beats Verbosity for AI Agent Recovery

AgentsDGX agent

arXiv:2606.05037v1 Announce Type: cross Abstract: When an AI agent calls an API and hits a validation error, it needs more than what went wrong -- it needs what to do next. A self-reflective API retur

Semantic Constraint Synthesis for Adaptive Trajectory Optimization via Large Language Models

AgentsDGX agent

arXiv:2606.04123v1 Announce Type: cross Abstract: Trajectory optimization is a critical component for enabling safe and reliable autonomous operations in space exploration. As space missions increase

SePO: Self-Evolving Prompt Agent for System Prompt Optimization

AgentsDGX agent

arXiv:2606.04465v1 Announce Type: cross Abstract: System prompt optimization improves agent behavior without modifying the underlying model, yielding human-readable, model-agnostic instructions. Exist

ShareVerse: Multi-Agent Consistent Video Generation for Shared World Modeling

AgentsDGX agent

arXiv:2603.02697v2 Announce Type: replace-cross Abstract: This paper presents ShareVerse, a video generation framework enabling multi-agent shared world modeling, addressing the gap in existing works

Sometimes you need to start over. But that decision is hard. @llama_index had to make that call: they built one of the most popular AI frame…

AgentsDGX agent

Sometimes you need to start over. But that decision is hard. @llama_index had to make that call: they built one of the most popular AI frameworks in the world, but saw the agent harness and frontier l

Stateful Visual Encoders for Vision-Language Models

AgentsDGX agent

arXiv:2606.04433v1 Announce Type: cross Abstract: Vision-language models (VLMs) are increasingly used in multi-image, multi-turn agentic settings where decisions depend on visual changes. However, in

Strabo: Declarative Specification and Implementation of Agentic Interaction Protocols

AgentsDGX agent

arXiv:2606.05043v1 Announce Type: new Abstract: The last few years have witnessed major advances in the modeling and implementation of multiagent systems based on declarative interaction protocols. Ou

SUSD: Structured Unsupervised Skill Discovery through State Factorization

AgentsDGX agent

arXiv:2602.01619v2 Announce Type: replace-cross Abstract: Unsupervised Skill Discovery (USD) aims to autonomously learn a diverse set of skills without relying on extrinsic rewards. One of the most co

This is the enterprise-agent security frame I like: treat agents like untrusted developers. They need the ability to call an API, not posses…

AgentsDGX agent

This is the enterprise-agent security frame I like: treat agents like untrusted developers. They need the ability to call an API, not possession of the credential. Control plane outside runtime. Fail

TIDE: Proactive Multi-Problem Discovery via Template-Guided Iteration

AgentsDGX agent

arXiv:2606.04743v1 Announce Type: cross Abstract: Agents are widely deployed as assistants over documents, tools, and code. However, they typically act only on explicit user requests, which surface on

Topology Matters: Measuring Memory Leakage in Multi-Agent LLMs

AgentsDGX agent

arXiv:2512.04668v4 Announce Type: replace-cross Abstract: Graph topology is a fundamental determinant of memory leakage in multi-agent LLM systems, yet its effects remain poorly quantified. We introdu

Toward Autonomous O-RAN: A Multi-Scale Agentic AI Framework for Real-Time Network Control and Management

AgentsDGX agent

arXiv:2602.14117v2 Announce Type: replace-cross Abstract: Open Radio Access Networks (O-RAN) promise flexible 6G network access through disaggregated, software-driven components and open interfaces, b

Trivium: Temporal Regret as a First-Class Objective for Causal-Memory Controllers

AgentsDGX agent

arXiv:2606.04421v1 Announce Type: new Abstract: Many current agentic systems and LLM pipelines correct mistakes by optimizing outcome reward. This addresses only the what of failure: when an outcome d

Vectorized Online POMDP Planning

AgentsDGX agent

arXiv:2510.27191v5 Announce Type: replace-cross Abstract: Planning under partial observability is an essential capability of autonomous robots. The Partially Observable Markov Decision Process (POMDP)

We have a Fleet agent called @docs_plz in our Slack that's made a very noticeable impact on our velocity of docs changes. In the chart below…

AgentsDGX agent

We have a Fleet agent called @docs_plz in our Slack that's made a very noticeable impact on our velocity of docs changes. In the chart below, you can clearly see that after it was added, the amount of

We partnered with Shopify so you can go from idea to live store in minutes Just tell Replit Agent what you want to sell. It will: - Build a …

AgentsDGX agent

We partnered with Shopify so you can go from idea to live store in minutes Just tell Replit Agent what you want to sell. It will: - Build a custom storefront - Create your Shopify store - Help you add

we're taking an early bet on open models specifically, because they're SO much cheaper one point of reference: an app outputting 10M tokens/…

AgentsDGX agent

we're taking an early bet on open models specifically, because they're SO much cheaper one point of reference: an app outputting 10M tokens/day costs roughly 250/day on Opus 4.8 versus ~24/day for Min

⚠️Why didn’t the hyperscalers wait until after the IPO in switching to pay-by-usage-charging? My guess is that *they literally could not aff…

AgentsDGX agent

⚠️Why didn’t the hyperscalers wait until after the IPO in switching to pay-by-usage-charging? My guess is that *they literally could not afford to* —because it would bankrupt them. My reasoning: in “a

Why the semantic layer is becoming the foundation for trusted agentic AI

AgentsDGX agent

Agentic AI is making the semantic layer an essential enterprise priority because headless agents asking thousands of questions simultaneously have zero tolerance for inconsistent data definitions. Thi

3 Jun 2026

A Training-Free Mixture-of-Agents Framework for Multi-Document Summarization using LLMs and Knowledge Graphs

AgentsDGX agent

arXiv:2606.03867v1 Announce Type: cross Abstract: Multi-Document Summarization (MDS) plays a critical role in distilling essential information from collections of textual data. Existing approaches oft

Adaptive Latent Agentic Reasoning

AgentsDGX agent

arXiv:2606.02871v1 Announce Type: cross Abstract: Large reasoning models improve performance by generating extended chain-of-thought (CoT) reasoning, but this behavior becomes inefficient when applied

Adding MCP Tools to Reachy Mini

AgentsDGX agent

This article describes how to integrate Model Context Protocol (MCP) tools with Reachy Mini, a small humanoid robot platform. It likely covers the technical steps for extending Reachy Mini's capabilit

Agentic Chain-of-Thought Steering for Efficient and Controllable LLM Reasoning

AgentsDGX agent

arXiv:2606.03965v1 Announce Type: cross Abstract: Large language models improve final-answer accuracy through extended chain-of-thought reasoning, but often spend tokens inefficiently and offer little

At @harvey, the engineering team integrated Spectre — their internal background agent — into Devin Desktop. Now Spectre's organizational con…

AgentsDGX agent

At @harvey, the engineering team integrated Spectre — their internal background agent — into Devin Desktop. Now Spectre's organizational context can live on every engineer's laptop and flow across the

AUGUSTE: Online-Learning dApp for Predictive URLLC Scheduling

AgentsDGX agent

arXiv:2606.03664v1 Announce Type: cross Abstract: Ultra Reliable and Low Latency Communications (URLLC) was one of the main motivations behind 5G, with 3GPP advertising 1-10 ms latency targets for app

Automata-Conditioned Cooperative Multi-Agent Reinforcement Learning

AgentsDGX agent

arXiv:2511.02304v2 Announce Type: replace-cross Abstract: We study learning multi-task, multi-agent policies for cooperative, temporal objectives, under centralized training, decentralized execution.

Autonomous Navigation System for Library Service Robot Based on Unitree Go2 Edu

AgentsDGX agent

arXiv:2606.03340v1 Announce Type: new Abstract: Libraries require autonomous robots to move quietly through narrow aisles while remaining safe around readers, chairs, bags, and carts. This paper prese

BotDirector: Robot Storytelling Across the Symmetrical Reality with Multi-modal Interactions

AgentsDGX agent

arXiv:2606.03223v1 Announce Type: cross Abstract: Robot storytelling offers a unique blend of technological innovation and creative expression that engages children in unprecedented ways. However, the

Bring your favorite agent into Devin Desktop using ACP: https://docs.devin.ai/desktop/acp

AgentsDGX agent

Devin Desktop now supports Agent Control Protocol (ACP), allowing users to integrate their preferred AI agents into the platform. This feature enables users to bring custom or third-party agents into

Build 2026: From observability to ROI for AI agents on any framework

AgentsDGX agent

9 min read · June 3, 2026 · Sebastian Kohlmeier Shipping an AI agent is the easy part. Keeping it accurate, safe, and accountable in production is where teams get stuck. Agents are non-deterministic.

Build smarter document workflows: What’s new in Azure Content Understanding at Build 2026

AgentsDGX agent

Azure Content Understanding (CU) in Foundry Tools is Microsoft’s comprehensive content AI service. It ingests diverse data types — documents, audio, images, and video — and extracts the most critical

Capability Advertisement as a Market for Lemons: A Trust Layer for Heterogeneous Agent Networks

AgentsDGX agent

arXiv:2606.03034v1 Announce Type: cross Abstract: Large language model (LLM) agents have begun to delegate work to one another. Protocols such as the Model Context Protocol (MCP) and the Agent2Agent p

CARVE: Certified Affordable Repair of Vetoed Maneuvers via Envelopes for Interactive Driving

AgentsDGX agent

arXiv:2606.02641v1 Announce Type: cross Abstract: Interactive driving exposes a failure mode that is easy to miss in rule-aware autonomous-driving stacks: a hard-rule margin can be negative for an ego

Chopped it up with @swyx on @latentspacepod and we ran the gamut on this one. We talked platform, how roles are evolving, the agentic era, t…

AgentsDGX agent

Chopped it up with @swyx on @latentspacepod and we ran the gamut on this one. We talked platform, how roles are evolving, the agentic era, the future of open source, and what we’re building next. Spoi

Closed-Loop Molecular Design with Calibrated Deference

AgentsDGX agent

arXiv:2606.02618v1 Announce Type: cross Abstract: We present Cognitive Loop via In-Situ Optimization (CLIO), an agent that couples a continuously-updated belief-state graph with a recursive plan-then-

Co-evolving Agent Architectures and Interpretable Reasoning for Automated Optimization

AgentsDGX agent

arXiv:2604.17708v2 Announce Type: replace Abstract: Automating operations research (OR) with large language models (LLMs) remains limited by hand-crafted reasoning--execution workflows. Complex OR tas

CodeHacker: Automated Test Case Generation for Detecting Vulnerabilities in Competitive Programming Solutions

AgentsDGX agent

arXiv:2602.20213v2 Announce Type: replace-cross Abstract: The evaluation of Large Language Models (LLMs) for code generation relies heavily on the quality and robustness of test cases. However, existi

Decentralized Stochastic Nonconvex Optimization under the (L_0,L_1)-Smoothness

AgentsDGX agent

arXiv:2509.08726v3 Announce Type: replace-cross Abstract: This paper focuses on the decentralized stochastic optimization problem f(mathbf{x})=frac{1}{m}sum_{i=1}^m f_i(mathbf{x}) over a connected net

DELTAMEM: Incremental Experience Memory for LLM Agents via Residual Trees

AgentsDGX agent

arXiv:2606.03083v1 Announce Type: new Abstract: Large Language Model (LLM)-based agents increasingly rely on memory to learn from experiences over continual interactions. However, storing experiences

DyaPlex: Full-Duplex Speech-Motion Model for Dyadic Interaction

AgentsDGX agent

arXiv:2606.03874v1 Announce Type: new Abstract: We present DyaPlex, a streaming, full-duplex speech-and-motion model designed for dyadic interaction. To capture the continuous and reciprocal nature of

← Previous
1…4748495051…121
Next →