AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,201 results
9 Jun 2026

Taming Perception Jitter: Uncertainty-Aware LiDAR Object Detection for Reliable Motion Classification

AgentsDGX agent

arXiv:2606.09350v1 Announce Type: cross Abstract: Reliable motion classification is critical for autonomous driving, as false dynamic predictions of static objects can cascade into unnecessary planner

// The Consistency Illusion // Multi-agent debate can make agents agree on the final answer while their underlying reasoning stays misaligne…

AgentsDGX agent

// The Consistency Illusion // Multi-agent debate can make agents agree on the final answer while their underlying reasoning stays misaligned. This work finds that consensus on the output hides disagr

The Token Not Taken: Sampling, State, and the Variability of AI Agent Outputs

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2606.08998v1 Announce Type: new Abstract: Agentic AI systems can behave differently across runs: the same request may produce a different plan, a different tool call, a different code edit, or a

this is sick haha.

AgentsDGX agent

this is sick haha. The Agent Open 🎾🏓 Everyone loves pickleball. We’re hosting a massive pickleball tournament during the AI Engineer World Fair. I’m so excited to see this event come together. This is

thought-to-image

AgentsDGX agent

thought-to-image Can you believe the image on the right was generated from brain activity? We showed a person the dog image on the left for 0.5 seconds while recording brain signals. Our model, OB1, d

TianJi-Environ: An Autonomous AI Scientist for Atmospheric Environmental Research

AgentsDGX agent

arXiv:2606.07697v1 Announce Type: cross Abstract: As atmospheric environmental prediction continues to improve, interpretable validation of pollution mechanisms and feedback processes has become a mai

To Nuke or Not to Nuke: LLMs' (Missing) Ethical Reasoning and Actions in a High-Stakes Decision-Making Simulation

AgentsDGX agent

arXiv:2606.08310v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed as long-horizon agents with decision-making capacities. While LLMs can show ethical competence on

Traxia: A Framework for Verifiable, Agent-Native Scientific Publishing

AgentsDGX agent

arXiv:2606.08256v1 Announce Type: new Abstract: Verifiability, attribution, and reproducibility are foundational requirements of scientific knowledge, yet current publishing infrastructure does not en

Trustworthy Smart Fabs via Professional Proxies: Scaling Safe and Sustainable by Design (SSbD) through Industrial Data Spaces

AgentsDGX agent

arXiv:2606.09227v1 Announce Type: cross Abstract: The convergence of the 2026 European Union Safe and Sustainable by Design (SSbD) framework, Corporate Sustainability Due Diligence Directive (CSDDD),

Try out Devin today! http://devin.ai

AgentsDGX agent

Devin is an AI software engineer developed by Cognition AI that automates coding tasks and assists with software development. The post appears to be a promotional announcement encouraging users to try

VGP-Nav: Metric-Aware Visual Geometric Perception for Robot Navigation

AgentsDGX agent

arXiv:2606.09268v1 Announce Type: new Abstract: Reliable robotic navigation necessitates the seamless integration of accurate global localization and dense, metric-consistent obstacle perception. A co

ViMax: Agentic Video Generation

AgentsDGX agent

arXiv:2606.07649v1 Announce Type: cross Abstract: Long-form video generation requires systematic narrative planning and visual consistency that current short-clip methods cannot provide. Existing meth

Virtual-point-based Solutions to Handle Generalized Absolute Pose Problem

AgentsDGX agent

arXiv:2606.09294v1 Announce Type: new Abstract: Multi-camera systems are increasingly adopted in robotics and autonomous navigation for their wide field of view, flexibility, and fault tolerance. Neve

Voting Protocols as Coordination Mechanisms for Role-Constrained Multi-Agent Tutoring Systems

AgentsDGX agent

arXiv:2606.08030v1 Announce Type: cross Abstract: Agentic tutoring systems introduce a coordination challenge: multiple agents may propose different but reasonable interventions, yet only one response

What are loops, and how do you build one? A 'loop' is the repeated process where some event or input kicks off an action. For example: 1. CI…

AgentsDGX agent

What are loops, and how do you build one? A 'loop' is the repeated process where some event or input kicks off an action. For example: 1. CI fails -> you fix it 2. CI fails again -> you fix it 3. CI p

When Should an AI Scientist Stop? Verifiable Experiment Steering and Refusal for Autonomous Discovery

AgentsDGX agent

arXiv:2606.07576v1 Announce Type: new Abstract: We present CARTOGRAPH, a verification layer for AI scientists that couples unresolved-subspace experiment steering (select), explicit ambiguity closure

When Video Misreads: Closed-Loop Distillation of Reading Heuristics for Exploratory Manipulation Trace QA

AgentsDGX agent

arXiv:2606.08542v1 Announce Type: cross Abstract: Exploratory manipulation often turns an apparent failed attempt into the key evidence for what to do next. For example, a robot pulls a locked cabinet

8 Jun 2026

A year ago the closest thing we had to an AI agent was o3.

AgentsDGX agent

One year prior to this post, o3 represented the most advanced AI agent available, marking a significant milestone in AI development. The statement reflects how rapidly AI capabilities have evolved, wi

AdMem: Advanced Memory for Task-solving Agents

AgentsDGX agent

arXiv:2606.06787v1 Announce Type: new Abstract: Large Language Models (LLMs) show promise as tool-using agents but remain limited in long-horizon tasks that require remembering, organizing, and reusin

Agent Swarm & Instant Document Delivery Kimi will automatically coordinate 300 sub-agents to break down and execute your tasks. Delivers pro…

AgentsDGX agent

Agent Swarm & Instant Document Delivery Kimi will automatically coordinate 300 sub-agents to break down and execute your tasks. Delivers production-ready output in PPTX, Word, PDF, and Excel, straight

Agentic Large Language Models for Automated Structural Analysis of 3D Frame Systems

AgentsDGX agent

arXiv:2606.06525v1 Announce Type: cross Abstract: Large language models (LLMs) have emerged as powerful foundation models with strong reasoning capabilities across domains. Beyond reactive text genera

Agentopia: Long-Term Life Simulation and Learning in Agent Societies

AgentsDGX agent

arXiv:2606.07513v1 Announce Type: new Abstract: Humans learn from social life. Simulating this process with LLM-powered agents represents a promising research direction, raising a natural question: wh

AMD-FCG: An Enhanced Function Call Graph Dataset with Integrated Topological Features for Malware Detection and Classification

AgentsDGX agent

arXiv:2606.06815v1 Announce Type: cross Abstract: As malware illustrates a complex structure and behavior, detection of these has been a significant challenge in the domain of cybersecurity along with

AnchorWorld: Embodied Egocentric World Simulation with View-based Evolution Customization

AgentsDGX agent

arXiv:2606.07326v1 Announce Type: new Abstract: Despite being a pivotal frontier, interactive world modeling remains underexplored in terms of the versatile controllability required by practical scena

Apple announces a new Foundation Models framework for developers, a new Core AI framework, and a set of Xcode enhancements aimed at agentic coding workflows (Hartley Charlton/MacRumors)

AgentsDGX agent

Hartley Charlton / MacRumors: Apple announces a new Foundation Models framework for developers, a new Core AI framework, and a set of Xcode enhancements aimed at agentic coding workflows — Apple today

Apple unveils an Apple Intelligence feature to automatically change compromised passwords, using agentic AI, and save them to the Passwords app (James Pero/Gizmodo)

AgentsDGX agent

James Pero / Gizmodo: Apple unveils an Apple Intelligence feature to automatically change compromised passwords, using agentic AI, and save them to the Passwords app — Agentic AI and security are norm

At @tryramp, engineers use Devin Desktop to bring their favorite agents into one place. With Devin Desktop they can dispatch, monitor, and j…

AgentsDGX agent

At @tryramp, engineers use Devin Desktop to bring their favorite agents into one place. With Devin Desktop they can dispatch, monitor, and jump between agents from a single surface with shared context

Autonomous computational catalysis through an agentic research system

AgentsDGX agent

arXiv:2601.13508v4 Announce Type: replace-cross Abstract: Autonomous agents are beginning to transform scientific research from tool-assisted workflows toward self-sustaining discovery processes. Comp

AutoTool: Dynamic Tool Selection and Integration for Agentic Reasoning

AgentsDGX agent

arXiv:2512.13278v2 Announce Type: replace Abstract: Agentic reinforcement learning has advanced large language models (LLMs) to reason through long chain-of-thought trajectories while interleaving ext

@cognition funny you should say that https://x.com/manuelsampedrop/status/2063746243180773656?s=20

AgentsDGX agent

@cognition funny you should say that https://x.com/manuelsampedrop/status/2063746243180773656?s=20 @swyx @cognition hope it scores scope discipline too, half my agent failures are perfectly fine code

CRAFT: Coaching Reinforcement Learning Autonomously using Foundation Models for Multi-Robot Coordination Tasks

AgentsDGX agent

arXiv:2509.14380v3 Announce Type: replace Abstract: Multi-Agent Reinforcement Learning (MARL) provides a powerful framework for learning coordination in multi-agent systems. However, applying MARL to

Deep Agents explained in <90 seconds by @sydneyrunkle

AgentsDGX agent

Deep Agents are an AI concept that Sydney Runkle explains concisely in under 90 seconds, likely covering how agents can be designed to operate with deeper reasoning and decision-making capabilities. T

deepagents in 90 seconds!

AgentsDGX agent

DeepAgents is a framework or tool designed to enable the creation of intelligent agents, likely demonstrated or explained in a brief 90-second format by Harrison Chase, a co-founder of LangChain. The

Do Coding Agents Deceive Us? Detecting and Preventing Cheating via Capped Evaluation with Randomized Tests

AgentsDGX agent

arXiv:2606.07379v1 Announce Type: cross Abstract: A growing failure mode in agent evaluation and training is that models can achieve high evaluation scores by exploiting shortcuts instead of solving t

Does a token buy you more or less now than it did a few months ago? We built a consumer price index (CPI) for AI coding output from Anthropi…

AgentsDGX agent

Does a token buy you more or less now than it did a few months ago? We built a consumer price index (CPI) for AI coding output from Anthropic's Opus 4.6 model in SWE-chat, Feb 5–Apr 15, 2026. What we

Dual Latent Memory for Visual Multi-agent System

AgentsDGX agent

arXiv:2602.00471v2 Announce Type: replace Abstract: While Visual Multi-Agent Systems (VMAS) promise to enhance comprehensive abilities through inter-agent collaboration, empirical evidence reveals a c

DuMate-DeepResearch: An Auditable Multi-Agent System with Recursive Search and Rubric-Grounded Reasoning

AgentsDGX agent

arXiv:2606.07299v1 Announce Type: new Abstract: Deep Research (DR) has emerged as a new agentic paradigm to tackle complex, open-ended research tasks, demanding systems that can iteratively frame prob

Efficient Coordination and Synchronization of Multi-Robot Systems Under Recurring Linear Temporal Logic

AgentsDGX agent

arXiv:2502.16531v2 Announce Type: replace Abstract: We consider multi-robot systems under recurring tasks formalized as linear temporal logic (LTL) specifications. To solve the planning problem effici

Evaluate your Amazon Nova Sonic voice agent at scale, no microphone required

AgentsDGX agent

In this post, we walk you through the Nova Sonic Test Harness, an open source framework that we built to solve both problems. It serves as a rapid iteration tool for tuning system prompts and tool con

Exploring Agentic Tool-Calling Decisions via Uncertainty-Aligned Reinforcement Learning

AgentsDGX agent

arXiv:2606.06976v1 Announce Type: new Abstract: Large language model (LLM)-based agents often make suboptimal tool-use decisions, including unsupported tool invocation and hallucinated direct response

Feasible Action Space Reduction for Quantifying Causal Responsibility in Continuous Spatial Interactions

AgentsDGX agent

arXiv:2505.17739v2 Announce Type: replace-cross Abstract: Understanding the causal influence of one agent on another agent is crucial for safely deploying artificially intelligent systems such as auto

For the 2nd time in weeks, Microsoft packages laced with credential stealer

AgentsDGX agent

Microsoft's durabletask PyPI package was poisoned in May 2026 with credential-harvesting malware that steals secrets from AWS, Azure, GCP, Kubernetes, and 90+ developer tools . In June, 73 Microsoft r

From Pixels to Shelf: An Integrated Robotic System for Autonomous Supermarket Stocking with a Mobile Manipulator

AgentsDGX agent

arXiv:2509.11740v2 Announce Type: replace Abstract: Autonomous stocking in retail environments, particularly supermarkets, presents challenges due to dynamic human interactions, constrained spaces, an

From Privacy to Workflow Integrity: Communication-Graph Metadata in Autonomous Agent Interoperability

AgentsDGX agent

arXiv:2606.07150v1 Announce Type: cross Abstract: Agent-interoperability protocols such as A2A and MCP standardize what agents say to one another, but assume address-based transport over HTTP(S). Such

FrontierCode has three task sets: Extended (150 tasks), Main (100 tasks) and Diamond (50 tasks). SOTA LLMs have significant room for improve…

AgentsDGX agent

FrontierCode has three task sets: Extended (150 tasks), Main (100 tasks) and Diamond (50 tasks). SOTA LLMs have significant room for improvement, with the top model earning a score of just 13.4/100 on

GOPAgen: Motion-Aware and Efficient Agentic Long-Video Understanding with Structural Memory and Hierarchical Reasoning

AgentsDGX agent

arXiv:2606.06532v1 Announce Type: new Abstract: Despite significant progress in agentic long video understanding, existing methods still lack detailed motion comprehension coupled with an efficient me

here's what it means to create a loop, and how we use them to automate our work

AgentsDGX agent

This post likely explains the concept of loops in programming or automation workflows, covering how they function and their practical applications for automating repetitive tasks. It probably includes

How AI Agents Reshape Knowledge Work: Autonomy, Efficiency, and Scope

AgentsDGX agent

arXiv:2606.07489v1 Announce Type: new Abstract: Frontier AI systems are bridging the gap between intelligence and utility by shifting from conversational assistants to autonomous agents that execute t

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a…

AgentsDGX agent

I really do think we'll see a lot of value accrue in AI startups building 'model routing as a service' Not just OpenRouter - this includes a much broader set of verticalized agents and infrastructure.

IDDMBSE: Integrating Data-Driven and Model-Based Systems Engineering for Trusted Autonomous Cyber-Physical Systems

AgentsDGX agent

arXiv:2606.06727v1 Announce Type: new Abstract: Autonomous cyber-physical systems (CPS) sit at the intersection of Model-Based Systems Engineering (MBSE) and data-driven Machine Learning and Artificia

Introducing FrontierCode: a coding eval that raises the bar for difficulty & quality. Each task took 40+ hrs of work by leading open-source …

AgentsDGX agent

Introducing FrontierCode: a coding eval that raises the bar for difficulty & quality. Each task took 40+ hrs of work by leading open-source maintainers. Models write sloppy code that works but isn’t m

IRAF: Interference-Resilient Adaptive Fusion for Noise-Robust End-to-End Full-Duplex Spoken Dialogue Systems

AgentsDGX agent

arXiv:2606.06559v1 Announce Type: cross Abstract: Full-duplex spoken dialogue models allow voice agents to listen and speak concurrently, enabling natural interaction with real-time overlap. However,

It's finally out!!! @METR_Evals found that more than half of SWEBench results is unmergeable slop. FrontierCode represents over 1000+ hours …

AgentsDGX agent

It's finally out!!! @METR_Evals found that more than half of SWEBench results is unmergeable slop. FrontierCode represents over 1000+ hours of maintainer validated software engineering work most front

Kimi Code, our open-source coding agent, just got a major upgrade! 🔹One-line CLI install, zero setup, fast startup​ 🔹Drag in videos as cod…

AgentsDGX agent

Kimi Code, our open-source coding agent, just got a major upgrade! 🔹One-line CLI install, zero setup, fast startup​ 🔹Drag in videos as coding context: reference-to-LUT, long-video-to-short, screen-rec

LangSmith Fleet lets you work with files directly. Create documents, presentations, and webpages inside a conversation, or upload your own f…

AgentsDGX agent

LangSmith Fleet introduces file handling capabilities that enable users to create and manipulate documents, presentations, and webpages directly within conversations, as well as upload existing files

LLM Agent-Assisted Reverse Engineering with Quantitative Readability Metrics

AgentsDGX agent

arXiv:2606.06838v1 Announce Type: cross Abstract: Automatic decompilers produce functionally correct but often unreadable C code. This paper addresses one stage of the reverse engineering workflow: im

Measuring Agents in Production

AgentsDGX agent

arXiv:2512.04123v4 Announce Type: replace-cross Abstract: LLM-based agents already operate in production across many industries, yet we lack an understanding of what technical methods make deployments

Model Context Protocols in Adaptive Transport Systems: A Survey

AgentsDGX agent

arXiv:2508.19239v2 Announce Type: replace Abstract: The rapid expansion of interconnected devices, autonomous systems, and AI applications has created severe fragmentation in adaptive transport system

Multi-Agent Reasoning with Consistency Verification Improves Uncertainty Calibration in Medical MCQA

AgentsDGX agent

arXiv:2603.24481v2 Announce Type: replace Abstract: Miscalibrated confidence scores are a practical obstacle to deploying AI in clinical settings. A model that is always overconfident offers no useful

New paper on how AI agents are reshaping knowledge work. This is a nice economic read on where agents actually change knowledge work to meet…

AgentsDGX agent

New paper on how AI agents are reshaping knowledge work. This is a nice economic read on where agents actually change knowledge work to meet that gap directly. (bookmark it) It studies agent adoption

← Previous
1…4344454647…121
Next →