AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,920 results
4 May 2026

Affordance Agent Harness: Verification-Gated Skill Orchestration

AgentsDGX agent

arXiv:2605.00663v1 Announce Type: cross Abstract: Affordance grounding requires identifying where and how an agent should interact in open-world scenes, where actionable regions are often small, occlu

Agentic development demands a multi-model strategy — and the governance to match

AgentsDGX agent

The rapid rise of agentic development has radically transformed the software engineering landscape, compelling enterprises to embrace a multi-model AI ecosystem. The pace of change in software develop

Cisco agrees to acquire Astrix Security, which helps companies monitor and control the permissions granted to AI agents, a source says for approximately $400M (Meir Orbach/CTech)

AgentsDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Meir Orbach / CTech: Cisco agrees to acquire Astrix Security, which helps companies monitor and control the permissions granted to AI agents, a source says for approximately $400M — The Israeli startu

Disentangled Control of Multi-Agent Systems

AgentsDGX agent

arXiv:2511.05900v3 Announce Type: replace-cross Abstract: This paper develops a general framework for multi-agent control synthesis, which applies to a wide range of problems with convergence guarante

Don't let Anthropic own your stack. @jerryjliu0 on why modular architecture is the only real hedge in the agent era. https://www.youtube.com…

AgentsDGX agent

Jerry Liu advocates for modular architecture as a strategy to avoid vendor lock-in with AI providers like Anthropic during the emergence of AI agents. He argues that building flexible, interchangeable

Group Cognition Learning: Making Everything Better Through Governed Two-Stage Agents Collaboration

AgentsDGX agent

arXiv:2605.00370v1 Announce Type: new Abstract: Centralized multimodal learning commonly compresses language, acoustic, and visual signals into a single fused representation for prediction. While effe

Introducing agent quality optimization in AgentCore, now in preview

AgentsDGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

Introducing the agent performance loop: AgentCore Optimization now in preview

AgentsDGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

Introducing the agent quality loop: AgentCore Optimization now in preview

AgentsDGX agent

Generate recommendations from production traces, validate them with batch evaluation and A/B testing, and ship with confidence. AI agents that perform well at launch don’t stay that way. As models evo

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.00444v1 Announce Type: new Abstract: Multi-modal large language models (MLLMs) advance vision language understanding but face inherent limitations in long-video tasks due to bounded percept

2 May 2026

We are gonna have genuinely useful long running agents in 2026. Somehow, this feels both entirely predictable, yet scarily futuristic

AgentsDGX agent

Jerry Liu predicts that genuinely useful long-running AI agents will become a reality by 2026, viewing this development as simultaneously inevitable and remarkable. The post reflects on how advances i

1 May 2026

Agentic Compilation: Mitigating the LLM Rerun Crisis for Minimized-Inference-Cost Web Automation

AgentsDGX agent

arXiv:2604.09718v2 Announce Type: cross Abstract: LLM-driven web agents operating through continuous inference loops -- repeatedly querying a model to evaluate browser state and select actions -- exhi

Agentic engineering startup JuliaHub lands $65M to automate design and testing of industrial products

AgentsDGX agent

JuliaHub Inc., an “agentic” industrial engineering startup that’s trying to automate complex manufacturing processes with artificial intelligence, said today it has raised 65 million in a Series B fun

AgenticRecTune: Multi-Agent with Self-Evolving Skillhub for Recommendation System Optimization

Model ReleasesDGX agent

arXiv:2604.26969v1 Announce Type: cross Abstract: Modern large-scale recommendation systems are typically constructed as multi-stage pipelines, encompassing pre-ranking, ranking, and re-ranking phases

From surveillance to signalling: escalation channels as environmental controls for agentic AI

SafetyDGX agent

arXiv:2510.05192v2 Announce Type: replace-cross Abstract: When AI agents operating with access to sensitive information encounter a conflict between completing an assigned task and following rules or

MM-StanceDet: Retrieval-Augmented Multi-modal Multi-agent Stance Detection

AgentsDGX agent

arXiv:2604.27934v1 Announce Type: new Abstract: Multimodal Stance Detection (MSD) is crucial for understanding public discourse, yet effectively fusing text and image, especially with conflicting sign

Replit Agent is free tomorrow for everyone starting at 5am PST Show use what you can build in 24 hours And Replit is turning10! A trip down …

AgentsDGX agent

Replit is offering free access to Replit Agent to all users for 24 hours starting at 5am PST, allowing people to demonstrate what they can build in that time period. The announcement coincides with Re

Should you use a sandbox for your agent? @ListenLabs Co-Founder & CTO @florian_jue shared what can go wrong on the Max Agency podcast hosted…

AgentsDGX agent

This post discusses the use of sandboxes for AI agents, featuring insights from Florian Jue, co-founder and CTO of Listen Labs, during an appearance on the Max Agency podcast. The discussion likely co

The US, UK, Australia, Canada, and New Zealand publish guidance on orgs' use of agentic AI systems, saying many give AI more access than can be safely monitored (Greg Otto/CyberScoop)

AgentsDGX agent

Greg Otto / CyberScoop: The US, UK, Australia, Canada, and New Zealand publish guidance on orgs' use of agentic AI systems, saying many give AI more access than can be safely monitored — The guidance

Think it, Run it: Autonomous ML pipeline generation via self-healing multi-agent AI

AgentsDGX agent

arXiv:2604.27096v1 Announce Type: new Abstract: The purpose of our paper is to develop a unified multi-agent architecture that automates end-to-end machine learning (ML) pipeline generation from datas

Trace-Level Analysis of Information Contamination in Multi-Agent Systems

AgentsDGX agent

arXiv:2604.27586v1 Announce Type: new Abstract: Reasoning over heterogeneous artifacts (PDFs, spreadsheets, slide decks, etc.) increasingly occurs within structured agent workflows that iteratively ex

30 Apr 2026

1/5 We hit SOTA performance on BrowseComp-Plus with 95.18% accuracy using AI21 Maestro’s agent optimization. Here’s how we automated the sea…

AgentsDGX agent

AI21 Labs achieved state-of-the-art performance on the BrowseComp-Plus benchmark with 95.18% accuracy using their AI21 Maestro model with agent optimization techniques. The post appears to discuss aut

DigiCert debuts AI Trust framework to secure agents, models and content

AgentsDGX agent

Digital security company DigiCert Inc. today introduced a new AI Trust framework to help organizations secure AI systems and their outputs, along with new capabilities to help secure autonomous agents

EvolvingAgent: Curriculum Self-evolving Agent with Continual World Model for Long-Horizon Tasks

AgentsDGX agent

arXiv:2502.05907v3 Announce Type: replace Abstract: Completing Long-Horizon (LH) tasks in open-ended worlds is an important yet difficult problem for embodied agents. Existing approaches suffer from t

Google Cloud is rebuilding the enterprise stack for the age of agents

AgentsDGX agent

Google Cloud is assembling an ambitious agentic enterprise stack it believes will close the gap between AI ambition and real business outcomes. But the transition from traditional, linear workflows to

How Vultr Enables Agentic AI Experiences with AMD

HardwareDGX agent

Vultr partnered with AMD to provide infrastructure and computing capabilities that support autonomous AI agents and agentic AI applications. The content likely discusses how Vultr's cloud platform, en

Introducing Hermes Curator! The new system built in to Hermes Agent now helps you keep your skills that the self improvement loop creates in…

AgentsDGX agent

Introducing Hermes Curator! The new system built in to Hermes Agent now helps you keep your skills that the self improvement loop creates in check, by consolidating and pruning automatically. The cura

OxyGent: Making Multi-Agent Systems Modular, Observable, and Evolvable via Oxy Abstraction

AgentsDGX agent

arXiv:2604.25602v2 Announce Type: replace Abstract: Deploying production-ready multi-agent systems (MAS) in complex industrial environments remains challenging due to limitations in scalability, obser

🎂 Replit turns 10 this weekend, and we're giving every Replit user free Agent for 24 hours on Saturday.🤯 Clear your Saturday. And tune in …

AgentsDGX agent

🎂 Replit turns 10 this weekend, and we're giving every Replit user free Agent for 24 hours on Saturday.🤯 Clear your Saturday. And tune in live tomorrow, Friday at 9 AM PT, when we go live to walk you

Training Computer Use Agents to Assess the Usability of Graphical User Interfaces

AgentsDGX agent

arXiv:2604.26020v1 Announce Type: cross Abstract: Usability testing with experts and potential users can assess the effectiveness, efficiency, and user satisfaction of graphical user interfaces (GUIs)

29 Apr 2026

'Agent optimization should be automatic, efficient, observable and future-proof' (Or Dagan, 2 minutes ago)

AgentsDGX agent

Agent optimization requires four key characteristics: automation to reduce manual intervention, efficiency to maximize performance with minimal resource waste, observability to enable monitoring and d

BenchGuard: Who Guards the Benchmarks? Automated Auditing of LLM Agent Benchmarks

Model ReleasesDGX agent

arXiv:2604.24955v1 Announce Type: new Abstract: As benchmarks grow in complexity, many apparent agent failures are not failures of the agent at all - they are failures of the benchmark itself: broken

Full Workshop: @OpenAI Codex masterclass The agent is no longer just one chat window. In this workshop, @reach_vb and @kagigz get into how c…

AgentsDGX agent

Full Workshop: @OpenAI Codex masterclass The agent is no longer just one chat window. In this workshop, @reach_vb and @kagigz get into how coding systems start to change when you can delegate work acr

.@MadrigalPharma’s multi-agent research and intelligence platform for pharma is powered by LangChain + LangSmith. When users ask research qu…

AgentsDGX agent

.@MadrigalPharma’s multi-agent research and intelligence platform for pharma is powered by LangChain + LangSmith. When users ask research questions, the orchestrator breaks them into sub-tasks. Multip

What to expect during the AI Agent Conference: Join theCUBE May 4-5

AgentsDGX agent

Agentic enterprise is moving artificial intelligence from isolated tools into systems that actively run business operations. The transition reflects a broader shift in enterprise strategy, where organ

You can now control both local and cloud ComfyUI with Hermes Agent with ease with the new built in skill `hermes update` and run /comfyui to…

AgentsDGX agent

You can now control both local and cloud ComfyUI with Hermes Agent with ease with the new built in skill `hermes update` and run /comfyui to get started ComfyUI is the most flexible, composable, and p

28 Apr 2026

Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents

SafetyDGX agent

arXiv:2604.23990v1 Announce Type: new Abstract: This paper presents PSA-Eval, a failure-centered runtime evaluation framework for deployed trilingual public-space agents. The central claim is that, wh

From Coarse to Fine: Self-Adaptive Hierarchical Planning for LLM Agents

TutorialsDGX agent

arXiv:2604.23194v1 Announce Type: new Abstract: Large language model-based agents have recently emerged as powerful approaches for solving dynamic and multi-step tasks. Most existing agents employ pla

Latency and Cost of Multi-Agent Intelligent Tutoring at Scale

Model ReleasesDGX agent

arXiv:2604.24110v1 Announce Type: cross Abstract: Multi-agent LLM tutoring systems improve response quality through agent specialization, but each student query triggers several concurrent API calls w

OS-SPEAR: A Toolkit for the Safety, Performance,Efficiency, and Robustness Analysis of OS Agents

SafetyDGX agent

arXiv:2604.24348v1 Announce Type: new Abstract: The evolution of Multimodal Large Language Models (MLLMs) has shifted the focus from text generation to active behavioral execution, particularly via OS

Reasonably reasoning AI agents can avoid game-theoretic failures in zero-shot, provably

Model ReleasesDGX agent

arXiv:2603.18563v2 Announce Type: replace Abstract: As autonomous AI agents increasingly mediate online platform markets, a fundamental question emerges: do these markets generate stable strategic out

24 Apr 2026

KompeteAI: Accelerated Autonomous Multi-Agent System for End-to-End Pipeline Generation for Machine Learning Problems

Model ReleasesDGX agent

arXiv:2508.10177v3 Announce Type: replace Abstract: Recent Large Language Model (LLM)-based AutoML systems demonstrate impressive capabilities but face significant limitations such as constrained expl

MATRAG: Multi-Agent Transparent Retrieval-Augmented Generation for Explainable Recommendations

Model ReleasesDGX agent

arXiv:2604.20848v1 Announce Type: cross Abstract: Large Language Model (LLM)-based recommendation systems have demonstrated remarkable capabilities in understanding user preferences and generating per

Nemobot Games: Crafting Strategic AI Gaming Agents for Interactive Learning with Large Language Models

Model ReleasesDGX agent

arXiv:2604.21896v1 Announce Type: new Abstract: This paper introduces a new paradigm for AI game programming, leveraging large language models (LLMs) to extend and operationalize Claude Shannon's taxo

23 Apr 2026

Chasing the Public Score: User Pressure and Evaluation Exploitation in Coding Agent Workflows

Model ReleasesDGX agent

arXiv:2604.20200v1 Announce Type: new Abstract: Frontier coding agents are increasingly used in workflows where users supervise progress primarily through repeated improvement of a public score, namel

Enhancing Research Idea Generation through Combinatorial Innovation and Multi-Agent Iterative Search Strategies

AgentsDGX agent

arXiv:2604.20548v1 Announce Type: cross Abstract: Scientific progress depends on the continual generation of innovative re-search ideas. However, the rapid growth of scientific literature has greatly

From Data to Theory: Autonomous Large Language Model Agents for Materials Science

Model ReleasesDGX agent

arXiv:2604.19789v1 Announce Type: new Abstract: We present an autonomous large language model (LLM) agent for end-to-end, data-driven materials theory development. The model can choose an equation for

Mol-Debate: Multi-Agent Debate Improves Structural Reasoning in Molecular Design

AgentsDGX agent

arXiv:2604.20254v1 Announce Type: new Abstract: Text-guided molecular design is a key capability for AI-driven drug discovery, yet it remains challenging to map sequential natural-language instruction

22 Apr 2026

ARGUS: Agentic GPU Optimization Guided by Data-Flow Invariants

HardwareDGX agent

arXiv:2604.18616v1 Announce Type: cross Abstract: LLM-based coding agents can generate functionally correct GPU kernels, yet their performance remains far below hand-optimized libraries on critical co

BAPO: Boundary-Aware Policy Optimization for Reliable Agentic Search

SafetyDGX agent

arXiv:2601.11037v2 Announce Type: replace Abstract: RL-based agentic search enables LLMs to solve complex questions via dynamic planning and external search. While this approach significantly enhances

benchmarking agents by their ability to play video games 🎮

AgentsDGX agent

benchmarking agents by their ability to play video games 🎮 We are excited to launch VideoGameBench on Antim Labs, created by @a1zhang, Thomas L. Griffiths (@cocosci_lab), @karthik_r_n, and @OfirPress

Debug2Fix: Can Interactive Debugging Help Coding Agents Fix More Bugs?

Model ReleasesDGX agent

arXiv:2602.18571v2 Announce Type: replace-cross Abstract: While significant progress has been made in automating various aspects of software development through coding agents, there is still significa

Four-Axis Decision Alignment for Long-Horizon Enterprise AI Agents

Model ReleasesDGX agent

arXiv:2604.19457v1 Announce Type: new Abstract: Long-horizon enterprise agents make high-stakes decisions (loan underwriting, claims adjudication, clinical review, prior authorization) under lossy mem

How Adversarial Environments Mislead Agentic AI?

Model ReleasesDGX agent

arXiv:2604.18874v1 Announce Type: new Abstract: Tool-integrated agents are deployed on the premise that external tools ground their outputs in reality. Yet this very reliance creates a critical attack

Human-Guided Harm Recovery for Computer Use Agents

Model ReleasesDGX agent

arXiv:2604.18847v1 Announce Type: new Abstract: As LM agents gain the ability to execute actions on real computer systems, we need ways to not only prevent harmful actions at scale but also effectivel

LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat te…

SafetyDGX agent

LiteParse: our open-source, layout-aware PDF parser for AI agents. The secret? Grid projection. Instead of heavy ML layout models or flat text extraction, it projects text onto a monospace grid so ali

21 Apr 2026

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

Model ReleasesDGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

Kimi K2.6 @Kimi_Moonshot is the new leading open-weights agent model, landing at #4 on Claw-Eval (Pass^3: 62.3%). Key takeaways: - 👑 Best o…

Model ReleasesDGX agent

Kimi K2.6 @Kimi_Moonshot is the new leading open-weights agent model, landing at #4 on Claw-Eval (Pass^3: 62.3%). Key takeaways: - 👑 Best open-source agent, period: Pass^3 of 62.3% is the highest of a

MTSQL-R1: Towards Long-Horizon Multi-Turn Text-to-SQL via Agentic Training

Model ReleasesDGX agent

arXiv:2510.12831v3 Announce Type: replace Abstract: Multi-turn Text-to-SQL aims to translate a user's conversational utterances into executable SQL while preserving dialogue coherence and grounding to

SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy

SafetyDGX agent

arXiv:2604.18557v1 Announce Type: new Abstract: Controllable cooperative humanoid manipulation is a fundamental yet challenging problem for embodied intelligence, due to severe data scarcity, complexi

← Previous
1…7475767778…299
Next →