AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Model Releases

Tiger Data launches PostgreSQL extension designed for AI agents

DGX agent

Tiger Data today introduced a managed PostgreSQL database service designed specifically for AI agents, saying conventional database architectures are poorly suited to a future in which software is inc

model-releasessiliconangle
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

Zscaler launches AI Broker and Endpoint AI Security for agents

DGX agent

Zscaler Inc. today unveiled a set of products designed to secure autonomous artificial intelligence agents, with the cybersecurity company claiming it has built the industry’s first complete zero-trus

model-releasessiliconangle
9 Jun 2026
Safety

CAF-Gen: A Multi-Agent System for Enriching Argumentation Structures

DGX agent

arXiv:2606.06646v1 Announce Type: cross Abstract: Formalizing complex reasoning from natural text is one of the central challenges in computational linguistics. It requires systems to understand not j

safetyarxiv-cs-ai
8 Jun 2026
Safety

EVA: Evolving Semantic Adversaries for Red-Teaming GUI Agents Against Environmental Injection Attacks

DGX agent

arXiv:2505.14289v2 Announce Type: replace Abstract: Graphical User Interface (GUI) agents powered by Multimodal Large Language Models (MLLMs) are increasingly deployed yet vulnerable to Environmental

safetyarxiv-cs-ai
8 Jun 2026
Agents

Evaluate your Amazon Nova Sonic voice agent at scale, no microphone required

DGX agent

In this post, we walk you through the Nova Sonic Test Harness, an open source framework that we built to solve both problems. It serves as a rapid iteration tool for tuning system prompts and tool con

agentsaws-ml-blog
8 Jun 2026
Model Releases

Lean4Agent: Formal Modeling and Verification for Agent Workflow and Trajectory

DGX agent

arXiv:2606.06523v1 Announce Type: new Abstract: Equipping Large Language Models (LLMs) to execute reliable multi-step workflows has become a central challenge in artificial intelligence. Despite recen

model-releasesarxiv-cs-ai
8 Jun 2026
Agents

Small Language Model Agents Enable Efficient and High-Quality Knowledge Mining

DGX agent

arXiv:2510.01427v3 Announce Type: replace Abstract: At the core of Deep Research is knowledge mining, the task of extracting structured information from massive unstructured text in response to user i

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

Tree-of-Experience: A Structured Experience-Management Solution for Self-Evolving Agents under Low-Repetition and Implicit-Reward Environments

DGX agent

arXiv:2606.06960v1 Announce Type: new Abstract: Experience-based self-evolution is crucial for LLM agents, but existing benchmarks often assume explicit goals, stable task patterns, and clear feedback

model-releasesarxiv-cs-cl
8 Jun 2026
Model Releases

we just shipped support for rubrics in deepagents ✅ give your agent a clear definition of what 'done' looks like, and force it to run in a l…

DGX agent

we just shipped support for rubrics in deepagents ✅ give your agent a clear definition of what 'done' looks like, and force it to run in a loop until said goal is complete this is similar to /goal in

model-releasesharrison-chase--x
8 Jun 2026
Agents

Agentic Molecular Recovery via Molecule-Aware Exploration

DGX agent

arXiv:2606.05847v1 Announce Type: new Abstract: Text-guided molecular generation with LLMs often yields invalid SMILES. We argue that invalid drafts should be addressed through a shift from validity-o

agentsarxiv-cs-ai
6 Jun 2026
Model Releases

AdaPlanBench: Evaluating Adaptive Planning in Large Language Model Agents under World and User Constraints

DGX agent

arXiv:2606.05622v1 Announce Type: new Abstract: Planning for real-world problems by language models often involves both world and user constraints, which may not be fully specified upfront and are pro

model-releasesarxiv-cs-cl
5 Jun 2026
Safety

Beyond tokens: a unified framework for latent communication in LLM-based multi-agent systems

DGX agent

arXiv:2606.05711v1 Announce Type: new Abstract: Multi-agent systems built on large language models (LLMs) have become a prevailing paradigm for tackling complex reasoning, planning, and tool-use tasks

safetyarxiv-cs-cl
5 Jun 2026
Model Releases

Coding with 'Enemy': Can Human Developers Detect AI Agent Sabotage?

DGX agent

arXiv:2606.05647v1 Announce Type: cross Abstract: AI coding agents are increasingly embedded in real-world software development, collaborating with human developers while gaining broader access to cod

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

EpiEvolve: Self-Evolving Agents for Streaming Pandemic Forecasting under Regime Shifts

DGX agent

arXiv:2606.05513v1 Announce Type: cross Abstract: Epidemic LLM forecasters are usually trained and evaluated as static supervised models, whereas operational pandemic forecasting is a streaming proces

agentsarxiv-cs-cl
5 Jun 2026
Local Ai

Statistical Priors for Implicit Preferences: Decoupling Skill Selection as a Local Harness in Personal Agents

DGX agent

arXiv:2606.05828v1 Announce Type: cross Abstract: As Large Language Model (LLM) capabilities advance, locally deployed personal agents relying on API-based remote models and external skills have emerg

local-aiarxiv-cs-cl
5 Jun 2026
Model Releases

Benchmarking Living-Screen-Native GUI Agents on Short-Video Platforms

DGX agent

arXiv:2606.04701v1 Announce Type: cross Abstract: GUI agents today assume a static screen, where the world is frozen between two actions. However, real interfaces such as short-video applications viol

model-releasesarxiv-cs-cl
4 Jun 2026
Agents

Deliberate Evolution: Agentic Reasoning for Sample-Efficient Symbolic Regression with LLMs

DGX agent

arXiv:2606.04360v1 Announce Type: new Abstract: Symbolic regression (SR) discovers compact mathematical expressions from data, yet recent LLM-based evolutionary methods remain sample-inefficient becau

agentsarxiv-cs-cl
4 Jun 2026
Model Releases

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-l…

DGX agent

Introducing two NVIDIA Nemotron models on Together AI: Nemotron 3 Ultra for high-throughput agentic workloads and Nemotron 3.5 ASR for low-latency multilingual speech recognition. AI natives can now b

model-releasestogether-ai--x
4 Jun 2026
Model Releases

Learning When to Act or Refuse: Guarding Agentic Reasoning Models for Safe Multi-Step Tool Use

DGX agent

arXiv:2603.03205v2 Announce Type: replace Abstract: Agentic language models operate in a fundamentally different safety regime than chat models: they must plan, call tools, and execute long-horizon ac

model-releasesarxiv-cs-cl
4 Jun 2026
Model Releases

Online Skill Learning for Web Agents via State-Grounded Dynamic Retrieval

DGX agent

arXiv:2606.04391v1 Announce Type: new Abstract: Language agents increasingly rely on reusable skills to improve multi-step web automation across related tasks. A growing line of work studies online sk

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

SMADE-IE: Sparse Multi-Agent Framework with Evidence-Driven Debate for Zero-Shot Information Extraction

DGX agent

arXiv:2606.04691v1 Announce Type: new Abstract: Zero-shot information extraction (IE) with large language models (LLMs) has attracted increasing attention due to its flexibility in adapting to new sch

model-releasesarxiv-cs-cl
4 Jun 2026
Safety

The Digital Apprentice: A Framework for Human-Directed Agentic AI Development

DGX agent

arXiv:2606.04321v1 Announce Type: new Abstract: Agentic AI deployments face a recurring design tension: heavy human oversight limits scale, while broad autonomy outruns accountability. Neither posture

safetyarxiv-cs-ai
4 Jun 2026
Model Releases

The Saturation Trap and the Subjectivity of Intervention Timing: Why Affect-Based Triggers and LLM Judges Fail to Time Interventions on Autonomous Agents

DGX agent

arXiv:2606.04296v1 Announce Type: new Abstract: As autonomous AI agents move from conversational systems to long-horizon software execution, runtime safety layers that decide when to interrupt an agen

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Toward Pre-Deployment Assurance for Enterprise AI Agents: Ontology-Grounded Simulation and Trust Certification

DGX agent

arXiv:2606.04037v1 Announce Type: new Abstract: Pre-deployment verification of enterprise artificial intelligence (AI) agents remains a critical gap between large language model (LLM) capability bench

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

Tree-Based Formalization of Multi-Agent Complementarity in Human-AI Interactions

DGX agent

arXiv:2606.04779v1 Announce Type: new Abstract: Complementarity is the case in which a human--AI interaction (HAI) outperforms the best prediction benchmark available among its members. Although this

model-releasesarxiv-cs-ai
4 Jun 2026
Model Releases

DMF: A Deterministic Memory Framework for Conversational AI Agents

DGX agent

arXiv:2606.03463v1 Announce Type: new Abstract: Conversational AI agents require memory systems that are both scalable and semantically coherent across long interaction horizons. Existing approaches r

model-releasesarxiv-cs-ai
3 Jun 2026
Research

MemVerse: Multimodal Memory for Lifelong Learning Agents

DGX agent

arXiv:2512.03627v2 Announce Type: replace Abstract: Despite rapid progress in large-scale language and vision models, AI agents still suffer from a fundamental limitation: they cannot remember. Withou

researcharxiv-cs-ai
3 Jun 2026
Agents

MUSE: A Unified Agentic Harness for MLLMs

DGX agent

arXiv:2606.03005v1 Announce Type: cross Abstract: Despite rapid progress, multimodal large language models (MLLMs) still fail on tasks that humans solve effortlessly, such as navigating a grid maze fr

agentsarxiv-cs-ai
3 Jun 2026
Model Releases

SCOPE: Real-Time Natural Language Camera Agent at the Edge

DGX agent

arXiv:2606.02951v1 Announce Type: cross Abstract: Deploying language-driven agents in robotics requires evaluations that reflect real-world task demands: natural-language instructions with reproducibl

model-releasesarxiv-cs-ai
3 Jun 2026
Research

SkillPyramid: A Hierarchical Skill Consolidation Framework for Self-Evolving Agents

DGX agent

arXiv:2606.03692v1 Announce Type: new Abstract: Recent AI agents can flexibly invoke skills to solve complex tasks, but their long-term improvement is fundamentally constrained by a lack of systematic

researcharxiv-cs-ai
3 Jun 2026
Agents

Snowflake CoWork brings the agentic enterprise to life for social media marketing

DGX agent

As AI moves from analytical tool to active participant in daily business operations, solutions like Snowflake CoWork are helping to close the gap between data insight and real-world execution faster t

agentssiliconangle
3 Jun 2026
Model Releases

TadA-Bench: A Million-Variant Benchmark for Future-Round Discovery Toward Agentic Protein Engineering

DGX agent

arXiv:2606.02624v1 Announce Type: cross Abstract: AI for scientific discovery is entering an agentic era, where protein-engineering systems are expected to prioritize future wet-lab experiments rather

model-releasesarxiv-cs-ai
3 Jun 2026
Model Releases

We partnered with @FireworksAI_HQ to train open-source models for legal. Here's what we found: 1) Hybrid legal agents can beat frontier mode…

DGX agent

We partnered with @FireworksAI_HQ to train open-source models for legal. Here's what we found: 1) Hybrid legal agents can beat frontier models on quality and cost by routing selectively to a frontier

model-releasessonya-huang--x
3 Jun 2026
Model Releases

A Machine-to-Machine Knowledge-Guided LLM Agent for Generalizable Radiotherapy Treatment Planning

DGX agent

arXiv:2606.00922v1 Announce Type: cross Abstract: In this work, we propose a prototype machine-to-machine (M2M) knowledge-guided Large Language Model (LLM) framework for automated radiotherapy treatme

model-releasesarxiv-cs-ro
2 Jun 2026
Model Releases

ADRA-Bank: A Modular Benchmark for Academic Deep Research Agents

DGX agent

arXiv:2512.00986v3 Announce Type: replace Abstract: A surge in academic publications calls for automated deep research (DR) systems, but accurately evaluating them is still an open problem. First, exi

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Agent Tools Orchestration Leaks More: Dataset, Benchmark, and Mitigation

DGX agent

arXiv:2512.16310v3 Announce Type: replace-cross Abstract: LLM-based agents increasingly use multiple external tools to complete complex tasks. We study Tools Orchestration Privacy Risk (TOP-R): an age

model-releasesarxiv-cs-ai
2 Jun 2026
Safety

BranPO: Scalable Contrastive Branch Sampling for Long-Horizon Agentic Reinforcement Learning

DGX agent

arXiv:2602.03719v2 Announce Type: replace Abstract: Agentic reinforcement learning enables large language models to perform multi-turn planning and tool use, but long-horizon training remains challeng

safetyarxiv-cs-cl
2 Jun 2026
Safety

CAPF: Guiding Search-Agent Rollouts with Credit-Attenuated Privileged Feedback

DGX agent

arXiv:2606.01830v1 Announce Type: new Abstract: Recent LLM search agents use reinforcement learning with verifiable rewards (RLVR) to learn search-augmented reasoning from outcome rewards. On hard pro

safetyarxiv-cs-ai
2 Jun 2026
Agents

Digital Twin-Assisted Adaptive Multi-Agent DRL for Intelligent Spectrum and Resource Management in Open-RAN UAV-Enabled 6G Networks

DGX agent

arXiv:2606.01324v1 Announce Type: cross Abstract: The evolution toward 6G wireless networks envisions a seamlessly intelligent, Open-RAN-enabled architecture where unmanned aerial vehicles (UAVs) play

agentsarxiv-cs-ai
2 Jun 2026
Safety

Dynamic Coordination Strategy Selection for Enterprise Multi-Agent Systems

DGX agent

arXiv:2606.00804v1 Announce Type: cross Abstract: Enterprise multi-agent systems increasingly expose multiple coordination patterns, but deployments often lack evidence for when to use consensus, deba

safetyarxiv-cs-ai
2 Jun 2026
Agents

Empathic and agentic artificial intelligence in nursing: perspectives on a human-centered framework for cancer care navigation in the United States

DGX agent

arXiv:2606.00010v1 Announce Type: cross Abstract: For patients experiencing cancer, nurse navigation can ease the burden of complex care by enhancing coordination of health services and patient outcom

agentsarxiv-cs-ai
2 Jun 2026
Local Ai

From Features to Actions: Explainability in Traditional and Agentic AI Systems

DGX agent

arXiv:2602.06841v4 Announce Type: replace Abstract: Over the last decade, Explainable AI has primarily focused on interpreting individual model predictions, producing post-hoc explanations that relate

local-aiarxiv-cs-ai
2 Jun 2026
Model Releases

K-BrowseComp: A Web Browsing Agent Benchmark Grounded in Korean Contexts

DGX agent

arXiv:2606.02404v1 Announce Type: new Abstract: Frontier model evaluations are shifting from foundational capabilities (e.g., instruction following and reasoning) toward compositional, agentic ones, b

model-releasesarxiv-cs-cl
2 Jun 2026
Model Releases

Learning to Retrieve: Dual-Level Long-Term Memory for Text-to-SQL Agents

DGX agent

arXiv:2606.00547v1 Announce Type: new Abstract: Interactive text-to-SQL agents solve database tasks through multi-turn interactions involving schema exploration, query execution, feedback interpretati

model-releasesarxiv-cs-cl
2 Jun 2026
Safety

Learning When Not to Act: Mitigating Tool Abuse in Agentic Reinforcement Learning

DGX agent

arXiv:2606.02132v1 Announce Type: new Abstract: Agentic reinforcement learning can induce tool abuse, where models overuse external tools even for queries solvable by internal reasoning. Existing appr

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

LocalSearchBench: Benchmarking Agentic Search in Real-World Local Life Services

DGX agent

arXiv:2512.07436v3 Announce Type: replace Abstract: Recent advances in large reasoning models LRMs have enabled agentic search systems to perform complex multi-step reasoning across multiple sources.

model-releasesarxiv-cs-ai
2 Jun 2026
Industry

Microsoft releases Web IQ, a search service for AI agents that is powered by Bing, currently used by Copilot, ChatGPT, and other platforms (Barry Schwartz/Search Engine Land)

DGX agent

Barry Schwartz / Search Engine Land: Microsoft releases Web IQ, a search service for AI agents that is powered by Bing, currently used by Copilot, ChatGPT, and other platforms — Table of Contents — Mi

industrytechmeme
2 Jun 2026
Model Releases

MOC: Multi-Order Communication in LLM-based Multi-Agent Systems

DGX agent

arXiv:2606.02359v1 Announce Type: new Abstract: Despite the remarkable progress of Large Language Model (LLM) based Multi-Agent Systems, most research focuses on optimizing coordination topology while

model-releasesarxiv-cs-ai
2 Jun 2026
← Previous
1…131132133134135…375
Next →