AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,959 results
Model Releases

Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

DGX agent

arXiv:2606.08960v1 Announce Type: cross Abstract: Agent benchmarks score submissions with outcome verifiers that are typically hand-written and brittle, leaving them open to reward hacking. We audit 1

model-releasesarxiv-cs-ai
9 Jun 2026
Agents
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

MAR:Multi-Agent Reflexion Improves Reasoning Abilities in LLMs

DGX agent

arXiv:2512.20845v2 Announce Type: replace Abstract: LLMs have shown the capacity to improve their performance on reasoning tasks through reflecting on their mistakes, and acting with these reflections

agentsarxiv-cs-ai
9 Jun 2026
Safety

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

DGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

safetyarxiv-cs-ai
9 Jun 2026
Safety

SpaceVLN: A Zero-Shot Vision-and-Language Navigation Agent with Online Spatial Cognitive Memory and Reasoning

DGX agent

arXiv:2606.08992v1 Announce Type: cross Abstract: Vision-and-Language Navigation in continuous environments requires agents to understand the spatial structure of previously unseen environments in ord

safetyarxiv-cs-ai
9 Jun 2026
Model Releases

Strained Coherence: A Pre-Failure Signal in Coding Agent Execution Trajectories

DGX agent

arXiv:2606.07889v1 Announce Type: cross Abstract: LLM-based coding agents sometimes acknowledge a problem in their own reasoning and then proceed anyway. We call this pattern strained coherence: a saf

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

Struct-Searcher: Agentic Structural Thinking Advances Multimodal Deep Information Seeking

DGX agent

arXiv:2606.07689v1 Announce Type: new Abstract: Deep research agents have attracted increasing attention for their ability to collect large-scale online information to acquire target knowledge, with r

model-releasesarxiv-cs-cv
9 Jun 2026
Model Releases

TAME: A Trustworthy Test-Time Evolution of Agent Memory with Systematic Benchmarking

DGX agent

arXiv:2602.03224v2 Announce Type: replace Abstract: Test-time evolution of agent memory represents a pivotal paradigm for advancing AGI, as it strengthens complex reasoning through experience accumula

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Web Agents Should Use Typed Actions Instead of Click-Based Browsing

DGX agent

arXiv:2602.17245v2 Announce Type: replace Abstract: This position paper argues that building a reliable agentic Web requires shifting from low-level interaction primitives to typed actions supported b

safetyarxiv-cs-ai
9 Jun 2026
Agents

PandaAI: A Practical Agent CQ2 for Neuro-symbolic Data Analysis And Integrated Decision-Making in Quantitative Finance

DGX agent

arXiv:2606.06823v1 Announce Type: cross Abstract: While deep learning has excelled in various domains, its application to sequential decision-making in finance remains challenging due to the low Signa

agentsarxiv-cs-ai
8 Jun 2026
Model Releases

ReclAIm: A Multi-Agent Framework for Monitoring and Correcting Performance Decline in Medical Imaging AI

DGX agent

arXiv:2510.17004v2 Announce Type: replace-cross Abstract: Purpose: To develop and evaluate a multi-agent framework (ReclAIm) for automated monitoring, detection, and correction of performance decline

model-releasesarxiv-cs-ai
8 Jun 2026
Safety

Silverfort brings runtime identity controls to Microsoft Copilot Studio agents

DGX agent

Identity security company Silverfort Inc. today launched an integration that applies its identity and access controls to artificial intelligence agents built into Microsoft Corp.’s Copilot Studio, enf

safetysiliconangle
8 Jun 2026
Agents

A2RAG: Adaptive Agentic Graph Retrieval for Cost-Aware and Reliable Reasoning

DGX agent

arXiv:2601.21162v2 Announce Type: replace-cross Abstract: Graph Retrieval-Augmented Generation (Graph-RAG) enhances multihop question answering by organizing corpora into knowledge graphs and routing

agentsarxiv-cs-ai
6 Jun 2026
Agents

The Virtual Roundtable: Multi-Agent Personas Simulating the Dynamics of Human Brainstorming

DGX agent

arXiv:2606.05178v1 Announce Type: cross Abstract: As AI-driven product development accelerates, the bottleneck is shifting from how we build to what we build. Traditional human brainstorming faces cha

agentsarxiv-cs-ai
6 Jun 2026
Agents

PathWISE: Multi-Agent Cancer Pathway Triaging Ontology Learning from Clinical Flowcharts

DGX agent

arXiv:2605.25970v2 Announce Type: replace Abstract: Clinical pathways are disseminated as visual flowcharts where spatial topology, arrow direction, colour coding, and font weight encode critical tria

agentsarxiv-cs-cv
5 Jun 2026
Model Releases

SubtleMemory: A Benchmark for Fine-Grained Relational Memory Discrimination in Long-Horizon AI Agents

DGX agent

arXiv:2606.05761v1 Announce Type: cross Abstract: Persistent AI assistants, such as OpenClaw, accumulate large collections of related memories over long-term interactions. As these memories grow, they

model-releasesarxiv-cs-cl
5 Jun 2026
Agents

Archi: Agentic Operations at the CMS Experiment

DGX agent

arXiv:2606.04755v1 Announce Type: cross Abstract: We present Archi, an open-source, end-to-end framework for scientific collaborations that combines the systematic ingestion and organization of hetero

agentsarxiv-cs-ai
4 Jun 2026
Model Releases

Asana launches AI-powered products to help organizations manage human and agent work

DGX agent

Asana Inc. announced today during the company’s Work Innovation Summit in London the launch of a new product suite that helps organizations manage work by humans and artificial intelligence agents usi

model-releasessiliconangle
4 Jun 2026
Agents

BRAINCELL-AID: An Agentic AI Created Brain Cell Type Resource for Community Annotation

DGX agent

arXiv:2510.17064v4 Announce Type: replace Abstract: Single-cell RNA sequencing has transformed our ability to identify diverse cell types and their transcriptomic signatures. However, annotating these

agentsarxiv-cs-ai
4 Jun 2026
Agents

MapAgent: An Industrial-Grade Agentic Framework for City-scale Lane-level Map Generation

DGX agent

arXiv:2606.04513v1 Announce Type: new Abstract: Lane-level maps are critical infrastructure for autonomous driving and lane-level navigation, yet constructing and maintaining standardized lane network

agentsarxiv-cs-ai
4 Jun 2026
Agents

MetaPoint: Unlocking Precise Spatial Control in Agentic Visual Generation

DGX agent

arXiv:2606.05031v1 Announce Type: new Abstract: Generative visual models fundamentally struggle with precise spatial control. This arises from a core disconnect: models can process textual description

agentsarxiv-cs-cv
4 Jun 2026
Safety

Scaling Self-Evolving Agents via Parametric Memory

DGX agent

arXiv:2606.04536v1 Announce Type: new Abstract: Existing memory-augmented LLM agents store past experience exclusively in prompt space, as textual summaries or retrieved passages, while keeping model

safetyarxiv-cs-ai
4 Jun 2026
Agents

A Training-Free Mixture-of-Agents Framework for Multi-Document Summarization using LLMs and Knowledge Graphs

DGX agent

arXiv:2606.03867v1 Announce Type: cross Abstract: Multi-Document Summarization (MDS) plays a critical role in distilling essential information from collections of textual data. Existing approaches oft

agentsarxiv-cs-ai
3 Jun 2026
Safety

Collab-REC: An LLM-based Agentic Framework for Balancing Recommendations in Tourism

DGX agent

arXiv:2508.15030v5 Announce Type: replace Abstract: We propose COLLAB-REC, a multi-agent framework designed to counteract popularity bias and improve diversity in tourism recommendations. In our setup

safetyarxiv-cs-ai
3 Jun 2026
Safety

CP-Agent: Context-Aware Multimodal Reasoning for Cellular Morphological Profiling under Chemical Perturbations

DGX agent

arXiv:2606.03435v1 Announce Type: new Abstract: Cell Painting combines multiplexed fluorescent staining, high-content imaging, and quantitative analysis to generate high-dimensional phenotypic readout

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

DeskCraft: Benchmarking Desktop Agents on Professional Workflows and Human-in-the-Loop Collaboration

DGX agent

arXiv:2606.03103v1 Announce Type: new Abstract: Real-world professional desktop workflows in specialized creative and engineering software unfold over long horizons and often require human-in-the-loop

model-releasesarxiv-cs-ai
3 Jun 2026
Local Ai

FederatedSkill: Federated Learning for Agentic Skill Evolution

DGX agent

arXiv:2606.03143v1 Announce Type: cross Abstract: Modern LLM agents increasingly rely on skill libraries to handle complex tasks, making skill evolution a primary driver of self-improvement. However,

local-aiarxiv-cs-cl
3 Jun 2026
Agents

FORGE: Multi-Agent Graduated Exploitation and Detection Engineering

DGX agent

arXiv:2606.03453v1 Announce Type: cross Abstract: Vulnerability disclosure volumes now far exceed organizational assessment capacity, yet three adjacent research communities (proof-of-concept generati

agentsarxiv-cs-ai
3 Jun 2026
Safety

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

DGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

safetyarxiv-cs-ai
3 Jun 2026
Model Releases

Absorbing Complexity: An Interaction-Native Knowledge Harness for Financial LLM Agents

DGX agent

arXiv:2606.01886v1 Announce Type: new Abstract: Financial AI agents often fail for a simple reason: they make users carry the complexity. A user must repeatedly restate goals, risk preferences, portfo

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

AgentRedBench: Dynamic Redteaming and Integration-Aware Defense for LLM Agents over SaaS Integrations

DGX agent

arXiv:2606.02240v1 Announce Type: cross Abstract: Indirect prompt injection in tool-use agents is a concrete production threat: LLM agents read from integrations (third-party services such as Gmail, S

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

APE: Agentic Prompt Enhancer for Image Generation and Editing

DGX agent

arXiv:2606.00204v1 Announce Type: new Abstract: Natural language has become a powerful interface for image generation and editing, yet text-guided visual systems remain highly sensitive to prompt form

model-releasesarxiv-cs-cv
2 Jun 2026
Model Releases

Business Utility of Large Language Models as Exploratory Data Analysis Agents

DGX agent

arXiv:2606.00051v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in analytical workflows, but their suitability as exploratory data analysis (EDA) agents in busines

model-releasesarxiv-cs-ai
2 Jun 2026
Applications

Computer-use agents are moving from the cloud to your local machine. Fast. When we launched Holo3 two months ago, the production feedback wa…

DGX agent

Computer-use agents are moving from the cloud to your local machine. Fast. When we launched Holo3 two months ago, the production feedback was clear: digital agents need to be blazing fast, cost-effect

applicationsclem-delangue--x
2 Jun 2026
Model Releases

HomeFlow: A Data Flywheel for Smart Home Agent Training with Verifiable Simulation

DGX agent

arXiv:2606.01230v1 Announce Type: new Abstract: Large language model agents are moving beyond text-only interaction toward physical-world control, with smart homes as a representative domain. Real dom

model-releasesarxiv-cs-ai
2 Jun 2026
Model Releases

LLM4Cov: Execution-Aware Agentic Learning for High-coverage Testbench Generation

DGX agent

arXiv:2602.16953v3 Announce Type: replace Abstract: Execution-aware LLM agents offer a promising paradigm for learning from tool feedback, but such feedback can be expensive and slow to obtain, making

model-releasesarxiv-cs-ai
2 Jun 2026
Agents

MIRROR: A Multi-Agent Framework with Iterative Adaptive Revision and Hierarchical Retrieval for Optimization Modeling in Operations Research

DGX agent

arXiv:2602.03318v3 Announce Type: replace Abstract: Operations Research (OR) relies on expert-driven modeling-a slow and fragile process ill-suited to novel scenarios. While large language models (LLM

agentsarxiv-cs-cl
2 Jun 2026
Hardware

NVIDIA Jetson Brings Agentic AI to the Physical World

DGX agent

Agentic AI is getting physical. At COMPUTEX on Tuesday, NVIDIA announced NVIDIA JetPack 7.2 and NVIDIA NemoClaw support on NVIDIA Jetson. JetPack 7.2 brings agentic AI skills, Yocto project support, N

hardwarenvidia-blog
2 Jun 2026
Agents

PLanAR: Planning-Language-Grounded Agentic Reasoning for Robot Manipulation

DGX agent

arXiv:2602.01662v4 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have enabled increasing progress in real-world robot manipulation. However, long-horizon manipulati

agentsarxiv-cs-ro
2 Jun 2026
Agents

RDA: Reward Design Agent for Reinforcement Learning

DGX agent

arXiv:2606.01672v1 Announce Type: new Abstract: Reinforcement learning has enabled the acquisition of impressive robotic skills, but typically requires hand-crafted reward functions that are slow to d

agentsarxiv-cs-lg
2 Jun 2026
Agents

RelationalAI beefs up its reasoning capabilities to enhance AI agent decision-making

DGX agent

Enterprise decision intelligence startup RelationalAI Inc. is advancing its capabilities for Snowflake Inc.’s AI Data Cloud platform. At Snowflake Summit 2026 today, it announced a series of updates t

agentssiliconangle
2 Jun 2026
Safety

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

DGX agent

arXiv:2606.01619v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) enables LLM agents to improve continuously from environment rewards, yet the resulting policies do not systematicall

safetyarxiv-cs-ai
2 Jun 2026
Safety

Skill or Skip? Learning Selective Skill Invocation in Agentic Tasks via Dual-Granularity Preference Learning

DGX agent

arXiv:2606.00510v1 Announce Type: cross Abstract: Agent skills are callable procedural modules that provide reusable knowledge and execution policies for complex agentic tasks. However, existing metho

safetyarxiv-cs-ai
2 Jun 2026
Safety

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that th…

DGX agent

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that the environment can maintain reliably, it probably doesn't bel

safetydair-ai--x
2 Jun 2026
Model Releases

TVIR: Building Deep Research Agents Towards Text--Visual Interleaved Report Generation

DGX agent

arXiv:2606.02320v1 Announce Type: new Abstract: Deep Research Agents have shown strong capability in multi-step information retrieval, reasoning, and long-form report generation, but existing benchmar

model-releasesarxiv-cs-cl
2 Jun 2026
Agents

Automatically Attacking Software Reverse Engineering AI Agents

DGX agent

arXiv:2605.30667v1 Announce Type: cross Abstract: Software tools for reverse engineering executable binary files, such as Ghidra, enable malware analysts to safely conduct robust static analysis witho

agentsarxiv-cs-ai
1 Jun 2026
Model Releases

Harness Updating Is Not Harness Benefit: Disentangling Evolution Capabilities in Self-Evolving LLM Agents

DGX agent

arXiv:2605.30621v1 Announce Type: new Abstract: LLM agents are increasingly deployed as systems built around editable external harnesses, including prompts, skills, memories and tools, that shape task

model-releasesarxiv-cs-ai
1 Jun 2026
Safety

HiPER: Hierarchical Reinforcement Learning with Explicit Credit Assignment for Large Language Model Agents

DGX agent

arXiv:2602.16165v2 Announce Type: replace-cross Abstract: Training LLMs as interactive agents for multi-turn decision-making remains challenging, particularly in long-horizon tasks with sparse and del

safetyarxiv-cs-ai
1 Jun 2026
Model Releases

LongDS-Bench: On the Failure of Long-Horizon Agentic Data Analysis

DGX agent

arXiv:2605.30434v1 Announce Type: cross Abstract: Real-world data analysis is inherently iterative, yet existing benchmarks mostly evaluate isolated or short interactive tasks, leaving agents' ability

model-releasesarxiv-cs-ai
1 Jun 2026
← Previous
1…118119120121122…375
Next →