AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,630
  • Agents7,271
  • Applications5,200
  • Concepts5
  • Hardware1,757
  • Industry6,101
  • Local Ai4,731
  • Model Releases22,603
  • Research19,194
  • Safety12,821
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,630Total entries
1Added by human
84,629Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,980 results
Model Releases

OmniGAIA: Towards Native Omni-Modal AI Agents

DGX agent

arXiv:2602.22897v3 Announce Type: replace Abstract: Human intelligence naturally intertwines omni-modal perception -- spanning vision, audio, and language -- with complex reasoning and tool usage to i

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

QFedAgent: Quantum-Enhanced Personalized Federated Learning for Multi-Agent Activity Recognition

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2607.02426v1 Announce Type: cross Abstract: Federated learning (FL) enables collaborative model training across distributed devices without sharing raw data, making it suitable for privacy-sensi

model-releasesarxiv-cs-ai
3 Jul 2026
Safety

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

DGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

safetyarxiv-cs-ai
3 Jul 2026
Safety

Exploring the Semantic Gap in Agentic Data Systems: A Formative Study of Operationalization Failures in Analytical Workflows

DGX agent

arXiv:2607.00828v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to generate queries, invoke tools, and construct analytical workflows. Although recent advances hav

safetyarxiv-cs-ai
2 Jul 2026
Model Releases

GEAR-Seg: A Grounded Explainable Agent for Reasoning Segmentation and Data Engine

DGX agent

arXiv:2607.00544v1 Announce Type: new Abstract: Reasoning segmentation requires localizing targets based on complex, implicit queries. Current end-to-end models typically entangle perception and deduc

model-releasesarxiv-cs-cv
2 Jul 2026
Model Releases

TerraBench: Can Agents Reason Over Heterogeneous Earth-System Data?

DGX agent

arXiv:2606.13148v2 Announce Type: replace Abstract: Climate and environmental decision-making increasingly requires reasoning across heterogeneous inputs, including gridded physical data, satellite im

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

A Self-Evolving Agentic System for Automated Generation and Execution of Biological Protocols

DGX agent

arXiv:2606.31763v1 Announce Type: new Abstract: Autonomous wet-lab experimentation requires more than plausible protocol text: biological intent, quantitative procedures, device constraints and experi

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

A Tutorial on Autonomous Fault-Tolerant Control Using Knowledge-Grounded LLM Agents

DGX agent

arXiv:2606.31635v1 Announce Type: cross Abstract: Fault recovery in process plants still relies heavily on plant operators, especially when faults fall outside predefined supervisory logic. Operators

safetyarxiv-cs-ai
1 Jul 2026
Agents

Contrastive Reflection for Iterative Prompt Optimization

DGX agent

arXiv:2606.30840v1 Announce Type: new Abstract: LLM agents are becoming central to information retrieval: they issue retrieval queries, synthesize answers, and increasingly serve as judges for IR eval

agentsarxiv-cs-ai
1 Jul 2026
Model Releases

Embodied CAD: Solver-Grounded LLM Agents for Parametric B-Rep Assembly Modeling

DGX agent

arXiv:2606.31252v1 Announce Type: new Abstract: Large language models can write plausible CAD scripts, but reliable industrial CAD modeling requires more than syntactically valid code: every feature,

model-releasesarxiv-cs-ai
1 Jul 2026
Model Releases

Human-Agent Collaborative Paper-to-Page Crafting

DGX agent

arXiv:2510.19600v2 Announce Type: replace-cross Abstract: In the quest for scientific progress, communicating research is as vital as the discovery itself. Yet, researchers are often sidetracked by th

model-releasesarxiv-cs-ai
1 Jul 2026
Safety

LLM-Empowered Agentic MAC Protocols: A Dynamic Stackelberg Game Approach

DGX agent

arXiv:2510.10895v2 Announce Type: replace Abstract: Medium Access Control (MAC) protocols, essential for wireless networks, are typically manually configured. While deep reinforcement learning (DRL)-b

safetyarxiv-cs-ai
1 Jul 2026
Model Releases

A Physics-Grounded Benchmark for Multi-Agent Dynamics in World Models

DGX agent

arXiv:2606.28757v1 Announce Type: new Abstract: Generative world models hold immense promise as scalable simulators for autonomous systems, particularly for synthesizing rare but safety-critical multi

model-releasesarxiv-cs-cv
30 Jun 2026
Model Releases

An AI agent for treatment reasoning over a biomedical tool universe

DGX agent

arXiv:2606.28692v1 Announce Type: new Abstract: Treatment reasoning underpins every therapeutic decision, integrating disease context, comorbidities, medications, contraindications, and evolving biome

model-releasesarxiv-cs-ai
30 Jun 2026
Agents

Automating the Design of Embodied AgentArchitectures

DGX agent

arXiv:2606.30111v1 Announce Type: cross Abstract: Embodied agents are typically built as hand-designed compositions of perception, memory, planning, and action modules. This modularity exposes a large

agentsarxiv-cs-ai
30 Jun 2026
Agents

CaveAgent: Transforming LLMs into Stateful Runtime Operators

DGX agent

arXiv:2601.01569v4 Announce Type: replace Abstract: LLM-based agents are increasingly capable of complex task execution, yet current agentic systems remain constrained by text-centric paradigms that s

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Customized Generative AI Agent for Transportation Engineering Practice: A Development and Continued Pre-training Guideline

DGX agent

arXiv:2606.29014v1 Announce Type: new Abstract: Recent advancements in generative artificial intelligence (AI) and large language models (LLMs) have shown significant promise in automating complex rea

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Digitizing Coaching Intelligence: An Agentic Framework for Holistic Athlete Profiling using VLM and RAG

DGX agent

arXiv:2606.28570v1 Announce Type: cross Abstract: Athlete assessment is a critical process for tracking physical progress and identifying elite talent. However, during mass recruitment drives, traditi

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Rehearsed Multi-Agent Live Product Demonstrations with Real-Time Voice Question Answering

DGX agent

arXiv:2606.30294v1 Announce Type: new Abstract: Live product demonstrations are a recurring, high-cost activity in software organizations: a human presenter must select features, dispatch the correspo

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

SpreadsheetBench 2: Evaluating Agents on End-to-End Business Spreadsheet Workflows

DGX agent

arXiv:2606.29955v1 Announce Type: cross Abstract: Spreadsheets are widely used for business analysis, financial modeling, reporting, and decision-making. However, most existing spreadsheet benchmarks

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

The Contagion Tensor: A Framework for Measuring Output-Distribution Coupling in Multi-Agent LLM Systems -- and Auditing the Claims It Enables

DGX agent

arXiv:2606.28839v1 Announce Type: new Abstract: We introduce the Contagion Tensor, a measurement framework for quantifying how large language model (LLM) output distributions couple across modalities,

model-releasesarxiv-cs-lg
30 Jun 2026
Model Releases

Towards Generalizable and Evidential Nuclear Magnetic Resonance-Based Molecular Structure Elucidation via Large Language Model Agent

DGX agent

arXiv:2606.29776v1 Announce Type: cross Abstract: Nuclear Magnetic Resonance (NMR) spectroscopy is the gold standard for molecular structure elucidation, yet interpreting complex spectra for unknown m

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Hybrid Fact-Checking that Integrates Knowledge Graphs, Large Language Models, and Search-Based Retrieval Agents Improves Interpretable Claim Verification

DGX agent

arXiv:2511.03217v2 Announce Type: replace-cross Abstract: Large language models (LLMs) excel in generating fluent utterances but can lack reliable grounding in verified information. At the same time,

model-releasesarxiv-cs-ai
29 Jun 2026
Agents

langgraph <3

DGX agent

langgraph <3 New category drop on the Arena: ✨AI agent frameworks✨ Agent Experience (AX) update as of 06/26: 1/ @LangChain LangGraph 2/ @vercel AI SDK 3/ @crewAIInc 4/ @mastra For context: we evaluate

agentsharrison-chase--x
27 Jun 2026
Model Releases

Boundary-Aware Context Grounding for A Low-Channel EEG Agent

DGX agent

arXiv:2606.26519v1 Announce Type: new Abstract: Large language models (LLMs) can make scientific software easier to use. However, a general model does not automatically know which measurements a parti

model-releasesarxiv-cs-ai
26 Jun 2026
Model Releases

Perception, Verdict, and Evolution: Hindsight-Driven Self-Refining Forensics Agent for AI-Generated Image Detection

DGX agent

arXiv:2606.26552v1 Announce Type: cross Abstract: The rapid advancement of generative models presents a significant challenge to existing deepfake detection methods, particularly given the widespread

model-releasesarxiv-cs-ai
26 Jun 2026
Applications

HEART: Coordination of Heterogeneous Expert Agents for Physically Grounded Robotic Task Planning

DGX agent

arXiv:2606.25404v1 Announce Type: new Abstract: Large Language Models (LLMs) can reason over complex instructions but often fail to satisfy the physical and spatial constraints required for robotic ta

applicationsarxiv-cs-ro
25 Jun 2026
Safety

Heuresis: Search Strategies for Autonomous AI Research Agents Across Quality, Diversity and Novelty

DGX agent

arXiv:2606.25198v1 Announce Type: new Abstract: Autonomous AI Research promises to accelerate the scientific progress of machine learning. To realise this goal, current Large Language Model (LLM)-base

safetyarxiv-cs-ai
25 Jun 2026
Agents

Is GraphRAG Needed? From Basic RAG to Graph-/Agentic Solutions with Context Optimization

DGX agent

arXiv:2606.25656v1 Announce Type: new Abstract: As advanced RAG variants like GraphRAG and Agentic RAG emerge, one leading question is when and how to use them. Here, we introduce a framework for diff

agentsarxiv-cs-cl
25 Jun 2026
Safety

MedGuards: Multi-Agent System for Reliable Medical Error Detection and Correction

DGX agent

arXiv:2606.25651v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed in healthcare settings, accurate error detection and correction in generated or existing text

safetyarxiv-cs-cl
25 Jun 2026
Agents

New research from Meta. Building synthetic training data has stayed a fixed pipeline that you hand-tune and then freeze. Autodata casts an A…

DGX agent

New research from Meta. Building synthetic training data has stayed a fixed pipeline that you hand-tune and then freeze. Autodata casts an AI agent as a data scientist that builds training and evaluat

agentsdair-ai--x
25 Jun 2026
Industry

Notion killing Skiff-influenced email app since most users use AI agents instead

DGX agent

Notion is shutting down Notion Mail, its AI-enhanced email client launched in April 2025 based on the Skiff acquisition, on September 22. More than 50% of Notion Mail users now manage their email usin

industryars-technica
25 Jun 2026
Model Releases

HelloTwin launches ‘Digital Authority’ to bring governed AI agents to the enterprise

DGX agent

HelloTwin.ai GmbH today announced what it calls an accountable artificial intelligence AI twin that holds business intelligence and goals in a single source of truth. HelloTwin said it built its data

model-releasessiliconangle
24 Jun 2026
Model Releases

MortarBench: Evaluating Mortgage Loan Origination Agents

DGX agent

arXiv:2606.19416v2 Announce Type: replace Abstract: Loan origination is the process by which a lender creates a new loan, from application and underwriting through approval and funding. This process s

model-releasesarxiv-cs-lg
24 Jun 2026
Model Releases

Privacy-Preserving RAG via Multi-Agent Semantic Rewriting: Achieving Confidentiality Without Compromising Contextual Fidelity

DGX agent

arXiv:2606.24623v1 Announce Type: cross Abstract: Retrieval-Augmented Generation enhances large language models by incorporating external knowledge, but deploying it in sensitive scenarios risks priva

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

SP-Mind: An Autonomous Reasoning Agent for Spatial Proteomics Analysis

DGX agent

arXiv:2606.24235v1 Announce Type: new Abstract: Spatial proteomics enables single-cell-resolution characterization of protein expression within tissue architecture, playing a critical role in understa

model-releasesarxiv-cs-ai
24 Jun 2026
Hardware

Nvidia bets on agentic AI to turbocharge biotech discovery

DGX agent

Artificial intelligence played a prominent role at this week’s Bio International Convention in San Diego, the largest biotech event with vendors spanning the full ecosystem of companies in this indust

hardwaresiliconangle
23 Jun 2026
Hardware

RAVEN: Agentic RAG for Automated Vulnerability Repair

DGX agent

arXiv:2606.22647v1 Announce Type: cross Abstract: Automated vulnerability repair has emerged as a promising direction to mitigate the growing number of software vulnerabilities. Recent advances in Lar

hardwarearxiv-cs-lg
23 Jun 2026
Model Releases

RS-Gen: A Multi-Stage Agentic Framework for Reasoning and Search-Augmented Image Generation

DGX agent

arXiv:2606.23221v1 Announce Type: new Abstract: Recent years have witnessed remarkable progress in image generation and editing, particularly regarding instruction following and visual fidelity. Howev

model-releasesarxiv-cs-cv
23 Jun 2026
Model Releases

When AUC 0.998 Is Not Enough: A Candidate Evaluation Protocol for Hidden-State Probes of Indirect Prompt Injection in Multimodal Computer-Use Agents

DGX agent

arXiv:2606.22864v1 Announce Type: new Abstract: Hidden-state probing -- a linear classifier on a frozen vision-language model's internal activations -- has emerged as an attractive evaluation tool for

model-releasesarxiv-cs-lg
23 Jun 2026
Model Releases

Embodied-BenchClaw: An Autonomous Multi-Agent System for Embodied Spatial Intelligence Benchmark Construction

DGX agent

arXiv:2606.11909v1 Announce Type: new Abstract: Benchmarks are essential for evaluating embodied spatial intelligence, yet their construction is labor-intensive, hard to reuse, and difficult to mainta

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Multi-Agent Reasoning with Adaptive Worker Allocation for Stance Detection

DGX agent

arXiv:2606.11609v1 Announce Type: new Abstract: Stance detection requires identifying an author's position toward a target, often from short-form texts where stance is implicit, indirect, or rhetorica

model-releasesarxiv-cs-cl
11 Jun 2026
Local Ai

My Chemical Harness: Evolutionary Molecular Design over Synthetic Pathways with Large Language Model Agents

DGX agent

arXiv:2606.11256v1 Announce Type: cross Abstract: Designing molecules with target properties is most useful when candidate structures are accompanied by feasible synthetic routes. We introduce My Chem

local-aiarxiv-cs-lg
11 Jun 2026
Model Releases

ParseFixer: An Agentic Framework for Document Parsing via Selective Multimodal Correction

DGX agent

arXiv:2606.11977v1 Announce Type: new Abstract: In this report, we present our third-place solution for the DataMFM Challenge Track 1: Document Parsing. This track requires models to recover structure

model-releasesarxiv-cs-cv
11 Jun 2026
Safety

Sketch-to-Layout: A Human-Centric Computational Agent for Constraint-Aware Synthesis of Modular Photobioreactors

DGX agent

arXiv:2606.09849v1 Announce Type: cross Abstract: Building-integrated photobioreactors (PBRs) offer a pathway for carbon-neutral architecture, yet deployment is hindered by configuration complexity an

safetyarxiv-cs-cv
10 Jun 2026
Model Releases

What Matters in Orchestrating Robot Policies: A Systematic Study of Hierarchical VLA Agents

DGX agent

arXiv:2606.10267v1 Announce Type: cross Abstract: Hierarchical vision-language-action (Hi-VLA) systems have emerged as a promising paradigm for complex robot manipulation, by using high-level VLM plan

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

A Multi-modal Agentic Co-pilot for Evidence Grounded Computational Pathology

DGX agent

arXiv:2606.08093v1 Announce Type: new Abstract: Pathology is the cornerstone of modern medicine, where accurate decision-making relies heavily on evidence-based practices. While artificial intelligenc

model-releasesarxiv-cs-ai
9 Jun 2026
Safety

Aligned but Not Partner-Specific: Distinguishing How Multimodal LLM Agents Succeed in Reference Games Without Human-Like Conventions

DGX agent

arXiv:2606.08081v1 Announce Type: cross Abstract: Repeated reference games test whether interlocutors replace their initially long descriptions with shorter, partner-specific conventions grounded in s

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…173174175176177…375
Next →