AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,958 results
Industry

Patronus AI, which builds simulated digital environments for evaluating AI agents, raised a 50M Series B led by Greenfield, bringing its total funding to 70M (Marina Temkin/TechCrunch)

DGX agent

Marina Temkin / TechCrunch: Patronus AI, which builds simulated digital environments for evaluating AI agents, raised a 50M Series B led by Greenfield, bringing its total funding to 70M — AI agents ar

industrytechmeme
25 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Agentic AI for Bilevel Long-Term Optimization of Policy-Driven Physical Layer Systems

DGX agent

arXiv:2606.24416v1 Announce Type: new Abstract: Network operators' changing policies, service requirements, and stringent real-time constraints render existing methods designed with fixed objectives a

safetyarxiv-cs-ai
24 Jun 2026
Model Releases

AGORA: An Archive-Grounded Benchmark for Agentic Workplace Document Reasoning

DGX agent

arXiv:2606.24526v1 Announce Type: new Abstract: Large language models are increasingly deployed as agents that reason over documents rather than answer from parametric knowledge. We study archive-grou

model-releasesarxiv-cs-cl
24 Jun 2026
Model Releases

ASALT: Adaptive State Alignment for Lateral Transfer in Multi-agent Reinforcement Learning

DGX agent

arXiv:2606.24601v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) addresses the problem of training multiple agents that pursue collaborative, competitive, or mixed objectives.

model-releasesarxiv-cs-ai
24 Jun 2026
Model Releases

EComAgentBench: Benchmarking Shopping Agents on Long-Horizon Tasks with Distributed Hidden Intent

DGX agent

arXiv:2606.17698v2 Announce Type: replace Abstract: As LLM-based shopping agents enter production, existing benchmarks fail to capture how a shopper's requirements arrive: stated implicitly in the que

model-releasesarxiv-cs-ai
24 Jun 2026
Safety

Safe and Generalizable Hierarchical Multi-Agent RL via Constraint Manifold Control

DGX agent

arXiv:2606.24010v1 Announce Type: new Abstract: Multi-agent systems are widely used in safety-critical applications that require coordinated behavior under strict safety constraints. Existing approach

safetyarxiv-cs-ai
24 Jun 2026
Agents

Varying Bundle Size Reactive Multi-Task Assignment using Selective Cost Estimation for Multi-Agent Systems

DGX agent

arXiv:2606.24462v1 Announce Type: new Abstract: This paper presents a scalable framework for multi-robot task allocation in complex environments where estimating task execution costs is computationall

agentsarxiv-cs-ro
24 Jun 2026
Agents

Hypothesis-Disciplined Multi-Agent Automated Formalization of Asymptotic Statistical Theory

DGX agent

arXiv:2606.20642v1 Announce Type: cross Abstract: Asymptotic statistical theory is a challenging domain for AI-assisted formalization: its central results mix convergence statements, asymptotic expans

agentsarxiv-cs-lg
23 Jun 2026
Agents

Log Analytics is now Observability Analytics: Query logs and traces with SQL

DGX agent

To effectively operate and troubleshoot applications, developers and site reliability engineers (SREs) need to understand the full context of their system's behavior, typically as part of their loggin

agentsgoogle-cloud-ai
23 Jun 2026
Agents

RAPID: A Reproducible Multi-Agent Pipeline for Interpretable Disaster Damage Assessment from Satellite and Street-View Imagery

DGX agent

arXiv:2606.21819v1 Announce Type: new Abstract: Due to the increasing frequency and intensity of extreme climate events, there is a clear demand for intelligent, scalable, and autonomous approaches to

agentsarxiv-cs-cv
23 Jun 2026
Agents

SPARC: A Multi-Agent System for Electrical Circuit Question Answering

DGX agent

arXiv:2606.20643v1 Announce Type: cross Abstract: Electrical circuit diagram QA tasks require complex mathematical reasoning, which remains challenging for multimodal LLMs. We present SPARC, a multi-a

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

Specialize Roles, Mix Deployments: Pushing the Cost-Accuracy Frontier of LLM Agent Teams

DGX agent

arXiv:2606.20629v1 Announce Type: cross Abstract: LLM agents are increasingly deployed as multi-role teams, where tasks are divided across specialized roles such as planner, executor, and verifier. In

model-releasesarxiv-cs-lg
23 Jun 2026
Agents

WebCryptoAgent: Agentic Crypto Trading with Web Informatics

DGX agent

arXiv:2601.04687v2 Announce Type: replace Abstract: Cryptocurrency trading increasingly depends on timely integration of heterogeneous web information and market microstructure signals to support shor

agentsarxiv-cs-cv
23 Jun 2026
Model Releases

Ai2 just released TMax 27B on Hugging Face A 27B terminal agent that hits 42.7% on Terminal Bench 2.0, rivaling models 40× its size.

DGX agent

AI2 released TMax 27B, a 27 billion parameter terminal agent model available on Hugging Face that achieves 42.7% performance on Terminal Bench 2.0, matching the capabilities of much larger models desp

model-releasesclem-delangue--x
22 Jun 2026
Hardware

Join us in our Discord this time tomorrow for Office Hours! Our topic will be the @NVIDIAAI × @stripe × @NousResearch Hermes Agent Accelerat…

DGX agent

Join us in our Discord this time tomorrow for Office Hours! Our topic will be the @NVIDIAAI × @stripe × @NousResearch Hermes Agent Accelerated Business Hackathon which concludes at the end of the mont

hardwarenous-research--x
22 Jun 2026
Model Releases

Temporary Cloudflare Accounts for AI agents

DGX agent

Temporary Cloudflare Accounts for AI agents The announcement says this is 'for AI agents' but (as is pretty common these days) the AI hook isn't really necessary, this is an interesting feature for ev

model-releasessimon-willison
21 Jun 2026
Model Releases

A Five-Plane Reference Architecture for Runtime Governance of Production AI Agents

DGX agent

arXiv:2606.12320v1 Announce Type: new Abstract: Enterprise security was built to govern data boundaries: the protected surface was data at rest and in transit, and the controls -- access control, data

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

Agent Skill Evaluation and Evolution: Frameworks and Benchmarks

DGX agent

arXiv:2606.11435v1 Announce Type: new Abstract: The growth of agent skills has transformed how agentic systems are built, evaluated, and deployed. As skill libraries continue to scale, rigorous evalua

model-releasesarxiv-cs-cl
11 Jun 2026
Local Ai

Can Open-Source LLM Agents Replace Static Application Security Testing Tools? An Empirical Assessment

DGX agent

arXiv:2606.11672v1 Announce Type: cross Abstract: This paper explores the value of agentic AI tools for cybersecurity purposes. We evaluate the efficacy of a general-purpose GenAI Large Language Model

local-aiarxiv-cs-ai
11 Jun 2026
Model Releases

Claw-SWE-Bench: A Benchmark for Evaluating OpenClaw-style Agent Harnesses on Coding Tasks

DGX agent

arXiv:2606.12344v1 Announce Type: cross Abstract: General-purpose agents such as OpenClaw are increasingly used as autonomous tool users, but their coding ability is difficult to measure under SWE-ben

model-releasesarxiv-cs-cl
11 Jun 2026
Safety

Diffusing to Coordinate: Efficient Online Multi-Agent Diffusion Policies

DGX agent

arXiv:2602.18291v2 Announce Type: replace Abstract: Online Multi-Agent Reinforcement Learning (MARL) is a prominent framework for efficient agent coordination. Crucially, enhancing policy expressivene

safetyarxiv-cs-ai
11 Jun 2026
Agents

DrivingAgent: Design and Scheduling Agents for Autonomous Driving Systems

DGX agent

arXiv:2606.12236v1 Announce Type: cross Abstract: Many autonomous driving systems are increasingly incorporating foundation models to improve generalization and handle long-tail scenarios. However, th

agentsarxiv-cs-cv
11 Jun 2026
Agents

Fanar-Sadiq: A Multi-Agent Architecture for Grounded Islamic QA

DGX agent

arXiv:2603.08501v3 Announce Type: replace Abstract: Large language models (LLMs) can answer religious knowledge queries fluently, yet they often hallucinate and misattribute sources, which is especial

agentsarxiv-cs-cl
11 Jun 2026
Model Releases

Human-Guided Agentic AI for Multimodal Clinical Prediction: Lessons from the AgentDS Healthcare Benchmark

DGX agent

arXiv:2602.19502v2 Announce Type: replace Abstract: Agentic AI systems are increasingly capable of autonomous data science workflows, yet clinical prediction tasks demand domain expertise that purely

model-releasesarxiv-cs-ai
11 Jun 2026
Hardware

INFRAMIND: Infrastructure-Aware Multi-Agent Orchestration

DGX agent

arXiv:2606.11440v1 Announce Type: new Abstract: Existing multi-agent LLM orchestration methods, ranging from brute-force ensembles to learned routers, select models and topologies based on task and mo

hardwarearxiv-cs-ai
11 Jun 2026
Model Releases

MARIC: Multi-Agent Reasoning for Image Classification

DGX agent

arXiv:2509.14860v2 Announce Type: replace-cross Abstract: Image classification has traditionally relied on parameter-intensive model training, requiring large-scale annotated datasets and extensive fi

model-releasesarxiv-cs-ai
11 Jun 2026
Model Releases

MedCTA: A Benchmark for Clinical Tool Agents

DGX agent

arXiv:2606.11702v1 Announce Type: cross Abstract: To make clinically grounded decisions, medical AI agents are expected to go beyond simple recognition and be capable of tool retrieval, evidence acqui

model-releasesarxiv-cs-ai
11 Jun 2026
Local Ai

PROJECTMEM: A Local-First, Event-Sourced Memory and Judgment Layer for AI Coding Agents

DGX agent

arXiv:2606.12329v1 Announce Type: new Abstract: AI coding assistants now support a growing share of software work, from quick scripts to production applications. Yet these agents remain largely statel

local-aiarxiv-cs-ai
11 Jun 2026
Agents

Sustainability assessment using multimodal AI agents

DGX agent

arXiv:2507.17012v2 Announce Type: replace Abstract: Reducing the rapidly growing environmental impact of the computing industry requires assessing the emissions of electronics at scale. However, a tra

agentsarxiv-cs-ai
11 Jun 2026
Model Releases

ABC-Bench: An Agentic Bio-Capabilities Benchmark for Biosecurity

DGX agent

arXiv:2606.11150v1 Announce Type: new Abstract: Large language models (LLMs) are rapidly acquiring capabilities relevant to biological research, from literature synthesis to interpretation of experime

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

Assessing Automated Prompt Injection Attacks in Agentic Environments

DGX agent

arXiv:2606.10525v1 Announce Type: cross Abstract: Indirect prompt injection poses a critical threat to LLM agents that interact with untrusted external data, yet automated attack methods--proven effec

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

au-Rec: A Verifiable Benchmark for Agentic Recommender Systems

DGX agent

arXiv:2606.10156v1 Announce Type: cross Abstract: As recommender systems transition toward agentic, multi-turn conversational interfaces, evaluation paradigms have struggled to keep pace. Current benc

model-releasesarxiv-cs-ai
10 Jun 2026
Agents

ChartAgent: A Multimodal Agent for Visually Grounded Reasoning in Complex Chart Question Answering

DGX agent

arXiv:2510.04514v3 Announce Type: replace Abstract: Recent multimodal LLMs have shown promise in chart-based visual question answering, but their performance declines sharply on unannotated charts-tho

agentsarxiv-cs-ai
10 Jun 2026
Model Releases

Deployment-Time Memorization in Foundation-Model Agents

DGX agent

arXiv:2606.10062v1 Announce Type: new Abstract: Foundation-model agents are increasingly long-lived systems that remember users across interactions, making memorization an explicit deployment-time fun

model-releasesarxiv-cs-ai
10 Jun 2026
Model Releases

everything your agents do, in a bucket you own. private by default. session traces, task claims, artifacts — claude code, hermes, whatever y…

DGX agent

everything your agents do, in a bucket you own. private by default. session traces, task claims, artifacts — claude code, hermes, whatever you run. and make your agents friends via messaging https://h

model-releasesclem-delangue--x
10 Jun 2026
Agents

Securing the AI workforce: Zscaler’s zero-trust play for agentic AI

DGX agent

Since Zscaler Inc.‘s launch, the company’s mission has been to disrupt traditional access and security with its Zero Trust platform. At its user event, Zenith Live, in Las Vegas, the company made its

agentssiliconangle
10 Jun 2026
Agents

Agentic Search for Counterfactual Recourse under Fixed LLM Budgets

DGX agent

arXiv:2606.08696v1 Announce Type: cross Abstract: Counterfactual recourse aims to provide actionable feature changes that would alter an unfavorable decision made by a predictive model. In practice, a

agentsarxiv-cs-ai
9 Jun 2026
Local Ai

AGENTSERVESIM: A Hardware-aware Simulator for Multi-Turn LLM Agent Serving

DGX agent

arXiv:2606.09613v1 Announce Type: cross Abstract: Multi-turn LLM agents interleave model calls with external tool invocations, shifting serving from stateless request processing to stateful program ex

local-aiarxiv-cs-ai
9 Jun 2026
Agents

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models

DGX agent

arXiv:2606.08952v1 Announce Type: new Abstract: Multimodal Foundation Models (MFMs) have made substantial progress, yet remain fragile in spatial reasoning over the physical world. A key bottleneck li

agentsarxiv-cs-ai
9 Jun 2026
Agents

An AI Security Agent for University ACMIS: Multi-Vector Threat Detection and Automated Response

DGX agent

arXiv:2606.08270v1 Announce Type: cross Abstract: University Academic Management Information Systems (ACMIS) are high-value targets for a wide spectrum of security threats including brute-force login

agentsarxiv-cs-ai
9 Jun 2026
Safety

Autonomous Incident Resolution at Hyperscale: An Agentic AI Architecture for Network Operations

DGX agent

arXiv:2606.09122v1 Announce Type: cross Abstract: Cloud network infrastructure at hyperscale presents unique operational challenges where traditional human-driven incident response cannot keep pace wi

safetyarxiv-cs-ai
9 Jun 2026
Hardware

BRAIN: Bayesian Reasoning via Active Inference for Agentic and Embodied Intelligence in Mobile Networks

DGX agent

arXiv:2602.14033v1 Announce Type: cross Abstract: Future sixth-generation (6G) mobile networks will demand artificial intelligence (AI) agents that are not only autonomous and efficient, but also capa

hardwarearxiv-cs-ai
9 Jun 2026
Safety

Brain-Prompt Injection: A Route-Safety Audit for BCI-LLM Agents

DGX agent

arXiv:2606.09315v1 Announce Type: cross Abstract: BCI-to-agent pipelines turn decoded neural activity into an authorization channel for tool-use agents, exposing a new attack surface we call brain-pro

safetyarxiv-cs-ai
9 Jun 2026
Agents

DTEX adds AI Risk Management to track how agents and employees use AI

DGX agent

Behavioral intelligence security company DTEX Systems Inc. today introduced an expanded AI Risk Management product that reads the intent behind how employees and autonomous artificial intelligence age

agentssiliconangle
9 Jun 2026
Model Releases

Experience Makes Skillful: Enabling Generalizable Medical Agent Reasoning via Self-Evolving Skill Memory

DGX agent

arXiv:2606.09365v1 Announce Type: new Abstract: Medical agent systems are increasingly expected to support interactive clinical decision making rather than only static question answering. In such sett

model-releasesarxiv-cs-ai
9 Jun 2026
Model Releases

From Human Guidance to Autonomy: Agent Skill System for End-to-End LLM Deployment on Spatial NPUs

DGX agent

arXiv:2606.07586v1 Announce Type: cross Abstract: Spatial neural processing units (NPUs) provide an energy-efficient platform for edge LLM inference, but efficiently deploying an LLM end-to-end on suc

model-releasesarxiv-cs-ai
9 Jun 2026
Industry

Hands-free first notice of loss: Using Strands Agents and Amazon Bedrock AgentCore Browser Tool for intelligent claims intake

DGX agent

In this post, we demonstrate how a hands-free FNOL intake system combines agents built with the Strands Agents SDK for domain reasoning with Amazon Bedrock AgentCore Browser Tool for live portal inter

industryaws-ml-blog
9 Jun 2026
Model Releases

Hardening Agent Benchmarks with Adversarial Hacker-Fixer Loops

DGX agent

arXiv:2606.08960v1 Announce Type: cross Abstract: Agent benchmarks score submissions with outcome verifiers that are typically hand-written and brittle, leaving them open to reward hacking. We audit 1

model-releasesarxiv-cs-ai
9 Jun 2026
← Previous
1…117118119120121…375
Next →