AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

agents

GridTimelineEvolution
7,201 results
10 Jun 2026

Regimes: An Auditable, Held-Out-Gated Improvement Loop Demonstrated on LongMemEval with ActiveGraph

AgentsDGX agent

arXiv:2606.10241v1 Announce Type: new Abstract: Autonomous improvement loops are hard to trust because the improvement process is usually external scaffolding bolted onto the agent: failures go unlogg

Resilient Navigation for Autonomous Farm Robots by Leveraging Jerk-Augmented Models with IMU-Only Disturbance Rejection

AgentsDGX agent

arXiv:2606.10971v1 Announce Type: new Abstract: Precise state estimation for navigation of autonomous agricultural robots is often compromised by sensor outages (GNSS/LiDAR/Visual) and high-frequency

Robust Deep Reinforcement Learning Through Adversarial Attacks and Training : A Survey

AgentsDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2403.00420v3 Announce Type: replace-cross Abstract: Deep Reinforcement Learning (DRL) is a subfield of machine learning for training autonomous agents that take sequential actions across complex

Securing the AI workforce: Zscaler’s zero-trust play for agentic AI

AgentsDGX agent

Since Zscaler Inc.‘s launch, the company’s mission has been to disrupt traditional access and security with its Zero Trust platform. At its user event, Zenith Live, in Las Vegas, the company made its

shoutouts: • why multi-agent LLM systems fail? (arXiv:2503.13657) — @mertcemri @melissapan + @istoica05 @matei_zaharia @profjoeyg @adityagp …

AgentsDGX agent

shoutouts: • why multi-agent LLM systems fail? (arXiv:2503.13657) — @mertcemri @melissapan + @istoica05 @matei_zaharia @profjoeyg @adityagp & team • DSPy (arXiv:2310.03714) — @lateinteraction + @hazyr

Stop hand-tuning kernels: How Neuron Agentic Development accelerates AWS Trainium optimizations

AgentsDGX agent

Today, we’re announcing the Neuron Agentic Development capabilities: a collection of AI agents and skills that make this possible for developers building on AWS Trainium and AWS Inferentia. In this po

TabClaw: An Interactive and Self-Evolving Agent for Spreadsheet Manipulation and Table Reasoning

AgentsDGX agent

arXiv:2606.10316v1 Announce Type: new Abstract: Spreadsheets and tables are widely used representations for structured data analysis, but effective analysis still requires substantial manual effort an

TaCarla: A comprehensive benchmarking dataset for end-to-end autonomous driving

AgentsDGX agent

arXiv:2602.23499v4 Announce Type: replace-cross Abstract: Collecting a high-quality dataset is a critical task that demands meticulous attention to detail, as overlooking certain aspects can render th

The Arbiter Agent: Continually Monitoring Multi-Agent Conversations to Detect Emergent Misalignment

AgentsDGX agent

arXiv:2606.10747v1 Announce Type: new Abstract: As AI systems built from multiple language-model agents become more common, they are increasingly used to make decisions together: discussing, negotiati

The Confident Liar: Diagnosing Multi-Agent Debate with Log-Probabilities and LLM-as-Judge

AgentsDGX agent

arXiv:2606.10296v1 Announce Type: cross Abstract: Multi-agent debate systems are typically evaluated only on whether the final answer is correct, overlooking the quality of the intermediate reasoning

the cost of fable is going to make smart model routing impossible to ignore

AgentsDGX agent

This post discusses how the pricing model of Fable (likely an AI/LLM service) creates economic incentives that make intelligent routing between different AI models a necessary optimization strategy ra

The Distributed Detectability Band Against Marginal-Preserving Attacks

AgentsDGX agent

arXiv:2606.10456v1 Announce Type: cross Abstract: AI-control monitors score individual agent actions to detect misbehavior, but real harm can be distributed across many benign-looking steps, each indi

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now…

AgentsDGX agent

The Model Lab vs Agent Lab distinction is one of the clearest frameworks I've seen for understanding where AI value actually lives right now. TL;DR from @latentspacepod: • Model Labs compete on capabi

the results are modest (esp compared to parallel & similar research GRASP: https://arxiv.org/abs/2605.29668 - recommended!) the contribution…

AgentsDGX agent

the results are modest (esp compared to parallel & similar research GRASP: https://arxiv.org/abs/2605.29668 - recommended!) the contribution here is not the self improvement approach itself (yet), but

They should name the next model Halo 6 then ninja gaiden 7 cover the entire xbox catalogue

AgentsDGX agent

This post suggests naming future models after Xbox game franchises, specifically proposing 'Halo 6' and 'Ninja Gaiden 7' as model names that would reference the broader Xbox catalog. The comment appea

This is just awesomeness from @cohere, @nickfrosst, and team. I so badly want a coding agent that just runs on my local machine. We are not …

AgentsDGX agent

This is just awesomeness from @cohere, @nickfrosst, and team. I so badly want a coding agent that just runs on my local machine. We are not too far now! Excited to get this to work with my @dair_ai co

Toward Secure LLM Agents: Threat Surfaces, Attacks, Defenses, and Evaluation

AgentsDGX agent

arXiv:2606.10749v1 Announce Type: cross Abstract: Large language model (LLM) agents are rapidly moving from conversational interfaces to software components that plan, invoke tools, maintain memory, a

Towards Autonomous Accelerator Design: FPGA Accelerator Generation with SECDA

AgentsDGX agent

arXiv:2606.11117v1 Announce Type: cross Abstract: Designing FPGA-based accelerators for modern artificial intelligence workloads requires exploring a large and complex hardware design space that invol

Trace Only What You Need: Structure-Aware On-Demand Hypergraph Memory for Long-Document Question Answering

AgentsDGX agent

arXiv:2606.10921v1 Announce Type: new Abstract: Long-document question answering (QA) requires large language models (LLMs) to reason over evidence scattered across lengthy documents, where answers of

two fun surprises from using activegraph: - the coding agent i was using would query the trace db to debug instead of looking at the logs li…

AgentsDGX agent

two fun surprises from using activegraph: - the coding agent i was using would query the trace db to debug instead of looking at the logs like they normally would (i didn't ask it to) - when long eval

Understanding and mitigating the risks of OpenClaw for non-technical users: A practical guide with Skill

AgentsDGX agent

arXiv:2606.11007v1 Announce Type: cross Abstract: OpenClaw has rapidly emerged as a transformative artificial intelligence (AI) agent framework, and its ability to autonomously execute complex, multi-

Vehicle Prediction Model for Enhanced MPC Path Tracking in Formula Student Driverless

AgentsDGX agent

arXiv:2606.10732v1 Announce Type: new Abstract: Autonomous race cars, such as in Formula Student Driverless, operate close to their physical handling limits. The resulting highly nonlinear vehicle beh

Visa partners with OpenAI to let AI agents make payments for users

AgentsDGX agent

Visa Inc. has struck a deal with OpenAI Group PBC to let artificial intelligence agents make payments for users, bringing one of the world’s largest payment networks into ChatGPT’s push toward agentic

VISTA: A Versatile Interactive User Simulation Toolkit for Agent Evaluation

AgentsDGX agent

arXiv:2606.11079v1 Announce Type: new Abstract: Evaluation remains a critical bottleneck for interactive agent development. Existing evaluation methods often rely on static benchmarks, which fail to c

We raised $6M led by Sequoia to build the future of travel. Watch me plan a perfect trip to Mexico City in 3 minutes. Flights, hotels and a …

AgentsDGX agent

We raised $6M led by Sequoia to build the future of travel. Watch me plan a perfect trip to Mexico City in 3 minutes. Flights, hotels and a full itinerary that matches my preferences. All bookable on

What Spatial Memory Must Store: Occlusion as the Test for Language-Agent Memory

AgentsDGX agent

arXiv:2606.10299v1 Announce Type: new Abstract: Language-agent 'memory palace' systems anchor each memory to a world coordinate, on the intuition that geometry adds something text cannot. We make that

Years ago someone asked me what comes after the AI-first enterprise. I said the autonomous enterprise. The core of it is a flywheel: four pa…

AgentsDGX agent

Years ago someone asked me what comes after the AI-first enterprise. I said the autonomous enterprise. The core of it is a flywheel: four parts (well, five including feedback) that feed each other, an

Zscaler unveils ZAgent Framework to automate zero-trust SASE operations

AgentsDGX agent

Zscaler Inc. today unveiled a major expansion of its zero-trust SASE platform, adding an agentic framework that lets administrators manage the system through natural-language prompts and extending its

9 Jun 2026

A case study of evaluating AI agents on a neuroscience data-to-discovery pipeline

AgentsDGX agent

arXiv:2606.07718v1 Announce Type: new Abstract: Agentic AI tools offer a promising path to automating software development bottlenecks in scientific research pipelines, particularly for stages that ta

A Multi-Agent System for IPMSM Design Optimization via an FEA-AI Hybrid Approach

AgentsDGX agent

arXiv:2606.09037v1 Announce Type: new Abstract: Interior permanent magnet synchronous motor (IPMSM) design requires balancing conflicting objectives and multi-physics constraints, while modern optimiz

A multi-agent system for spine MRI report generation from multi-sequence imaging

AgentsDGX agent

arXiv:2606.08897v1 Announce Type: cross Abstract: Spinal pathology is a leading cause of pain and disability worldwide. Spine MRI is central to clinical evaluation, yet its interpretation remains comp

A Resilience-as-a-Service assessment framework for coordinated disruption response in interdependent urban transit systems

AgentsDGX agent

arXiv:2606.08849v1 Announce Type: new Abstract: Urban public transport disruptions require rapid response strategies, yet existing studies rarely provide a decision support framework to compare altern

A Survey on Deep Multi-Task Learning in Connected Autonomous Vehicles

AgentsDGX agent

arXiv:2508.00917v2 Announce Type: replace-cross Abstract: Connected autonomous vehicles (CAVs) must simultaneously perform multiple tasks, such as perception, prediction, planning, and control, to ens

A Survey on Large Language Model-Based Game Agents

AgentsDGX agent

arXiv:2404.02039v5 Announce Type: replace Abstract: Game environments provide rich, controllable settings that stimulate many aspects of real-world complexity. As such, game agents offer a valuable te

Advancing Mathematics Research with AI-Driven Formal Proof Search

AgentsDGX agent

arXiv:2605.22763v2 Announce Type: replace Abstract: Large language models (LLMs) increasingly excel at mathematical reasoning, but their unreliability limits their utility in mathematics research. A m

Agent filesystems are the new RAG It seems like this pattern is around to stay and will only get more robust over time. Agents need tools to…

AgentsDGX agent

Agent filesystems are the new RAG It seems like this pattern is around to stay and will only get more robust over time. Agents need tools to not only read and search over documents, but an entire infr

Agentic multi-fidelity learning of quasiparticle and excitonic properties

AgentsDGX agent

arXiv:2606.07836v1 Announce Type: cross Abstract: Many-body GW-Bethe-Salpeter equation calculations are essential for accurate simulations of electronic structure and optical properties in modern low-

Agentic Neuro-Symbolic Planning and Commissioning for Human-in-the-Loop Industrial Robotics with Digital Twins

AgentsDGX agent

arXiv:2606.08214v1 Announce Type: new Abstract: Flexible robotic automation requires systems that interpret operator intent, verify physical feasibility, and recover from execution failures across bot

Agentic Search for Counterfactual Recourse under Fixed LLM Budgets

AgentsDGX agent

arXiv:2606.08696v1 Announce Type: cross Abstract: Counterfactual recourse aims to provide actionable feature changes that would alter an unfavorable decision made by a predictive model. In practice, a

AgentTrust: A Self-Improving Trust Layer for AI-Agent Actions

AgentsDGX agent

arXiv:2606.08539v1 Announce Type: new Abstract: AI agents increasingly take consequential actions -- shell commands, cloud operations, and arbitrary tool-calls -- so a trust layer must decide, per act

AlloSpatial: Agentic Harness Framework for Spatial Reasoning in Foundation Models

AgentsDGX agent

arXiv:2606.08952v1 Announce Type: new Abstract: Multimodal Foundation Models (MFMs) have made substantial progress, yet remain fragile in spatial reasoning over the physical world. A key bottleneck li

An AI Security Agent for University ACMIS: Multi-Vector Threat Detection and Automated Response

AgentsDGX agent

arXiv:2606.08270v1 Announce Type: cross Abstract: University Academic Management Information Systems (ACMIS) are high-value targets for a wide spectrum of security threats including brute-force login

An Information-Theoretic Definition for Open-Ended Learning

AgentsDGX agent

arXiv:2606.08369v1 Announce Type: cross Abstract: A growing body of work points to the great promise of AI systems that can continually expand their capabilities as they operate in an open-ended envir

Anything2Skill: Compiling External Knowledge into Reusable Skills for Agents

AgentsDGX agent

arXiv:2606.09316v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) enables agents to access external knowledge at inference time, but it primarily retrieves fragmented declarative ev

As frontier models (e.g. Fable 5) continue to push the task horizon of knowledge work automation, it becomes ever more important for humans …

AgentsDGX agent

As frontier models (e.g. Fable 5) continue to push the task horizon of knowledge work automation, it becomes ever more important for humans to be able to audit decisions back to the source context. It

(Auto)formalization is supposed to be easy: Trellis process semantics for spelling out rigorous proofs

AgentsDGX agent

arXiv:2606.09674v1 Announce Type: new Abstract: We present Trellis: an autoformalization system that leverages LLM agents in a deterministically constrained workflow to enforce incremental progress in

Beyond Agent Architecture: Execution Assumptions and Reproducibility in LLM-Based Trading Systems

AgentsDGX agent

arXiv:2606.08285v1 Announce Type: new Abstract: Large language models (LLMs) and agentic systems are increasingly proposed for financial trading, yet their reported performance remains difficult to co

Build an agentic incident triage assistant with Amazon Quick and New Relic

AgentsDGX agent

This post shows engineering teams how to apply that principle to one of the most time-sensitive workflows in engineering: incident triage. You will build a custom incident triage assistant agent using

Building at the speed of research: Lambda at CVPR 2026

AgentsDGX agent

Every year, CVPR draws the researchers defining what AI can see, understand, and act on. This year in Denver, more than 9,000 attendees showed up with over 4,000 accepted papers, and one shared proble

Byzantine Cheap Talk: Adversarial Resilience and Topology Effects in LLM Coordination Games

AgentsDGX agent

arXiv:2606.07790v1 Announce Type: new Abstract: Multi-agent LLM systems increasingly rely on communication protocols for coordination, yet their robustness under adversarial and structural constraints

Collaborative Human-Agent Protocol (CHAP)

AgentsDGX agent

arXiv:2606.09751v1 Announce Type: new Abstract: Foundation models are moving from response generation into operational roles. They plan across steps, call tools, request human input, coordinate with o

ConMem: Structured Memory-Guided Adaptation in Training-Free Multi-Agent Systems

AgentsDGX agent

arXiv:2606.08702v1 Announce Type: new Abstract: Recent advances have improved the adaptive capabilities of LLM-based multi-agent systems (MAS) through memory-, skill-, and learning-based approaches, y

Connect to your messaging apps Connect an agent to Telegram, Discord, Slack, WhatsApp, Signal, Email, and more. One agent, one memory, every…

AgentsDGX agent

Ollama enables users to connect AI agents to multiple messaging platforms including Telegram, Discord, Slack, WhatsApp, Signal, and Email through a single unified agent with persistent memory across a

Context-Aware Deep Learning for Defect Classification in Atomic-Resolution STEM

AgentsDGX agent

arXiv:2606.09419v1 Announce Type: cross Abstract: Artificial intelligence is rapidly advancing materials characterization, yet most applications in electron microscopy rely solely on image contrast, o

Continual Quadruped Robots Coordination via Semantic Skill Discovery

AgentsDGX agent

arXiv:2606.08102v1 Announce Type: cross Abstract: Multi-quadruped coordination has attracted increasing attention due to its enhanced payload capacity, broader contact coverage, and improved adaptabil

Contract2Tool: Learning Preconditions and Effects for Reliable Tool-Augmented LLM Agents

AgentsDGX agent

arXiv:2606.07904v1 Announce Type: new Abstract: Tool-augmented large language model agents increasingly rely on external APIs, but standard tool schemas describe how to call a tool, not when the tool

Coop-WD: Cooperative Perception with Weighting and Denoising for Robust V2V Communication

AgentsDGX agent

arXiv:2505.03528v2 Announce Type: replace Abstract: Cooperative perception, leveraging shared information from multiple vehicles via vehicle-to-vehicle (V2V) communication, plays a vital role in auton

Cost-Aware Speculative Execution for LLM-Agent Workflows: An Integrated Five-Dimension Method

AgentsDGX agent

arXiv:2606.07846v1 Announce Type: cross Abstract: LLM-agent workflows chain model calls and tool invocations, and spend most of their wall-clock time waiting on upstream operations before downstream o

Data center modernization unlocks AI budget headroom as enterprises fund new AI workloads

AgentsDGX agent

As enterprise AI budgets hit their limits earlier each year, the pressure to fund new agentic and inference workloads without expanding total spend is forcing a fundamental rethink of infrastructure a

Deep reinforcement learning for process design: Review and perspective

AgentsDGX agent

arXiv:2308.07822v2 Announce Type: replace Abstract: The transformation towards renewable energy and feedstock supply in the chemical industry requires new conceptual process design approaches. Recentl

← Previous
1…4142434445…121
Next →