AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
17,951 results
12 Aug 2026

Self-Correcting Long-Horizon Search Agents via Tree-Structured Memory

ResearchDGX agent

arXiv:2608.10676v1 Announce Type: new Abstract: Large language model (LLM)-based search agents answer questions through multi-step interactions with external environments. However, providing complete

Skan AI raises $63M to give AI agents a map of enterprise work

ApplicationsDGX agent

Process intelligence company Skan AI said today it raised 63 million in a Series C round to help further develop a platform that records how enterprise work actually gets done and feeds that record to

SkillZip: Evaluation-Free Skill Compression for Self-Evolving Agents by Discovering Reusable Structure

ResearchDGX agent

arXiv:2608.11079v1 Announce Type: new Abstract: Self-evolving agents accumulate reusable skills by appending successful procedures and failure fixes. Over time, the same requirement is often restated

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Toward the Cognitive--Physical Limits of Embodied Intelligence through a World-Model-Centric Autonomous Racing Agent

Local AiDGX agent

arXiv:2608.10618v1 Announce Type: new Abstract: Embodied artificial intelligence aims to develop agents that perceive, reason, and act through continuous interaction with the physical world. However,

11 Aug 2026

ActBench: Self-Evolving Benchmark of Behavioral Safety in Cowork Agents

Model ReleasesDGX agent

arXiv:2608.09476v1 Announce Type: cross Abstract: Cowork agents may complete benign tasks while disclosing protected data, manipulating unauthorized state, invocate unauthorized API. We define behavio

Agentic AI-driven Immersive Simulation: A Knowledge-Aware Virtual Training Platform forHigh Dose Rate (HDR) Brachytherapy

Local AiDGX agent

arXiv:2608.08163v1 Announce Type: new Abstract: The convergence of the Metaverse and Large Language Model (LLM)-based AI agent is catalyzing a shift toward autonomous, immersive, and personalized peda

AIVV: Neuro-Symbolic LLM Agent-Integrated Verification and Validation for Trustworthy Autonomous Systems

AgentsDGX agent

arXiv:2604.02478v2 Announce Type: replace Abstract: Deep learning models excel at detecting anomaly patterns in normal data. However, they do not provide a direct solution for anomaly classification a

Beyond Pixels: Exploring DOM Downsampling for LLM-Based Web Agents

Model ReleasesDGX agent

arXiv:2508.04412v3 Announce Type: replace Abstract: The advent of large language models (LLMs) has sparked an evolution of autonomous web browsing agents: given a web browsing task and serialised user

Bidirectional Context Self-Distillation for Reinforcement Learning of Skill-Based LLM Agents

Model ReleasesDGX agent

arXiv:2608.09555v1 Announce Type: new Abstract: External natural-language skills provide large language model (LLM) agents with reusable and editable guidance for solving complex tasks. Yet their effe

DarwinX: Evolving Agent Harnesses Through Natural Selection

Model ReleasesDGX agent

arXiv:2608.07545v1 Announce Type: cross Abstract: An LLM agent's capability depends not only on model weights but on its harness: prompts, tools, skills, and control flow. Self-improvement loops alrea

Discovering Diverse Planning Policies for Multimodal Embodied Agents with Quality-Diversity Optimization

Model ReleasesDGX agent

arXiv:2608.08523v1 Announce Type: new Abstract: Multimodal embodied agents are increasingly required to solve long-horizon tasks by integrating visual observations, textual goals, and interaction hist

Experience-Sensitive Game Learning: A Behavioral Study of Humans and Language Agents

TutorialsDGX agent

arXiv:2608.07490v1 Announce Type: cross Abstract: Large language model agents are increasingly evaluated through games, but most benchmarks emphasize final outcomes rather than how players learn from

IDRAAK: From Multi-Agent NLP to Few-Shot Prompting for Semantic Drift Detection in Technical Requirements

AgentsDGX agent

arXiv:2608.08801v1 Announce Type: new Abstract: Translating technical requirements across languages can introduce semantic drift, altering numerical constraints, polarities, modalities, or other speci

InfMem: Learning System-2 Memory Control for Long-Context Agent

AgentsDGX agent

arXiv:2602.02704v2 Announce Type: replace Abstract: Reasoning over ultra-long documents requires synthesizing sparse evidence scattered across distant segments under strict memory constraints. While s

Knowing You Is Everything: LLM Agents Achieve Near-Perfect Profile-Consistent Reaction Prediction in Social Media Simulation

Model ReleasesDGX agent

arXiv:2608.07498v1 Announce Type: cross Abstract: Autonomous AI agents in social media present concrete risks to democratic discourse and platform governance, while also offering tools for pre-deploym

Legal Responsibilities Using Autonomous Agents For Artificial Intelligence

SafetyDGX agent

arXiv:2608.08022v1 Announce Type: new Abstract: Recent incidents involving Artificial Intelligence (AI) agents, which were reported escaping their containment `unintentionally' to gain unauthorized ac

Metanormative Theory for RL-Based Moral Agents

SafetyDGX agent

arXiv:2608.08220v1 Announce Type: new Abstract: The overlapping disciplines of machine ethics and value alignment are concerned with designing artificial agents that are aligned with human values and

NVIDIA JetPack 7.2.1 Adds Agentic Video Skills and T3000 Emulation

HardwareDGX agent

NVIDIA JetPack 7.2.1 adds agentic video skills with the unified jetson‑videosdk, allowing programmable, device-aware video workflows that link developer intent to live device discovery and performance

NVIDIA Nemotron 3.5 Lightning and NeMo Switchyard Deliver Faster, Smarter, More Efficient Agentic AI

Model ReleasesDGX agent

As AI shifts from chatbots to autonomous agents, open models are serving market demands for full control over where AI runs and how it’s deployed and evolves. Today, NVIDIA is expanding its Nemotron 3

Open Evaluation Agent: Efficient and Promptable Evaluation of Visual Generative Models

Model ReleasesDGX agent

arXiv:2608.09666v1 Announce Type: new Abstract: Recent advances in visual generative models have enabled high-quality image and video generation, but evaluating these models often demands sampling hun

Ouroboros: A Self-Developing Frontier Coding Agent with Reviewed Core Evolution

Model ReleasesDGX agent

arXiv:2608.08311v1 Announce Type: cross Abstract: We present Ouroboros, a self-developing agent harness whose tools, prompts, context assembly, and core implementation improve through reviewed commits

QuantumMind: Constraint-Grounded Agentic Reasoning for Speedup Analysis in Quantum Computing

AgentsDGX agent

arXiv:2608.07743v1 Announce Type: new Abstract: Identifying a meaningful quantum speedup requires more than matching a classical problem to a familiar quantum primitive: the claim must preserve the ta

Reading is not Reasoning: Bridging the Agentic Policy Gap in Vision-Text Compression

SafetyDGX agent

arXiv:2608.08960v1 Announce Type: new Abstract: Multi-step language-model agents repeatedly process growing interaction histories, leading to substantial context costs. Vision--text compression reduce

Sparks of Cooperative Reasoning: LLMs as Strategic Hanabi Agents

Model ReleasesDGX agent

arXiv:2601.18077v3 Announce Type: replace Abstract: Cooperative reasoning under incomplete information remains challenging for both humans and multi-agent systems. The card game Hanabi embodies this c

STEMMA: An Adversarial Multi-Agent Framework for Evaluating Self-Identity Consistency in LLMs

AgentsDGX agent

arXiv:2608.08164v1 Announce Type: cross Abstract: Knowledge Distillation is a widely adopted technique in the training and fine-tuning of large language models (LLMs) enabling transfer of structured i

SWE-Bench ProMax: Benchmarking Agents on Large-Scale Multilingual Code Refactoring

Model ReleasesDGX agent

arXiv:2608.09802v1 Announce Type: new Abstract: As AI coding agents take on increasingly complex, long-horizon software engineering tasks, existing benchmarks are rapidly saturating and their evaluati

WebChoreArena: Evaluating Web Browsing Agents on Realistic Tedious Web Tasks

Model ReleasesDGX agent

arXiv:2506.01952v2 Announce Type: replace-cross Abstract: Powered by large language models (LLMs), web browsing agents operate graphical user interfaces in a human-like manner, offering a transparent

When LLM Agents Negotiate: Private Information and Dynamic Bargaining in Supply Chains

Model ReleasesDGX agent

arXiv:2608.07538v1 Announce Type: new Abstract: As LLM agents move from decision support to autonomous procurement, firms need to know whether delegated negotiators create value, divide it predictably

You Don't Need To Stay in The Loop: An Agentic Robotics Loop for Robot-Policy Improvement

Model ReleasesDGX agent

arXiv:2608.07555v1 Announce Type: new Abstract: Coding agents such as Claude Code and Codex close the software loop: a main agent manages the loop, subagents analyze and execute, tools do the work. We

10 Aug 2026

HLSmith: An Expert-Guided Agentic Framework for C/C++-to-HLS Translation

Model ReleasesDGX agent

arXiv:2608.06791v1 Announce Type: cross Abstract: Application-specific FPGA accelerators offer substantial performance and energy-efficiency gains across many application domains, but developing them

Homebot: A Personal AI Agent for Conversational Home Assistance and Automation

Local AiDGX agent

arXiv:2608.02254v2 Announce Type: replace Abstract: exttt{Homebot} is a locally deployable AI agent for conversational household assistance and automation. It accepts voice and instant-messaging reque

How Much, Then Where: Credit-Conserving Action-to-Token Allocation for Multi-Turn Agent Reinforcement Learning

SafetyDGX agent

arXiv:2608.07118v1 Announce Type: new Abstract: Credit assignment in multi-turn agent reinforcement learning operates at two levels: assigning trajectory-level credit to actions and distributing each

MemOPD: On-Policy Distillation through Memory State Alignment for Long-Horizon Agents

SafetyDGX agent

arXiv:2608.07068v1 Announce Type: new Abstract: Long-horizon agents accumulate growing contexts during interaction, impairing performance and stability. Compact memory mitigates this problem by compre

Run Local Agentic AI Workflows with Meta’s Muse Glimmer on NVIDIA

Model ReleasesDGX agent

Meta released Muse Glimmer, a 30‑billion‑parameter dense language model with a context window exceeding 120 K tokens, designed for local, long‑running agentic AI workloads. The model is optimized to r

9 Aug 2026

I Turned My Underused Gaming Laptop Into a Local AI Workstation

Local AiDGX agent

TL;DR: I am building a Windows-first local AI setup for people who want to try local LLMs without spending days choosing models, setting up Ollama, Docker, WSL, Open WebUI, agents, and tool permission

8 Aug 2026

I tested a fresh GitHub download → Ollama → first local coding-agent task (72 seconds, no cloud API)

Model ReleasesDGX agent

I’m building DesktopLab, an open-source local-first control plane for development agents. I recorded the setup boundary that most agent demos skip: DesktopLab detects the host, proposes the supported

7 Aug 2026

APQF: Agentic Profiling-Guided Structured Pruning and Mixed-Precision Quantization with Adaptive Fine-Tuning

AgentsDGX agent

arXiv:2608.05499v1 Announce Type: cross Abstract: Modern deep neural networks achieve strong performance, but their scale makes them costly and slow, especially on resource-constrained edge devices. P

DoctorAgents: an agentic framework to iteratively refine AutoML pipeline for small clinical temporal data

AgentsDGX agent

arXiv:2608.05375v1 Announce Type: new Abstract: Clinical machine learning (ML) has the potential to support high-stakes medical decision-making, but reliable deployment is often constrained by scarce,

EnvACE: Internalizing Environment Dynamics via World Rehearsal for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2608.06197v1 Announce Type: new Abstract: Training large language model agents for long-horizon tool use typically relies on interactions with real or synthesized executable environments, whose

FinEvo-Bench: A Longitudinal Benchmark for Self-Evolving Agents in Professional Financial Workflows

Model ReleasesDGX agent

arXiv:2608.06144v1 Announce Type: new Abstract: Most agent benchmarks evaluate tasks independently and cannot measure whether experience from one task helps with later tasks. Existing self-evolution b

Learning Context-Free Grammars for Grammar-Constrained Decoding via Declarative Agentic Programming with Guarantees

AgentsDGX agent

arXiv:2608.05493v1 Announce Type: cross Abstract: Language models (LMs) are increasingly used to interact with external services via programs written in domain-specific languages (DSLs). Unfortunately

6 Aug 2026

ArtAnno: Annotating Implicit Semantics in Artworks through LLM Agent-Driven Bidirectional Human-AI Augmentation

AgentsDGX agent

arXiv:2608.05026v1 Announce Type: cross Abstract: High-quality annotation of artworks is essential for computational art research, yet extracting implicit semantics remains challenging due to the reli

Calibrating Artificial Guilt: Neurally Grounded Reward Shaping for Prosocial Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2608.04663v1 Announce Type: new Abstract: Cooperative multi-agent reinforcement learning often adds social terms to individual rewards, yet the scale of those terms is usually chosen by hand. We

Diagnosing Tool-Selection Reasoning in LLM Agents with Canary Tools

Model ReleasesDGX agent

arXiv:2608.04719v1 Announce Type: new Abstract: Agent evaluations tell us that a model picked the wrong tool, but rarely why. We introduce canary tools: diagnostic probe tools planted in an agent's Mo

HALT: Verification-Aware Stopping for Retrieval-Augmented Search Agents

SafetyDGX agent

arXiv:2608.02009v2 Announce Type: replace Abstract: Retrieval-augmented search agents answer multi-hop questions by repeatedly issuing search queries and accumulating evidence. This creates a stopping

SafeCommit: Certifying When Memory-Grounded Agents May Safely Act

SafetyDGX agent

arXiv:2608.04289v1 Announce Type: new Abstract: Long-horizon agents increasingly use persistent memory and tools to take actions with external side effects. A central failure mode is premature commitm

5 Aug 2026

AgenticVAU: Multi-Agent Explore-Verify Reasoning for Video Anomaly Understanding

Local AiDGX agent

arXiv:2608.03779v1 Announce Type: new Abstract: Video anomaly understanding (VAU) focuses on comprehensively interpreting abnormal events in videos, requiring models to identify anomalous occurrences,

Asking Questions the Right Way: A Multi-Agent Conversational System for Prompt Formulation in Complex Task Resolution

AgentsDGX agent

arXiv:2608.01366v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are integral to complex intellectual tasks, yet output quality remains constrained by user-provided prompts. Iter

Don't Regenerate, Debug: A Domain-Specific Agent for Repairing Near-Miss Hardware Operators

AgentsDGX agent

arXiv:2608.02712v1 Announce Type: cross Abstract: Kernel generation for hardware accelerators such as GPUs and NPUs has become a proving ground for large language models (LLMs), and state-of-the-art s

EduClaw-Bench: A Long-Horizon Benchmark for Pedagogical LLM Agents with Simulated Learners

Model ReleasesDGX agent

arXiv:2608.03206v1 Announce Type: cross Abstract: Large language models (LLMs) power educational applications from tutoring to essay scoring, but each is a point solution to a single task, and only re

Vectra AI launches Vectra AI Pro to feed AI agents better attack signals

Model ReleasesDGX agent

Threat detection company Vectra AI Inc. today launched Vectra AI Pro, a product built to give the artificial intelligence agents now working inside security operations centers a more reliable read on

When Agents Learn to Be You: Benchmarking Privacy Leakage, Impersonation Risk, and Defenses in Persona Skills

Model ReleasesDGX agent

arXiv:2608.03700v1 Announce Type: cross Abstract: Persona skills distill personal interaction histories into portable and executable artifacts for downstream agents. While enabling flexible personaliz

4 Aug 2026

Agentic Graph Token Reasoning

AgentsDGX agent

arXiv:2608.00542v1 Announce Type: new Abstract: Graphs model relational data throughout science and industry, from citation networks to product co-purchase graphs. Because the nodes of many such graph

Announcing Cloudflare Wallets: the programmable wallet for the agentic Internet

SafetyDGX agent

Cloudflare Wallets will provide AI agents with native payments and verifiable identity on the web. Using the x402 protocol, agents can autonomously purchase APIs and content within clear safety guardr

AWS launches Kiro Crew, an autonomous agentic orchestrator for 24/7 code development

Model ReleasesDGX agent

Amazon Web Services Inc. today launched Kiro Crew, an autonomous workspace that keeps artificial intelligence coding agents running all day and all night. The company said it developed Crew to allow d

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major…

Model ReleasesDGX agent

DeepSeek V4 Flash is now live on Together AI. Frontier agent performance is getting dramatically cheaper. DSV4 on Together AI brings a major jump in coding, tool use, and long-running agent performanc

FactorJEPA: Factorizing Monolithic Futures into Layout-Agent-Interaction Channels for Crowded and Chaotic Global South Urban Worlds

AgentsDGX agent

arXiv:2608.01049v1 Announce Type: cross Abstract: World models have attracted significant attention for their ability to capture and predict the structure and dynamics of the physical world. In this e

Long-Horizon Embodied Decision-Making via Multimodal Memory Compression

Model ReleasesDGX agent

arXiv:2608.01456v1 Announce Type: cross Abstract: Agents are increasingly expected to act not only as task executors, but also as decision-makers on behalf of human users. This shift requires agents t

MemSIF: From Structured Interactions to Dual-Track Fact Memory for LLM Agents

ResearchDGX agent

arXiv:2608.01742v1 Announce Type: cross Abstract: Long-term memory is critical for LLM agents operating over long-horizon interactions. However, several persistent limitations of existing memory syste

3 Aug 2026

Distilling Knowledge from Large Language Models into Lightweight Reinforcement Learning Agents for Autonomous Cyber Operations

Model ReleasesDGX agent

arXiv:2607.28826v1 Announce Type: new Abstract: Autonomous Cyber Operations (ACO) are increasingly important for defending enterprise networks as cyber threats continue to evolve in sophistication. AC

← Previous
1…9091929394…300
Next →