AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Agents

The Emerging Paradigm of Geospatial Foundation Models: From Pre-Training to Agentic Reasoning

DGX agent

arXiv:2607.12177v1 Announce Type: new Abstract: The analysis of satellite and aerial imagery has entered a new era with the advent of foundation models. This paper describes the concept of Geospatial

agentsarxiv-cs-ai
15 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Who Grades the Grader? Co-Evolving Evaluation Metrics and Skills for Self-Improving LLM Agents

DGX agent

arXiv:2607.12790v1 Announce Type: new Abstract: Self-evolving agent systems improve by creating, revising, and retiring their own skills, but every such loop rests on a hidden assumption: a reliable e

safetyarxiv-cs-ai
15 Jul 2026
Model Releases

Context Graphs for Proactive Enterprise Agents

DGX agent

arXiv:2607.07721v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) and agentic frameworks have advanced enterprise AI considerably, yet agents remain fundamentally reactive: they wai

model-releasesarxiv-cs-ai
10 Jul 2026
Model Releases

DeepSWE: Measuring Frontier Coding Agents on Original, Long-Horizon Engineering Tasks

DGX agent

arXiv:2607.07946v1 Announce Type: cross Abstract: DeepSWE is a benchmark of 113 original, long-horizon software engineering tasks for evaluating coding agents. Most public agentic coding benchmarks fo

model-releasesarxiv-cs-lg
10 Jul 2026
Agents

MASTE: A Multi-Agent Pipeline for Zero-Shot Aspect Sentiment Triplet Extraction

DGX agent

arXiv:2607.08080v1 Announce Type: new Abstract: Aspect Sentiment Triplet Extraction (ASTE) requires jointly identifying (aspect, opinion, sentiment) triples from a given review sentence. While large l

agentsarxiv-cs-cl
10 Jul 2026
Safety

Open-ended Multi-agent Autocurricula via Visual Inspection of Policies with Multi-modal LLMs

DGX agent

arXiv:2607.08193v1 Announce Type: cross Abstract: Open-ended curricula in Reinforcement Learning (RL) aim to train generally-capable agents by identifying tasks that facilitate learning increasingly c

safetyarxiv-cs-ai
10 Jul 2026
Local Ai

Token-Flow Firewall: Semantic Runtime Auditing for Persistent AI Agents

DGX agent

arXiv:2607.08395v1 Announce Type: cross Abstract: Persistent AI agents extend large language models (LLMs) beyond single-turn interaction into long-lived software systems. Unlike traditional chat assi

local-aiarxiv-cs-cl
10 Jul 2026
Model Releases

Beyond Attack-Success Rate: Action-Graded Severity Scale for Tool-Using AI Agents

DGX agent

arXiv:2607.07474v1 Announce Type: cross Abstract: Agentic red-teaming benchmarks report whether an injected agent was compromised as a single bit: the attack succeeded, or it did not. We argue that th

model-releasesarxiv-cs-ai
9 Jul 2026
Safety

Entropy Pacing Policy Optimization for Multi-Task Agentic Reinforcement Learning

DGX agent

arXiv:2607.07178v1 Announce Type: cross Abstract: Recent breakthroughs of Reinforcement Learning (RL) have highlighted its potential for complex agentic Large Language Model (LLM) tasks. However, exis

safetyarxiv-cs-ai
9 Jul 2026
Safety

Progressive Crystallization: Turning Agent Exploration into Deterministic, Lower-Cost Workflows in Production

DGX agent

arXiv:2607.07052v1 Announce Type: cross Abstract: AI agents deployed for IT operations are typically permanent cost centers because every execution requires full LLM inference, even for previously sol

safetyarxiv-cs-ai
9 Jul 2026
Agents

Security and Privacy in Agentic AI: Grand Challenges and Future Directions

DGX agent

arXiv:2607.06608v1 Announce Type: cross Abstract: We present key challenges and future research directions in the security and privacy of agentic AI, based on a horizon-scanning exercise that brought

agentsarxiv-cs-ai
9 Jul 2026
Agents

SpaCellAgent: A Self-Evolving LLM-Based Multi-Agent Framework for Trajectory Analysis

DGX agent

arXiv:2607.07467v1 Announce Type: new Abstract: Spatial and Single-cell transcriptomics are transformative in deciphering cellular dynamics. As the fundamental paradigm for reconstructing cell develop

agentsarxiv-cs-ai
9 Jul 2026
Model Releases

aiAuthZ: Off-Host, Identity-Bound Authorization for AI Agents

DGX agent

arXiv:2607.05518v1 Announce Type: cross Abstract: AI agents issue tool calls on the basis of text they cannot verify, so any party who controls part of the context can forge the appearance of authorit

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

Information Gain-based Rollout Policy Optimization: An Adaptive Tree-Structured Rollout Approach for Multi-Turn LLM Agents

DGX agent

arXiv:2607.06223v1 Announce Type: new Abstract: Reinforcement learning has become a promising paradigm for improving large language model (LLM) agents on long-horizon search tasks, where the agent mus

safetyarxiv-cs-ai
8 Jul 2026
Safety

MASCA: LLM based-Multi Agents System for Credit Assessment

DGX agent

arXiv:2507.22758v2 Announce Type: replace Abstract: Recent advancements in financial problem-solving have leveraged LLMs and agent-based systems, with a primary focus on trading and financial modeling

safetyarxiv-cs-cl
8 Jul 2026
Agents

MCP-Enabled Agentic AI for Autonomous IPoDWDM Network Lifecycle Automation

DGX agent

arXiv:2607.05975v1 Announce Type: cross Abstract: This demo presents an MCP-enabled agentic AI architecture for autonomous control of vendor-agnostic IPoDWDM networks. We demonstrate live end-to-end l

agentsarxiv-cs-ai
8 Jul 2026
Agents

Reward as An Agent for Embodied World Models

DGX agent

arXiv:2606.19990v2 Announce Type: replace Abstract: While RL has become a promising tool for refining world models, existing methods largely rely on conservative rollouts near the training distributio

agentsarxiv-cs-ai
8 Jul 2026
Agents

Task Decomposition-Guided Reranking for Adaptive Agent Skill Retrieval

DGX agent

arXiv:2607.06283v1 Announce Type: new Abstract: Skill usage can significantly enhance the ability of modern agent systems to complete complex tasks. However, the growing scale of skill libraries makes

agentsarxiv-cs-ai
8 Jul 2026
Model Releases

The Balkanization of Execution-Security Research for AI Coding Agents: Isolation, Access Control, and Time-of-Check-to-Time-of-Use Vulnerabilities

DGX agent

arXiv:2607.05743v1 Announce Type: cross Abstract: AI coding agents now read repositories, call tools, and execute shell commands with limited human oversight, and a fast-growing body of work studies w

model-releasesarxiv-cs-ai
8 Jul 2026
Safety

When Lower Privileges Suffice: Investigating Over-Privileged Tool Selection in LLM Agents

DGX agent

arXiv:2606.20023v2 Announce Type: replace-cross Abstract: As LLM agents increasingly select tools autonomously, their choices among tools with different privileges become safety-relevant. However, pri

safetyarxiv-cs-ai
8 Jul 2026
Agents

Agentic and Generative AI for Open-Source Intelligence and Cyber Investigations: Taxonomy, Evaluation, Challenges, and Future Directions

DGX agent

arXiv:2607.03233v1 Announce Type: cross Abstract: The rapid growth of publicly available digital information has rendered manual open-source intelligence (OSINT) analysis insufficient for modern intel

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

EvoAgentBench: Benchmarking Agent Self-Evolution via Ability Transfer

DGX agent

arXiv:2607.05202v1 Announce Type: new Abstract: Agent self-evolution in long-horizon LLM systems is largely procedural: useful experience is not merely stored information, but reusable procedures for

model-releasesarxiv-cs-ai
7 Jul 2026
Local Ai

Hierarchical Multi-Agent Reinforcement Learning for Carbon-Aware AI Data Centers in Power Distribution Systems

DGX agent

arXiv:2607.03324v1 Announce Type: cross Abstract: Eco-friendly energy management for artificial intelligence data centers (AIDCs) is crucial because of the significant increase in energy consumption-i

local-aiarxiv-cs-ai
7 Jul 2026
Local Ai

Memory-Orchestrated Semantic System (MOSS): An Auditable Agentic Memory Architecture

DGX agent

arXiv:2607.04391v1 Announce Type: new Abstract: Long-term memory remains a structural weakness of AI agents. The dominant approach, retrieval-augmented generation (RAG), relies on embedding-based simi

local-aiarxiv-cs-cl
7 Jul 2026
Applications

P^3: Toward Versatile Embodied Agents

DGX agent

arXiv:2508.07033v2 Announce Type: replace Abstract: Embodied agents have shown promising generalization capabilities across diverse physical environments, making them essential for a wide range of rea

applicationsarxiv-cs-ro
7 Jul 2026
Model Releases

SovereignPA-Bench: Evaluating User-Owned Personal Agents under Evolving Intent, Platform Mediation, and Consent Constraints

DGX agent

arXiv:2607.05363v1 Announce Type: new Abstract: Personal agents are becoming persistent user-owned intermediaries: they remember preferences, filter platform-mediated information, use tools, and negot

model-releasesarxiv-cs-ai
7 Jul 2026
Safety

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training

DGX agent

arXiv:2607.04963v1 Announce Type: new Abstract: Reinforcement Learning (RL) is the dominant paradigm for training Large Language Model (LLM) agents on long-horizon tasks. However, sparse and delayed r

safetyarxiv-cs-ai
7 Jul 2026
Agents

TACTIC-KG: Toward Small Agent Teams for Cyber Threat Intelligence Knowledge Graph Construction

DGX agent

arXiv:2607.05001v1 Announce Type: cross Abstract: Cyber Threat Intelligence (CTI) reports are predominantly unstructured, heterogeneous, and noisy, which limits their direct usability for automated an

agentsarxiv-cs-ai
7 Jul 2026
Model Releases

TRACE: Capability-Targeted Agentic Training

DGX agent

arXiv:2604.05336v2 Announce Type: replace Abstract: Models often fail to complete agentic tasks because they lack core capabilities required by the target environment. However, mainstream approaches f

model-releasesarxiv-cs-ai
7 Jul 2026
Agents

Atomic Task Graph: A Unified Framework for Agentic Planning and Execution

DGX agent

arXiv:2607.01942v1 Announce Type: new Abstract: LLM-based agents have shown strong potential for solving complex multi-step tasks, yet existing performance improvements often rely on either scaling to

agentsarxiv-cs-ai
3 Jul 2026
Model Releases

BuilderBench: The Building Blocks of Intelligent Agents

DGX agent

arXiv:2510.06288v4 Announce Type: replace Abstract: Today's AI models learn primarily through mimicry and refining, so it is not surprising that they struggle to solve problems beyond the limits set b

model-releasesarxiv-cs-ai
3 Jul 2026
Model Releases

Controllable Sim Agents with Behavior Latents

DGX agent

arXiv:2607.02496v1 Announce Type: cross Abstract: Realistic traffic simulation requires agents that imitate logged behavior and can also be steered along interpretable axes. Such controllability enabl

model-releasesarxiv-cs-lg
3 Jul 2026
Safety

Criticality-Based Guard Rail Validation for AI Agent Decisions in Autonomous Telecom Networks

DGX agent

arXiv:2607.02210v1 Announce Type: new Abstract: The evolution toward fully autonomous telecommunications networks (Autonomous Network Levels 4-5) requires AI/ML agents to make real-time network decisi

safetyarxiv-cs-ai
3 Jul 2026
Safety

ElephantAgent: Contextual State Continuity in Agentic Systems

DGX agent

arXiv:2607.01919v1 Announce Type: new Abstract: Agentic systems enhance their capabilities by invoking external tools and maintaining persistent memory. However, these external dependencies introduce

safetyarxiv-cs-ai
3 Jul 2026
Local Ai

Who Gets the Reward & Who Gets the Blame? Evaluation-Aligned Training Signals for Multi-LLM Agents

DGX agent

arXiv:2511.10687v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) in multi-agent systems (MAS) have shown promise for complex tasks, yet current training methods lack principled w

local-aiarxiv-cs-ai
3 Jul 2026
Agents

From Personas to Plot: Character-Grounded Multi-Agent Story Generation for Long-Form Narratives

DGX agent

arXiv:2607.00918v1 Announce Type: cross Abstract: Although large language models (LLMs) have demonstrated impressive creative fiction generation, they struggle to maintain narrative consistency and co

agentsarxiv-cs-ai
2 Jul 2026
Model Releases

PHREEQC-MCQ-200: A Diagnostic Benchmark for Tool-Augmented Scientific Simulator Agents

DGX agent

arXiv:2607.00436v1 Announce Type: new Abstract: Large language model agents are increasingly connected to scientific software, yet it remains unclear when tool access makes scientific computation more

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

Skills Are Not Islands: Measuring Dependency and Risk in Agent Skill Supply Chains

DGX agent

arXiv:2607.01136v1 Announce Type: cross Abstract: Agent skills package reusable operational knowledge for Large Language Model (LLM) agents, yet as they grow in scope, they become dependency-bearing a

model-releasesarxiv-cs-ai
2 Jul 2026
Model Releases

FinPersona-Bench: A Benchmark for Longitudinal Psychometric Stability of Autonomous Financial Agents

DGX agent

arXiv:2606.31522v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly deployed as autonomous financial agents initialized with explicit behavioral mandates such as 'preserve

model-releasesarxiv-cs-ai
1 Jul 2026
Research

LUMOS: A Semantic Operating-System Layer for Accessibility-Grounded AI Agents

DGX agent

arXiv:2606.30697v1 Announce Type: cross Abstract: Current operating systems expose interfaces optimized for human users but not for AI agents. Humans benefit from pixels, icons, windows, visual groupi

researcharxiv-cs-ai
1 Jul 2026
Agents

Robust Autonomous UAV Landing on Maritime Platforms via Multimodal Agentic AI and Active Wave Compensation

DGX agent

arXiv:2606.31613v1 Announce Type: new Abstract: Autonomous aerial inspection of marine infrastructure is frequently compromised by stochastic sea states, introducing risks of high-kinetic impacts, pos

agentsarxiv-cs-cv
1 Jul 2026
Model Releases

What Memory Do GUI Agents Really Need? From Passive Records to Active Task-Driving States

DGX agent

arXiv:2606.31612v1 Announce Type: new Abstract: Mobile GUI agents increasingly face long-horizon tasks that require reading, updating, and reusing task-relevant data across pages and applications. Exi

model-releasesarxiv-cs-cv
1 Jul 2026
Agents

AutoB2G: Agentic Simulation and Reinforcement Learning for Spatio-Temporal Grid-Interactive Building Control

DGX agent

arXiv:2603.26005v2 Announce Type: replace Abstract: Grid-interactive building control has emerged as a promising approach for improving demand-side flexibility in modern power systems. Realistic studi

agentsarxiv-cs-ai
30 Jun 2026
Agents

Choose Your Agent: Tradeoffs in Adopting AI Advisors, Coaches, and Delegates in Multi-Party Negotiation

DGX agent

arXiv:2602.12089v3 Announce Type: replace-cross Abstract: As AI usage becomes more prevalent in social contexts, understanding agent-user interaction is critical to designing systems that imp rove bot

agentsarxiv-cs-ai
30 Jun 2026
Model Releases

Collective cooperation without individual fidelity in LLM agents

DGX agent

arXiv:2606.30454v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as agents in simulations of social systems, yet it remains unclear when their behavior can be inter

model-releasesarxiv-cs-ai
30 Jun 2026
Safety

Complementary RL: Towards Efficient Experience-Driven Agent Learning

DGX agent

arXiv:2603.17621v2 Announce Type: replace-cross Abstract: Reinforcement Learning (RL) has emerged as a powerful paradigm for training LLM-based agents, yet remains limited by low sample efficiency, st

safetyarxiv-cs-cl
30 Jun 2026
Model Releases

Echoes of Human Malice in Agents: Benchmarking LLMs for Multi-Turn Online Harassment Attacks

DGX agent

arXiv:2510.14207v3 Announce Type: replace Abstract: Large Language Model (LLM) agents are powering a growing share of interactive web applications, yet remain vulnerable to misuse and harm. Prior jail

model-releasesarxiv-cs-ai
30 Jun 2026
Model Releases

Governance Decay: How Context Compaction Silently Erases Safety Constraints in Long-Horizon LLM Agents

DGX agent

arXiv:2606.22528v2 Announce Type: replace Abstract: Modern LLM agents increasingly rely on context compaction, summarization, or eviction to keep long-running sessions within a token budget. We show t

model-releasesarxiv-cs-ai
30 Jun 2026
← Previous
1…5556575859…233
Next →