AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Safety

Policy Gradient Methods for Non-Markovian Reinforcement Learning

DGX agent

arXiv:2605.10816v1 Announce Type: cross Abstract: We study policy gradient methods for reinforcement learning in non-Markovian decision processes (NMDPs), where observations and rewards depend on the

safetyarxiv-cs-ai
12 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Preventing Prompt Injection with Type-Directed Privilege Separation

DGX agent

arXiv:2509.25926v2 Announce Type: replace-cross Abstract: Modern language models have enabled the development of agentic systems that achieve strong performance on reasoning-intensive tasks. Unfortuna

agentsarxiv-cs-lg
12 May 2026
Agents

SAR-RAG: ATR Visual Question Answering by Semantic Search, Retrieval, and MLLM Generation

DGX agent

arXiv:2602.04712v2 Announce Type: replace-cross Abstract: We present a visual-context image-retrieval-augmented generation (ImageRAG)- assisted AI agent for automatic target recognition (ATR) of synth

agentsarxiv-cs-ai
12 May 2026
Agents

SecureForge: Finding and Preventing Vulnerabilities in LLM-Generated Code via Prompt Optimization

DGX agent

arXiv:2605.08382v1 Announce Type: cross Abstract: LLM coding agents now generate code at an unprecedented scale, yet LLM-generated code introduces cybersecurity vulnerabilities into codebases without

agentsarxiv-cs-cl
12 May 2026
Agents

AGWM: Affordance-Grounded World Models for Environments with Compositional Prerequisites

DGX agent

arXiv:2605.06841v1 Announce Type: new Abstract: In model-based learning, the agent learns behaviors by simulating trajectories based on world model predictions. Standard world models typically learn a

agentsarxiv-cs-ai
11 May 2026
Model Releases

EnvSimBench: A Benchmark for Evaluating and Improving LLM-Based Environment Simulation

DGX agent

arXiv:2605.07247v1 Announce Type: new Abstract: Scalable AI agents training relies on interactive environments that faithfully simulate the consequences of agent actions. Manually crafted environments

model-releasesarxiv-cs-ai
11 May 2026
Safety

From Assistance to Agency: Rethinking Autonomy and Control in CI/CD Pipelines

DGX agent

arXiv:2605.07062v1 Announce Type: cross Abstract: AI agents are assuming active roles in Continuous Integration and Continuous Deployment (CI/CD) workflows, yet the research community lacks a shared v

safetyarxiv-cs-ai
11 May 2026
Agents

GameGen-Verifier: Parallel Keypoint-Based Verification for LLM-Generated Games via Runtime State Injection

DGX agent

arXiv:2605.07442v1 Announce Type: new Abstract: LLM-based game generation promises to turn natural-language specifications into executable games, but progress is limited by the lack of reliable automa

agentsarxiv-cs-lg
11 May 2026
Agents

CyberAId: AI-Driven Cybersecurity for Financial Service Providers

DGX agent

arXiv:2605.01892v1 Announce Type: new Abstract: European financial institutions face mounting regulatory pressure while their security operations centres remain constrained not by data or staffing but

agentsarxiv-cs-ai
6 May 2026
Safety

Safety in Embodied AI: A Survey of Risks, Attacks, and Defenses

DGX agent

arXiv:2605.02900v1 Announce Type: cross Abstract: Embodied Artificial Intelligence (Embodied AI) integrates perception, cognition, planning, and interaction into agents that operate in open-world, saf

safetyarxiv-cs-cv
6 May 2026
Agents

Taming the Curses of Multiagency in Robust Markov Games with Large State Space through Linear Function Approximation

DGX agent

arXiv:2605.03125v1 Announce Type: new Abstract: Multi-agent reinforcement learning (MARL) holds great potential but faces robustness challenges due to environmental uncertainty. To address this, distr

agentsarxiv-cs-lg
6 May 2026
Agents

End-to-end autonomous scientific discovery on a real optical platform

DGX agent

arXiv:2604.27092v1 Announce Type: new Abstract: Scientific research has long been human-led, driving new knowledge and transformative technologies through the continual revision of questions, methods

agentsarxiv-cs-ai
1 May 2026
Agents

EmoTransCap: Dataset and Pipeline for Emotion Transition-Aware Speech Captioning in Discourses

DGX agent

arXiv:2604.26417v1 Announce Type: new Abstract: Emotion perception and adaptive expression are fundamental capabilities in human-agent interaction. While recent advances in speech emotion captioning (

agentsarxiv-cs-cl
30 Apr 2026
Agents

SODA-CitrON: Static Object Data Association by Clustering Multi-Modal Sensor Detections Online

DGX agent

arXiv:2602.22243v2 Announce Type: replace Abstract: The online fusion and tracking of static objects from heterogeneous sensor detections is a fundamental problem in robotics, autonomous systems, and

agentsarxiv-cs-ro
29 Apr 2026
Agents

Accelerating Reinforcement Learning for Wind Farm Control via Expert Demonstrations

DGX agent

arXiv:2604.22794v1 Announce Type: cross Abstract: Reinforcement learning (RL) offers a promising approach for adaptive wind farm flow control, yet its practical deployment is hindered by slow training

agentsarxiv-cs-lg
28 Apr 2026
Model Releases

Benchmarking Emergent Coordination in Large-Scale LLM Populations: An Evaluation Framework on the MoltBook Archive

DGX agent

arXiv:2603.03555v2 Announce Type: replace-cross Abstract: As multi-agent Large Language Model (LLM) systems scale, evaluating their emergent coordination dynamics becomes increasingly critical. Howeve

model-releasesarxiv-cs-ai
28 Apr 2026
Agents

Reconstructive Authority Model: Runtime Execution Validity Under Partial Observability

DGX agent

arXiv:2604.22898v1 Announce Type: cross Abstract: Autonomous systems increasingly operate under partial observability where execution-relevant state is never fully accessible. Existing governance mech

agentsarxiv-cs-ai
28 Apr 2026
Agents

Think Anywhere in Code Generation

DGX agent

arXiv:2603.29957v3 Announce Type: replace-cross Abstract: Recent advances in reasoning Large Language Models (LLMs) have primarily relied on upfront thinking, where reasoning occurs before final answe

agentsarxiv-cs-lg
28 Apr 2026
Safety

ZenBrain: A Neuroscience-Inspired 7-Layer Memory Architecture for Autonomous AI Systems

DGX agent

arXiv:2604.23878v1 Announce Type: new Abstract: Despite a century of empirical memory research, existing AI agent memory systems rely on system-engineering metaphors (virtual-memory paging, flat LLM s

safetyarxiv-cs-ai
28 Apr 2026
Agents

QDTraj: Exploration of Diverse Trajectory Primitives for Articulated Objects Robotic Manipulation

DGX agent

arXiv:2604.22551v1 Announce Type: cross Abstract: Thanks to the latest advances in learning and robotics, domestic robots are beginning to enter homes, aiming to execute household chores autonomously.

agentsarxiv-cs-ai
27 Apr 2026
Agents

Source-Modality Monitoring in Vision-Language Models

DGX agent

arXiv:2604.22038v1 Announce Type: new Abstract: We define and investigate source-modality monitoring -- the ability of multimodal models to track and communicate the input source from which pieces of

agentsarxiv-cs-cl
27 Apr 2026
Agents

OpInf-LLM: Parametric PDE Solving with LLMs via Operator Inference

DGX agent

arXiv:2602.01493v2 Announce Type: replace-cross Abstract: Solving diverse partial differential equations (PDEs) is fundamental in science and engineering. Large language models (LLMs) have demonstrate

agentsarxiv-cs-ai
24 Apr 2026
Agents

Decision-Focused Federated Learning Under Heterogeneous Objectives and Constraints

DGX agent

arXiv:2604.20031v1 Announce Type: cross Abstract: We consider what we refer to as {Decision-Focused Federated Learning (DFFL)} framework, i.e., a predict-then-optimize approach employed by a collectio

agentsarxiv-cs-lg
23 Apr 2026
Agents

DeVI: Physics-based Dexterous Human-Object Interaction via Synthetic Video Imitation

DGX agent

arXiv:2604.20841v1 Announce Type: new Abstract: Recent advances in video generative models enable the synthesis of realistic human-object interaction videos across a wide range of scenarios and object

agentsarxiv-cs-cv
23 Apr 2026
Research

On Accelerating Grounded Code Development for Research

DGX agent

arXiv:2604.19022v1 Announce Type: new Abstract: A major challenge for niche scientific and technical domains in leveraging coding agents is the lack of access to up-to-date, domain- specific knowledge

researcharxiv-cs-ai
22 Apr 2026
Safety

The PROPER Approach to Proactivity: Benchmarking and Advancing Knowledge Gap Navigation

DGX agent

arXiv:2601.09926v3 Announce Type: replace Abstract: Current approaches to proactive assistance move beyond the ask-and-respond paradigm by anticipating user needs. In practice, they either burden user

safetyarxiv-cs-lg
22 Apr 2026
Safety

VimRAG: Navigating Massive Visual Context in Retrieval-Augmented Generation via Multimodal Memory Graph

DGX agent

arXiv:2602.12735v2 Announce Type: replace-cross Abstract: Effectively retrieving, reasoning, and understanding multimodal information remains a critical challenge for agentic systems. Traditional Retr

safetyarxiv-cs-cl
22 Apr 2026
Safety

Negative Advantage Is a Double-Edged Sword: Calibrating Advantage in GRPO for Deep Search

DGX agent

arXiv:2604.18235v1 Announce Type: new Abstract: Deep search agents can autonomously initiate multi-turn interactions with search engines, thereby exhibiting strong question-answering capabilities. Suc

safetyarxiv-cs-cl
21 Apr 2026
Agents

Hijacking online reviews: sparse manipulation and behavioral buffering in popularity-biased rating systems

DGX agent

arXiv:2604.13049v1 Announce Type: cross Abstract: Online reviews and recommendation systems help users navigate overwhelming choice, but they are vulnerable to self-reinforcing distortions. This paper

agentsarxiv-cs-ai
17 Apr 2026
Safety

UniDoc-RL: Coarse-to-Fine Visual RAG with Hierarchical Actions and Dense Rewards

DGX agent

arXiv:2604.14967v1 Announce Type: new Abstract: Retrieval-Augmented Generation (RAG) extends Large Vision-Language Models (LVLMs) with external visual knowledge. However, existing visual RAG systems t

safetyarxiv-cs-cv
17 Apr 2026
Agents

Optimized Human-Robot Co-Dispatch Planning for Petro-Site Surveillance under Varying Criticalities

DGX agent

arXiv:2602.07924v2 Announce Type: replace Abstract: Securing petroleum infrastructure requires balancing autonomous system efficiency with human judgment for threat escalation, a challenge unaddressed

agentsarxiv-cs-ro
16 Apr 2026
Safety

GUIDE: Guided Updates for In-context Decision Evolution in LLM-Driven Spacecraft Operations

DGX agent

arXiv:2603.27306v2 Announce Type: replace-cross Abstract: Large language models (LLMs) have been proposed as supervisory agents for spacecraft operations, but existing approaches rely on static prompt

safetyarxiv-cs-ai
15 Apr 2026
Agents

Evolving Many Worlds: Towards Open-Ended Discovery in Petri Dish NCA via Population-Based Training

DGX agent

arXiv:2604.11248v1 Announce Type: cross Abstract: The generation of sustained, open-ended complexity from local interactions remains a fundamental challenge in artificial life. Differentiable multi-ag

agentsarxiv-cs-ai
14 Apr 2026
Agents

Micro-Dexterity in Biological Micromanipulation: Embodiment, Perception, and Control

DGX agent

arXiv:2604.11640v1 Announce Type: new Abstract: Microscale manipulation has advanced substantially in controlled locomotion and targeted transport, yet many biomedical applications require precise and

agentsarxiv-cs-ro
14 Apr 2026
Agents

NimbusGuard: A Novel Framework for Proactive Kubernetes Autoscaling Using Deep Q-Networks

DGX agent

arXiv:2604.11017v1 Announce Type: cross Abstract: Cloud native architecture is about building and running scalable microservice applications to take full advantage of the cloud environments. Managed K

agentsarxiv-cs-ai
14 Apr 2026
Agents

Fast-dVLM: Efficient Block-Diffusion VLM via Direct Conversion from Autoregressive VLM

DGX agent

arXiv:2604.06832v2 Announce Type: replace Abstract: Vision-language models (VLMs) predominantly rely on autoregressive decoding, which generates tokens one at a time and fundamentally limits inference

agentsarxiv-cs-cl
13 Apr 2026
Model Releases

Towards Knowledgeable Deep Research: Framework and Benchmark

DGX agent

arXiv:2604.07720v2 Announce Type: replace Abstract: Deep Research (DR) requires LLM agents to autonomously perform multi-step information seeking, processing, and reasoning to generate comprehensive r

model-releasesarxiv-cs-ai
13 Apr 2026
Local Ai

Heterogeneity-Aware Belief Synchronization for Semantic Communication in AI-Native 6G Networks

DGX agent

arXiv:2608.13394v1 Announce Type: cross Abstract: 6G networks will not be serving as communication infrastructures only; rather, they are expected to evolve into intelligent systems, where thousands o

local-aiarxiv-cs-ai
14 Aug 2026
Agents

MLLM-Routed Heterogeneous Ensembles for Robust Cross-Dataset Image Classification

DGX agent

arXiv:2608.13463v1 Announce Type: cross Abstract: Modern image classification models excel when trained on single task-specific datasets but often struggle to generalize across domains and difficulty

agentsarxiv-cs-ai
14 Aug 2026
Model Releases

TsuGO: Probing Search Efficiency in LLM Reasoning via Go Life-and-Death Problems

DGX agent

arXiv:2608.13221v1 Announce Type: new Abstract: The evaluation of LLM reasoning is moving from final-answer accuracy to process-level assessment, yet existing methods still fail to capture how models

model-releasesarxiv-cs-ai
14 Aug 2026
Agents

GUIDE: Governed Unified Intelligence for Document-to-Artifact Generation in Enterprise Settings

DGX agent

arXiv:2608.12133v1 Announce Type: new Abstract: Enterprise guideline documents are heterogeneous and multimodal, combining narrative text, complex tables, and embedded images. Existing LLM and VLM sys

agentsarxiv-cs-ai
13 Aug 2026
Agents

OpenAg: Democratizing Agricultural Intelligence

DGX agent

arXiv:2506.04571v3 Announce Type: replace Abstract: Agriculture is undergoing a major transformation driven by artificial intelligence (AI), machine learning, and knowledge representation technologies

agentsarxiv-cs-ai
13 Aug 2026
Agents

Preference Tree Optimization: Enhancing Goal-Oriented Dialogue with Look-Ahead Simulations

DGX agent

arXiv:2608.12062v1 Announce Type: cross Abstract: Developing dialogue systems capable of engaging in multi-turn, goal-oriented conversations remains a significant challenge, especially in specialized

agentsarxiv-cs-ai
13 Aug 2026
Agents

REVERE: Reflective Evolving Research Engineer

DGX agent

arXiv:2603.20667v2 Announce Type: replace-cross Abstract: Existing prompt-optimization techniques rely on local signals, causing poor generalization across tasks. In addition, they also rely on weak u

agentsarxiv-cs-ai
13 Aug 2026
Model Releases

Self-Harness: Harnesses That Improve Themselves

DGX agent

arXiv:2606.09498v2 Announce Type: replace Abstract: The performance of LLM-based agents is jointly shaped by their base models and the harnesses that mediate their interaction with the environment. Be

model-releasesarxiv-cs-cl
13 Aug 2026
Agents

Synchronizing Beliefs with Second-Order Theory-of-Mind in Human-Autonomy Teams (Extended Version)

DGX agent

arXiv:2608.11229v1 Announce Type: new Abstract: Comparative feedback, asking people which of two behaviors they prefer, has become a standard way to align robot and agent behavior with human intent wh

agentsarxiv-cs-ai
13 Aug 2026
Agents

Inferential Capability Does Not Determine Legal Scope

DGX agent

arXiv:2608.10601v1 Announce Type: cross Abstract: Two instruments of EU digital law place inference at their centre and mean different things by it. Article 3(1) of the AI Act uses the capability to i

agentsarxiv-cs-ai
12 Aug 2026
Safety

Operationalising Relative Causal Knowledge: Backbone Identifiability from Private Reports on a Shared Outcome

DGX agent

arXiv:2608.10664v1 Announce Type: new Abstract: The Relativity of Causal Knowledge (RCK) explains how a network of agents with different structural causal models can exchange causal knowledge through

safetyarxiv-cs-ai
12 Aug 2026
← Previous
1…126127128129130…236
Next →