AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Optimized Three-Dimensional Photovoltaic Structures with LLM guided Tree Search

DGX agent

arXiv:2605.16191v1 Announce Type: new Abstract: We present a case study for how AI coding systems can be used to generate novel scientific hypotheses. We combine a generic coding agent (Google's AntiG

agentsarxiv-cs-cl
18 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

Distributionally Robust Multi-Task Reinforcement Learning via Adaptive Task Sampling

DGX agent

arXiv:2605.14350v1 Announce Type: new Abstract: Multi-task reinforcement learning (MTRL) aims to train a single agent to efficiently optimize performance across multiple tasks simultaneously. However,

agentsarxiv-cs-lg
15 May 2026
Agents

Cognifold: Always-On Proactive Memory via Cognitive Folding

DGX agent

arXiv:2605.13438v1 Announce Type: new Abstract: Existing agent memory remains predominantly reactive and retrieval-based, lacking the capacity to autonomously organize experience into persistent cogni

agentsarxiv-cs-ai
14 May 2026
Agents

Scaling Retrieval-Augmented Reasoning with Parallel Search and Explicit Merging

DGX agent

arXiv:2605.13534v1 Announce Type: new Abstract: Deep search agents have proven effective in enhancing LLMs by retrieving external knowledge during multi-step reasoning. However, existing methods often

agentsarxiv-cs-ai
14 May 2026
Agents

Encore: Conditioning Trajectory Forecasting via Biased Ego Rehearsals

DGX agent

arXiv:2605.11463v1 Announce Type: new Abstract: Learning and representing the subjectivities of agents has become a challenging but crucial problem in the trajectory prediction task. Such subjectiviti

agentsarxiv-cs-cv
13 May 2026
Agents

Finite and Corruption-Robust Regret Bounds in Online Inverse Linear Optimization under M-Convex Action Sets

DGX agent

arXiv:2602.01682v2 Announce Type: replace Abstract: We study online inverse linear optimization, also known as contextual recommendation, where a learner sequentially infers an agent's hidden objectiv

agentsarxiv-cs-lg
13 May 2026
Agents

Shaping Zero-Shot Coordination via State Blocking

DGX agent

arXiv:2605.11688v1 Announce Type: new Abstract: Zero-shot coordination (ZSC) aims to enable agents to cooperate with independently trained partners without prior interaction, a key requirement for rea

agentsarxiv-cs-lg
13 May 2026
Agents

Controllability in preference-conditioned multi-objective reinforcement learning

DGX agent

arXiv:2605.10585v1 Announce Type: new Abstract: Multi-objective reinforcement learning (MORL) allows a user to express preference over outcomes in terms of the relative importance of the objectives, b

agentsarxiv-cs-lg
12 May 2026
Agents

Robust Remote Reinforcement Learning over Unreliable Communication Channels using Homomorphic State Encoding

DGX agent

arXiv:2508.07722v2 Announce Type: replace Abstract: Traditional Reinforcement Learning (RL) frameworks generally assume that the agent perceives the state of the underlying Markov process instantaneou

agentsarxiv-cs-lg
12 May 2026
Agents

Zero-shot Imitation Learning by Latent Topology Mapping

DGX agent

arXiv:2605.08450v1 Announce Type: cross Abstract: Imitation learning is effective for training agents when expert demonstrations are available, but collecting demonstrations for every complex task in

agentsarxiv-cs-ai
12 May 2026
Agents

A Differentiable Bayesian Relaxation for Latent Partial-Order Inference

DGX agent

arXiv:2605.06976v1 Announce Type: cross Abstract: Many ranking and agent trace datasets are recorded as linear orders even though their latent structure is only partially ordered. This is especially c

agentsarxiv-cs-lg
11 May 2026
Agents

One World, Dual Timeline: Decoupled Spatio-Temporal Gaussian Scene Graph for 4D Cooperative Driving Reconstruction

DGX agent

arXiv:2605.07910v1 Announce Type: new Abstract: Reconstructing dynamic scenes from Vehicle-to-Infrastructure Cooperative Autonomous Driving (VICAD) data is fundamentally complicated by temporal asynch

agentsarxiv-cs-cv
11 May 2026
Safety

Sequential Strategic Classification with Multi-Stage Selective Classifiers

DGX agent

arXiv:2605.04202v1 Announce Type: new Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outcomes

safetyarxiv-cs-lg
7 May 2026
Agents

Memory in the LLM Era: Modular Architectures and Strategies in a Unified Framework

DGX agent

arXiv:2604.01707v2 Announce Type: replace Abstract: Memory emerges as the core module in the large language model (LLM)-based agents for long-horizon complex tasks (e.g., multi-turn dialogue, game pla

agentsarxiv-cs-cl
4 May 2026
Agents

Knowledge Affordances for Hybrid Human-AI Information Seeking

DGX agent

arXiv:2604.27539v1 Announce Type: cross Abstract: As information ecosystems grow more heterogeneous, both humans and artificial agents increasingly face a simple yet unresolved question: when seeking

agentsarxiv-cs-ai
1 May 2026
Agents

R^3-SQL: Ranking Reward and Resampling for Text-to-SQL

DGX agent

arXiv:2604.25325v1 Announce Type: cross Abstract: Modern Text-to-SQL systems generate multiple candidate SQL queries and rank them to judge a final prediction. However, existing methods face two limit

agentsarxiv-cs-cl
29 Apr 2026
Agents

SciDER: Scientific Data-centric End-to-end Researcher

DGX agent

arXiv:2603.01421v2 Announce Type: replace-cross Abstract: Automated scientific discovery with large language models is transforming the research lifecycle from ideation to experimentation, yet existin

agentsarxiv-cs-cl
29 Apr 2026
Safety

Analytica: Soft Propositional Reasoning for Robust and Scalable LLM-Driven Analysis

DGX agent

arXiv:2604.23072v1 Announce Type: new Abstract: Large language model (LLM) agents are increasingly tasked with complex real-world analysis (e.g., in financial forecasting, scientific discovery), yet t

safetyarxiv-cs-ai
28 Apr 2026
Agents

Collaborative Trajectory Prediction via Late Fusion

DGX agent

arXiv:2604.22973v1 Announce Type: new Abstract: Predicting future trajectories of surrounding traffic agents is critical for safe autonomous navigation and collision avoidance. Despite all advances in

agentsarxiv-cs-ro
28 Apr 2026
Agents

Flickering Multi-Armed Bandits

DGX agent

arXiv:2602.17315v2 Announce Type: replace-cross Abstract: We introduce Flickering Multi-Armed Bandits (FMAB) to model sequential decision-making in environments with changing action availability, wher

agentsarxiv-cs-ai
28 Apr 2026
Agents

Information-Theoretic Measures in AI: A Practical Decision Guide

DGX agent

arXiv:2604.23716v1 Announce Type: new Abstract: Information-theoretic (IT) measures are ubiquitous in artificial intelligence: entropy drives decision-tree splits and uncertainty quantification, cross

agentsarxiv-cs-ai
28 Apr 2026
Safety

Quantifying Divergence in Inter-LLM Communication Through API Retrieval and Ranking

DGX agent

arXiv:2604.22760v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly operate as autonomous agents that reason over external APIs to perform complex tasks. However, their reliabi

safetyarxiv-cs-ai
28 Apr 2026
Agents

X-NegoBox: An Explainable Privacy-Budget Negotiation Framework for Secure Peer-to-Peer Energy Data Exchange

DGX agent

arXiv:2604.24326v1 Announce Type: cross Abstract: The decentralization of modern energy systems is transforming consumers into prosumers who continuously exchange data with aggregators, peers, and mar

agentsarxiv-cs-ai
28 Apr 2026
Agents

Do Not Imitate, Reinforce: Iterative Classification via Belief Refinement

DGX agent

arXiv:2604.22110v1 Announce Type: new Abstract: Standard supervised classification trains models to imitate the exact labels provided by a perfect oracle. This imitation happens in a single pass, rest

agentsarxiv-cs-lg
27 Apr 2026
Agents

When Quotes Crumble: Detecting Transient Mechanical Liquidity Erosion in Limit Order Books

DGX agent

arXiv:2604.21993v1 Announce Type: new Abstract: We study the detection of transient liquidity erosion ('crumbling quotes') in electronic limit order books, where observable quote deterioration may ref

agentsarxiv-cs-lg
27 Apr 2026
Agents

A Hierarchical MARL-Based Approach for Coordinated Retail P2P Trading and Wholesale Market Participation of DERs

DGX agent

arXiv:2604.20586v1 Announce Type: new Abstract: The ongoing shift towards decentralization of the electric energy sector, driven by the growing electrification across end-use sectors, and widespread a

agentsarxiv-cs-lg
23 Apr 2026
Agents

Visual Reasoning through Tool-supervised Reinforcement Learning

DGX agent

arXiv:2604.19945v1 Announce Type: new Abstract: In this paper, we investigate the problem of how to effectively master tool-use to solve complex visual reasoning tasks for Multimodal Large Language Mo

agentsarxiv-cs-cv
23 Apr 2026
Agents

MAGICIAN: Efficient Long-Term Planning with Imagined Gaussians for Active Mapping

DGX agent

arXiv:2603.22650v2 Announce Type: replace Abstract: Active mapping aims to determine how an agent should move to efficiently reconstruct unknown environments. Most existing approaches rely on greedy n

agentsarxiv-cs-cv
22 Apr 2026
Agents

GRAIL: Autonomous Concept Grounding for Neuro-Symbolic Reinforcement Learning

DGX agent

arXiv:2604.16871v1 Announce Type: cross Abstract: Neuro-symbolic Reinforcement Learning (NeSy-RL) combines symbolic reasoning with gradient-based optimization to achieve interpretable and generalizabl

agentsarxiv-cs-lg
21 Apr 2026
Agents

Live LTL Progress Tracking: Towards Task-Based Exploration

DGX agent

arXiv:2604.17106v1 Announce Type: new Abstract: Motivated by the challenge presented by non-Markovian objectives in reinforcement learning (RL), we present a novel framework to track and represent the

agentsarxiv-cs-lg
21 Apr 2026
Agents

MoRI: Learning Motivation-Grounded Reasoning for Scientific Ideation in Large Language Models

DGX agent

arXiv:2603.19044v2 Announce Type: replace Abstract: Scientific ideation aims to propose novel solutions within a given scientific context. Existing LLM-based agentic approaches emulate human research

agentsarxiv-cs-cl
21 Apr 2026
Agents

Scaling Beyond Context: A Survey of Multimodal Retrieval-Augmented Generation for Document Understanding

DGX agent

arXiv:2510.15253v3 Announce Type: replace Abstract: Document understanding is critical for applications from financial analysis to scientific discovery. Current approaches, whether OCR-based pipelines

agentsarxiv-cs-cl
21 Apr 2026
Agents

The Thin Line Between Comprehension and Persuasion in LLMs

DGX agent

arXiv:2507.01936v3 Announce Type: replace Abstract: Large language models (LLMs) are excellent at maintaining high-level, convincing dialogue, but it remains unclear whether their persuasive success r

agentsarxiv-cs-cl
21 Apr 2026
Agents

Will People Enjoy a Robot Trainer? A Case Study with Snoopie the Pacerbot

DGX agent

arXiv:2604.18331v1 Announce Type: new Abstract: The physicality of exercise makes the role of athletic trainers unique. Their physical presence allows them to guide a student through a motion, demonst

agentsarxiv-cs-ro
21 Apr 2026
Agents

CSLE: A Reinforcement Learning Platform for Autonomous Security Management

DGX agent

arXiv:2604.15590v1 Announce Type: cross Abstract: Reinforcement learning is a promising approach to autonomous and adaptive security management in networked systems. However, current reinforcement lea

agentsarxiv-cs-ai
20 Apr 2026
Model Releases

Dynamic Tool Dependency Retrieval for Lightweight Function Calling

DGX agent

arXiv:2512.17052v4 Announce Type: replace Abstract: Function calling agents powered by Large Language Models (LLMs) select external tools to automate complex tasks. On-device agents typically use a re

model-releasesarxiv-cs-lg
20 Apr 2026
Model Releases

Mind DeepResearch Technical Report

DGX agent

arXiv:2604.14518v2 Announce Type: replace Abstract: We present Mind DeepResearch (MindDR), an efficient multi-agent deep research framework that achieves leading performance with only ~30B-parameter m

model-releasesarxiv-cs-ai
20 Apr 2026
Agents

VeriGraph: Scene Graphs for Execution Verifiable Robot Planning

DGX agent

arXiv:2411.10446v3 Announce Type: replace-cross Abstract: Recent progress in vision-language models (VLMs) has opened new possibilities for robot task planning, but these models often produce incorrec

agentsarxiv-cs-ai
20 Apr 2026
Agents

Beyond Literal Mapping: Benchmarking and Improving Non-Literal Translation Evaluation

DGX agent

arXiv:2601.07338v2 Announce Type: replace Abstract: Large Language Models (LLMs) have significantly advanced Machine Translation (MT), applying them to linguistically complex domains-such as Social Ne

agentsarxiv-cs-cl
17 Apr 2026
Agents

Interpretable and Explainable Surrogate Modeling for Simulations: A State-of-the-Art Survey and Perspectives on Explainable AI for Decision-Making

DGX agent

arXiv:2604.14240v1 Announce Type: cross Abstract: The simulation of complex systems increasingly relies on sophisticated but fundamentally opaque computational black-box simulators. Surrogate models p

agentsarxiv-cs-lg
17 Apr 2026
Safety

Optimistic Policy Learning under Pessimistic Adversaries with Regret and Violation Guarantees

DGX agent

arXiv:2604.14243v1 Announce Type: new Abstract: Real-world decision-making systems operate in environments where state transitions depend not only on the agent's actions, but also on extbf{exogenous f

safetyarxiv-cs-lg
17 Apr 2026
Agents

Towards AI-assisted Neutrino Flavor Theory Design

DGX agent

arXiv:2506.08080v2 Announce Type: replace-cross Abstract: Particle physics theories, such as those which explain neutrino flavor mixing, arise from a vast landscape of model-building possibilities. A

agentsarxiv-cs-lg
17 Apr 2026
Agents

WybeCoder: Verified Imperative Code Generation

DGX agent

arXiv:2603.29088v2 Announce Type: replace-cross Abstract: Recent progress in large language models (LLMs) has substantially advanced automatic code generation and formal theorem proving, yet software

agentsarxiv-cs-ai
17 Apr 2026
Research

A closer look at how large language models trust humans: patterns and biases

DGX agent

arXiv:2504.15801v2 Announce Type: replace Abstract: As large language models (LLMs) and LLM-based agents increasingly interact with humans in decision-making contexts, understanding the trust dynamics

researcharxiv-cs-cl
16 Apr 2026
Model Releases

CollabCoder: Plan-Code Co-Evolution via Collaborative Decision-Making for Efficient Code Generation

DGX agent

arXiv:2604.13946v1 Announce Type: cross Abstract: Automated code generation remains a persistent challenge in software engineering, as conventional multi-agent frameworks are often constrained by stat

model-releasesarxiv-cs-cl
16 Apr 2026
Safety

From Instruction to Event: Sound-Triggered Mobile Manipulation

DGX agent

arXiv:2601.21667v2 Announce Type: replace-cross Abstract: Current mobile manipulation research predominantly follows an instruction-driven paradigm, where agents rely on predefined textual commands to

safetyarxiv-cs-cv
16 Apr 2026
Model Releases

MERRIN: A Benchmark for Multimodal Evidence Retrieval and Reasoning in Noisy Web Environments

DGX agent

arXiv:2604.13418v1 Announce Type: new Abstract: Motivated by the underspecified, multi-hop nature of search queries and the multimodal, heterogeneous, and often conflicting nature of real-world web re

model-releasesarxiv-cs-cl
16 Apr 2026
Agents

Numerical Instability and Chaos: Quantifying the Unpredictability of Large Language Models

DGX agent

arXiv:2604.13206v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly integrated into agentic workflows, their unpredictability stemming from numerical instability has eme

agentsarxiv-cs-lg
16 Apr 2026
← Previous
1…122123124125126…236
Next →