AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Agents

Graph of States: Solving Abductive Tasks with Large Language Models

DGX agent

arXiv:2603.21250v2 Announce Type: replace Abstract: Logical reasoning encompasses deduction, induction, and abduction. However, while Large Language Models (LLMs) have effectively mastered the former

agentsarxiv-cs-ai
15 May 2026
Agents
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Learning Direct Control Policies with Flow Matching for Autonomous Driving

DGX agent

arXiv:2605.14832v1 Announce Type: cross Abstract: We present a flow-matching planner for autonomous driving that directly outputs actionable control trajectories defined by acceleration and curvature

agentsarxiv-cs-cv
15 May 2026
Agents

Not All Symbols Are Equal: Importance-Aware Constellation Design for Semantic Communication

DGX agent

arXiv:2605.14940v1 Announce Type: cross Abstract: Semantic communication systems for goal-oriented transmission must protect task-relevant information not only through source compression but also via

agentsarxiv-cs-ai
15 May 2026
Agents

S-AI-Recursive: A Bio-Inspired and Temporal Sparse AI Architecture for Iterative, Introspective, and Energy-Frugal Reasoning

DGX agent

arXiv:2605.13872v1 Announce Type: cross Abstract: This article introduces S-AI-Recursive, a bio-inspired Sparse Artificial Intelligence architecture in which reasoning is operationalized as a hormonal

agentsarxiv-cs-ai
15 May 2026
Agents

Silent Collapse in Recursive Learning Systems

DGX agent

arXiv:2605.14588v1 Announce Type: new Abstract: Recursive learning -- where models are trained on data generated by previous versions of themselves -- is increasingly common in large language models,

agentsarxiv-cs-lg
15 May 2026
Safety

GAGPO: Generalized Advantage Grouped Policy Optimization

DGX agent

arXiv:2605.13217v1 Announce Type: cross Abstract: Reinforcement learning has become a powerful paradigm for post-training large language model agents, yet credit assignment in multi-turn environments

safetyarxiv-cs-lg
14 May 2026
Agents

Identifying AI Web Scrapers Using Canary Tokens

DGX agent

arXiv:2605.13706v1 Announce Type: cross Abstract: From pre-training to query-time augmentation, web-scraped data helps to improve the quality and contextual relevancy of content generated by large lan

agentsarxiv-cs-ai
14 May 2026
Agents

Semantic knowledge guides innovation and drives cultural evolution

DGX agent

arXiv:2510.12837v3 Announce Type: replace-cross Abstract: Cultural evolution allows ideas and technologies to accumulate across generations, reaching their most complex and open-ended form in humans.

agentsarxiv-cs-ai
14 May 2026
Agents

Training Long-Context Vision-Language Models Effectively with Generalization Beyond 128K Context

DGX agent

arXiv:2605.13831v1 Announce Type: new Abstract: Long-context modeling is becoming a core capability of modern large vision-language models (LVLMs), enabling sustained context management across long-do

agentsarxiv-cs-cv
14 May 2026
Model Releases

Useful Memories Become Faulty When Continuously Updated by LLMs

DGX agent

arXiv:2605.12978v1 Announce Type: new Abstract: Learning from past experience benefits from two complementary forms of memory: episodic traces -- raw trajectories of what happened -- and consolidated

model-releasesarxiv-cs-ai
14 May 2026
Agents

When Should an AI Workflow Release? Always-Valid Inference for Black-Box Generate-Verify Systems

DGX agent

arXiv:2605.12947v1 Announce Type: cross Abstract: LLM-enabled AI workflows increasingly produce outputs through iterative generate-evaluate-revise loops. Each iteration can improve the candidate, but

agentsarxiv-cs-ai
14 May 2026
Local Ai

Grounding by Remembering: Cross-Scene and In-Scene Memory for 3D Functional Affordances

DGX agent

arXiv:2605.11616v1 Announce Type: new Abstract: Functional affordance grounding requires more than recognizing an object: an agent must localize the specific region that supports an interaction, such

local-aiarxiv-cs-cv
13 May 2026
Agents

LDDR: Linear-DPP-Based Dynamic-Resolution Frame Sampling for Video MLLMs

DGX agent

arXiv:2605.11477v1 Announce Type: new Abstract: Video understanding in multimodal large language models requires selecting informative frames from long, redundant videos under limited visual-token bud

agentsarxiv-cs-cv
13 May 2026
Agents

LychSim: A Controllable and Interactive Simulation Framework for Vision Research

DGX agent

arXiv:2605.12449v1 Announce Type: new Abstract: While self-supervised pretraining has reduced vision systems' reliance on synthetic data, simulation remains an indispensable tool for closed-loop optim

agentsarxiv-cs-cv
13 May 2026
Agents

RoboBlockly Studio: Conversational Block Programming with Embodied Robot Feedback for Computational Thinking

DGX agent

arXiv:2605.12059v1 Announce Type: cross Abstract: Computational thinking (CT) is increasingly promoted as a core literacy, yet learners and teachers face challenges in connecting abstract program logi

agentsarxiv-cs-ro
13 May 2026
Agents

SenseNova-U1: Unifying Multimodal Understanding and Generation with NEO-unify Architecture

DGX agent

arXiv:2605.12500v1 Announce Type: new Abstract: Recent large vision-language models (VLMs) remain fundamentally constrained by a persistent dichotomy: understanding and generation are treated as disti

agentsarxiv-cs-cv
13 May 2026
Agents

Adaptive Action Chunking via Multi-Chunk Q Value Estimation

DGX agent

arXiv:2605.10044v1 Announce Type: cross Abstract: Action chunking emerged as a pivotal technique in imitation learning, enabling policies to predict cohesive action sequences rather than single action

agentsarxiv-cs-ai
12 May 2026
Agents

Batch-of-Thought: Cross-Instance Learning for Enhanced LLM Reasoning

DGX agent

arXiv:2601.02950v3 Announce Type: replace Abstract: Current Large Language Model reasoning systems process queries independently, discarding valuable cross-instance signals such as shared reasoning pa

agentsarxiv-cs-ai
12 May 2026
Model Releases

Beyond Isolation: A Unified Benchmark for General-Purpose Navigation

DGX agent

arXiv:2605.09441v1 Announce Type: new Abstract: The pursuit of general-purpose embodied agents is hindered by fragmented evaluation protocols that isolate navigation skills and fixate on specific robo

model-releasesarxiv-cs-ro
12 May 2026
Safety

Beyond Self-Play: Hierarchical Reasoning for Continuous Motion in Closed-Loop Traffic Simulation

DGX agent

arXiv:2605.09153v1 Announce Type: cross Abstract: Closed-loop traffic simulation requires agents that are both scalable and behaviorally realistic. Recent self-play reinforcement learning approaches d

safetyarxiv-cs-ai
12 May 2026
Model Releases

Computer Use at the Edge of the Statistical Precipice

DGX agent

arXiv:2605.08261v1 Announce Type: cross Abstract: Evaluating Computer Use Agents (CUAs) on interactive environments is fraught with methodological pitfalls that the field has yet to systematically add

model-releasesarxiv-cs-ai
12 May 2026
Agents

ConsistNav: Closing the Action Consistency Gap in Zero-Shot Object Navigation with Semantic Executive Control

DGX agent

arXiv:2605.09869v1 Announce Type: cross Abstract: Zero-shot object navigation has advanced rapidly with open-vocabulary detectors, image--text models, and language-guided exploration. However, even af

agentsarxiv-cs-cv
12 May 2026
Safety

Core-Halo Decomposition: Decentralizing Large-Scale Fixed-Point Problems

DGX agent

arXiv:2605.08681v1 Announce Type: cross Abstract: We study solving large-scale fixed-point equation (x^star=ar F(x^star)) with decomposition. Standard strict decomposition assigns each agent a disjoin

safetyarxiv-cs-ai
12 May 2026
Agents

FocuSFT: Bilevel Optimization for Dilution-Aware Long-Context Fine-Tuning

DGX agent

arXiv:2605.09932v1 Announce Type: new Abstract: Large language models can now process increasingly long inputs, yet their ability to effectively use information spread across long contexts remains lim

agentsarxiv-cs-cl
12 May 2026
Agents

GRC: Unifying Reasoning-Driven Generation, Retrieval and Compression

DGX agent

arXiv:2605.09100v1 Announce Type: new Abstract: Text embedding and generative tasks are usually trained separately based on large language models (LLMs) nowadays. This causes a large amount of trainin

agentsarxiv-cs-cl
12 May 2026
Local Ai

Kintsugi: Learning Policies by Repairing Executable Knowledge Bases

DGX agent

arXiv:2605.09487v1 Announce Type: new Abstract: Modern embodied agents achieve impressive performance, but their task knowledge is often stored in neural weights, latent state, or prompt-bound memory,

local-aiarxiv-cs-lg
12 May 2026
Agents

LCGNav: Local Candidate-Aware Geometric Enhancement for General Topological Planning in Vision-Language Navigation

DGX agent

arXiv:2605.09053v1 Announce Type: new Abstract: Online topological planning has become an effective paradigm for Vision-Language Navigation in Continuous Environments (VLN-CE), but existing methods st

agentsarxiv-cs-cv
12 May 2026
Agents

Omni-scale Learning-based Sequential Decision Framework for Order Fulfillment of Tote-handling Robotic Systems

DGX agent

arXiv:2605.08758v1 Announce Type: cross Abstract: Driven by the rapid expansion of e-commerce and small-batch production, the size of the intralogistics load unit of finished goods, semi-finished good

agentsarxiv-cs-ai
12 May 2026
Agents

PnP-Corrector: A Universal Correction Framework for Coupled Spatiotemporal Forecasting

DGX agent

arXiv:2605.08935v1 Announce Type: new Abstract: Coupled spatiotemporal forecasting is important for predicting the future evolution of multiple interacting dynamical systems, such as in climate models

agentsarxiv-cs-ai
12 May 2026
Safety

Reward-Conditioned Reinforcement Learning

DGX agent

arXiv:2603.05066v2 Announce Type: replace Abstract: Single-task RL agents are typically trained under a fixed reward function, which limits their robustness to reward misspecification and their abilit

safetyarxiv-cs-lg
12 May 2026
Agents

SoK: A Systematic Bidirectional Literature Review of AI & DLT Convergence

DGX agent

arXiv:2605.10515v1 Announce Type: cross Abstract: The integration of Artificial Intelligence (AI) with Distributed Ledger Technology (DLT) has become a growing research area, yet contributions tend to

agentsarxiv-cs-ai
12 May 2026
Agents

Towards Autonomous Railway Operations: A Semi-Hierarchical Deep Reinforcement Learning Approach to the Vehicle Rescheduling Problem

DGX agent

arXiv:2605.10257v1 Announce Type: new Abstract: Managing disruptions in railway traffic management is a major challenge. Rising traffic density and infrastructure limits increase complexity, making th

agentsarxiv-cs-ai
12 May 2026
Model Releases

Benchmarking World-Model Learning with Environment-Level Queries

DGX agent

arXiv:2510.19788v4 Announce Type: replace Abstract: World models are central to building AI agents capable of flexible reasoning and planning. Yet current evaluations (i) test only properties measurab

model-releasesarxiv-cs-ai
11 May 2026
Model Releases

Bridging the Last Mile of Circuit Design: PostEDA-Bench, a Hierarchical Benchmark for PPA Convergence and DRC Fixing

DGX agent

arXiv:2605.06936v1 Announce Type: cross Abstract: LLM-based agents are increasingly applied to the 'last mile' of Electronic Design Automation (EDA): repairing residual sign-off Design Rule Check (DRC

model-releasesarxiv-cs-ai
11 May 2026
Safety

Discovering Multiagent Learning Algorithms with Large Language Models

DGX agent

arXiv:2602.16928v3 Announce Type: replace-cross Abstract: Much of the advancement in Multi-Agent Reinforcement Learning (MARL) for imperfect-information games has historically depended on the manual,

safetyarxiv-cs-ai
11 May 2026
Agents

mathsf{VISTA}: Decentralized Machine Learning in Adversary Dominated Environments

DGX agent

arXiv:2605.07841v1 Announce Type: cross Abstract: Decentralized machine learning often relies on outsourcing computations, such as gradient evaluations, to untrusted worker nodes. Existing robust aggr

agentsarxiv-cs-ai
11 May 2026
Model Releases

ProactiveMobile: A Comprehensive Benchmark for Boosting Proactive Intelligence on Mobile Devices

DGX agent

arXiv:2602.21858v4 Announce Type: replace Abstract: Multimodal large language models (MLLMs) have made significant progress in mobile agent development, yet their capabilities are predominantly confin

model-releasesarxiv-cs-ai
11 May 2026
Agents

See Tomorrow, Act Today: Foresight-Driven Autonomous Driving

DGX agent

arXiv:2605.07195v1 Announce Type: new Abstract: Current end-to-end autonomous driving planners are fundamentally reactive: they condition on historical and present observations to predict future actio

agentsarxiv-cs-cv
11 May 2026
Agents

State Representation and Termination for Recursive Reasoning Systems

DGX agent

arXiv:2605.06690v1 Announce Type: new Abstract: Recursive reasoning systems alternate between acquiring new evidence and refining an accumulated understanding. Two design choices are typically left im

agentsarxiv-cs-ai
11 May 2026
Agents

The AI-Native Large-Scale Agile Software Development Manifesto

DGX agent

arXiv:2605.07717v1 Announce Type: cross Abstract: Despite the widespread adoption of agile methods, achieving true agility at scale remains elusive. Large-scale agile frameworks remain largely human-c

agentsarxiv-cs-ai
11 May 2026
Safety

The Moltbook Files: A Harmless Slopocalypse or Humanity's Last Experiment

DGX agent

arXiv:2605.07462v1 Announce Type: cross Abstract: Moltbook is a Reddit-like platform where OpenClaw agents post, comment, and vote at scale - a so far unprecedented incident that comes with serious sa

safetyarxiv-cs-ai
11 May 2026
Agents

Designing a double deep reinforcement learning selection tool for resilient demand prediction

DGX agent

arXiv:2605.04068v1 Announce Type: new Abstract: The use of artificial intelligence in supply chain forecasting has attracted many scientific studies for several decades. However, the process of select

agentsarxiv-cs-lg
7 May 2026
Agents

From Knowledge to Action: Outcomes of the 2025 Large Language Model (LLM) Hackathon for Applications in Materials Science and Chemistry

DGX agent

arXiv:2605.03205v1 Announce Type: cross Abstract: Large language models (LLMs) are rapidly changing how researchers in materials science and chemistry discover, organize, and act on scientific knowled

agentsarxiv-cs-ai
7 May 2026
Agents

Human-Provenance Verification should be Treated as Labor Infrastructure in AI-Saturated Markets

DGX agent

arXiv:2605.03210v1 Announce Type: cross Abstract: We argue that AI-saturated markets are likely to create Veblen-good premiums, which we term human-provenance premiums, for verified human presence, an

agentsarxiv-cs-ai
7 May 2026
Agents

LAWS: Learning from Actual Workloads Symbolically -- A Self-Certifying Parametrized Cache Architecture for Neural Inference, Robotics, and Edge Deployment

DGX agent

arXiv:2605.04069v1 Announce Type: new Abstract: We introduce LAWS (Learning from Actual Workloads Symbolically), a self-certifying inference caching architecture that builds a growing library of certi

agentsarxiv-cs-lg
7 May 2026
Agents

Modular Reinforcement Learning For Cooperative Swarms

DGX agent

arXiv:2605.04939v1 Announce Type: new Abstract: A cooperative robot swarm is a collective of computationally-limited robots that share a common goal. Each robot can only interact with a small subset o

agentsarxiv-cs-ro
7 May 2026
Safety

Rollout Pass-Rate Control: Steering Binary-Reward RL Toward Its Most Informative Regime

DGX agent

arXiv:2605.05112v1 Announce Type: new Abstract: SWE-bench-style agentic reinforcement learning relies on expensive stateful trajectories, yet substantial compute is wasted on sampled rollout groups wi

safetyarxiv-cs-lg
7 May 2026
Agents

RouteFormer: A Transformer-Based Routing Framework for Autonomous Vehicles

DGX agent

arXiv:2504.05407v2 Announce Type: replace-cross Abstract: Autonomous surveillance missions in Internet of Things (IoT) networks often involve solving NP-hard combinatorial optimization problems to ens

agentsarxiv-cs-lg
7 May 2026
← Previous
1…148149150151152…236
Next →