AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,289 results
Model Releases

SVI-Bench: A Dynamic Microworld for Strategic Video Intelligence

DGX agent

arXiv:2605.31529v1 Announce Type: new Abstract: True video intelligence demands more than recognizing what is visible: it requires reasoning about why events unfold, predicting what would change under

model-releasesarxiv-cs-cv
1 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Autonomous Driving

DGX agent

arXiv:2605.30576v1 Announce Type: new Abstract: Exploration in reinforcement learning for autonomous driving is inherently unsafe: agents must experience novel behaviors to learn, yet exploration can

safetyarxiv-cs-ai
1 Jun 2026
Agents

Code-QA-Bench: Separating Code Reasoning from Documentation Memorization in Repository-Level QA

DGX agent

arXiv:2605.29277v1 Announce Type: cross Abstract: We present Code-QA-Bench, a fully automated framework for synthesizing repository-level code understanding benchmarks that separates genuine code comp

agentsarxiv-cs-ai
29 May 2026
Agents

CompilerDream: Learning a Compiler World Model for General Code Optimization

DGX agent

arXiv:2404.16077v4 Announce Type: replace-cross Abstract: Effective code optimization in compilers is crucial for computer and software engineering. The success of these optimizations primarily depend

agentsarxiv-cs-lg
29 May 2026
Agents

Croissant Tasks: A Metadata Format for Reproducible Machine Learning Evaluations

DGX agent

arXiv:2605.29786v1 Announce Type: new Abstract: Reproducibility is fundamental to the scientific method, yet remains a critical challenge in machine learning. Contributing factors include underspecifi

agentsarxiv-cs-ai
29 May 2026
Safety

Discovering Cooperative Pipelines: Autoresearch for Sequential Social Dilemmas

DGX agent

arXiv:2605.30003v1 Announce Type: cross Abstract: We study two-level autoresearch for cooperation: an outer-loop AI agent autonomously redesigns the inner-loop pipeline of an LLM policy-synthesis syst

safetyarxiv-cs-ai
29 May 2026
Agents

Error as a Lens: Probing LLM Reasoning through Synthetic Misconception Generation

DGX agent

arXiv:2605.29007v1 Announce Type: new Abstract: Personalized tutoring, teacher training, and education research need access to targeted synthetic misconceptions, but privacy and IRB constraints make l

agentsarxiv-cs-cl
29 May 2026
Agents

PhyGenHOI: Physically-Aware 4D Generation of Dynamic Human-Object Interactions

DGX agent

arXiv:2605.30268v1 Announce Type: cross Abstract: We address the task of generating physically accurate and visually faithful 4D Human-Object Interaction (HOI). Given a static 3D human and target obje

agentsarxiv-cs-ai
29 May 2026
Model Releases

Training Deliberative Monitors for Black-Box Scheming Detection

DGX agent

arXiv:2605.29601v1 Announce Type: cross Abstract: As autonomous agents become more capable of performing real-world tasks, distinguishing scheming behavior from benign task pursuit may become a centra

model-releasesarxiv-cs-ai
29 May 2026
Agents

unix-ctf: Procedural Environments for Unix-Competence Reinforcement Learning

DGX agent

arXiv:2605.29115v1 Announce Type: cross Abstract: Unix competence is the ability to use shell and operating-system primitives as first-class tools, not merely to write programs through a terminal. Cur

agentsarxiv-cs-ai
29 May 2026
Agents

An LLM-Based Assistance System for Intuitive and Flexible Capability-Based Planning

DGX agent

arXiv:2605.28666v1 Announce Type: new Abstract: In modern industry, dynamic environments and the complexity of modular and reconfigurable resources require automated planning of process sequences. Cap

agentsarxiv-cs-ai
28 May 2026
Agents

Falsification-driven reinforcement learning for maritime motion planning

DGX agent

arXiv:2510.06970v2 Announce Type: replace-cross Abstract: Compliance with maritime traffic rules is essential for the safe operation of autonomous vessels, yet training reinforcement learning (RL) age

agentsarxiv-cs-lg
28 May 2026
Model Releases

Personalized Observation Normalization for Federated Reinforcement Learning in Simulation Environments with Heterogeneity

DGX agent

arXiv:2605.27385v1 Announce Type: cross Abstract: Federated reinforcement learning (FedRL) enables multiple agents to collaboratively train a global policy without sharing raw data, making it ideal fo

model-releasesarxiv-cs-ai
28 May 2026
Model Releases

You Live More Than Once: Towards Hierarchical Skill Meta-Evolving

DGX agent

arXiv:2605.28390v1 Announce Type: new Abstract: Test-time skill evolving is regarded as a new paradigm for enhancing deployed agentic systems. Existing works mainly focus on hard-coded skill evolving

model-releasesarxiv-cs-ai
28 May 2026
Safety

Constrained Meta Reinforcement Learning with Provable Test-Time Safety

DGX agent

arXiv:2601.21845v2 Announce Type: replace Abstract: Meta reinforcement learning (RL) allows agents to leverage experience across a distribution of tasks on which the agent can train at will, enabling

safetyarxiv-cs-lg
27 May 2026
Agents

Cordon-MAS: Defending RAG against Knowledge Poisoning via Information-Flow Control

DGX agent

arXiv:2605.26754v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) increasingly underpins high-stakes applications, yet remains vulnerable to Confundo-style poisoning where adversa

agentsarxiv-cs-ai
27 May 2026
Agents

FoundObj: Self-supervised Foundation Models as Rewards for Label-free 3D Object Segmentation

DGX agent

arXiv:2605.27178v1 Announce Type: cross Abstract: We address the challenging task of 3D object segmentation in complex scene point clouds without relying on any scene-level human annotations during tr

agentsarxiv-cs-ai
27 May 2026
Model Releases

Knowledge Graphs as the Missing Data Layer for LLM-Based Industrial Asset Operations

DGX agent

arXiv:2605.26874v1 Announce Type: cross Abstract: LLM-based agents for industrial asset operations show limited accuracy when reasoning over flat document stores. AssetOpsBench (KDD 2026) establishes

model-releasesarxiv-cs-ai
27 May 2026
Agents

AutoSOTA: An End-to-End Automated Research System for State-of-the-Art AI Model Discovery

DGX agent

arXiv:2604.05550v2 Announce Type: replace Abstract: Artificial intelligence research increasingly depends on prolonged cycles of reproduction, debugging, and iterative refinement to achieve State-Of-T

agentsarxiv-cs-cl
26 May 2026
Agents

EfficientGraph-RAG: Structured Retrieval-State Management for Cross-Task Retrieval-Augmented Generation

DGX agent

arXiv:2605.25379v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) has become the standard way to ground large language models in external knowledge, but many systems still organize

agentsarxiv-cs-cl
26 May 2026
Applications

Hypothesis Generation and Inductive Inference in Children and Language Models

DGX agent

arXiv:2605.24528v1 Announce Type: new Abstract: Real world decision-making requires constructing mental models under uncertainty over evidence, over the underlying causal rules, and over the state of

applicationsarxiv-cs-ai
26 May 2026
Agents

Meta-Engineering Harnesses for AI-Native Software Production: A Contract-Driven Adversarial Verification Architecture with Early Deployment Report

DGX agent

arXiv:2605.25665v1 Announce Type: cross Abstract: AI-native software development is often evaluated at the level of individual models, prompts, or generated artifacts. This framing is insufficient for

agentsarxiv-cs-ai
26 May 2026
Agents

Multi-Persona Debate System for Automated Scientific Hypothesis Generation

DGX agent

arXiv:2605.23917v1 Announce Type: new Abstract: Modern scientific discovery is bottlenecked not by data scarcity, but by the inability to synthesize fragmented knowledge into actionable hypotheses. Th

agentsarxiv-cs-cl
26 May 2026
Agents

PCGRLLM: Large Language Model-Driven Reward Design for Procedural Content Generation Reinforcement Learning

DGX agent

arXiv:2502.10906v2 Announce Type: replace Abstract: Reward design plays a pivotal role in the training of game AIs, requiring substantial domain-specific knowledge and human effort. In recent years, s

agentsarxiv-cs-ai
26 May 2026
Agents

Towards Multi-Turn Dialog Systems for Industrial Asset Operations and Maintenance

DGX agent

arXiv:2605.24953v1 Announce Type: new Abstract: Industrial asset operations and maintenance question answering is inherently multi-turn, iterative, and highly dependent on external tool invocation. Ho

agentsarxiv-cs-ai
26 May 2026
Agents

Why We Need World Models for AGI: Where LLMs Fail and How World Models May Outperform

DGX agent

arXiv:2605.23972v1 Announce Type: new Abstract: Large language models achieve strong performance in language generation and knowledge-intensive tasks, yet remain limited in settings requiring causal r

agentsarxiv-cs-ai
26 May 2026
Agents

EvalVerse: Pipeline-Aware and Expert-Calibrated Benchmarking for Professional Cinematic Video Generation

DGX agent

arXiv:2605.23271v1 Announce Type: cross Abstract: The rapid evolution of generative video foundation models has propelled the field toward professional-grade cinematic synthesis. To achieve such deman

agentsarxiv-cs-ai
25 May 2026
Model Releases

Inductive Deductive Synthesis: Enabling AI to Generate Formally Verified Systems

DGX agent

arXiv:2605.23109v1 Announce Type: new Abstract: AI agents increasingly excel at generating, testing, and refining code. However, they fall short on tasks requiring formal guarantees of full coverage t

model-releasesarxiv-cs-ai
25 May 2026
Safety

Understanding Goal Generalisation in Sequential Reinforcement Learning

DGX agent

arXiv:2605.23565v1 Announce Type: cross Abstract: Reinforcement learning agents often exhibit unintended goal-directed behaviour outside their training distribution, but we currently lack a principled

safetyarxiv-cs-ai
25 May 2026
Local Ai

Remember to be Curious: Episodic Context and Persistent Worlds for 3D Exploration

DGX agent

arXiv:2605.22814v1 Announce Type: new Abstract: Exploration is a prerequisite for learning useful behaviors in sparse-reward, long-horizon tasks, particularly within 3D environments. Curiosity-driven

local-aiarxiv-cs-lg
23 May 2026
Model Releases

Dissecting Embodied Abilities in Multimodal Language Models through Skill-level Evaluation and Diagnosis

DGX agent

arXiv:2510.08759v2 Announce Type: replace Abstract: Understanding the capability bottlenecks of embodied multimodal large language models (MLLMs) is crucial for improving embodied agents. However, exi

model-releasesarxiv-cs-cv
22 May 2026
Agents

Psy-Chronicle:A Structured Pipeline for Synthesizing Long-Horizon Campus Psychological Counseling Dialogues

DGX agent

arXiv:2605.22140v1 Announce Type: new Abstract: In recent years, large language models have shown substantial potential in psychological support tasks. However, existing psychological counseling data

agentsarxiv-cs-cl
22 May 2026
Model Releases

Hack-Verifiable Environments: Towards Evaluating Reward Hacking at Scale

DGX agent

arXiv:2605.20744v1 Announce Type: new Abstract: Aligning autonomous agents with human intent remains a central challenge in modern AI. A key manifestation of this challenge is reward hacking, whereby

model-releasesarxiv-cs-lg
21 May 2026
Agents

Towards Resilient and Autonomous Networks: A BlueSky Vision on AI-Native 6G

DGX agent

arXiv:2605.21395v1 Announce Type: cross Abstract: The proliferation of emerging applications, such as autonomous driving and immersive experiences, demands cellular networks that are not only faster,

agentsarxiv-cs-lg
21 May 2026
Model Releases

ZEBRA: Zero-shot Budgeted Resource Allocation for LLM Orchestration

DGX agent

arXiv:2605.20485v1 Announce Type: new Abstract: As autonomous agents increasingly execute end-to-end tasks under fixed monetary budgets, the pressing open question shifts from whether the budget is re

model-releasesarxiv-cs-lg
21 May 2026
Agents

Adaptive Threshold-Driven Continuous Greedy Method for Scalable Submodular Optimization

DGX agent

arXiv:2604.03419v2 Announce Type: replace Abstract: Submodular maximization under matroid constraints is a fundamental problem in combinatorial optimization with applications in sensing, data summariz

agentsarxiv-cs-lg
20 May 2026
Tutorials

Beyond Rational Illusion: Behaviorally Realistic Strategic Classification

DGX agent

arXiv:2605.19674v1 Announce Type: new Abstract: Strategic classification(SC) studies the interaction between decision models and agents who strategically manipulate their features for favorable outcom

tutorialsarxiv-cs-ai
20 May 2026
Safety

Distributional AGI Safety

DGX agent

arXiv:2512.16856v2 Announce Type: replace Abstract: AI safety and alignment research has predominantly been focused on methods for safeguarding individual AI systems, resting on the assumption of an e

safetyarxiv-cs-ai
20 May 2026
Safety

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning

DGX agent

arXiv:2605.19235v1 Announce Type: new Abstract: Competitive multi-agent reinforcement learning in imperfect-information games requires agents to act under partial observability and against adversarial

safetyarxiv-cs-lg
20 May 2026
Agents

PASC: Pipeline-Aware Conformal Prediction with Joint Coverage Guarantees for Multi-Stage NLP and LLM Pipelines

DGX agent

arXiv:2605.18812v1 Announce Type: cross Abstract: Modern NLP and LLM systems are pipelines: named entity recognition (NER) -> entity disambiguation (NED) -> entity typing, retrieval-augmented generati

agentsarxiv-cs-cl
20 May 2026
Agents

The Wikidata Query Logs Dataset

DGX agent

arXiv:2602.14594v2 Announce Type: replace Abstract: We present the Wikidata Query Logs (WDQL) dataset, a dataset consisting of 335k question-query pairs over the Wikidata knowledge graph. It is over 1

agentsarxiv-cs-cl
20 May 2026
Model Releases

A Machine With Human-Like Memory Systems

DGX agent

arXiv:2204.01611v3 Announce Type: replace Abstract: Inspired by the cognitive science theory, we explicitly model an agent with both semantic and episodic memory systems, and show that it is better th

model-releasesarxiv-cs-ai
19 May 2026
Local Ai

Beyond the Cartesian Illusion: Testing Two-Stage Multi-Modal Theory of Mind under Perceptual Bottlenecks

DGX agent

arXiv:2605.18194v1 Announce Type: new Abstract: While Multi-Modal Large Language Models (MLLMs) demonstrate impressive capabilities in general reasoning, their embodied spatial intelligence remains ha

local-aiarxiv-cs-ai
19 May 2026
Research

Counterparty Modeling is Not Strategy: The Limits of LLM Negotiators

DGX agent

arXiv:2605.16575v1 Announce Type: new Abstract: Negotiation requires more than inferring what the other side wants: it requires using that information to make advantageous offers and counteroffers ove

researcharxiv-cs-ai
19 May 2026
Model Releases

From Imitation to Interaction: Mastering Game of Schnapsen with Shallow Reinforcement Learning

DGX agent

arXiv:2605.17162v1 Announce Type: new Abstract: This paper investigates whether shallow neural network agents can master the card game Schnapsen and challenge a strong search-based baseline, RdeepBot,

model-releasesarxiv-cs-ai
19 May 2026
Agents

Generative AI and Two-Tiered Online Mental Health Communities

DGX agent

arXiv:2605.16279v1 Announce Type: cross Abstract: Online mental health communities (OMHCs) are tiered platforms that connect patients with licensed counselors through public Q&A forums and paid privat

agentsarxiv-cs-ai
19 May 2026
Agents

Genflow Ad Studio: A Compound AI Architecture for Brand-Aligned, Self-Correcting Video Generation

DGX agent

arXiv:2605.16748v1 Announce Type: cross Abstract: Recent advancements in generative video models demonstrate high visual fidelity, yet their integration into enterprise environments is restricted by t

agentsarxiv-cs-ai
19 May 2026
Local Ai

Incentive-Aware Federated Averaging with Performance Guarantees under Strategic Participation

DGX agent

arXiv:2603.20873v2 Announce Type: replace Abstract: Federated learning (FL) is a communication-efficient collaborative learning framework that enables model training across multiple agents with privat

local-aiarxiv-cs-lg
19 May 2026
← Previous
1…136137138139140…236
Next →