AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,040 results
Research

A Hormone-inspired Emotion Layer for Transformer language models (HELT)

DGX agent

arXiv:2605.13858v1 Announce Type: cross Abstract: Large Language Models have demonstrated remarkable capabilities in generating contextually relevant and grammatically correct text. However, they fund

researcharxiv-cs-cl
15 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Research

Angel or Demon: Investigating the Plasticity Interventions' Impact on Backdoor Threats in Deep Reinforcement Learning

DGX agent

arXiv:2605.14587v1 Announce Type: cross Abstract: Extensive research has highlighted the severe threats posed by backdoor attacks to deep reinforcement learning (DRL). However, prior studies primarily

researcharxiv-cs-ai
15 May 2026
Model Releases

Beyond Binary: Reframing GUI Critique as Continuous Semantic Alignment

DGX agent

arXiv:2605.14311v1 Announce Type: cross Abstract: Test-Time Scaling (TTS), which samples multiple candidate actions and ranks them via a Critic Model, has emerged as a promising paradigm for generalis

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Chinese Short-Form Creative Content Generation via Explanation-Oriented Multi-Objective Optimization

DGX agent

arXiv:2511.15408v2 Announce Type: replace-cross Abstract: Chinese demonstrates high semantic compactness and rich metaphorical expressiveness, enabling limited text to convey dense meanings while incr

model-releasesarxiv-cs-ai
15 May 2026
Research

Confidence Estimation for LLMs in Multi-turn Interactions

DGX agent

arXiv:2601.02179v2 Announce Type: replace Abstract: While confidence estimation is a promising direction for mitigating hallucinations in Large Language Models (LLMs), current research overwhelmingly

researcharxiv-cs-cl
15 May 2026
Model Releases

Data-Augmented Game Starts for Accelerating Self-Play Exploration in Imperfect Information Games

DGX agent

arXiv:2605.14379v1 Announce Type: cross Abstract: Finding approximate equilibria for large-scale imperfect-information competitive games such as StarCraft, Dota, and CounterStrike remains computationa

model-releasesarxiv-cs-ai
15 May 2026
Model Releases

Descriptor: Distance-Annotated Traffic Perception Question Answering (DTPQA)

DGX agent

arXiv:2511.13397v2 Announce Type: replace-cross Abstract: The remarkable progress of Vision-Language Models (VLMs) on a variety of tasks has raised interest in their application to automated driving.

model-releasesarxiv-cs-ai
15 May 2026
Safety

Distributions as Actions: A Unified Framework for Diverse Action Spaces

DGX agent

arXiv:2506.16608v3 Announce Type: replace-cross Abstract: We introduce a novel reinforcement learning (RL) framework that treats parameterized action distributions as actions, redefining the boundary

safetyarxiv-cs-ai
15 May 2026
Safety

DIVER: Reinforced Diffusion Breaks Imitation Bottlenecks in End-to-End Autonomous Driving

DGX agent

arXiv:2507.04049v4 Announce Type: replace Abstract: Most end-to-end autonomous driving methods rely on imitation learning from single expert demonstrations, often leading to conservative and homogeneo

safetyarxiv-cs-cv
15 May 2026
Safety

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

DGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

safetyarxiv-cs-ai
15 May 2026
Safety

Native Parallel Reasoner: Reasoning in Parallelism via Self-Distilled Reinforcement Learning

DGX agent

arXiv:2512.07461v3 Announce Type: replace Abstract: We introduce Native Parallel Reasoner (NPR), a teacher-free framework that enables Large Language Models (LLMs) to self-evolve genuine parallel reas

safetyarxiv-cs-cl
15 May 2026
Model Releases

Peng's Q(lambda) for Conservative Value Estimation in Offline Reinforcement Learning

DGX agent

arXiv:2605.14779v1 Announce Type: new Abstract: We propose a model-free offline multi-step reinforcement learning (RL) algorithm, Conservative Peng's Q(lambda) (CPQL). Our algorithm adapts the Peng's

model-releasesarxiv-cs-lg
15 May 2026
Safety

Position: Behavioural Assurance Cannot Verify the Safety Claims Governance Now Demands

DGX agent

arXiv:2605.15164v1 Announce Type: cross Abstract: This position paper argues that behavioural assurance, even when carefully designed, is being asked to carry safety claims it cannot verify. AI govern

safetyarxiv-cs-ai
15 May 2026
Model Releases

QOuLiPo: What a quantum computer sees when it reads a book

DGX agent

arXiv:2605.14188v1 Announce Type: cross Abstract: What does a book look like to a quantum computer? This paper takes eight classical works of the Renaissance and its late-antique inheritance -- from A

model-releasesarxiv-cs-cl
15 May 2026
Model Releases

SToRe3D: Sparse Token Relevance in ViTs for Efficient Multi-View 3D Object Detection

DGX agent

arXiv:2605.14110v1 Announce Type: new Abstract: Vision Transformers (ViTs) enable strong multi-view 3D detection but are limited by high inference latency from dense token and query processing across

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

SVAG-Bench: A Large-Scale Benchmark for Multi-Instance Spatio-temporal Video Action Grounding

DGX agent

arXiv:2510.13016v3 Announce Type: replace Abstract: A truly capable AI system must do more than detect objects or recognize activities in isolation. It must form unified, grounded representations of w

model-releasesarxiv-cs-cv
15 May 2026
Model Releases

Test-Time Learning with an Evolving Library

DGX agent

arXiv:2605.14477v1 Announce Type: new Abstract: We introduce EvoLib, a test-time learning framework that enables large language models to accumulate, reuse, and evolve knowledge across problem instanc

model-releasesarxiv-cs-lg
15 May 2026
Safety

Active Sensing with Meta-Reinforcement Learning for Emitter Localization from RF Observations

DGX agent

arXiv:2605.12569v1 Announce Type: cross Abstract: Global navigation satellite system (GNSS) interference poses a serious threat to reliable positioning, especially in indoor and multipath-rich environ

safetyarxiv-cs-ai
14 May 2026
Safety

AdaptNC: Adaptive Nonconformity Scores for Conformal Prediction under Distribution Shift

DGX agent

arXiv:2602.01629v2 Announce Type: replace Abstract: Rigorous uncertainty quantification is essential for the safe deployment of autonomous systems in unconstrained environments. Conformal Prediction (

safetyarxiv-cs-lg
14 May 2026
Safety

Asynchronous Reasoning: Training-Free Interactive Thinking LLMs

DGX agent

arXiv:2512.10931v3 Announce Type: replace Abstract: Many state-of-the-art LLMs are trained to think before giving their answer. Reasoning can greatly improve language model capabilities, but it also m

safetyarxiv-cs-lg
14 May 2026
Safety

Automated Rubrics for Reliable Evaluation of Medical Dialogue Systems

DGX agent

arXiv:2601.15161v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used for clinical decision support, where hallucinations and unsafe suggestions may pose direct

safetyarxiv-cs-ai
14 May 2026
Model Releases

BEAVER: An Enterprise Benchmark for Text-to-SQL

DGX agent

arXiv:2409.02038v3 Announce Type: replace-cross Abstract: Existing text-to-SQL benchmarks have largely been constructed from public databases with well-structured schemas and simplistic question-SQL p

model-releasesarxiv-cs-ai
14 May 2026
Safety

BEHAVE: A Hybrid AI Framework for Real-Time Modeling of Collective Human Dynamics

DGX agent

arXiv:2605.12730v1 Announce Type: new Abstract: Existing AI systems for modeling human behavior operate at the level of individuals or detect events after they occur. As a result, they systematically

safetyarxiv-cs-ai
14 May 2026
Model Releases

CodeClash: Benchmarking Goal-Oriented Software Engineering

DGX agent

arXiv:2511.00839v2 Announce Type: replace-cross Abstract: Current benchmarks for coding evaluate language models (LMs) on concrete, well-specified tasks such as fixing specific bugs or writing targete

model-releasesarxiv-cs-ai
14 May 2026
Local Ai

Contextual Bandits for Resource-Constrained Devices using Probabilistic Learning

DGX agent

arXiv:2605.13346v1 Announce Type: new Abstract: Contextual bandits (CB) are online sequential decision-making problems under partial feedback that underpin many adaptive services. There is a growing d

local-aiarxiv-cs-lg
14 May 2026
Model Releases

D-VLA: A High-Concurrency Distributed Asynchronous Reinforcement Learning Framework for Vision-Language-Action Models

DGX agent

arXiv:2605.13276v1 Announce Type: new Abstract: The rapid evolution of Embodied AI has enabled Vision-Language-Action (VLA) models to excel in multimodal perception and task execution. However, applyi

model-releasesarxiv-cs-ai
14 May 2026
Safety

Decoupling Exploration and Policy Optimization: Uncertainty Guided Tree Search for Hard Exploration

DGX agent

arXiv:2603.22273v4 Announce Type: replace Abstract: The process of discovery requires active exploration -- the act of collecting new and informative data. However, efficient autonomous exploration re

safetyarxiv-cs-lg
14 May 2026
Local Ai

GRIP-VLM: Group-Relative Importance Pruning for Efficient Vision-Language Models

DGX agent

arXiv:2605.13375v1 Announce Type: cross Abstract: In Vision-Language Models (VLMs), processing a massive number of visual tokens incurs prohibitive computational overhead. While recent training-aware

local-aiarxiv-cs-ai
14 May 2026
Model Releases

HCSG: Human-Centric Semantic-Geometric Reasoning for Vision-Language Navigation

DGX agent

arXiv:2605.13321v1 Announce Type: new Abstract: VLN has achieved remarkable progress by scaling data and model capacity. However, the assumption of a static environment breaks down in real-world indoo

model-releasesarxiv-cs-ro
14 May 2026
Model Releases

Large Language Models Lack Temporal Awareness of Medical Knowledge

DGX agent

arXiv:2605.13045v1 Announce Type: new Abstract: The existing methods for evaluating the medical knowledge of Large Language Models (LLMs) are largely based on atemporal examination-style benchmarks, w

model-releasesarxiv-cs-lg
14 May 2026
Research

Limits of Personalizing Differential Privacy Budgets

DGX agent

arXiv:2605.13503v1 Announce Type: cross Abstract: A key technical difficulty in differential privacy is selecting a privacy budget that satisfies privacy requirements while maximizing utility. A natur

researcharxiv-cs-lg
14 May 2026
Research

Prismatic World Model: Learning Compositional Dynamics for Planning in Hybrid Systems

DGX agent

arXiv:2512.08411v2 Announce Type: replace Abstract: Model-based planning in robotic domains is challenged by the hybrid nature of physical dynamics, where continuous motion is punctuated by discrete e

researcharxiv-cs-ai
14 May 2026
Local Ai

PROMETHEUS: Automating Deep Causal Research Integrating Text, Data and Models

DGX agent

arXiv:2605.12835v1 Announce Type: new Abstract: Large language models can extract local causal claims from text, but those claims become more useful when organized as persistent, navigable world model

local-aiarxiv-cs-ai
14 May 2026
Research

Rigel3D: Rig-aware Latents for Animation-Ready 3D Asset Generation

DGX agent

arXiv:2605.13129v1 Announce Type: cross Abstract: Recent 3D generative models can synthesize high-quality assets, but their outputs are typically static: they lack the skeletal rigs, joint hierarchies

researcharxiv-cs-cv
14 May 2026
Model Releases

scShapeBench: Discovering geometry from high dimensional scRNAseq data

DGX agent

arXiv:2605.12662v1 Announce Type: new Abstract: High-dimensional point cloud data arise across many scientific domains, especially single-cell biology. The shapes or topologies of these datasets deter

model-releasesarxiv-cs-lg
14 May 2026
Model Releases

Senses Wide Shut: A Representation-Action Gap in Omnimodal LLMs

DGX agent

arXiv:2605.13737v1 Announce Type: new Abstract: When an omnimodal large language model accepts a question whose textual premise contradicts what it actually sees or hears, does the failure lie in perc

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

SupChain-Bench: Benchmarking Large Language Models for Real-World Supply Chain Management

DGX agent

arXiv:2602.07342v2 Announce Type: replace Abstract: Large language models (LLMs) have shown promise in complex reasoning and tool-based decision making, motivating their application to real-world supp

model-releasesarxiv-cs-ai
14 May 2026
Model Releases

TiCo: Time-Controllable Spoken Dialogue Model

DGX agent

arXiv:2603.22267v2 Announce Type: replace-cross Abstract: We introduce TiCo, a time-controllable spoken dialogue model (SDM) that follows time-constrained instructions (e.g., 'Please generate a respon

model-releasesarxiv-cs-ai
14 May 2026
Safety

Unweighted ranking for value-based decision making with uncertainty

DGX agent

arXiv:2605.13601v1 Announce Type: new Abstract: As intelligent systems are increasingly implemented in our society to make autonomous decisions, their commitment to human values raises serious concern

safetyarxiv-cs-ai
14 May 2026
Safety

A Survey of On-Policy Distillation for Large Language Models

DGX agent

arXiv:2604.00626v2 Announce Type: replace-cross Abstract: As Large Language Models (LLMs) continue to grow in both capability and cost, transferring frontier capabilities into smaller, deployable stud

safetyarxiv-cs-cl
13 May 2026
Safety

Characterizing the Robustness of Black-Box LLM Planners Under Perturbed Observations with Adaptive Stress Testing

DGX agent

arXiv:2505.05665v4 Announce Type: replace-cross Abstract: Large language models (LLMs) have recently demonstrated success in decision-making tasks including planning, control, and prediction, but thei

safetyarxiv-cs-cl
13 May 2026
Safety

Entropy Polarity in Reinforcement Fine-Tuning: Direction, Asymmetry, and Control

DGX agent

arXiv:2605.11775v1 Announce Type: cross Abstract: Policy entropy has emerged as a fundamental measure for understanding and controlling exploration in reinforcement learning with verifiable rewards (R

safetyarxiv-cs-cl
13 May 2026
Local Ai

Hierarchical LLM-Driven Control for HAPS-Assisted UAV Networks: Joint Optimization of Flight and Connectivity

DGX agent

arXiv:2605.11509v1 Announce Type: cross Abstract: Uncrewed aerial vehicles (UAVs) are increasingly deployed in complex networked environments, yet the joint optimization of multi-UAV motion control an

local-aiarxiv-cs-lg
13 May 2026
Model Releases

Intention-Conditioned Flow Occupancy Models

DGX agent

arXiv:2506.08902v4 Announce Type: replace Abstract: Large-scale pre-training has fundamentally changed how machine learning research is done today: large foundation models are trained once, and then c

model-releasesarxiv-cs-lg
13 May 2026
Safety

Internalizing Curriculum Judgment for LLM Reinforcement Fine-Tuning

DGX agent

arXiv:2605.11235v1 Announce Type: new Abstract: In LLM Reinforcement Fine-Tuning (RFT), curriculum learning drives both efficiency and performance. Yet, current methods externalize curriculum judgment

safetyarxiv-cs-lg
13 May 2026
Model Releases

KV-Fold: One-Step KV-Cache Recurrence for Long-Context Inference

DGX agent

arXiv:2605.12471v1 Announce Type: cross Abstract: We introduce KV-Fold, a simple, training-free long-context inference protocol that treats the key-value (KV) cache as the accumulator in a left fold o

model-releasesarxiv-cs-cl
13 May 2026
Safety

Looking and Listening Inside and Outside: Multimodal Artificial Intelligence Systems for Driver Safety Assessment and Intelligent Vehicle Decision-Making

DGX agent

arXiv:2602.07668v2 Announce Type: replace Abstract: The looking-in-looking-out (LILO) framework has enabled intelligent vehicle applications that understand both the outside scene and the driver state

safetyarxiv-cs-cv
13 May 2026
Applications

PrivacySIM: Evaluating LLM Simulation of User Privacy Behavior

DGX agent

arXiv:2605.12147v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used to simulate human behavior, but their ability to simulate individual privacy decisions is not well

applicationsarxiv-cs-lg
13 May 2026
← Previous
1…217218219220221…230
Next →