AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Safety

Multi-Gait Learning for Humanoid Robots Using Reinforcement Learning with Selective Adversarial Motion Prior

DGX agent

arXiv:2604.19102v1 Announce Type: cross Abstract: Learning diverse locomotion skills for humanoid robots in a unified reinforcement learning framework remains challenging due to the conflicting requir

safetyarxiv-cs-ai
22 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Reasoning Over Space: Enabling Geographic Reasoning for LLM-Based Generative Next POI Recommendation

DGX agent

arXiv:2601.04562v2 Announce Type: replace Abstract: Generative recommendation with large language models (LLMs) reframes prediction as sequence generation, yet existing LLM-based recommenders remain l

local-aiarxiv-cs-ai
22 Apr 2026
Safety

VoteGCL: Enhancing Graph-based Recommendations with Majority-Voting LLM-Rerank Augmentation

DGX agent

arXiv:2507.21563v4 Announce Type: replace-cross Abstract: Recommendation systems often suffer from data sparsity caused by limited user-item interactions, which degrade their performance and amplify p

safetyarxiv-cs-lg
22 Apr 2026
Model Releases

Chain Of Interaction Benchmark (COIN): When Reasoning meets Embodied Interaction

DGX agent

arXiv:2604.16886v1 Announce Type: new Abstract: Generalist embodied agents must perform interactive, causally-dependent reasoning, continually interacting with the environment, acquiring information,

model-releasesarxiv-cs-ro
21 Apr 2026
Safety

Efficient Federated RLHF via Zeroth-Order Policy Optimization

DGX agent

arXiv:2604.17747v1 Announce Type: new Abstract: This paper considers reinforcement learning from human feedback in a federated learning setting with resource-constrained agents, such as edge devices.

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Freshness-Aware Prioritized Experience Replay for LLM/VLM Reinforcement Learning

DGX agent

arXiv:2604.16918v1 Announce Type: new Abstract: Reinforcement Learning (RL) has achieved impressive success in post-training Large Language Models (LLMs) and Vision-Language Models (VLMs), with on-pol

model-releasesarxiv-cs-cl
21 Apr 2026
Local Ai

Fringe Projection Based Vision Pipeline for Autonomous Hard Drive Disassembly

DGX agent

arXiv:2604.17231v1 Announce Type: new Abstract: Unrecovered e-waste represents a significant economic loss. Hard disk drives (HDDs) comprise a valuable e-waste stream necessitating robotic disassembly

local-aiarxiv-cs-cv
21 Apr 2026
Safety

HAVEN: Hierarchical Adversary-aware Visibility-Enabled Navigation with Cover Utilization using Deep Transformer Q-Networks

DGX agent

arXiv:2512.00592v2 Announce Type: replace Abstract: Autonomous navigation in partially observable environments requires agents to reason beyond immediate sensor input, exploit occlusion, and ensure sa

safetyarxiv-cs-ro
21 Apr 2026
Research

LVLMs and Humans Ground Differently in Referential Communication

DGX agent

arXiv:2601.19792v3 Announce Type: replace Abstract: For generative AI agents to partner effectively with human users, the ability to accurately predict human intent is critical. But this ability to co

researcharxiv-cs-cl
21 Apr 2026
Hardware

PiERN: Token-Level Routing for Integrating High-Precision Computation and Reasoning

DGX agent

arXiv:2509.18169v3 Announce Type: replace-cross Abstract: Tasks on complex systems require high-precision numerical computation to support decisions, but current large language models (LLMs) cannot in

hardwarearxiv-cs-cl
21 Apr 2026
Model Releases

PRISM: Probing Reasoning, Instruction, and Source Memory in LLM Hallucinations

DGX agent

arXiv:2604.16909v1 Announce Type: new Abstract: As large language models (LLMs) evolve from conversational assistants into agents capable of handling complex tasks, they are increasingly deployed in h

model-releasesarxiv-cs-cl
21 Apr 2026
Model Releases

Progressive Online Video Understanding with Evidence-Aligned Timing and Transparent Decisions

DGX agent

arXiv:2604.18459v1 Announce Type: new Abstract: Visual agents operating in the wild must respond to queries precisely when sufficient evidence first appears in a video stream, a critical capability th

model-releasesarxiv-cs-cv
21 Apr 2026
Safety

Rethinking Jailbreak Detection of Large Vision Language Models with Representational Contrastive Scoring

DGX agent

arXiv:2512.12069v3 Announce Type: replace-cross Abstract: Large Vision-Language Models (LVLMs) are vulnerable to a growing array of multimodal jailbreak attacks, necessitating defenses that are both g

safetyarxiv-cs-cl
21 Apr 2026
Safety

Robust Tool Use via Fission-GRPO: Learning to Recover from Execution Errors

DGX agent

arXiv:2601.15625v2 Announce Type: replace Abstract: Large language models (LLMs) can call tools effectively, yet they remain brittle in multi-turn execution: after a tool-call error, smaller models of

safetyarxiv-cs-lg
21 Apr 2026
Model Releases

Rule-VLN: Bridging Perception and Compliance via Semantic Reasoning and Geometric Rectification

DGX agent

arXiv:2604.16993v1 Announce Type: cross Abstract: As embodied AI transitions to real-world deployment, the success of the Vision-and-Language Navigation (VLN) task tends to evolve from mere reachabili

model-releasesarxiv-cs-cv
21 Apr 2026
Research

StageMem: Lifecycle-Managed Memory for Language Models

DGX agent

arXiv:2604.16774v1 Announce Type: new Abstract: Long-horizon language model systems increasingly rely on persistent memory, yet many current designs still treat memory primarily as a static store: wri

researcharxiv-cs-cl
21 Apr 2026
Safety

Support Sufficiency as Consequence-Sensitive Compression in Belief Arbitration

DGX agent

arXiv:2604.16434v1 Announce Type: cross Abstract: When a system commits to a hypothesis, much of the evidential structure behind that commitment is lost to compression. Standard accounts assume that s

safetyarxiv-cs-lg
21 Apr 2026
Research

Synthetic Data Generation for Training Diversified Commonsense Reasoning Models

DGX agent

arXiv:2603.18361v2 Announce Type: replace Abstract: Conversational agents are required to respond to their users not only with high quality (i.e. commonsense bearing) responses, but also considering m

researcharxiv-cs-cl
21 Apr 2026
Safety

Training Language Models to Use Prolog as a Tool

DGX agent

arXiv:2512.07407v2 Announce Type: replace Abstract: Language models frequently produce plausible yet incorrect reasoning traces that are difficult to verify. We investigate fine-tuning models to use P

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

WorldDB: A Vector Graph-of-Worlds Memory Engine with Ontology-Aware Write-Time Reconciliation

DGX agent

arXiv:2604.18478v1 Announce Type: cross Abstract: Persistent memory is the bottleneck separating stateless chatbots from long-running agentic systems. Retrieval-augmented generation (RAG) over flat ve

model-releasesarxiv-cs-cl
21 Apr 2026
Safety

ZoFia: Zero-Shot Fake News Detection with Entity-Guided Retrieval and Multi-LLM Interaction

DGX agent

arXiv:2511.01188v2 Announce Type: replace Abstract: The rapid spread of fake news threatens social stability and public trust, highlighting the urgent need for its effective detection. Although large

safetyarxiv-cs-cl
21 Apr 2026
Model Releases

AscendKernelGen: A Systematic Study of LLM-Based Kernel Generation for Neural Processing Units

DGX agent

arXiv:2601.07160v2 Announce Type: replace Abstract: To meet the ever-increasing demand for computational efficiency, Neural Processing Units (NPUs) have become critical in modern AI infrastructure. Ho

model-releasesarxiv-cs-ai
20 Apr 2026
Safety

Cognitive Agency Surrender: Defending Epistemic Sovereignty via Scaffolded AI Friction

DGX agent

arXiv:2603.21735v2 Announce Type: replace-cross Abstract: The proliferation of Generative Artificial Intelligence has transformed benign cognitive offloading into a systemic risk of cognitive agency s

safetyarxiv-cs-ai
20 Apr 2026
Local Ai

DataCenterGym: A Physics-Grounded Simulator for Multi-Objective Data Center Scheduling

DGX agent

arXiv:2604.15594v1 Announce Type: cross Abstract: Modern datacenters schedule heterogeneous workloads across geo-distributed sites with diverse compute capacities, electricity prices, and thermal cond

local-aiarxiv-cs-ai
20 Apr 2026
Model Releases

FineCog-Nav: Integrating Fine-grained Cognitive Modules for Zero-shot Multimodal UAV Navigation

DGX agent

arXiv:2604.16298v1 Announce Type: new Abstract: UAV vision-language navigation (VLN) requires an agent to navigate complex 3D environments from an egocentric perspective while following ambiguous mult

model-releasesarxiv-cs-cv
20 Apr 2026
Research

Imperfectly Cooperative Human-AI Interactions: Comparing the Impacts of Human and AI Attributes in Simulated and User Studies

DGX agent

arXiv:2604.15607v1 Announce Type: cross Abstract: AI design characteristics and human personality traits each impact the quality and outcomes of human-AI interactions. However, their relative and join

researcharxiv-cs-ai
20 Apr 2026
Local Ai

KRONE: Scalable LLM-Augmented Log Anomaly Detection via Hierarchical Abstraction

DGX agent

arXiv:2602.07303v3 Announce Type: replace-cross Abstract: Log anomaly detection is crucial for uncovering system failures and security risks. Although logs originate from nested component executions w

local-aiarxiv-cs-ai
20 Apr 2026
Model Releases

PRL-Bench: A Comprehensive Benchmark Evaluating LLMs' Capabilities in Frontier Physics Research

DGX agent

arXiv:2604.15411v1 Announce Type: cross Abstract: The paradigm of agentic science requires AI systems to conduct robust reasoning and engage in long-horizon, autonomous exploration. However, current s

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Reasoning over Video: Evaluating How MLLMs Extract, Integrate, and Reconstruct Spatiotemporal Evidence

DGX agent

arXiv:2603.13091v2 Announce Type: replace Abstract: The growing interest in embodied agents increases the demand for spatiotemporal video understanding, yet existing benchmarks largely emphasize extra

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Robust Synchronisation for Federated Learning in The Face of Correlated Device Failure

DGX agent

arXiv:2604.16090v1 Announce Type: cross Abstract: Probabilistic Synchronous Parallel (PSP) is a technique in distributed learning systems to reduce synchronization bottlenecks by sampling a subset of

safetyarxiv-cs-ai
20 Apr 2026
Model Releases

The Reasoning Trap: How Enhancing LLM Reasoning Amplifies Tool Hallucination

DGX agent

arXiv:2510.22977v2 Announce Type: replace-cross Abstract: Enhancing the reasoning capabilities of Large Language Models (LLMs) is a key strategy for building Agents that 'think then act.' However, rec

model-releasesarxiv-cs-ai
20 Apr 2026
Model Releases

Watching Movies Like a Human: Egocentric Emotion Understanding for Embodied Companions

DGX agent

arXiv:2604.15823v1 Announce Type: new Abstract: Embodied robotic agents often perceive movies through an egocentric screen-view interface rather than native cinematic footage, introducing domain shift

model-releasesarxiv-cs-cv
20 Apr 2026
Safety

Building Trust in the Skies: A Knowledge-Grounded LLM-based Framework for Aviation Safety

DGX agent

arXiv:2604.13101v1 Announce Type: cross Abstract: The integration of Large Language Models (LLMs) into aviation safety decision-making represents a significant technological advancement, yet their sta

safetyarxiv-cs-ai
17 Apr 2026
Model Releases

CaptionQA: Is Your Caption as Useful as the Image Itself?

DGX agent

arXiv:2511.21025v2 Announce Type: replace Abstract: Image captions serve as efficient surrogates for visual content in multimodal systems such as retrieval, recommendation, and multi-step agentic infe

model-releasesarxiv-cs-cv
17 Apr 2026
Safety

Direct Preference Optimization for Primitive-Enabled Hierarchical RL: A Bilevel Approach

DGX agent

arXiv:2411.00361v4 Announce Type: replace Abstract: Hierarchical reinforcement learning (HRL) enables agents to solve complex, long-horizon tasks by decomposing them into manageable sub-tasks. However

safetyarxiv-cs-lg
17 Apr 2026
Safety

[Emerging Ideas] Artificial Tripartite Intelligence: A Bio-Inspired, Sensor-First Architecture for Physical AI

DGX agent

arXiv:2604.13959v1 Announce Type: new Abstract: As AI moves from data centers to robots and wearables, scaling ever-larger models becomes insufficient. Physical AI operates under tight latency, energy

safetyarxiv-cs-ai
17 Apr 2026
Research

Listening Alone, Understanding Together: Collaborative Context Recovery for Privacy-Aware AI

DGX agent

arXiv:2604.13348v1 Announce Type: new Abstract: We introduce CONCORD, a privacy-aware asynchronous assistant-to-assistant (A2A) framework that leverages collaboration between proactive speech-based AI

researcharxiv-cs-ai
17 Apr 2026
Model Releases

MemGround: Long-Term Memory Evaluation Kit for Large Language Models in Gamified Scenarios

DGX agent

arXiv:2604.14158v1 Announce Type: new Abstract: Current evaluations of long-term memory in LLMs are fundamentally static. By fixating on simple retrieval and short-context inference, they neglect the

model-releasesarxiv-cs-cl
17 Apr 2026
Safety

Doc-V*:Coarse-to-Fine Interactive Visual Reasoning for Multi-Page Document VQA

DGX agent

arXiv:2604.13731v1 Announce Type: new Abstract: Multi-page Document Visual Question Answering requires reasoning over semantics, layouts, and visual elements in long, visually dense documents. Existin

safetyarxiv-cs-cl
16 Apr 2026
Model Releases

LongVideo-R1: Smart Navigation for Low-cost Long Video Understanding

DGX agent

arXiv:2602.20913v2 Announce Type: replace Abstract: This paper addresses the critical and underexplored challenge of long video understanding with low computational budgets. We propose LongVideo-R1, a

model-releasesarxiv-cs-cv
16 Apr 2026
Model Releases

Dreamer-CDP: Improving Reconstruction-free World Models Via Continuous Deterministic Representation Prediction

DGX agent

arXiv:2603.07083v2 Announce Type: replace Abstract: Model-based reinforcement learning (MBRL) agents operating in high-dimensional observation spaces, such as Dreamer, rely on learning abstract repres

model-releasesarxiv-cs-lg
15 Apr 2026
Research

From Kinematics to Dynamics: Learning to Refine Hybrid Plans for Physically Feasible Execution

DGX agent

arXiv:2604.12474v1 Announce Type: cross Abstract: In many robotic tasks, agents must traverse a sequence of spatial regions to complete a mission. Such problems are inherently mixed discrete-continuou

researcharxiv-cs-ai
15 Apr 2026
Model Releases

Memory as Metabolism: A Design for Companion Knowledge Systems

DGX agent

arXiv:2604.12034v1 Announce Type: new Abstract: Retrieval-Augmented Generation remains the dominant pattern for giving LLMs persistent memory, but a visible cluster of personal wiki-style memory archi

model-releasesarxiv-cs-ai
15 Apr 2026
Model Releases

ProbeLogits: Kernel-Level LLM Inference Primitives for AI-Native Operating Systems

DGX agent

arXiv:2604.11943v1 Announce Type: cross Abstract: An OS kernel that runs LLM inference internally can read logit distributions before any text is generated -- and act on them as a governance primitive

model-releasesarxiv-cs-lg
15 Apr 2026
Safety

Progress-Think: Semantic Progress Reasoning for Vision-Language Navigation

DGX agent

arXiv:2511.17097v2 Announce Type: replace Abstract: Vision-Language Navigation requires agents to act coherently over long horizons by understanding not only local visual context but also how far they

safetyarxiv-cs-ro
15 Apr 2026
Model Releases

Reading Between the Pixels: Linking Text-Image Embedding Alignment to Typographic Attack Success on Vision-Language Models

DGX agent

arXiv:2604.12371v1 Announce Type: new Abstract: We study typographic prompt injection attacks on vision-language models (VLMs), where adversarial text is rendered as images to bypass safety mechanisms

model-releasesarxiv-cs-cv
15 Apr 2026
Model Releases

Round-Trip Translation Reveals What Frontier Multilingual Benchmarks Miss

DGX agent

arXiv:2604.12911v1 Announce Type: cross Abstract: Multilingual benchmarks guide the development of frontier models. Yet multilingual evaluations reported by frontier models are structured similar to p

model-releasesarxiv-cs-ai
15 Apr 2026
Safety

Scaffold-Conditioned Preference Triplets for Controllable Molecular Optimization with Large Language Models

DGX agent

arXiv:2604.12350v1 Announce Type: cross Abstract: Molecular property optimization is central to drug discovery, yet many deep learning methods rely on black-box scoring and offer limited control over

safetyarxiv-cs-ai
15 Apr 2026
← Previous
1…183184185186187…233
Next →