AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
11,153 results
Model Releases

Value Explicit Pretraining for Learning Transferable Representations

DGX agent

arXiv:2312.12339v3 Announce Type: replace Abstract: Understanding visual inputs for a given task amidst varied changes is a key challenge posed by visual reinforcement learning agents. We propose exti

model-releasesarxiv-cs-lg
4 May 2026
Local Ai

Attractor FCM

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2604.27947v1 Announce Type: cross Abstract: In this paper an attractor FCM is created, tested, and analyzed. This FCM is neither a hebbian based nor agentic, nor a hybrid; it rather is a gradien

local-aiarxiv-cs-ai
1 May 2026
Safety

D3-Gym: Constructing Real-World Verifiable Environments for Data-Driven Discovery

DGX agent

arXiv:2604.27977v1 Announce Type: new Abstract: Despite recent progress in language models and agents for scientific data-driven discovery, further advancing their capabilities is held back by the abs

safetyarxiv-cs-ai
1 May 2026
Local Ai

Detecting is Easy, Adapting is Hard: Local Expert Growth for Visual Model-Based Reinforcement Learning under Distribution Shift

DGX agent

arXiv:2604.27411v1 Announce Type: new Abstract: Visual model-based reinforcement learning (MBRL) agents can perform well on the training distribution, but often break down once the test environment sh

local-aiarxiv-cs-lg
1 May 2026
Hardware

EdgeFM: Efficient Edge Inference for Vision-Language Models

DGX agent

arXiv:2604.27476v1 Announce Type: new Abstract: Vision-language models (VLMs) have demonstrated strong applicability in edge industrial applications, yet their deployment remains severely constrained

hardwarearxiv-cs-cv
1 May 2026
Tutorials

Graph World Models: Concepts, Taxonomy, and Future Directions

DGX agent

arXiv:2604.27895v1 Announce Type: new Abstract: As one of the mainstream models of artificial intelligence, world models allow agents to learn the representation of the environment for efficient predi

tutorialsarxiv-cs-ai
1 May 2026
Model Releases

Learning to Forget: Continual Learning with Adaptive Weight Decay

DGX agent

arXiv:2604.27063v1 Announce Type: new Abstract: Continual learning agents with finite capacity must balance acquiring new knowledge with retaining the old. This requires controlled forgetting of knowl

model-releasesarxiv-cs-lg
1 May 2026
Model Releases

NashPG: A Policy Gradient Method with Iteratively Refined Regularization for Finding Nash Equilibria

DGX agent

arXiv:2510.18183v2 Announce Type: replace Abstract: Finding Nash equilibria in two-player zero-sum imperfect-information games remains a central challenge in multi-agent reinforcement learning. Recent

model-releasesarxiv-cs-lg
1 May 2026
Research

Simulating Infant First-Person Sensorimotor Experience via Motion Retargeting from Babies to Humanoids

DGX agent

arXiv:2604.27583v1 Announce Type: cross Abstract: Motion retargeting from humans to human-like artificial agents is becoming increasingly important as humanoid robots grow more capable. However, most

researcharxiv-cs-ro
1 May 2026
Model Releases

What Suppresses Nash Equilibrium Play in Large Language Models? Mechanistic Evidence and Causal Control

DGX agent

arXiv:2604.27167v1 Announce Type: cross Abstract: LLM agents are known to deviate from Nash equilibria in strategic interactions, but nobody has looked inside the model to understand why, or asked whe

model-releasesarxiv-cs-ai
1 May 2026
Research

A Decision-Theoretic Formalisation of Steganography With Applications to LLM Monitoring

DGX agent

arXiv:2602.23163v3 Announce Type: replace Abstract: Large language models are beginning to show steganographic capabilities. Such capabilities could allow misaligned models to evade oversight mechanis

researcharxiv-cs-ai
30 Apr 2026
Research

Beyond Screenshots: Evaluating VLMs' Understanding of UI Animations

DGX agent

arXiv:2604.26148v1 Announce Type: cross Abstract: AI agents operating on user interfaces must understand how interfaces communicate state and feedback to act reliably. As a core communicative modality

researcharxiv-cs-cl
30 Apr 2026
Local Ai

Distill-Belief: Closed-Loop Inverse Source Localization and Characterization in Physical Fields

DGX agent

arXiv:2604.26095v1 Announce Type: new Abstract: {Closed-loop inverse source localization and characterization (ISLC) requires a mobile agent to select measurements that localize sources and infer late

local-aiarxiv-cs-ai
30 Apr 2026
Model Releases

Inferix: A Block-Diffusion based Next-Generation Inference Engine for World Simulation

DGX agent

arXiv:2511.20714v2 Announce Type: replace-cross Abstract: World models serve as core simulators for fields such as agentic AI, embodied AI, and gaming, capable of generating long, physically realistic

model-releasesarxiv-cs-ai
30 Apr 2026
Model Releases

LLM Psychosis: A Theoretical and Diagnostic Framework for Reality-Boundary Failures in Large Language Models

DGX agent

arXiv:2604.25934v1 Announce Type: cross Abstract: The deployment of large language models (LLMs) as interactive agents has exposed a category of behavioral failure that prevailing terminology, princip

model-releasesarxiv-cs-ai
30 Apr 2026
Safety

Thinking About Thinking: Evaluating Reasoning in Post-Trained Language Models

DGX agent

arXiv:2510.16340v2 Announce Type: replace Abstract: Recent advances in post-training techniques have endowed Large Language Models (LLMs) with enhanced capabilities for tackling complex, logic-intensi

safetyarxiv-cs-cl
29 Apr 2026
Local Ai

Benchmarking Source-Sensitive Reasoning in Turkish: Humans and LLMs under Evidential Trust Manipulation

DGX agent

arXiv:2604.24665v1 Announce Type: cross Abstract: This paper investigates whether source trustworthiness shapes Turkish evidential morphology and whether large language models (LLMs) track this sensit

local-aiarxiv-cs-ai
28 Apr 2026
Local Ai

Domain-Filtered Knowledge Graphs from Sparse Autoencoder Features

DGX agent

arXiv:2604.23829v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) extract millions of interpretable features from a language model, but flat feature inventories aren't very useful on their ow

local-aiarxiv-cs-ai
28 Apr 2026
Model Releases

Evaluating whether AI models would sabotage AI safety research

DGX agent

arXiv:2604.24618v1 Announce Type: new Abstract: We evaluate the propensity of frontier models to sabotage or refuse to assist with safety research when deployed as AI research agents within a frontier

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

Grammar-Constrained Refinement of Safety Operational Rules Using Language in the Loop: What Could Go Wrong

DGX agent

arXiv:2604.23523v1 Announce Type: cross Abstract: Safety specifications in cyber-physical systems (CPS) capture the operational conditions the system must satisfy to operate safely within its intended

safetyarxiv-cs-ai
28 Apr 2026
Safety

InCoM: Intent-Driven Perception and Structured Coordination for Mobile Manipulation

DGX agent

arXiv:2602.23024v2 Announce Type: replace Abstract: Mobile manipulation is a fundamental capability for general-purpose robotic agents, requiring both coordinated control of the mobile base and manipu

safetyarxiv-cs-ro
28 Apr 2026
Safety

IndustryAssetEQA: A Neurosymbolic Operational Intelligence System for Embodied Question Answering in Industrial Asset Maintenance

DGX agent

arXiv:2604.23446v1 Announce Type: new Abstract: Industrial maintenance environments increasingly rely on AI systems to assist operators in understanding asset behavior, diagnosing failures, and evalua

safetyarxiv-cs-ai
28 Apr 2026
Safety

Intervention-Aware Multiscale Representation Learning from Imaging Phenomics and Perturbation Transcriptomics

DGX agent

arXiv:2604.22832v1 Announce Type: cross Abstract: Microscopy-based phenotypic profiling is scalable for drug discovery but lacks the mechanistic depth of transcriptomics, which remains costly and scar

safetyarxiv-cs-ai
28 Apr 2026
Safety

Learning Selective LLM Autonomy from Copilot Feedback in Enterprise Customer Support Workflows

DGX agent

arXiv:2604.23855v1 Announce Type: new Abstract: We present a deployed system that automates end-to-end customer support workflows inside an enterprise Business Process Management (BPM) platform. The a

safetyarxiv-cs-cl
28 Apr 2026
Model Releases

RAT: RunAnyThing via Fully Automated Environment Configuration

DGX agent

arXiv:2604.23190v1 Announce Type: cross Abstract: Automating repository-level software engineering tasks is a foundational challenge for autonomous code agents, largely due to the difficulty of config

model-releasesarxiv-cs-ai
28 Apr 2026
Safety

SMP: Reusable Score-Matching Motion Priors for Physics-Based Character Control

DGX agent

arXiv:2512.03028v3 Announce Type: replace-cross Abstract: Data-driven motion priors that can guide agents toward producing naturalistic behaviors play a pivotal role in creating life-like virtual char

safetyarxiv-cs-ai
28 Apr 2026
Tutorials

Context-Sensitive Abstractions for Reinforcement Learning with Parameterized Actions

DGX agent

arXiv:2512.20831v2 Announce Type: replace Abstract: Real-world sequential decision-making often involves parameterized action spaces that require both, decisions regarding discrete actions and decisio

tutorialsarxiv-cs-ai
27 Apr 2026
Model Releases

When Does LLM Self-Correction Help? A Control-Theoretic Markov Diagnostic and Verify-First Intervention

DGX agent

arXiv:2604.22273v1 Announce Type: new Abstract: Iterative self-correction is widely used in agentic LLM systems, but when repeated refinement helps versus hurts remains unclear. We frame self-correcti

model-releasesarxiv-cs-ai
27 Apr 2026
Model Releases

Do MLLMs Understand Pointing? Benchmarking and Enhancing Referential Reasoning in Egocentric Vision

DGX agent

arXiv:2604.21461v1 Announce Type: new Abstract: Egocentric AI agents, such as smart glasses, rely on pointing gestures to resolve referential ambiguities in natural language commands. However, despite

model-releasesarxiv-cs-cv
24 Apr 2026
Local Ai

Frozen LLMs as Map-Aware Spatio-Temporal Reasoners for Vehicle Trajectory Prediction

DGX agent

arXiv:2604.21479v1 Announce Type: new Abstract: Large language models (LLMs) have recently demonstrated strong reasoning capabilities and attracted increasing research attention in the field of autono

local-aiarxiv-cs-cv
24 Apr 2026
Model Releases

Hyperloop Transformers

DGX agent

arXiv:2604.21254v1 Announce Type: cross Abstract: LLM architecture research generally aims to maximize model quality subject to fixed compute/latency budgets. However, many applications of interest su

model-releasesarxiv-cs-cl
24 Apr 2026
Safety

Inferring High-Level Events from Timestamped Data: Complexity and Medical Applications

DGX agent

arXiv:2604.21793v1 Announce Type: new Abstract: In this paper, we develop a novel logic-based approach to detecting high-level temporally extended events from timestamped data and background knowledge

safetyarxiv-cs-ai
24 Apr 2026
Local Ai

Measure Twice, Click Once: Co-evolving Proposer and Visual Critic via Reinforcement Learning for GUI Grounding

DGX agent

arXiv:2604.21268v1 Announce Type: cross Abstract: Graphical User Interface (GUI) grounding requires mapping natural language instructions to precise pixel coordinates. However, due to visually homogen

local-aiarxiv-cs-ai
24 Apr 2026
Model Releases

Rectified Schrodinger Bridge Matching for Few-Step Visual Navigation

DGX agent

arXiv:2604.05673v2 Announce Type: replace-cross Abstract: Visual navigation is a core challenge in Embodied AI, requiring autonomous agents to translate high-dimensional sensory observations into cont

model-releasesarxiv-cs-ai
24 Apr 2026
Model Releases

Reinforcing 3D Understanding in Point-VLMs via Geometric Reward Credit Assignment

DGX agent

arXiv:2604.21160v1 Announce Type: new Abstract: Point-Vision-Language Models promise to empower embodied agents with executable spatial reasoning, yet they frequently succumb to geometric hallucinatio

model-releasesarxiv-cs-cv
24 Apr 2026
Model Releases

Reinforcing privacy reasoning in LLMs via normative simulacra from fiction

DGX agent

arXiv:2604.20904v1 Announce Type: cross Abstract: Information handling practices of LLM agents are broadly misaligned with the contextual privacy expectations of their users. Contextual Integrity (CI)

model-releasesarxiv-cs-ai
24 Apr 2026
Safety

DRIV-EX: Counterfactual Explanations for Driving LLMs

DGX agent

arXiv:2603.00696v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly used as reasoning engines in autonomous driving, yet their decision-making remains opaque. We propose

safetyarxiv-cs-cl
23 Apr 2026
Safety

EmbodiedMidtrain: Bridging the Gap between Vision-Language Models and Vision-Language-Action Models via Mid-training

DGX agent

arXiv:2604.20012v1 Announce Type: cross Abstract: Vision-Language-Action Models (VLAs) inherit their visual and linguistic capabilities from Vision-Language Models (VLMs), yet most VLAs are built from

safetyarxiv-cs-ai
23 Apr 2026
Safety

Frictionless Love: Associations Between AI Companion Roles and Behavioral Addiction

DGX agent

arXiv:2604.20011v1 Announce Type: cross Abstract: AI companion chatbots increasingly shape how people seek social and emotional connection, sometimes substituting for relationships with romantic partn

safetyarxiv-cs-ai
23 Apr 2026
Safety

Rethinking Reinforcement Fine-Tuning in LVLM: Convergence, Reward Decomposition, and Generalization

DGX agent

arXiv:2604.19857v1 Announce Type: cross Abstract: Reinforcement fine-tuning with verifiable rewards (RLVR) has emerged as a powerful paradigm for equipping large vision-language models (LVLMs) with ag

safetyarxiv-cs-cl
23 Apr 2026
Model Releases

SciCoQA: Quality Assurance for Scientific Paper--Code Alignment

DGX agent

arXiv:2601.12910v3 Announce Type: replace-cross Abstract: Discrepancies between scientific papers and their code undermine reproducibility, a concern that grows as automated research agents scale scie

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Where and What: Reasoning Dynamic and Implicit Preferences in Situated Conversational Recommendation

DGX agent

arXiv:2604.20749v1 Announce Type: new Abstract: Situated conversational recommendation (SCR), which utilizes visual scenes grounded in specific environments and natural language dialogue to deliver co

safetyarxiv-cs-ai
23 Apr 2026
Model Releases

WorkflowGen:an adaptive workflow generation mechanism driven by trajectory experience

DGX agent

arXiv:2604.19756v1 Announce Type: cross Abstract: Large language model (LLM) agents often suffer from high reasoning overhead, excessive token consumption, unstable execution, and inability to reuse p

model-releasesarxiv-cs-ai
23 Apr 2026
Safety

Benchmarking Misuse Mitigation Against Covert Adversaries

DGX agent

arXiv:2506.06414v2 Announce Type: replace-cross Abstract: Existing language model safety evaluations focus on overt attacks and low-stakes tasks. In reality, an attacker can easily subvert existing sa

safetyarxiv-cs-ai
22 Apr 2026
Model Releases

Evaluating Answer Leakage Robustness of LLM Tutors against Adversarial Student Attacks

DGX agent

arXiv:2604.18660v1 Announce Type: cross Abstract: Large Language Models (LLMs) are increasingly used in education, yet their default helpfulness often conflicts with pedagogical principles. Prior work

model-releasesarxiv-cs-ai
22 Apr 2026
Safety

Hybrid Task and Motion Planning with Reactive Collision Handling for Multi-Robot Disassembly of Complex Products: Application to EV Batteries

DGX agent

arXiv:2509.21020v2 Announce Type: replace Abstract: This paper addresses the problem of multi-robot coordination for complex manipulation task sequences. We present a vision-driven task-and-motion pla

safetyarxiv-cs-ro
22 Apr 2026
Applications

InHabit: Leveraging Image Foundation Models for Scalable 3D Human Placement

DGX agent

arXiv:2604.19673v1 Announce Type: new Abstract: Training embodied agents to understand 3D scenes as humans do requires large-scale data of people meaningfully interacting with diverse environments, ye

applicationsarxiv-cs-cv
22 Apr 2026
Safety

Machine individuality: Separating genuine idiosyncrasy from response bias in large language models

DGX agent

arXiv:2604.16755v2 Announce Type: replace Abstract: As large language models (LLMs) are increasingly integrated into daily life, in roles ranging from high-stakes decision support to companionship, un

safetyarxiv-cs-ai
22 Apr 2026
← Previous
1…182183184185186…233
Next →