AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
All
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,812 results
Safety

it’s funny how people here just make stuff up.

DGX agent

it’s funny how people here just make stuff up. @GaryMarcus Take away LLM & every single AI application today goes back to the stone age, driverless cars will immediately break down, all the apps will

safetygary-marcus--x
18 May 2026
Safety

just as i feared

DGX agent

just as i feared AI can be a powerful tool for opening people up to new perspectives, yet we find that people actually prefer to use “sycophantic” AI systems that reinforce their pre-existing beliefs.

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
safetygary-marcus--x
18 May 2026
Safety

Learn Where Outcomes Diverge: Efficient VLA RL via Probabilistic Chunk Masking

DGX agent

arXiv:2605.16154v1 Announce Type: new Abstract: Reinforcement learning (RL) allows vision-language-action (VLA) policies to generalize beyond their training distribution by optimizing directly for tas

safetyarxiv-cs-lg
18 May 2026
Safety

Learning Bilevel Policies over Symbolic World Models for Long-Horizon Planning

DGX agent

arXiv:2605.15975v1 Announce Type: new Abstract: We tackle the challenge of building embodied AI agents that can reliably solve long-horizon planning problems. Imitation learning from demonstrations ha

safetyarxiv-cs-ai
18 May 2026
Safety

Learning Context-conditioned Gaussian Overbounds for Convolution-Based Uncertainty Propagation

DGX agent

arXiv:2605.15789v1 Announce Type: new Abstract: Uncertainty quantification is essential in safety-critical settings--from autonomous driving to aviation, finance, and health--where decisions must rely

safetyarxiv-cs-lg
18 May 2026
Safety

Learning in Structured Stackelberg Games

DGX agent

arXiv:2504.09006v4 Announce Type: replace-cross Abstract: We initiate the study of structured Stackelberg games, a novel form of strategic interaction between a leader and a follower where contextual

safetyarxiv-cs-lg
18 May 2026
Safety

Learning Sim-Grounded Policies for Bimanual Rope Manipulation from Human Teleoperation Data

DGX agent

arXiv:2605.16043v1 Announce Type: cross Abstract: Deformable Linear Objects (DLOs) such as ropes and cables are widely encountered in both household and industrial applications, yet remain challenging

safetyarxiv-cs-ai
18 May 2026
Safety

Linked Multi-Model Data on Russian Domestic and Foreign Policy Speeches

DGX agent

arXiv:2605.15886v1 Announce Type: new Abstract: This paper introduces a dataset of interlinked multimodal political communications from the Russian government, addressing persistent deficiencies in th

safetyarxiv-cs-cl
18 May 2026
Safety

LRCP: Low-Rank Compressibility Guided Visual Token Pruning for Efficient LVLMs

DGX agent

arXiv:2605.15621v1 Announce Type: new Abstract: Large vision-language models (LVLMs) achieve strong multimodal understanding, but their inference cost grows rapidly with the number of visual tokens, e

safetyarxiv-cs-cv
18 May 2026
Safety

LUIVITON: Learned Universal Interoperable VIrtual Try-ON

DGX agent

arXiv:2509.05030v2 Announce Type: replace Abstract: To enable large-scale reuse of real-world 3D assets, where garments and characters rarely share skeletons, templates, or dense correspondences, we p

safetyarxiv-cs-cv
18 May 2026
Safety

MAgSeg: Segmentation of Agricultural Landscapes in High-Resolution Satellite Imagery using Multimodal Large Language Models

DGX agent

arXiv:2605.16179v1 Announce Type: new Abstract: Agricultural landscape segmentation in the Global South is challenging as it is characterized by fragmented plots, high intra-class variance, and a scar

safetyarxiv-cs-cv
18 May 2026
Safety

MaTe: Images Are All You Need for Material Transfer via Diffusion Transformer

DGX agent

arXiv:2605.15660v1 Announce Type: new Abstract: Recent diffusion-based methods for material transfer rely on image fine-tuning or complex architectures with assistive networks, but face challenges inc

safetyarxiv-cs-cv
18 May 2026
Safety

Mind Dreamer: Untethering Imagination via Active Latent Intervention on Latent Manifolds

DGX agent

arXiv:2605.16030v1 Announce Type: new Abstract: Model-Based Reinforcement Learning (MBRL) leverages latent imagination for sample efficiency, yet remains constrained by Historical Tethering: imaginati

safetyarxiv-cs-lg
18 May 2026
Safety

Missing-Data-Induced Phase Transitions in Spectral PLS for Multimodal Learning

DGX agent

arXiv:2601.21294v2 Announce Type: replace Abstract: Partial Least Squares (PLS) learns shared structure from paired data via the top singular vectors of the empirical cross-covariance (PLS-SVD), but m

safetyarxiv-cs-lg
18 May 2026
Safety

Monotone and Separable Set Functions: Characterizations and Neural Models

DGX agent

arXiv:2510.23634v3 Announce Type: replace-cross Abstract: Motivated by applications for set containment problems, we consider the following fundamental problem: can we design set-to-vector functions s

safetyarxiv-cs-ai
18 May 2026
Safety

NavRL++: A System-Level Framework for Improving Sim-to-Real Transfer in Reinforcement Learning-Based Robot Navigation

DGX agent

arXiv:2605.15559v1 Announce Type: new Abstract: Recent years have witnessed significant progress in autonomous navigation using reinforcement learning. However, existing approaches largely emphasize r

safetyarxiv-cs-ro
18 May 2026
Safety

Neutral-Reference Prompting for Vision-Language Models

DGX agent

arXiv:2605.15615v1 Announce Type: new Abstract: Efficient transfer learning of vision-language models (VLMs) commonly suffers from a Base-New Trade-off (BNT): improving performance on unseen (new) cla

safetyarxiv-cs-cv
18 May 2026
Safety

Note that as of last summer Waymo said that LLMs/VLMs were experimental, and that their foundation model then could “process only a small am…

DGX agent

Note that as of last summer Waymo said that LLMs/VLMs were experimental, and that their foundation model then could “process only a small amount of image frames, does not incorporate accurate 3D sensi

safetygary-marcus--x
18 May 2026
Safety

Nudging Beyond the Comfort Zone: Efficient Strategy-Guided Exploration for RLVR

DGX agent

arXiv:2605.15726v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a scalable paradigm for improving the reasoning capabilities of large language mode

safetyarxiv-cs-ai
18 May 2026
Safety

Offline Reinforcement Learning with Universal Horizon Models

DGX agent

arXiv:2605.15603v1 Announce Type: cross Abstract: Model-based reinforcement learning (RL) offers a compelling approach to offline RL by enabling value learning on imagined on-policy trajectories. Howe

safetyarxiv-cs-ai
18 May 2026
Safety

OHP-RL: Online Human Preference as Guidance in Reinforcement Learning for Robot Manipulation

DGX agent

arXiv:2605.15971v1 Announce Type: new Abstract: While reinforcement learning (RL) enables robots to acquire skills autonomously, its real-world deployment is severely limited by inefficient and unsafe

safetyarxiv-cs-ro
18 May 2026
Safety

OmniVL-Guard: Towards Unified Vision-Language Forgery Detection and Grounding via Balanced RL

DGX agent

arXiv:2602.10687v3 Announce Type: replace-cross Abstract: Existing forgery detection methods are often limited to uni-modal or bi-modal settings, failing to handle the interleaved text, images, and vi

safetyarxiv-cs-ai
18 May 2026
Safety

On Kernel Eigen-alignments of KRR: Reconstruction and Generalization

DGX agent

arXiv:2605.15240v1 Announce Type: cross Abstract: This paper investigates the critical role of eigenalignments between the kernel matrix and learning targets in achieving robust generalization in lear

safetyarxiv-cs-lg
18 May 2026
Safety

OpenFrontier: General Navigation with Visual-Language Grounded Frontiers

DGX agent

arXiv:2603.05377v2 Announce Type: replace-cross Abstract: Open-world navigation requires robots to make decisions in complex everyday environments while adapting to flexible task requirements. Convent

safetyarxiv-cs-cv
18 May 2026
Safety

Opponent State Inference Under Partial Observability: An HMM-POMDP Framework for 2026 Formula 1 Energy Strategy

DGX agent

arXiv:2603.01290v3 Announce Type: replace Abstract: The 2026 Formula 1 technical regulations introduce a fundamental change to energy strategy: under a 50/50 internal combustion engine / battery power

safetyarxiv-cs-ai
18 May 2026
Safety

parallelcbf: A composable safety-filter and auditability framework for tensor-parallel reinforcement learning

DGX agent

arXiv:2605.15509v1 Announce Type: new Abstract: While Isaac Lab provides massive parallel UAV simulation, OmniSafe and safe-control-gym provide constrained-RL benchmarks, and CBFKit provides control-b

safetyarxiv-cs-lg
18 May 2026
Safety

PCASim: Promptable Closed-loop Adversarial Simulation for Urban Traffic Environment

DGX agent

arXiv:2605.15654v1 Announce Type: new Abstract: Real-world autonomous driving, particularly in urban environments with numerous corner cases, requires rigorous testing to ensure product safety and rob

safetyarxiv-cs-ro
18 May 2026
Safety

Pessimistic Risk-Aware Policy Learning in Contextual Bandits

DGX agent

arXiv:2605.15620v1 Announce Type: cross Abstract: We study risk-aware offline policy learning, aiming to learn a decision rule from logged data that is optimal under general risk criteria. This proble

safetyarxiv-cs-lg
18 May 2026
Safety

phi-Balancing for Mixture-of-Experts Training

DGX agent

arXiv:2605.15403v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models rely on balanced expert utilization to fully realize their scalability. However, existing load-balancing methods are lar

safetyarxiv-cs-lg
18 May 2026
Safety

Polynomial Neural Sheaf Diffusion: A Spectral Filtering Approach on Cellular Sheaves

DGX agent

arXiv:2512.00242v3 Announce Type: replace-cross Abstract: Sheaf Neural Networks equip graph structures with a cellular sheaf: a geometric structure which assigns local vector spaces (stalks) and a lin

safetyarxiv-cs-ai
18 May 2026
Safety

Preconditioned Regularized Wasserstein Proximal Sampling

DGX agent

arXiv:2509.01685v2 Announce Type: replace-cross Abstract: We consider sampling from a Gibbs distribution by evolving finitely many particles. We propose a preconditioned version of a recently proposed

safetyarxiv-cs-lg
18 May 2026
Safety

Propagating Unsafe Actions in LLM Controlled Multi-Robot Collaboration via Single Robot Compromise

DGX agent

arXiv:2605.15641v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as general planners in embodied intelligence, enabling high level coordination and low level task pla

safetyarxiv-cs-ro
18 May 2026
Safety

PSD: Pushing the Pareto Frontier of Diffusion LLMs via Parallel Speculative Decoding

DGX agent

arXiv:2605.15609v1 Announce Type: new Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising masked token sequences. Although dLLMs can predict all masked positions i

safetyarxiv-cs-cl
18 May 2026
Safety

Quantum Artificial Intelligence for Mission-Critical Systems: Foundations, Architectural Elements, and Future Directions

DGX agent

arXiv:2511.09884v2 Announce Type: replace Abstract: Mission critical (MC) applications such as defense operations, energy management, cybersecurity, and aerospace control require reliable, determinist

safetyarxiv-cs-ai
18 May 2026
Safety

RanSOM: Second-Order Momentum with Randomized Scaling for Constrained and Unconstrained Optimization

DGX agent

arXiv:2602.06824v2 Announce Type: replace-cross Abstract: Momentum methods, such as Polyak's Heavy Ball, are the standard for training deep networks but suffer from curvature-induced bias in stochasti

safetyarxiv-cs-lg
18 May 2026
Safety

RE-SAC: Disentangling aleatoric and epistemic risks in bus fleet control: A stable and robust ensemble DRL approach

DGX agent

arXiv:2603.18396v3 Announce Type: replace Abstract: Bus holding control is challenging due to stochastic traffic and passenger demand. While deep reinforcement learning (DRL) shows promise, standard a

safetyarxiv-cs-lg
18 May 2026
Safety

Reactive Robot-Centric Safety for Autonomous Navigation in Constrained and Dynamic Environments

DGX agent

arXiv:2605.15782v1 Announce Type: new Abstract: In this work, we address the problem of ensuring real-time safety in autonomous robot navigation, in spatially constrained dynamic environments, by util

safetyarxiv-cs-ro
18 May 2026
Safety

ReactiveGWM: Steering NPC in Reactive Game World Models

DGX agent

arXiv:2605.15256v1 Announce Type: new Abstract: Current game world models simulate environments from a subjective, player-centric perspective. However, by treating the Non-Player Character (NPC) merel

safetyarxiv-cs-cv
18 May 2026
Safety

ReAlign: Generalizable Image Forgery Detection via Reasoning-Aligned Representation

DGX agent

arXiv:2605.16080v1 Announce Type: new Abstract: The rise of AI-generated images (AIGIs) poses growing challenges for digital authenticity, prompting the need for efficient, generalizable image forgery

safetyarxiv-cs-cv
18 May 2026
Safety

Reference-Free Reinforcement Learning Fine-Tuning for MT: A Seq2Seq Perspective

DGX agent

arXiv:2605.15976v1 Announce Type: cross Abstract: Production machine translation relies overwhelmingly on encoder-decoder Seq2Seq models, yet reinforcement learning approaches to MT fine-tuning have l

safetyarxiv-cs-ai
18 May 2026
Safety

Reference Games as a Testbed for the Alignment of Model Uncertainty and Clarification Requests

DGX agent

arXiv:2601.07820v2 Announce Type: replace Abstract: In human conversation, both interlocutors play an active role in maintaining mutual understanding. When listeners are uncertain about what speakers

safetyarxiv-cs-cl
18 May 2026
Safety

Res^2CLIP: Few-Shot Generalist Anomaly Detection with Residual-to-Residual Alignment

DGX agent

arXiv:2605.16171v1 Announce Type: new Abstract: Few-shot Generalist Anomaly Detection requires models to generalize to novel categories without retraining, posing significant challenges in real-world

safetyarxiv-cs-cv
18 May 2026
Safety

Residual Reinforcement Learning for Robot Teleoperation under Stochastic Delays

DGX agent

arXiv:2605.15480v1 Announce Type: cross Abstract: Stochastic communication delays in teleoperation introduce signal discontinuities that undermine control stability and degrade control performance. Co

safetyarxiv-cs-ai
18 May 2026
Safety

Response-Conditioned Parallel-to-Sequential Orchestration for Multi-Agent Systems

DGX agent

arXiv:2605.15573v1 Announce Type: new Abstract: Multi-agent systems can solve complex tasks through collaboration between multiple Large Language Model agents. Existing collaboration frameworks typica

safetyarxiv-cs-cl
18 May 2026
Safety

Reversing the Flow: Generation-to-Understanding Synergy in Large Multimodal Models

DGX agent

arXiv:2605.15792v1 Announce Type: new Abstract: The long-standing goal of multimodal AI is to build unified models in which visual understanding and visual generation mutually enhance one another. Des

safetyarxiv-cs-cv
18 May 2026
Safety

RoPE Distinguishes Neither Positions Nor Tokens in Long Contexts, Provably

DGX agent

arXiv:2605.15514v1 Announce Type: cross Abstract: We identify intrinsic limitations of Rotary Positional Embeddings (RoPE) in Transformer-based long-context language models. Our theoretical analysis a

safetyarxiv-cs-ai
18 May 2026
Safety

SAFE Quantum Machine Learning with Variational Quantum Classifiers

DGX agent

arXiv:2605.16067v1 Announce Type: new Abstract: We propose a variational quantum classifier operating on high dimensional deep representations via amplitude encoding, stabilized by a learnable classic

safetyarxiv-cs-lg
18 May 2026
Safety

SafeGPT: Preventing Data Leakage and Unethical Outputs in Enterprise LLM Use

DGX agent

arXiv:2601.06366v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are transforming enterprise workflows but introduce security and ethics challenges when employees inadvertently s

safetyarxiv-cs-ai
18 May 2026
← Previous
1…172173174175176…267
Next →