AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
30 Jun 2026

The Two Genie Game: Adoption and Welfare in Audit-Grounded AI Governance

SafetyDGX agent

arXiv:2606.28710v1 Announce Type: new Abstract: We ask under what conditions an agent with a harm-minimizing policy can displace an approval-seeking (RLHF) agent in a competitive market, and when that

The Undecidability of Artificial General Intelligence (AGI) Alignment

SafetyDGX agent

arXiv:2606.28639v1 Announce Type: cross Abstract: This article establishes the foundational mathematical limits of Artificial General Intelligence (AGI) safety, proving that the core barrier is not th

Theory of Continual Learning Against Data Poisoning Attacks

SafetyDGX agent

arXiv:2606.29841v1 Announce Type: new Abstract: Continual learning (CL), where a model is trained on a sequence of data tasks, is increasingly being adopted across key fields such as large language mo


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Timesteps of Mamba Align with Human Reading Times

SafetyDGX agent

arXiv:2606.29904v1 Announce Type: new Abstract: This study demonstrates an alignment of per-word processing time in a popular state-space language model Mamba and human readers. In Mamba, the recurren

To Use or not to Use Muon: How Simplicity Bias in Optimizers Matters

SafetyDGX agent

arXiv:2603.00742v2 Announce Type: replace Abstract: While Adam has long been the ubiquitous default optimizer for deep neural networks, Muon has recently seen rapid adoption due to its superior traini

Tool Use Enables Undetectable Steganography in Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2606.28425v1 Announce Type: cross Abstract: Increasingly autonomous agentic AI systems pose novel multi-agent risks, such as secret collusion via covert communication channels. The natural defen

Towards Physical Intuitions for Alignment Dynamics: A Case Study With Randomness Crystallization

SafetyDGX agent

arXiv:2606.29933v1 Announce Type: new Abstract: The alignment of language models is typically studied through the lens of capability benchmarks, but the dynamics of how models change during post-train

TrafficAlign: Aligning Large Language Models for Traffic Scenario Generation

SafetyDGX agent

arXiv:2606.29097v1 Announce Type: new Abstract: Recent research has investigated the use of large language models (LLMs) to generate traffic scenarios for autonomous driving. However, pretrained LLMs

TrajRS: Towards Certified Robustness in Pedestrian Trajectory Prediction

SafetyDGX agent

arXiv:2606.28716v1 Announce Type: new Abstract: The robustness of trajectory prediction models is crucial for developing safe autonomous driving systems. Adversarial attacks on trajectory prediction c

Transformer-Based Active Learning for Data-Efficient Vaccine Epitope Selection in PRRS

SafetyDGX agent

arXiv:2606.28659v1 Announce Type: cross Abstract: High-fidelity molecular docking simulations can produce biologically relevant estimates of epitope-receptor binding affinity but are computationally e

Trust Your Instincts: Confidence-Driven Test-Time RL for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.29892v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become indispensable for pushing Vision-Language-Action Models (VLAs) beyond static imitation learning. However, exist

Uncovering Salience-Driven Dynamics in Consumer Confidence with Generative Social Simulation

SafetyDGX agent

arXiv:2606.30395v1 Announce Type: cross Abstract: Consumer confidence is typically modeled as a persistent macroeconomic index, yet its movements arise from households that interpret economic informat

UniGP: Taming Diffusion Transformer for Prior-Preserved Unified Generation and Perception

SafetyDGX agent

arXiv:2606.30332v1 Announce Type: new Abstract: Recent advances in diffusion models have shown impressive performance in controllable image generation and dense prediction tasks. However, existing app

UniMotion: A Unified Framework for Motion-Text-Vision Understanding and Generation

SafetyDGX agent

arXiv:2603.22282v2 Announce Type: replace-cross Abstract: We present UniMotion, to our knowledge the first unified framework for simultaneous understanding and generation of human motion, natural lang

Using Large Language Models as Low-Cost Statistical Estimators for Human-Response Data

SafetyDGX agent

arXiv:2606.30372v1 Announce Type: new Abstract: Quantitative research across the social and behavioral sciences depends on human subject experiments that are expensive, slow, and subject to sampling b

Value-Action Alignment in Large Language Models under Privacy-Prosocial Conflict

SafetyDGX agent

arXiv:2601.03546v2 Announce Type: replace-cross Abstract: Large language models (LLMs) are increasingly used to simulate decision-making tasks involving personal data sharing, where privacy concerns a

Vision-driven Preference Synthesis for Mitigating Hallucinations in VLMs

SafetyDGX agent

arXiv:2606.28401v1 Announce Type: new Abstract: Vision-Language Models (VLMs) have shown strong performance in visual understanding, yet they still suffer from hallucinations, generating content that

Vision-Language-Action Models: Experimental Insights from a Real-World UR5 Platform

SafetyDGX agent

arXiv:2606.30456v1 Announce Type: cross Abstract: This project investigates whether recent Vision-Language-Action (VLA) models can be transferred from controlled research benchmarks to a real-world ro

Vision-Language Models for Deployable Social Robot Navigation: Bridging Semantic Reasoning and Low-Level Control

SafetyDGX agent

arXiv:2606.28760v1 Announce Type: new Abstract: Social robot navigation (SRN) requires more than geometric path planning; it demands understanding human intentions, social norms, and contextual cues t

VISTA-DZ: Visual Semantic Trajectory Adaptation for Personalized Dilemma Zone Prediction

SafetyDGX agent

arXiv:2606.29548v1 Announce Type: cross Abstract: Driver decision making in the dilemma zone at signalized intersections is safety critical, as vehicles approaching a yellow signal must decide whether

Vivid-VR: Distilling Concepts from Text-to-Video Diffusion Transformer for Photorealistic Video Restoration

SafetyDGX agent

arXiv:2508.14483v4 Announce Type: replace Abstract: We present Vivid-VR, a DiT-based generative video restoration method built upon an advanced T2V foundation model, where ControlNet is leveraged to c

VLK: Learning Humanoid Loco-Manipulation from Synthetic Interactions in Reconstructed Scenes

SafetyDGX agent

arXiv:2606.30645v1 Announce Type: cross Abstract: Perception-based humanoid loco-manipulation requires connecting egocentric observations and task instructions to whole-body motion. Learning this mapp

VLOD-TTA: Test-Time Adaptation of Vision-Language Object Detectors

SafetyDGX agent

arXiv:2510.00458v3 Announce Type: replace Abstract: Vision-language object detectors (VLODs) such as YOLO-World and Grounding DINO exhibit strong zero-shot generalization, but their performance degrad

When Does Online Imitation Learning Help in LLM Post-Training? The Role of (Non-)Realizability Beyond Horizon

SafetyDGX agent

arXiv:2606.30445v1 Announce Type: new Abstract: Online imitation learning (IL), particularly on-policy distillation, has emerged as a strong LLM post-training approach, often outperforming offline sup

When Stopping Fails: Rethinking Minimal Risk Conditions through Human-Interactive Autonomous Driving for Safe Transportation Systems

SafetyDGX agent

arXiv:2606.29115v1 Announce Type: cross Abstract: Autonomous vehicles (AVs) are increasingly deployed in urban environments, yet their safety frameworks remain primarily designed around collision avoi

Words Speak Louder Than Code: Investigating Cognitive Heuristics in LLM-Based Code Vulnerability Detection

SafetyDGX agent

arXiv:2606.30587v1 Announce Type: cross Abstract: Researchers and practitioners increasingly apply Large Language Models (LLMs) for automated vulnerability detection. Recent work has shown that LLMs a

WoVR: World Models as Reliable Simulators for Post-Training VLA Policies with RL

SafetyDGX agent

arXiv:2602.13977v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) promises to unlock capabilities beyond imitation learning for Vision--Language--Action (VLA) models, but its requi

X-Mind: Efficient Visual Chain-of-Thought via Predictive World Model for End-to-End Driving

SafetyDGX agent

arXiv:2606.28758v1 Announce Type: cross Abstract: Predicting future states is essential for autonomous agents, yet current Vision-Language-Action (VLA) models fundamentally lack this capability, relyi

X-Morph: Human Motion Priors for Scalable Robot Learning Across Morphologies

SafetyDGX agent

arXiv:2606.30290v1 Announce Type: new Abstract: Recent progress in humanoid behavior models has been driven in large part by abundant human motion data, but comparable motion data is scarce for non-hu

X-Tokenizer: A Multimodal Action Tokenizer for Vision-Language-Action Pretraining

SafetyDGX agent

arXiv:2606.14752v2 Announce Type: replace-cross Abstract: Modern Vision-Language-Action (VLA) models must bridge pretrained vision-language reasoning and precise continuous robot control. Existing act

Zero-Label Driving Scenario Complexity Detection via Joint Embedding Predictive Architecture

SafetyDGX agent

arXiv:2606.28383v1 Announce Type: new Abstract: Identifying complex and safety-critical driving scenarios in large unlabelled datasets is an important but expensive problem. Existing approaches rely o

29 Jun 2026

Agent-Native Immune System: Architecture, Taxonomy, and Engineering

SafetyDGX agent

arXiv:2606.28270v1 Announce Type: new Abstract: The transition from static chat bots to autonomous agents--equipped with persistent memory, tool-use protocols, and multi-agent collaboration--has funda

Algorithms for Deciding the Safety of States in Fully Observable Non-deterministic Problems: Technical Report

SafetyDGX agent

arXiv:2603.15282v2 Announce Type: replace Abstract: Learned action policies are increasingly popular in sequential decision-making, but suffer from a lack of safety guarantees. Recent work introduced

Artificial Intelligence Index Report 2026

SafetyDGX agent

arXiv:2606.15708v2 Announce Type: replace Abstract: Welcome to the ninth edition of the AI Index report. As AI continues to advance rapidly, the question becomes whether the systems built around it ca

ATOD: Annealed Turn-aware On-policy Distillation for Multi-turn Autonomous Agents

SafetyDGX agent

arXiv:2606.27814v1 Announce Type: new Abstract: Training small language-model agents for long-horizon interactive tasks requires both fast imitation and reward-driven improvement. On-policy distillati

BiDeMem: Bidirectional Degradation Memory for Explainable Image Restoration

SafetyDGX agent

arXiv:2606.28112v1 Announce Type: cross Abstract: Degradation-aware prompts, conditions, and latent priors are increasingly used in image restoration, yet they are usually judged by a single endpoint:

CacheMPC: Certified Cached Model Predictive Control for Quadruped Locomotion

SafetyDGX agent

arXiv:2606.28300v1 Announce Type: new Abstract: Model Predictive Control (MPC) is the standard predictive layer in hierarchical quadruped controllers, but the per-cycle QP solve limits the update rate

Can Generative Artificial Intelligence Survive Data Contamination? Theoretical Guarantees under Contaminated Recursive Training

SafetyDGX agent

arXiv:2602.16065v2 Announce Type: replace-cross Abstract: As artificial intelligence (AI)-generated content proliferates, models are increasingly trained on their own outputs, risking progressive degr

Check Yourself Before You Wreck Yourself: Selectively Quitting Improves LLM Agent Safety

SafetyDGX agent

arXiv:2510.16492v4 Announce Type: replace Abstract: As Large Language Model (LLM) agents increasingly operate in complex environments with real-world consequences, their safety becomes critical. While

Cognitive Episodes in LLM Reasoning Traces Enable Interpretable Human Item Difficulty Prediction

SafetyDGX agent

arXiv:2606.28186v1 Announce Type: cross Abstract: Predicting human item difficulty is central to educational assessment, where reliable estimates support fairness and effective test construction. Exis

Conservative Equilibrium Discovery in Offline Game-Theoretic Multiagent Reinforcement Learning

SafetyDGX agent

arXiv:2603.00374v2 Announce Type: replace Abstract: Offline learning of strategies takes data efficiency to its extreme by restricting algorithms to a fixed dataset of state-action trajectories. We co

Contrastive Language-Colored Pointmap Pretraining for Unified 3D Scene Understanding

SafetyDGX agent

arXiv:2604.02546v2 Announce Type: replace Abstract: Pretraining 3D encoders by aligning with Contrastive Language Image Pretraining (CLIP) has emerged as a promising direction to learn generalizable r

CPAgents: Agentic Composite Phenotype Generation for Cardiac Disease Association

SafetyDGX agent

arXiv:2606.28179v1 Announce Type: cross Abstract: Identifying robust associations between cardiac imaging phenotypes and clinical diseases is fundamental to population-scale cardiovascular research an

CSD: Content-aware Speculative Decoding for Efficient Image Generation

SafetyDGX agent

arXiv:2606.27829v1 Announce Type: new Abstract: Speculative decoding (SD) has emerged as a key solution to accelerate the inference of autoregressive models. However, in the field of image generation,

CWI: Composite Humanoid Whole-Body Imitation System for Loco-manipulation

SafetyDGX agent

arXiv:2606.27676v1 Announce Type: new Abstract: Achieving everyday tasks with humanoid robots requires coordinating stable locomotion with versatile manipulation. However, existing whole-body controll

Dangerous Liaisons of Convex Learning and Non-Affine Aggregation

SafetyDGX agent

arXiv:2606.28123v1 Announce Type: new Abstract: Last-iterate convergence and generalization guarantees in first-order convex learning hinge on the monotonicity of the update operator. While linear ave

Data Scaling Laws in Imitation Learning for Robotic Manipulation

SafetyDGX agent

arXiv:2410.18647v4 Announce Type: replace Abstract: Data scaling has revolutionized fields like natural language processing and computer vision, providing models with remarkable generalization capabil

Democratic ICAI: Debating Our Way to Steering Principles from Preferences

SafetyDGX agent

arXiv:2606.28294v1 Announce Type: new Abstract: Preference-based alignment often struggles to capture the reasoning that underlies human judgments. Many evaluations rely on multiple interacting criter

Deployment-Side Adaptiveness in Multi-Horizon Volatility Forecasting

SafetyDGX agent

arXiv:2606.27688v1 Announce Type: cross Abstract: In financial forecasting, predictive performance depends not only on which model is trained, but also on how the trained model is deployed. We study t

DexCompose: Reusing Dexterous Policies for Multi-Task Manipulation with a Single Hand

SafetyDGX agent

arXiv:2606.28323v1 Announce Type: cross Abstract: Dexterous manipulation policies can solve individual skills, but composing them to perform multiple tasks with a single hand remains challenging. Addi

Directing the World: Fast Autoregressive Video Generation with Compositional Human-Camera Control

SafetyDGX agent

arXiv:2606.27964v1 Announce Type: new Abstract: Building interactive world models requires generating realistic videos while maintaining controllable dynamics over long horizons. Autoregressive video

Drifting in the Future: Stabilizing Path Following Drifting on High-Latency Vehicle Systems

SafetyDGX agent

arXiv:2606.27914v1 Announce Type: new Abstract: Autonomously controlling and handling a vehicle at and beyond its stability limit is a mathematically and computationally demanding task. Prior demonstr

Dual-Learning based Penalized Multi-Align Clustering for Multi-View Incomplete and Disorderly Data

SafetyDGX agent

arXiv:2606.27984v1 Announce Type: new Abstract: Multimodal feature fusion can effectively capture complex patterns in real-world data by integrating complementary information from different modalities

DysLexLens: A Low-Resource LLM Framework for Analysing Dyslexic Learners Insights from Online Forums

SafetyDGX agent

arXiv:2606.27619v1 Announce Type: new Abstract: Dyslexic learners increasingly use artificial intelligence (AI) tools to support reading, writing, organisation, and study-related tasks. However, their

EchoSonar-R: A Multi-View Reasoning-Enabled Model for Disease Classification and Report Generation in Echocardiography

SafetyDGX agent

arXiv:2606.28164v1 Announce Type: new Abstract: Echocardiography is the most widely used non-invasive cardiac imaging modality, providing essential information for cardiovascular diagnosis. Interpreti

Exposure Bias Can Alleviate Itself via Directional and Frequency Rectification in Flow Matching

SafetyDGX agent

arXiv:2606.28226v1 Announce Type: cross Abstract: Flow Matching (FM) has achieved remarkable generative performance, yet it suffers from exposure bias due to discrepancies between training and inferen

Fair Classification with Efficient and Post-hoc Controllable Fairness-Accuracy Trade-off

SafetyDGX agent

arXiv:2606.28097v1 Announce Type: new Abstract: Post-hoc controllability of fair machine learning models, the ability to control the trade-off between fairness and accuracy after training, is valuable

GeoFace: Consistent Multi-View Face Generation with Geometry-Constrained Diffusion

SafetyDGX agent

arXiv:2606.27659v1 Announce Type: new Abstract: We present GeoFace, a geometry-constrained multi-view diffusion framework for consistent face generation from a single input. % While recent multi-view

Graph Unfolding and Sampling for Transitory Video Keyframe Selection via Gershgorin Disc Alignment

SafetyDGX agent

arXiv:2408.01859v2 Announce Type: replace Abstract: User-generated videos (UGVs) uploaded from mobile phones to social media sites like YouTube and TikTok are short and non-repetitive. We summarize a

Halt Fast! Early Stopping for Certified Robustness

SafetyDGX agent

arXiv:2606.27694v1 Announce Type: cross Abstract: Randomized Smoothing (RS) provides rigorous robustness guarantees for neural networks without architectural constraints, yet its adoption is limited b

← Previous
1…5657585960…212
Next →