AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
30 Jun 2026

Is Muon as good as they say? We looked beyond training speed and found a hidden cost: Muon loses the simplicity bias of older optimizers lik…

SafetyDGX agent

Muon optimizer shows faster training speeds compared to traditional optimizers, but analysis reveals it sacrifices the simplicity bias that older optimizers maintain, potentially impacting model gener

ITSPACE: Monotone Gaussian Optimal Transport Updates

SafetyDGX agent

arXiv:2606.30523v1 Announce Type: new Abstract: Covariance matrices serve as compact descriptors of feature distributions in many machine-learning pipelines, including domain adaptation and Gaussian e

Keypose Exploration: Efficient Automatic Trajectory Labelling and Cross-Embodiment Policy Transfer

SafetyDGX agent

arXiv:2606.29028v1 Announce Type: new Abstract: Keypose-based manipulation decomposes tasks into critical waypoints to simplify policy learning for long-horizon tasks, but existing approaches rely on


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Knowing Bias, Doing Better: Mitigating Social Bias in LLMs via Know-Bias Neuron Enhancement

SafetyDGX agent

arXiv:2601.21864v2 Announce Type: replace Abstract: Large language models (LLMs) exhibit social biases that reinforce harmful stereotypes, limiting their safe deployment. Most existing debiasing metho

KYON: Semi-Modular Wheel-Legged Quadruped With Agile Bimanual Capability

SafetyDGX agent

arXiv:2606.30243v1 Announce Type: new Abstract: This paper presents KYON, a hybrid wheel-legged quadruped robot equipped with a bimanual upper body for loco-manipulation tasks. The platform features a

L2D2-GS: Learning to Densify for Feedforward Dynamic Gaussian Scene Reconstruction

SafetyDGX agent

arXiv:2606.29374v1 Announce Type: new Abstract: High-fidelity reconstruction of dynamic urban environments is a cornerstone of autonomous driving simulation and large-scale world modeling. While 3D Ga

LaGen: Towards Autoregressive LiDAR Scene Generation

SafetyDGX agent

arXiv:2511.21256v2 Announce Type: replace Abstract: Generative world models for autonomous driving (AD) are of great value in applications such as data augmentation, closed-loop simulation, and safety

Langshaw: Declarative Interaction Protocols Based on Sayso and Conflict

SafetyDGX agent

arXiv:2606.29601v1 Announce Type: cross Abstract: Current languages for specifying multiagent protocols either over-constrain protocol enactments or complicate capturing their meanings. We propose Lan

Latent Actions from Factorized Transition Effects under Agent Ambiguity

SafetyDGX agent

arXiv:2606.30544v1 Announce Type: new Abstract: Latent Action Models (LAMs) learn action-like proxies from observation transitions. However, in multi-object or distractor-rich scenes, these visual eff

LatentRevise: Learning from Zero-Hit Reasoning

SafetyDGX agent

arXiv:2606.29938v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) is bottlenecked by hard prompts on which correct trajectories have low probability, so sampling mi

Lateral String Stability for Vehicle Platoons

SafetyDGX agent

arXiv:2606.29677v1 Announce Type: new Abstract: Connected and automated vehicle (CAV) platooning promises gains in energy efficiency and traffic throughput and, most critically, in safety. These safet

Leader Reward for POMO-Based Neural Combinatorial Optimization

SafetyDGX agent

arXiv:2405.13947v2 Announce Type: replace Abstract: Deep neural networks based on reinforcement learning (RL) for solving combinatorial optimization (CO) problems are developing rapidly and have shown

Learned Coordination Conventions in Cooperative MARL: Measuring the Translation Gap Between Theory-Informed Roles and Learned Routing

SafetyDGX agent

arXiv:2606.29541v1 Announce Type: new Abstract: Role-semantic assignments provide priors over how heterogeneous agents may coordinate, but cooperative MARL systems instead settle on conventions throug

Learning from Mistakes: Rollout-Retrieval Lifelong Policy Learning for Autonomous Driving

SafetyDGX agent

arXiv:2606.30537v1 Announce Type: cross Abstract: Autonomous driving policies should be able to improve continually as deployment exposes them to increasingly diverse and long-tail traffic situations.

Learning Transferable Dynamics Priors from Action to World Modeling

SafetyDGX agent

arXiv:2606.29501v1 Announce Type: new Abstract: We study action-conditioned world modeling as a scalable way to learn transferable dynamics priors for robot learning. By pretraining a model to predict

LeVo 2: Stable and Melodious Song Generation via Hierarchical Representation Modeling and Progressive Post-Training

SafetyDGX agent

arXiv:2606.30642v1 Announce Type: cross Abstract: Full-length song generation must preserve coherence and musicality, render detailed vocal and accompaniment acoustics, and follow lyrics and prompts.

LLM Semantic Signaling Game and Mechanism Design: Systematic Blindness, Awareness Shaping, and Mindset Dynamics

SafetyDGX agent

arXiv:2606.29113v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate strategic interactions through natural language, making semantic control a critical element of commu

LNN-Fly: Continuous-Time UAV Navigation for Robust Obstacle Avoidance under Timing Mismatch

SafetyDGX agent

arXiv:2606.28827v1 Announce Type: new Abstract: End-to-end unmanned aerial vehicle (UAV) navigation can achieve impressive agility in simulation, yet its obstacle-avoidance behavior often degrades aft

LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing

SafetyDGX agent

arXiv:2507.07056v2 Announce Type: replace-cross Abstract: The proliferation of Low-Rank Adaptation (LoRA) models has democratized personalized text-to-image generation, enabling users to share lightwe

MARS: A neurosymbolic approach for interpretable drug discovery

SafetyDGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

Masked Diffusion Decoding as x-Prediction Flow

SafetyDGX agent

arXiv:2606.29066v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens, but their standard decoder reduces each step to a binary action:

Mechanistically Eliciting Latent Behaviors in Language Models

SafetyDGX agent

arXiv:2606.29604v1 Announce Type: cross Abstract: We aim to discover diverse, generalizable perturbations of LLM internals that can surface hidden behavioral modes. Such perturbations could help resha

Meta-learning as a principle for human-like visual representations

SafetyDGX agent

arXiv:2606.28399v1 Announce Type: new Abstract: The structure of human visual representations underpins our capacity for adaptive behaviour. While pretrained neural networks model human visual represe

Metric Aggregation Divergence: A Hidden Validity Threat in Agent-Based Policy Optimization and a Contractual Remedy

SafetyDGX agent

arXiv:2606.29038v1 Announce Type: cross Abstract: Metric aggregation divergence (MAD) is the silent inconsistency that arises when distinct pipeline stages in an agent-based model coupled with a multi

MIRI Newsletter #126

SafetyDGX agent

Announcing: AI StopWatch In our last update, we mentioned we had something new in the works: a dedicated channel for news and analysis about AI. Subscribe to AI StopWatch An experiment from the writer

MIRROR: Aligning Semantic Relations from Language to Image via Gromov--Wasserstein

SafetyDGX agent

arXiv:2606.29462v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) inherit rich relational priors from their language backbones, yet often fail when asked to apply these relation

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

SafetyDGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning

SafetyDGX agent

arXiv:2602.13562v2 Announce Type: replace-cross Abstract: While reasoning models have achieved remarkable success in complex reasoning tasks, their increasing power necessitates stringent safety measu

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning

SafetyDGX agent

arXiv:2510.03142v2 Announce Type: replace-cross Abstract: Visual navigation policy is widely regarded as a promising direction, as it mimics humans by using egocentric visual observations for navigati

MOAR Planner: Multi-Objective and Adaptive Risk-Aware Path Planning for Infrastructure Inspection with a UAV

SafetyDGX agent

arXiv:2606.30575v1 Announce Type: new Abstract: The problem of autonomous navigation for UAV inspection remains challenging as it requires effectively navigating in close proximity to obstacles, while

Modelling Human Values for Value-Aware Multi-Agent Systems

SafetyDGX agent

arXiv:2402.06359v2 Announce Type: replace Abstract: One of today's most pressing societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacti

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

SafetyDGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training

SafetyDGX agent

arXiv:2606.30406v1 Announce Type: new Abstract: Modern large language models (LLMs) rely on reinforcement learning during post-training to push specific capabilities, yet integrating multiple capabili

MOSAIC: Orchestrating Collaborative Knowledge Tracing with Hierarchical Semantic Alignment

SafetyDGX agent

arXiv:2606.29049v1 Announce Type: new Abstract: Knowledge Tracing (KT) is important for personalized education but traditionally suffers from two key limitations: a reliance on shallow ID-based repres

MR-IQA: A Unified Margin View of Regression and Ranking for Blind Image Quality Assessment

SafetyDGX agent

arXiv:2606.29760v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) is commonly built on two basic learning paradigms: regression and ranking. Regression calibrates absolute scores,

muFlow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors

SafetyDGX agent

arXiv:2606.30528v1 Announce Type: new Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy a

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

SafetyDGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

Multimodal Large Language Model driven Radiology Report Generation with Clinical Knowledge Enhancement

SafetyDGX agent

arXiv:2403.06728v2 Announce Type: replace Abstract: Radiology report generation (RRG) has attracted significant attention due to its potential to reduce the workload of radiologists. The performance o

Multimodal Representation Alignment for Cross-modal Information Retrieval

SafetyDGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

Neuromorphic Energy-Aware Learning for Adaptive Deep Brain Stimulation

SafetyDGX agent

arXiv:2606.28600v1 Announce Type: cross Abstract: Neuromorphic and edge computing research has focused on reducing the inference cost of neural network controllers, yet in physical closed-loop systems

Node-to-Neighborhood Semantic Consistency: Text-Topology Alignment for TAGs Anomaly Detection

SafetyDGX agent

arXiv:2606.30009v1 Announce Type: new Abstract: Graph anomaly detection (GAD) on text-attributed graphs (TAGs) is vital for applications such as fraud detection and academic integrity verification. Ex

NoiseTilt: Noise-Tilted Reverse Kernels for Diffusion Reward Alignment

SafetyDGX agent

arXiv:2606.18066v2 Announce Type: replace Abstract: We introduce the Noise-Tilted Reverse Kernel (NTRK), a reward-guided diffusion sampler that injects reward gradients through the noise term, leaving

Not Just How Much, But Where: Decomposing Epistemic Uncertainty into Per-Class Contributions

SafetyDGX agent

arXiv:2602.21160v4 Announce Type: replace-cross Abstract: In safety-critical classification, the cost of failure is often asymmetric, yet Bayesian deep learning summarises epistemic uncertainty with a

Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment

SafetyDGX agent

arXiv:2601.22823v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning of style-conditioned policies using explicit style supervision via subtrajectory labeling functions. I

OGM-CBF: Occupancy Grid Map-based Control Barrier Function for Safe Mobile Robot Control with Memory of out of View Obstacles

SafetyDGX agent

arXiv:2405.10703v4 Announce Type: replace Abstract: Safe control in unknown environments is a key challenge in mobile robotics. Control Barrier Functions (CBFs) provide a principled framework for guar

Online Experiential Learning for Language Models

SafetyDGX agent

arXiv:2603.16856v2 Announce Type: replace Abstract: The prevailing paradigm for improving large language models relies on offline training with human annotations or simulated environments, leaving the

OWMDrive: Causality-Aware End-to-End Autonomous Driving via 4D Occupancy World Model

SafetyDGX agent

arXiv:2606.30421v1 Announce Type: new Abstract: Autonomous driving systems are steadily moving toward end-to-end paradigms to mitigate the limited adaptability of rule-based pipelines in complex traff

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

SafetyDGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

Pessimism's Paradox: Conservative Offline Training Amplifies Reward Hacking During Online Adaptation in Reasoning Models

SafetyDGX agent

arXiv:2606.30627v1 Announce Type: cross Abstract: Conservative offline training is widely advocated as a safe foundation for subsequent online adaptation: if a policy stays close to well-supported beh

PHF: Privileged Hidden Flow for On-Policy Self-Distillation

SafetyDGX agent

arXiv:2606.29340v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a reasoning model on rollouts sampled from its own policy by matching a privileged teacher that also sees veri

Phonological Perception of Sign Language Models

SafetyDGX agent

arXiv:2606.28667v1 Announce Type: new Abstract: Sign languages are compositional systems where meaning arises by combining sublexical phonological parameters, such as handshape, location, and movement

Physics Models for Sim-to-Real Transfer in Professional-Level Robot Table Tennis

SafetyDGX agent

arXiv:2606.28805v1 Announce Type: new Abstract: At competitive speeds and spins, a table tennis ball follows complex, counterintuitive trajectories that a robot must track and precisely counter within

PL-LIT: A LiDAR-Inertial-Thermal SLAM Using Point-Line Features and Thermographic Mapping

SafetyDGX agent

arXiv:2606.29259v1 Announce Type: new Abstract: Thermal imaging is resilient to adverse conditions, such as intense illumination, low-light operation, and fog, and can therefore mitigate odometry degr

PolarAPP: Beyond Polarization Demosaicking for Polarimetric Applications

SafetyDGX agent

arXiv:2603.23071v2 Announce Type: replace Abstract: Polarimetric imaging enables advanced vision applications such as normal estimation and de-reflection by capturing unique surface-material interacti

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

SafetyDGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

Pose-Based Fall Detection System: Efficient Monitoring on Standard CPUs

SafetyDGX agent

arXiv:2503.19501v2 Announce Type: replace-cross Abstract: Falls among elderly residents in assisted living homes pose significant health risks, often leading to injuries and a decreased quality of lif

Predictive Objectives Discard Exogenous Control-Relevant Features: A Controlled Mechanistic Study

SafetyDGX agent

arXiv:2606.30068v1 Announce Type: new Abstract: Joint-embedding predictive (JEPA-style) objectives learn representations by predicting future latents. In doing so they can discard features that are ex

Preserving Fairness and Safety in Quantized LLMs Through Critical Weight Protection

SafetyDGX agent

arXiv:2601.12033v2 Announce Type: replace Abstract: Quantization is widely adopted to reduce the computational cost of large language models (LLMs); however, its implications for fairness and safety,

Priced Motion Through Optimal Faces: A Normal-Fan Geometry for Non-Stationary Adversarial MDPs

SafetyDGX agent

arXiv:2606.29092v1 Announce Type: cross Abstract: In a changing decision problem, standard dynamic-regret analyses have often equated the cost of non-stationarity to how far loss moves. However, it is

Process Advantage Signal Shaping: A Paradigm-Agnostic Middleware for Process-Supervised RL in LLM Reasoners

SafetyDGX agent

arXiv:2606.29296v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a default recipe for process-supervised reinforcement learning of LLM reasoners, and dense process supervis

← Previous
1…5455565758…212
Next →