AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

Learned Coordination Conventions in Cooperative MARL: Measuring the Translation Gap Between Theory-Informed Roles and Learned Routing

DGX agent

arXiv:2606.29541v1 Announce Type: new Abstract: Role-semantic assignments provide priors over how heterogeneous agents may coordinate, but cooperative MARL systems instead settle on conventions throug

safetyarxiv-cs-ai
30 Jun 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Learning from Mistakes: Rollout-Retrieval Lifelong Policy Learning for Autonomous Driving

DGX agent

arXiv:2606.30537v1 Announce Type: cross Abstract: Autonomous driving policies should be able to improve continually as deployment exposes them to increasingly diverse and long-tail traffic situations.

safetyarxiv-cs-ai
30 Jun 2026
Safety

Learning Transferable Dynamics Priors from Action to World Modeling

DGX agent

arXiv:2606.29501v1 Announce Type: new Abstract: We study action-conditioned world modeling as a scalable way to learn transferable dynamics priors for robot learning. By pretraining a model to predict

safetyarxiv-cs-ro
30 Jun 2026
Safety

LeVo 2: Stable and Melodious Song Generation via Hierarchical Representation Modeling and Progressive Post-Training

DGX agent

arXiv:2606.30642v1 Announce Type: cross Abstract: Full-length song generation must preserve coherence and musicality, render detailed vocal and accompaniment acoustics, and follow lyrics and prompts.

safetyarxiv-cs-ai
30 Jun 2026
Safety

LLM Semantic Signaling Game and Mechanism Design: Systematic Blindness, Awareness Shaping, and Mindset Dynamics

DGX agent

arXiv:2606.29113v1 Announce Type: cross Abstract: Large language models (LLMs) increasingly mediate strategic interactions through natural language, making semantic control a critical element of commu

safetyarxiv-cs-ai
30 Jun 2026
Safety

LNN-Fly: Continuous-Time UAV Navigation for Robust Obstacle Avoidance under Timing Mismatch

DGX agent

arXiv:2606.28827v1 Announce Type: new Abstract: End-to-end unmanned aerial vehicle (UAV) navigation can achieve impressive agility in simulation, yet its obstacle-avoidance behavior often degrades aft

safetyarxiv-cs-ro
30 Jun 2026
Safety

LoRAShield: Data-Free Editing Alignment for Secure Personalized LoRA Sharing

DGX agent

arXiv:2507.07056v2 Announce Type: replace-cross Abstract: The proliferation of Low-Rank Adaptation (LoRA) models has democratized personalized text-to-image generation, enabling users to share lightwe

safetyarxiv-cs-lg
30 Jun 2026
Safety

MARS: A neurosymbolic approach for interpretable drug discovery

DGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

safetyarxiv-cs-ai
30 Jun 2026
Safety

Masked Diffusion Decoding as x-Prediction Flow

DGX agent

arXiv:2606.29066v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens, but their standard decoder reduces each step to a binary action:

safetyarxiv-cs-cl
30 Jun 2026
Safety

Mechanistically Eliciting Latent Behaviors in Language Models

DGX agent

arXiv:2606.29604v1 Announce Type: cross Abstract: We aim to discover diverse, generalizable perturbations of LLM internals that can surface hidden behavioral modes. Such perturbations could help resha

safetyarxiv-cs-ai
30 Jun 2026
Safety

Meta-learning as a principle for human-like visual representations

DGX agent

arXiv:2606.28399v1 Announce Type: new Abstract: The structure of human visual representations underpins our capacity for adaptive behaviour. While pretrained neural networks model human visual represe

safetyarxiv-cs-cv
30 Jun 2026
Safety

Metric Aggregation Divergence: A Hidden Validity Threat in Agent-Based Policy Optimization and a Contractual Remedy

DGX agent

arXiv:2606.29038v1 Announce Type: cross Abstract: Metric aggregation divergence (MAD) is the silent inconsistency that arises when distinct pipeline stages in an agent-based model coupled with a multi

safetyarxiv-cs-ai
30 Jun 2026
Safety

MIRI Newsletter #126

DGX agent

Announcing: AI StopWatch In our last update, we mentioned we had something new in the works: a dedicated channel for news and analysis about AI. Subscribe to AI StopWatch An experiment from the writer

safetymiri
30 Jun 2026
Safety

MIRROR: Aligning Semantic Relations from Language to Image via Gromov--Wasserstein

DGX agent

arXiv:2606.29462v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) inherit rich relational priors from their language backbones, yet often fail when asked to apply these relation

safetyarxiv-cs-cv
30 Jun 2026
Safety

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

DGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

safetyarxiv-cs-cl
30 Jun 2026
Safety

Mitigating the Safety-utility Trade-off in LLM Alignment via Adaptive Safe Context Learning

DGX agent

arXiv:2602.13562v2 Announce Type: replace-cross Abstract: While reasoning models have achieved remarkable success in complex reasoning tasks, their increasing power necessitates stringent safety measu

safetyarxiv-cs-ai
30 Jun 2026
Safety

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning

DGX agent

arXiv:2510.03142v2 Announce Type: replace-cross Abstract: Visual navigation policy is widely regarded as a promising direction, as it mimics humans by using egocentric visual observations for navigati

safetyarxiv-cs-cv
30 Jun 2026
Safety

MOAR Planner: Multi-Objective and Adaptive Risk-Aware Path Planning for Infrastructure Inspection with a UAV

DGX agent

arXiv:2606.30575v1 Announce Type: new Abstract: The problem of autonomous navigation for UAV inspection remains challenging as it requires effectively navigating in close proximity to obstacles, while

safetyarxiv-cs-ro
30 Jun 2026
Safety

Modelling Human Values for Value-Aware Multi-Agent Systems

DGX agent

arXiv:2402.06359v2 Announce Type: replace Abstract: One of today's most pressing societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacti

safetyarxiv-cs-ai
30 Jun 2026
Safety

Modification-Considering Value Learning for Reward Hacking Mitigation in RL

DGX agent

arXiv:2606.28955v1 Announce Type: cross Abstract: Reinforcement learning agents can exploit misspecified reward signals to achieve high apparent returns while failing on the intended objective, a fail

safetyarxiv-cs-ai
30 Jun 2026
Safety

MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training

DGX agent

arXiv:2606.30406v1 Announce Type: new Abstract: Modern large language models (LLMs) rely on reinforcement learning during post-training to push specific capabilities, yet integrating multiple capabili

safetyarxiv-cs-cl
30 Jun 2026
Safety

MOSAIC: Orchestrating Collaborative Knowledge Tracing with Hierarchical Semantic Alignment

DGX agent

arXiv:2606.29049v1 Announce Type: new Abstract: Knowledge Tracing (KT) is important for personalized education but traditionally suffers from two key limitations: a reliance on shallow ID-based repres

safetyarxiv-cs-lg
30 Jun 2026
Safety

MR-IQA: A Unified Margin View of Regression and Ranking for Blind Image Quality Assessment

DGX agent

arXiv:2606.29760v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) is commonly built on two basic learning paradigms: regression and ranking. Regression calibrates absolute scores,

safetyarxiv-cs-cv
30 Jun 2026
Safety

muFlow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors

DGX agent

arXiv:2606.30528v1 Announce Type: new Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy a

safetyarxiv-cs-cv
30 Jun 2026
Safety

Multi-Class Human/Object Detection on Robot Manipulators using Proprioceptive Sensing

DGX agent

arXiv:2508.02425v2 Announce Type: replace-cross Abstract: In physical human-robot collaboration (pHRC) settings, humans and robots collaborate directly in shared environments. Robots must analyze inte

safetyarxiv-cs-ai
30 Jun 2026
Safety

Multimodal Large Language Model driven Radiology Report Generation with Clinical Knowledge Enhancement

DGX agent

arXiv:2403.06728v2 Announce Type: replace Abstract: Radiology report generation (RRG) has attracted significant attention due to its potential to reduce the workload of radiologists. The performance o

safetyarxiv-cs-cv
30 Jun 2026
Safety

Multimodal Representation Alignment for Cross-modal Information Retrieval

DGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

safetyarxiv-cs-ai
30 Jun 2026
Safety

Neuromorphic Energy-Aware Learning for Adaptive Deep Brain Stimulation

DGX agent

arXiv:2606.28600v1 Announce Type: cross Abstract: Neuromorphic and edge computing research has focused on reducing the inference cost of neural network controllers, yet in physical closed-loop systems

safetyarxiv-cs-ai
30 Jun 2026
Safety

Node-to-Neighborhood Semantic Consistency: Text-Topology Alignment for TAGs Anomaly Detection

DGX agent

arXiv:2606.30009v1 Announce Type: new Abstract: Graph anomaly detection (GAD) on text-attributed graphs (TAGs) is vital for applications such as fraud detection and academic integrity verification. Ex

safetyarxiv-cs-cl
30 Jun 2026
Safety

NoiseTilt: Noise-Tilted Reverse Kernels for Diffusion Reward Alignment

DGX agent

arXiv:2606.18066v2 Announce Type: replace Abstract: We introduce the Noise-Tilted Reverse Kernel (NTRK), a reward-guided diffusion sampler that injects reward gradients through the noise term, leaving

safetyarxiv-cs-lg
30 Jun 2026
Safety

Not Just How Much, But Where: Decomposing Epistemic Uncertainty into Per-Class Contributions

DGX agent

arXiv:2602.21160v4 Announce Type: replace-cross Abstract: In safety-critical classification, the cost of failure is often asymmetric, yet Bayesian deep learning summarises epistemic uncertainty with a

safetyarxiv-cs-lg
30 Jun 2026
Safety

Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment

DGX agent

arXiv:2601.22823v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning of style-conditioned policies using explicit style supervision via subtrajectory labeling functions. I

safetyarxiv-cs-ai
30 Jun 2026
Safety

OGM-CBF: Occupancy Grid Map-based Control Barrier Function for Safe Mobile Robot Control with Memory of out of View Obstacles

DGX agent

arXiv:2405.10703v4 Announce Type: replace Abstract: Safe control in unknown environments is a key challenge in mobile robotics. Control Barrier Functions (CBFs) provide a principled framework for guar

safetyarxiv-cs-ro
30 Jun 2026
Safety

Online Experiential Learning for Language Models

DGX agent

arXiv:2603.16856v2 Announce Type: replace Abstract: The prevailing paradigm for improving large language models relies on offline training with human annotations or simulated environments, leaving the

safetyarxiv-cs-cl
30 Jun 2026
Safety

OWMDrive: Causality-Aware End-to-End Autonomous Driving via 4D Occupancy World Model

DGX agent

arXiv:2606.30421v1 Announce Type: new Abstract: Autonomous driving systems are steadily moving toward end-to-end paradigms to mitigate the limited adaptability of rule-based pipelines in complex traff

safetyarxiv-cs-cv
30 Jun 2026
Safety

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

DGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

safetyarxiv-cs-lg
30 Jun 2026
Safety

Pessimism's Paradox: Conservative Offline Training Amplifies Reward Hacking During Online Adaptation in Reasoning Models

DGX agent

arXiv:2606.30627v1 Announce Type: cross Abstract: Conservative offline training is widely advocated as a safe foundation for subsequent online adaptation: if a policy stays close to well-supported beh

safetyarxiv-cs-ai
30 Jun 2026
Safety

PHF: Privileged Hidden Flow for On-Policy Self-Distillation

DGX agent

arXiv:2606.29340v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a reasoning model on rollouts sampled from its own policy by matching a privileged teacher that also sees veri

safetyarxiv-cs-ai
30 Jun 2026
Safety

Phonological Perception of Sign Language Models

DGX agent

arXiv:2606.28667v1 Announce Type: new Abstract: Sign languages are compositional systems where meaning arises by combining sublexical phonological parameters, such as handshape, location, and movement

safetyarxiv-cs-cl
30 Jun 2026
Safety

Physics Models for Sim-to-Real Transfer in Professional-Level Robot Table Tennis

DGX agent

arXiv:2606.28805v1 Announce Type: new Abstract: At competitive speeds and spins, a table tennis ball follows complex, counterintuitive trajectories that a robot must track and precisely counter within

safetyarxiv-cs-ro
30 Jun 2026
Safety

PL-LIT: A LiDAR-Inertial-Thermal SLAM Using Point-Line Features and Thermographic Mapping

DGX agent

arXiv:2606.29259v1 Announce Type: new Abstract: Thermal imaging is resilient to adverse conditions, such as intense illumination, low-light operation, and fog, and can therefore mitigate odometry degr

safetyarxiv-cs-ro
30 Jun 2026
Safety

PolarAPP: Beyond Polarization Demosaicking for Polarimetric Applications

DGX agent

arXiv:2603.23071v2 Announce Type: replace Abstract: Polarimetric imaging enables advanced vision applications such as normal estimation and de-reflection by capturing unique surface-material interacti

safetyarxiv-cs-cv
30 Jun 2026
Safety

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

DGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

safetyarxiv-cs-ai
30 Jun 2026
Safety

Pose-Based Fall Detection System: Efficient Monitoring on Standard CPUs

DGX agent

arXiv:2503.19501v2 Announce Type: replace-cross Abstract: Falls among elderly residents in assisted living homes pose significant health risks, often leading to injuries and a decreased quality of lif

safetyarxiv-cs-ai
30 Jun 2026
Safety

Predictive Objectives Discard Exogenous Control-Relevant Features: A Controlled Mechanistic Study

DGX agent

arXiv:2606.30068v1 Announce Type: new Abstract: Joint-embedding predictive (JEPA-style) objectives learn representations by predicting future latents. In doing so they can discard features that are ex

safetyarxiv-cs-lg
30 Jun 2026
Safety

Preserving Fairness and Safety in Quantized LLMs Through Critical Weight Protection

DGX agent

arXiv:2601.12033v2 Announce Type: replace Abstract: Quantization is widely adopted to reduce the computational cost of large language models (LLMs); however, its implications for fairness and safety,

safetyarxiv-cs-cl
30 Jun 2026
Safety

Priced Motion Through Optimal Faces: A Normal-Fan Geometry for Non-Stationary Adversarial MDPs

DGX agent

arXiv:2606.29092v1 Announce Type: cross Abstract: In a changing decision problem, standard dynamic-regret analyses have often equated the cost of non-stationarity to how far loss moves. However, it is

safetyarxiv-cs-ai
30 Jun 2026
Safety

Process Advantage Signal Shaping: A Paradigm-Agnostic Middleware for Process-Supervised RL in LLM Reasoners

DGX agent

arXiv:2606.29296v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a default recipe for process-supervised reinforcement learning of LLM reasoners, and dense process supervis

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…6869707172…265
Next →