AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

LNN-Fly: Continuous-Time UAV Navigation for Robust Obstacle Avoidance under Timing Mismatch

DGX agent

arXiv:2606.28827v1 Announce Type: new Abstract: End-to-end unmanned aerial vehicle (UAV) navigation can achieve impressive agility in simulation, yet its obstacle-avoidance behavior often degrades aft

safetyarxiv-cs-ro
30 Jun 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

MARS: A neurosymbolic approach for interpretable drug discovery

DGX agent

arXiv:2410.05289v4 Announce Type: replace Abstract: Background: Neurosymbolic (NeSy) artificial intelligence describes the combination of logic or rule-based techniques with neural networks. Compared

safetyarxiv-cs-ai
30 Jun 2026
Safety

Masked Diffusion Decoding as x-Prediction Flow

DGX agent

arXiv:2606.29066v1 Announce Type: new Abstract: Masked diffusion language models (MDLMs) generate text by iteratively unmasking tokens, but their standard decoder reduces each step to a binary action:

safetyarxiv-cs-cl
30 Jun 2026
Safety

Meta-learning as a principle for human-like visual representations

DGX agent

arXiv:2606.28399v1 Announce Type: new Abstract: The structure of human visual representations underpins our capacity for adaptive behaviour. While pretrained neural networks model human visual represe

safetyarxiv-cs-cv
30 Jun 2026
Safety

Metric Aggregation Divergence: A Hidden Validity Threat in Agent-Based Policy Optimization and a Contractual Remedy

DGX agent

arXiv:2606.29038v1 Announce Type: cross Abstract: Metric aggregation divergence (MAD) is the silent inconsistency that arises when distinct pipeline stages in an agent-based model coupled with a multi

safetyarxiv-cs-ai
30 Jun 2026
Safety

MIRI Newsletter #126

DGX agent

Announcing: AI StopWatch In our last update, we mentioned we had something new in the works: a dedicated channel for news and analysis about AI. Subscribe to AI StopWatch An experiment from the writer

safetymiri
30 Jun 2026
Safety

MIRROR: Aligning Semantic Relations from Language to Image via Gromov--Wasserstein

DGX agent

arXiv:2606.29462v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) inherit rich relational priors from their language backbones, yet often fail when asked to apply these relation

safetyarxiv-cs-cv
30 Jun 2026
Safety

MIThinker: A Plug-and-Play Policy-Optimized Thinker For Motivational Interviewing Counseling

DGX agent

arXiv:2606.29265v1 Announce Type: new Abstract: Reasoning large language models (LLMs) have recently made much progress in complex problem-solving, leveraging internal reasoning (or thought) to guide

safetyarxiv-cs-cl
30 Jun 2026
Safety

MM-Nav: Multi-View VLA Model for Robust Visual Navigation via Multi-Expert Learning

DGX agent

arXiv:2510.03142v2 Announce Type: replace-cross Abstract: Visual navigation policy is widely regarded as a promising direction, as it mimics humans by using egocentric visual observations for navigati

safetyarxiv-cs-cv
30 Jun 2026
Safety

Modelling Human Values for Value-Aware Multi-Agent Systems

DGX agent

arXiv:2402.06359v2 Announce Type: replace Abstract: One of today's most pressing societal challenges is building AI systems whose behaviour, or the behaviour it enables within communities of interacti

safetyarxiv-cs-ai
30 Jun 2026
Safety

MOPD: Multi-Teacher On-Policy Distillation for Capability Integration in LLM Post-Training

DGX agent

arXiv:2606.30406v1 Announce Type: new Abstract: Modern large language models (LLMs) rely on reinforcement learning during post-training to push specific capabilities, yet integrating multiple capabili

safetyarxiv-cs-cl
30 Jun 2026
Safety

MOSAIC: Orchestrating Collaborative Knowledge Tracing with Hierarchical Semantic Alignment

DGX agent

arXiv:2606.29049v1 Announce Type: new Abstract: Knowledge Tracing (KT) is important for personalized education but traditionally suffers from two key limitations: a reliance on shallow ID-based repres

safetyarxiv-cs-lg
30 Jun 2026
Safety

MR-IQA: A Unified Margin View of Regression and Ranking for Blind Image Quality Assessment

DGX agent

arXiv:2606.29760v1 Announce Type: new Abstract: Blind image quality assessment (BIQA) is commonly built on two basic learning paradigms: regression and ranking. Regression calibrates absolute scores,

safetyarxiv-cs-cv
30 Jun 2026
Safety

muFlow: Leveraging Average Images for Improving Generalisation of Deepfake Faces Detectors

DGX agent

arXiv:2606.30528v1 Announce Type: new Abstract: Current generative models, including GANs and diffusion models, have reached an outstanding level of photorealism, posing significant risks to privacy a

safetyarxiv-cs-cv
30 Jun 2026
Safety

Multimodal Large Language Model driven Radiology Report Generation with Clinical Knowledge Enhancement

DGX agent

arXiv:2403.06728v2 Announce Type: replace Abstract: Radiology report generation (RRG) has attracted significant attention due to its potential to reduce the workload of radiologists. The performance o

safetyarxiv-cs-cv
30 Jun 2026
Safety

Multimodal Representation Alignment for Cross-modal Information Retrieval

DGX agent

arXiv:2506.08774v2 Announce Type: replace-cross Abstract: Different machine learning models can represent the same underlying concept in different ways. This variability is particularly valuable for i

safetyarxiv-cs-ai
30 Jun 2026
Safety

Neuromorphic Energy-Aware Learning for Adaptive Deep Brain Stimulation

DGX agent

arXiv:2606.28600v1 Announce Type: cross Abstract: Neuromorphic and edge computing research has focused on reducing the inference cost of neural network controllers, yet in physical closed-loop systems

safetyarxiv-cs-ai
30 Jun 2026
Safety

Node-to-Neighborhood Semantic Consistency: Text-Topology Alignment for TAGs Anomaly Detection

DGX agent

arXiv:2606.30009v1 Announce Type: new Abstract: Graph anomaly detection (GAD) on text-attributed graphs (TAGs) is vital for applications such as fraud detection and academic integrity verification. Ex

safetyarxiv-cs-cl
30 Jun 2026
Safety

NoiseTilt: Noise-Tilted Reverse Kernels for Diffusion Reward Alignment

DGX agent

arXiv:2606.18066v2 Announce Type: replace Abstract: We introduce the Noise-Tilted Reverse Kernel (NTRK), a reward-guided diffusion sampler that injects reward gradients through the noise term, leaving

safetyarxiv-cs-lg
30 Jun 2026
Safety

Offline Reinforcement Learning of High-Quality Behaviors Under Robust Style Alignment

DGX agent

arXiv:2601.22823v2 Announce Type: replace-cross Abstract: We study offline reinforcement learning of style-conditioned policies using explicit style supervision via subtrajectory labeling functions. I

safetyarxiv-cs-ai
30 Jun 2026
Safety

Online Experiential Learning for Language Models

DGX agent

arXiv:2603.16856v2 Announce Type: replace Abstract: The prevailing paradigm for improving large language models relies on offline training with human annotations or simulated environments, leaving the

safetyarxiv-cs-cl
30 Jun 2026
Safety

Persona-Trained Monte Carlo: Estimating Market-Outcome Distributions via Swarms of Persona-Conditioned Neural Policy Bots in a Limit Order Book

DGX agent

arXiv:2606.29556v1 Announce Type: new Abstract: We propose Persona-Trained Monte Carlo (PTMC), a method for estimating distributions of market-outcome statistics by repeatedly simulating limit-order-b

safetyarxiv-cs-lg
30 Jun 2026
Safety

Pessimism's Paradox: Conservative Offline Training Amplifies Reward Hacking During Online Adaptation in Reasoning Models

DGX agent

arXiv:2606.30627v1 Announce Type: cross Abstract: Conservative offline training is widely advocated as a safe foundation for subsequent online adaptation: if a policy stays close to well-supported beh

safetyarxiv-cs-ai
30 Jun 2026
Safety

PHF: Privileged Hidden Flow for On-Policy Self-Distillation

DGX agent

arXiv:2606.29340v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) trains a reasoning model on rollouts sampled from its own policy by matching a privileged teacher that also sees veri

safetyarxiv-cs-ai
30 Jun 2026
Safety

Phonological Perception of Sign Language Models

DGX agent

arXiv:2606.28667v1 Announce Type: new Abstract: Sign languages are compositional systems where meaning arises by combining sublexical phonological parameters, such as handshape, location, and movement

safetyarxiv-cs-cl
30 Jun 2026
Safety

Physics Models for Sim-to-Real Transfer in Professional-Level Robot Table Tennis

DGX agent

arXiv:2606.28805v1 Announce Type: new Abstract: At competitive speeds and spins, a table tennis ball follows complex, counterintuitive trajectories that a robot must track and precisely counter within

safetyarxiv-cs-ro
30 Jun 2026
Safety

PolarAPP: Beyond Polarization Demosaicking for Polarimetric Applications

DGX agent

arXiv:2603.23071v2 Announce Type: replace Abstract: Polarimetric imaging enables advanced vision applications such as normal estimation and de-reflection by capturing unique surface-material interacti

safetyarxiv-cs-cv
30 Jun 2026
Safety

Pondering the Way: Spatial-perceiving World Action Model for Embodied Navigation

DGX agent

arXiv:2606.29908v1 Announce Type: cross Abstract: Existing world model-based planners for visual navigation typically follow a verification-centric paradigm, decoupling goal intent from trajectory syn

safetyarxiv-cs-ai
30 Jun 2026
Safety

Predictive Objectives Discard Exogenous Control-Relevant Features: A Controlled Mechanistic Study

DGX agent

arXiv:2606.30068v1 Announce Type: new Abstract: Joint-embedding predictive (JEPA-style) objectives learn representations by predicting future latents. In doing so they can discard features that are ex

safetyarxiv-cs-lg
30 Jun 2026
Safety

Priced Motion Through Optimal Faces: A Normal-Fan Geometry for Non-Stationary Adversarial MDPs

DGX agent

arXiv:2606.29092v1 Announce Type: cross Abstract: In a changing decision problem, standard dynamic-regret analyses have often equated the cost of non-stationarity to how far loss moves. However, it is

safetyarxiv-cs-ai
30 Jun 2026
Safety

Process Advantage Signal Shaping: A Paradigm-Agnostic Middleware for Process-Supervised RL in LLM Reasoners

DGX agent

arXiv:2606.29296v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is a default recipe for process-supervised reinforcement learning of LLM reasoners, and dense process supervis

safetyarxiv-cs-ai
30 Jun 2026
Safety

PromptGNN-sim: Deep Fusion and Alignment of GNN and LLMs for Text-Attributed Graph Learning

DGX agent

arXiv:2606.30291v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs) combine textual semantics with graph structure and are central to many graph learning tasks. However, existing fusion meth

safetyarxiv-cs-ai
30 Jun 2026
Safety

ProSpec RL: Plan Ahead, then Execute

DGX agent

arXiv:2407.21359v2 Announce Type: replace-cross Abstract: Imagining potential outcomes of actions before execution helps agents make more informed decisions, a prospective thinking ability fundamental

safetyarxiv-cs-ai
30 Jun 2026
Safety

PS-PPO: Prefix-Sampling PPO for Critic-Free RLHF

DGX agent

arXiv:2606.29758v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) for Large Language Models increasingly relies on critic-free methods as a practical alternative to a

safetyarxiv-cs-ai
30 Jun 2026
Safety

Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

DGX agent

arXiv:2606.29464v1 Announce Type: cross Abstract: Vision-language dataset distillation (VLDD) compresses a large image-text paired dataset into a small set of synthetic pairs that can efficiently trai

safetyarxiv-cs-ai
30 Jun 2026
Safety

ReactiveBFM: Reactive Closed-Loop Motion Planning Towards Universal Humanoid Whole-Body Control

DGX agent

arXiv:2606.30362v1 Announce Type: cross Abstract: While current Behavior Foundation Models (BFMs) provide robust control priors for humanoids, they only execute pre-defined reference motions. As a res

safetyarxiv-cs-ai
30 Jun 2026
Safety

REAR: Test-time Preference Realignment through Reward Decomposition

DGX agent

arXiv:2606.30339v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse user preferences is a critical yet challenging task. While post-training methods can adapt models to

safetyarxiv-cs-cl
30 Jun 2026
Safety

Regime-Aware Peer Specialization for Robust RAG under Heterogeneous Knowledge Conflicts

DGX agent

arXiv:2606.30518v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language models by grounding generation in external context. However, it can be fragile when the retrieved

safetyarxiv-cs-cl
30 Jun 2026
Safety

ReGuide: From Test-Time Guidance to Self-Improving Diffusion Policies

DGX agent

arXiv:2606.28939v1 Announce Type: new Abstract: Behavior-cloned diffusion policies are expressive but remain vulnerable to covariate shift: small deviations from demonstrated states can compound into

safetyarxiv-cs-lg
30 Jun 2026
Safety

ReMAP-PET: Beyond Visual Understanding -- Learning Region-Guided Metabolic Alignment Semantics from Brain PET

DGX agent

arXiv:2606.29577v1 Announce Type: cross Abstract: Positron Emission Tomography (PET) reveals brain metabolism and is clinically central to neurodegenerative disease assessment, yet existing 3D brain f

safetyarxiv-cs-ai
30 Jun 2026
Safety

RePer-360: Releasing Perspective Priors for 360^irc Depth Estimation via Self-Modulation

DGX agent

arXiv:2603.05999v2 Announce Type: replace Abstract: Recent depth foundation models trained on perspective imagery achieve strong performance, yet generalize poorly to 360^irc images due to the substan

safetyarxiv-cs-cv
30 Jun 2026
Safety

Reproducing FACTER: Fairness via Conformal Thresholding and Prompt Repair

DGX agent

arXiv:2606.28620v1 Announce Type: cross Abstract: Fayyazi et al. (2025) recently proposed FACTER, a model-agnostic framework designed to jointly enforce fairness and statistical coverage in LLM-based

safetyarxiv-cs-lg
30 Jun 2026
Safety

Resolution Thresholds in VLM Detection of Harmful ASCII Art Across Construction Modes and Languages

DGX agent

arXiv:2606.29649v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) are increasingly deployed as content moderation tools, yet they remain vulnerable to jailbreak attacks in which harm

safetyarxiv-cs-cl
30 Jun 2026
Safety

RetrDex: Efficient Object Retrieval in Cluttered Scenes with a Dexterous Hand

DGX agent

arXiv:2502.18423v3 Announce Type: replace Abstract: Retrieving objects buried beneath clutter is both challenging and time-consuming, as complex support relationships make manipulation particularly di

safetyarxiv-cs-ro
30 Jun 2026
Safety

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation

DGX agent

arXiv:2602.09305v2 Announce Type: replace Abstract: Large Language Models (LLMs) demonstrate transformative potential, yet their reasoning remains inconsistent and unreliable. Reinforcement learning (

safetyarxiv-cs-lg
30 Jun 2026
Safety

Rigel: Self-Distilled Score Adaptation for Image and Video Captioning Evaluation

DGX agent

arXiv:2606.29997v1 Announce Type: new Abstract: Automatic evaluation of image and video captioning is essential for benchmarking multimodal systems, although standard evaluation metrics show limited a

safetyarxiv-cs-cv
30 Jun 2026
Safety

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

DGX agent

arXiv:2606.29934v1 Announce Type: new Abstract: Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinf

safetyarxiv-cs-ro
30 Jun 2026
Safety

RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought

DGX agent

arXiv:2606.15753v2 Announce Type: replace Abstract: Embodied reasoning requires models to perceive task-relevant objects and spaces in physical environments and maintain consistent visual grounding th

safetyarxiv-cs-ai
30 Jun 2026
← Previous
1…120121122123124…302
Next →