AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
Safety

PromptGNN-sim: Deep Fusion and Alignment of GNN and LLMs for Text-Attributed Graph Learning

DGX agent

arXiv:2606.30291v1 Announce Type: new Abstract: Text-Attributed Graphs (TAGs) combine textual semantics with graph structure and are central to many graph learning tasks. However, existing fusion meth

safetyarxiv-cs-ai
30 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Proof-of-Guardrail in AI Agents and What (Not) to Trust from It

DGX agent

arXiv:2603.05786v2 Announce Type: replace-cross Abstract: As AI agents become widely deployed as online services, users often rely on an agent developer's claim about how safety is enforced, which int

safetyarxiv-cs-ai
30 Jun 2026
Safety

Propagation of~Interval Belief Structures and~Imprecise Copulas for~Neural Network Verification

DGX agent

arXiv:2606.30105v1 Announce Type: new Abstract: Quantitative verification of neural networks requires reasoning about probabilities under substantial uncertainty in both input distributions and their

safetyarxiv-cs-ai
30 Jun 2026
Safety

ProSpec RL: Plan Ahead, then Execute

DGX agent

arXiv:2407.21359v2 Announce Type: replace-cross Abstract: Imagining potential outcomes of actions before execution helps agents make more informed decisions, a prospective thinking ability fundamental

safetyarxiv-cs-ai
30 Jun 2026
Safety

PS-PPO: Prefix-Sampling PPO for Critic-Free RLHF

DGX agent

arXiv:2606.29758v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) for Large Language Models increasingly relies on critic-free methods as a practical alternative to a

safetyarxiv-cs-ai
30 Jun 2026
Safety

Rank-Aware Hyperbolic Alignment for Vision-Language Dataset Distillation

DGX agent

arXiv:2606.29464v1 Announce Type: cross Abstract: Vision-language dataset distillation (VLDD) compresses a large image-text paired dataset into a small set of synthetic pairs that can efficiently trai

safetyarxiv-cs-ai
30 Jun 2026
Safety

ReactiveBFM: Reactive Closed-Loop Motion Planning Towards Universal Humanoid Whole-Body Control

DGX agent

arXiv:2606.30362v1 Announce Type: cross Abstract: While current Behavior Foundation Models (BFMs) provide robust control priors for humanoids, they only execute pre-defined reference motions. As a res

safetyarxiv-cs-ai
30 Jun 2026
Safety

REAR: Test-time Preference Realignment through Reward Decomposition

DGX agent

arXiv:2606.30339v1 Announce Type: new Abstract: Aligning large language models (LLMs) with diverse user preferences is a critical yet challenging task. While post-training methods can adapt models to

safetyarxiv-cs-cl
30 Jun 2026
Safety

Regime-Aware Peer Specialization for Robust RAG under Heterogeneous Knowledge Conflicts

DGX agent

arXiv:2606.30518v1 Announce Type: new Abstract: Retrieval-augmented generation (RAG) improves language models by grounding generation in external context. However, it can be fragile when the retrieved

safetyarxiv-cs-cl
30 Jun 2026
Safety

ReGuide: From Test-Time Guidance to Self-Improving Diffusion Policies

DGX agent

arXiv:2606.28939v1 Announce Type: new Abstract: Behavior-cloned diffusion policies are expressive but remain vulnerable to covariate shift: small deviations from demonstrated states can compound into

safetyarxiv-cs-lg
30 Jun 2026
Safety

ReMAP-PET: Beyond Visual Understanding -- Learning Region-Guided Metabolic Alignment Semantics from Brain PET

DGX agent

arXiv:2606.29577v1 Announce Type: cross Abstract: Positron Emission Tomography (PET) reveals brain metabolism and is clinically central to neurodegenerative disease assessment, yet existing 3D brain f

safetyarxiv-cs-ai
30 Jun 2026
Safety

RePer-360: Releasing Perspective Priors for 360^irc Depth Estimation via Self-Modulation

DGX agent

arXiv:2603.05999v2 Announce Type: replace Abstract: Recent depth foundation models trained on perspective imagery achieve strong performance, yet generalize poorly to 360^irc images due to the substan

safetyarxiv-cs-cv
30 Jun 2026
Safety

Reproducing FACTER: Fairness via Conformal Thresholding and Prompt Repair

DGX agent

arXiv:2606.28620v1 Announce Type: cross Abstract: Fayyazi et al. (2025) recently proposed FACTER, a model-agnostic framework designed to jointly enforce fairness and statistical coverage in LLM-based

safetyarxiv-cs-lg
30 Jun 2026
Safety

Resolution Thresholds in VLM Detection of Harmful ASCII Art Across Construction Modes and Languages

DGX agent

arXiv:2606.29649v1 Announce Type: new Abstract: Large Vision-Language Models (VLMs) are increasingly deployed as content moderation tools, yet they remain vulnerable to jailbreak attacks in which harm

safetyarxiv-cs-cl
30 Jun 2026
Safety

RetrDex: Efficient Object Retrieval in Cluttered Scenes with a Dexterous Hand

DGX agent

arXiv:2502.18423v3 Announce Type: replace Abstract: Retrieving objects buried beneath clutter is both challenging and time-consuming, as complex support relationships make manipulation particularly di

safetyarxiv-cs-ro
30 Jun 2026
Safety

Reward Modeling for Reinforcement Learning-Based LLM Reasoning: Design, Challenges, and Evaluation

DGX agent

arXiv:2602.09305v2 Announce Type: replace Abstract: Large Language Models (LLMs) demonstrate transformative potential, yet their reasoning remains inconsistent and unreliable. Reinforcement learning (

safetyarxiv-cs-lg
30 Jun 2026
Safety

Rigel: Self-Distilled Score Adaptation for Image and Video Captioning Evaluation

DGX agent

arXiv:2606.29997v1 Announce Type: new Abstract: Automatic evaluation of image and video captioning is essential for benchmarking multimodal systems, although standard evaluation metrics show limited a

safetyarxiv-cs-cv
30 Jun 2026
Safety

RoamFlow: Reinforcement-Aligned One-Step Action MeanFlow Policy for Image-Goal Navigation

DGX agent

arXiv:2606.29934v1 Announce Type: new Abstract: Image-goal navigation is a key challenge in embodied robotics, where an agent must reach a target specified solely by a goal image. While existing reinf

safetyarxiv-cs-ro
30 Jun 2026
Safety

RoboPIN: Grounded Embodied Reasoning via Pinned Chain-of-Thought

DGX agent

arXiv:2606.15753v2 Announce Type: replace Abstract: Embodied reasoning requires models to perceive task-relevant objects and spaces in physical environments and maintain consistent visual grounding th

safetyarxiv-cs-ai
30 Jun 2026
Safety

Robotic Arm-Based Spectral Sensing for Strawberry Positioning and Non-Destructive Sweetness Measurement

DGX agent

arXiv:2606.28555v1 Announce Type: new Abstract: Accurate assessment of sweetness is essential for quality control in agriculture, yet conventional methods rely on destructive sampling and are difficul

safetyarxiv-cs-ro
30 Jun 2026
Safety

Robust Extended Kalman Filter for Land Navigation Using Massive Array of MEMS IMUs

DGX agent

arXiv:2606.29271v1 Announce Type: cross Abstract: We propose a robust Extended Kalman Filter (EKF) architecture for land navigation using an array of hundreds of low-cost micro-electromechanical syste

safetyarxiv-cs-ro
30 Jun 2026
Safety

Robust Strategic Classification under Decision-Dependent Cost Uncertainty

DGX agent

arXiv:2606.30136v1 Announce Type: new Abstract: Humans facing algorithmic decision systems have been found to ``game'' them by altering their input data (at a cost to them) in order to favorably chang

safetyarxiv-cs-lg
30 Jun 2026
Safety

Robust Trajectory Distillation: Hybrid Reweighting Meets Teacher-Inspired Targets

DGX agent

arXiv:2606.29837v1 Announce Type: new Abstract: Dataset distillation (DD) condenses large corpora into compact, information-rich subsets for efficient training and reuse. However, under noisy supervis

safetyarxiv-cs-cv
30 Jun 2026
Safety

Robust Zero-shot Anomaly Detection under Limited Auxiliary Anomaly Priors

DGX agent

arXiv:2606.29428v1 Announce Type: new Abstract: Zero-shot anomaly detection aims to identify defects in arbitrary novel domains; however, existing models assume that the auxiliary data contains a rich

safetyarxiv-cs-cv
30 Jun 2026
Safety

Robustness and Structure Preservation in Flow-Based Generative Models via Wasserstein Path-Space Divergences

DGX agent

arXiv:2410.01244v2 Announce Type: replace-cross Abstract: We introduce a novel Wasserstein-1 (W_1) path-space divergence for stochastic and deterministic dynamics and establish a Wasserstein Uncertain

safetyarxiv-cs-lg
30 Jun 2026
Safety

SA-VLA: State-aware tokenizer for improving Vision-Language-Action Models' performance

DGX agent

arXiv:2606.30113v1 Announce Type: cross Abstract: Discrete action tokenization provides a compact interface for autoregressive VLA policies, but accurately recovering continuous robot actions from dis

safetyarxiv-cs-ai
30 Jun 2026
Safety

SAFE-DiT: Semantics-Aware Fast-path Execution for High-Resolution Diffusion Transformers

DGX agent

arXiv:2606.29360v1 Announce Type: new Abstract: High-resolution Diffusion Transformer (DiT) inference contains substantial spatial redundancy, but many spatially adaptive implementations encode region

safetyarxiv-cs-cv
30 Jun 2026
Safety

Safety from Honesty in a Disinterested AI Predictor

DGX agent

arXiv:2606.29657v1 Announce Type: new Abstract: As AI systems become more capable, training procedures that optimize for downstream outcomes risk introducing implicit agency: goal-directed behavior th

safetyarxiv-cs-ai
30 Jun 2026
Safety

SARLO-80: Worldwide Slant SAR Language Optic Dataset 80cm

DGX agent

arXiv:2606.20523v2 Announce Type: replace-cross Abstract: Multimodal foundation models have advanced rapidly thanks to large optical benchmarks, but comparable resources for synthetic aperture radar (

safetyarxiv-cs-ai
30 Jun 2026
Safety

SEAD: Competence-Aware On-Policy Distillation via Entropy-Guided Supervision

DGX agent

arXiv:2606.28562v1 Announce Type: new Abstract: On-policy distillation (OPD) has a property absent in offline distillation and RL: teacher supervision quality depends on student competence. Incoherent

safetyarxiv-cs-cl
30 Jun 2026
Safety

Seeing Touch from Motion: A Unified Modality-Aware Visuo-Tactile Policy with Tactile Motion Correlation

DGX agent

arXiv:2606.29941v1 Announce Type: cross Abstract: Visuo-Tactile policies leveraging optical tactile sensors have shown great promise in contact-rich manipulation. These sensors achieve high spatial re

safetyarxiv-cs-cv
30 Jun 2026
Safety

Self-Organized Conformal Prediction: Reducing Regional Coverage Gaps with Unsupervised Group Discovery

DGX agent

arXiv:2606.29403v1 Announce Type: cross Abstract: Conformal prediction guarantees marginal coverage, but pooled calibration averages over heterogeneous regions and can mask regional undercoverage in s

safetyarxiv-cs-ai
30 Jun 2026
Safety

Sequential Fairness Auditing with Limited Output Access

DGX agent

arXiv:2606.30338v1 Announce Type: new Abstract: External evaluations are becoming increasingly central to the governance of AI systems. In practice, however, independent auditors often have limited ac

safetyarxiv-cs-ai
30 Jun 2026
Safety

SGD Provably Prioritizes a Shortcut Spurious Feature in the XOR Model

DGX agent

arXiv:2606.30444v1 Announce Type: cross Abstract: Neural networks are known to be susceptible to over-reliance on spurious correlations. However, the precise mechanism by which models exploit shortcut

safetyarxiv-cs-lg
30 Jun 2026
Safety

SIR: Structured Image Representations for Explainable Robot Learning

DGX agent

arXiv:2606.30101v1 Announce Type: cross Abstract: Existing robot policies based on learned visual embeddings lack explicit structure and are sensitive to visual distractions. Thus, the representations

safetyarxiv-cs-cv
30 Jun 2026
Safety

“Sometimes I wonder, maybe everything is conscious, or nothing is conscious. I prefer the former; it just seems more fun.” – Elon Musk

DGX agent

“Sometimes I wonder, maybe everything is conscious, or nothing is conscious. I prefer the former; it just seems more fun.” – Elon Musk Media Neuralink has solved through-dura electrode implantation! T

safetyelon-musk--x
30 Jun 2026
Safety

SOTAlign: Semi-Supervised Alignment of Unimodal Vision and Language Models via Optimal Transport

DGX agent

arXiv:2602.23353v2 Announce Type: replace-cross Abstract: The Platonic Representation Hypothesis posits that neural networks trained on different modalities converge toward a shared statistical model

safetyarxiv-cs-ai
30 Jun 2026
Safety

SP-CACW: Convergence-Aware Client Weighting for Selfish Personalized Learning

DGX agent

arXiv:2606.29322v1 Announce Type: new Abstract: Collaborative learning is sustainable only when it benefits each participant. Standard federated learning optimizes a global average objective, which ca

safetyarxiv-cs-lg
30 Jun 2026
Safety

SPACE: Swarm Pheromone Fields for Adaptive Collision-Aware Exploration

DGX agent

arXiv:2606.29372v1 Announce Type: new Abstract: Massive robot swarms can explore unknown environments quickly, but adding robots eventually stops helping. Doorways and dense traffic create congestion,

safetyarxiv-cs-ro
30 Jun 2026
Safety

SPARC: Scalable Path-Specific Counterfactual Fairness via Causal Conditional Independence

DGX agent

arXiv:2412.04739v2 Announce Type: replace Abstract: Deep learning models exhibit fairness concerns when predictions are inadvertently influenced by sensitive attributes. However, existing attempts to

safetyarxiv-cs-cv
30 Jun 2026
Safety

Sparse Autoencoders are Capable LLM Jailbreak Mitigators

DGX agent

arXiv:2602.12418v2 Announce Type: replace-cross Abstract: Jailbreak attacks remain a persistent threat to large language model safety. We propose Context-Conditioned Delta Steering (CC-Delta), an SAE-

safetyarxiv-cs-cl
30 Jun 2026
Safety

SpecMind: Cognitively Inspired, Interactive Multi-Turn Framework for Postcondition Inference

DGX agent

arXiv:2602.20610v3 Announce Type: replace-cross Abstract: Specifications are vital for ensuring program correctness, yet writing them manually remains challenging and time-intensive. Recent large lang

safetyarxiv-cs-cl
30 Jun 2026
Safety

Spectral phase transitions and trainability in neural network learning dynamics

DGX agent

arXiv:2606.28486v1 Announce Type: cross Abstract: The emergence of low-dimensional structures in the spectra of neural network weight matrices is a common empirical feature of trained models, but the

safetyarxiv-cs-lg
30 Jun 2026
Safety

Sphere-VIO: Fast and Robust Visual-Inertial Odometry via Unified Spherical Representation for Heterogeneous Multi-Camera Systems

DGX agent

arXiv:2606.29910v1 Announce Type: new Abstract: Multi-camera visual-inertial odometry (VIO) overcomes the inherent limitations of pure visual systems by expanding the field of view. However, existing

safetyarxiv-cs-ro
30 Jun 2026
Safety

SRUM: Fine-Grained Self-Rewarding for Unified Multimodal Models

DGX agent

arXiv:2510.12784v2 Announce Type: replace-cross Abstract: Recently, remarkable progress has been made in Unified Multimodal Models (UMMs), which integrate vision-language generation and understanding

safetyarxiv-cs-cl
30 Jun 2026
Safety

StackingNet: Collective Inference Across Independent AI Foundation Models

DGX agent

arXiv:2602.13792v2 Announce Type: replace Abstract: Artificial intelligence built on large foundation models has transformed language understanding, computer vision, and reasoning, yet these systems r

safetyarxiv-cs-ai
30 Jun 2026
Safety

Staged Hybridisation for Visual Quantum Reinforcement Learning via Knowledge Distillation

DGX agent

arXiv:2606.30520v1 Announce Type: cross Abstract: Visual environments are a demanding setting for quantum reinforcement learning (QRL): high-dimensional observations, unstable RL optimisation, and con

safetyarxiv-cs-lg
30 Jun 2026
Safety

STEAM: Self-Supervised Temporal Ensemble Advantage Modeling for Real-World Robot Learning

DGX agent

arXiv:2606.29834v1 Announce Type: new Abstract: Real-world robot learning increasingly relies on heterogeneous data, but demonstrations and rollouts often mix useful progress with stalls, corrections,

safetyarxiv-cs-ro
30 Jun 2026
← Previous
1…6970717273…265
Next →