AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
7 Jul 2026

AnchorDream: Repurposing Video Diffusion for Embodiment-Aware Robot Data Synthesis

SafetyDGX agent

arXiv:2512.11797v2 Announce Type: replace-cross Abstract: The collection of large-scale and diverse robot demonstrations remains a major bottleneck for imitation learning, as real-world data acquisiti

Anticipatory Reinforcement Learning for Trajectory Tracking

SafetyDGX agent

arXiv:2607.03132v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) in industrial control often suffers from lag and overshoot due to purely reactive control based on the current trackin

AquaStereo: Enabling Underwater Stereo Matching via Depth-Conditioned Diffusion and Geometry Self-Distillation

SafetyDGX agent

arXiv:2607.04303v1 Announce Type: new Abstract: Learning-based stereo matching models struggle in underwater environments due to scarce in-domain data and the difficulty of extracting discriminative c

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ARLArena: A Unified Framework for Stable Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2602.21534v3 Announce Type: replace Abstract: Agentic reinforcement learning (ARL) has rapidly gained attention as a promising paradigm for training agents to solve complex, multi-step interacti

As the @UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications fo…

SafetyDGX agent

As the @UN’s Global Dialogue on AI Governance wraps up today, I’ve been encouraged by the discussions surrounding AI and its implications for global collaboration and policymaking. I’m hopeful that we

ASK in the Dark: Uncertainty-Gated LLM Assistance under Partial Observability

SafetyDGX agent

arXiv:2607.02686v1 Announce Type: new Abstract: Reinforcement learning agents operating under partial observability must act on incomplete information, making them natural candidates for guidance from

Athena-WBC: Capability-Aligned Policy Experts for Long-Tail Humanoid Whole-Body Control

SafetyDGX agent

arXiv:2607.04837v1 Announce Type: new Abstract: Large-scale humanoid motion-tracking controllers are commonly improved by reallocating training effort: difficult motions are sampled more often, isolat

Attention Limited Reward Learning

SafetyDGX agent

arXiv:2607.04590v1 Announce Type: new Abstract: Pairwise human comparisons are a primary interface through which modern AI systems learn human preferences. RLHF and related alignment pipelines typical

Attributing Emergence in Million-Agent Systems

SafetyDGX agent

arXiv:2605.11404v2 Announce Type: replace Abstract: Large language models (LLMs) can simulate human-like reasoning and decision-making in individual agents. LLM-powered multi-agent systems (MAS) combi

Aura: Consistent Multi-Subject Video Generation via VLM-Grounded Semantic Alignment

SafetyDGX agent

arXiv:2607.04311v1 Announce Type: new Abstract: Subject-driven and multi-element video generation are central to controllable video synthesis, but existing methods still struggle to preserve identity

Autonomous Information Seeking: A Roadmap for Agentic Recommender Systems

SafetyDGX agent

arXiv:2607.04433v1 Announce Type: cross Abstract: The rapid integration of large language model-based agents into recommender systems has driven a shift from static, ranking-based pipelines toward aut

AViS-Mamba: Adaptive Visual Steering of Audio State-Space Dynamics for Violence Detection

SafetyDGX agent

arXiv:2604.03329v2 Announce Type: replace-cross Abstract: Automatic violence detection from video is challenging because violent interactions may be distant, occluded, or only partially visible. Audio

Benign Overfitting Does Not Occur in Diffusion Models

SafetyDGX agent

arXiv:2607.02671v1 Announce Type: cross Abstract: Benign overfitting and double descent have come to shape our understanding of generalization in deep learning, establishing that overfitting is not on

Beyond Independent Labels: Schwartz-Geometry Decoding for Human Value Detection

SafetyDGX agent

arXiv:2607.05052v1 Announce Type: cross Abstract: Human value detection is commonly formulated as sentence-level multi-label classification over the 19 refined Schwartz values, typically predicted as

Beyond Point-Attached Semantics: Object-Centric Semantic Fields for Generalizable Manipulation

SafetyDGX agent

arXiv:2607.03163v1 Announce Type: new Abstract: Generalizable robot manipulation requires stable 3D understanding of functional object parts, such as handles, tool heads, openings, and graspable regio

Beyond Random Sampling: Distribution-Aware Alignment for Semi-Supervised Medical Image Segmentation

SafetyDGX agent

arXiv:2607.04249v1 Announce Type: new Abstract: Precise medical image segmentation is crucial for clinical diagnosis and treatment planning, yet relies heavily on expensive expert annotations. Semi-su

BGP route policies: Top 3 use cases by customer demand

SafetyDGX agent

When we first made BGP route policies for Cloud Router generally available over a year ago, our goal was to give network administrators deep, programmable control over how network paths are evaluated

BiSLW: Bi-Spectral Latent Watermarking for Generative Diffusion Models

SafetyDGX agent

arXiv:2607.02643v1 Announce Type: new Abstract: Diffusion-based generative models have transformed visual content synthesis, yet they remain vulnerable to unauthorized usage and lack reliable attribut

Bootstrap Flow-Map Tree Sampling Enables Online Feedback Driven Search

SafetyDGX agent

arXiv:2607.02915v1 Announce Type: cross Abstract: In many scientific and engineering domains, maximizing discovery within a limited sampling budget demands strategic, observation-guided exploration. W

BoRP: Bootstrapped Regression Probing for Scalable and Human-Aligned LLM Evaluation

SafetyDGX agent

arXiv:2601.18253v2 Announce Type: replace-cross Abstract: Accurate evaluation of user satisfaction is critical for iterative development of conversational AI. However, for open-ended assistants, tradi

Bridging Interleaved Multi-Modal Reasoning as a Unified Decision Process

SafetyDGX agent

arXiv:2607.03748v1 Announce Type: new Abstract: Unified multi-modal models (UMMs) have shown promising interleaved text-image reasoning capabilities, yet effectively optimizing such multi-turn generat

Can temporal article-level credibility signals improve domain-level credibility prediction?

SafetyDGX agent

arXiv:2607.04560v1 Announce Type: new Abstract: Web domain credibility evaluation is vital for combating misinformation. It is conducted by examining factors such as domain type, transparency, and ove

CanniUplift: A Holistic Framework for Mitigating Seller and Incentive Cannibalization in E-commerce Uplift Modeling

SafetyDGX agent

arXiv:2607.05242v1 Announce Type: cross Abstract: Personalized incentive allocation is vital for e-commerce, where uplift modeling is the standard for estimating Individual Treatment Effects (ITE). Ho

Causal Mechanism Reduction: Mechanism Replacement for Neural Network Pruning and Abstraction

SafetyDGX agent

arXiv:2602.24266v2 Announce Type: replace-cross Abstract: Which internal mechanisms of a neural network can be replaced while preserving the computation it performs? Structured pruning asks for smalle

Channel-Adaptive Robust Aggregation for Over-the-Air Federated Learning in Heterogeneous Networks

SafetyDGX agent

arXiv:2607.04218v1 Announce Type: new Abstract: The growing demand for privacy-preserving, data-intensive applications such as IoT, augmented reality, and autonomous systems positions Federated Learni

Claim-Level Rubric Rewards for Video Caption Reinforcement Learning

SafetyDGX agent

arXiv:2607.05150v1 Announce Type: new Abstract: In this paper, we introduce Claim-Level Rubric Rewards (CuRe), a structured reward framework designed to address the reward-design bottleneck in reinfor

CLEAR: Closed-Loop Reinforcement Learning at Scale for End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2607.02841v1 Announce Type: cross Abstract: End-to-end autonomous driving (E2E-AD) aims to directly map raw sensor information to driving actions. Recently, with the rapid advancement of multi-m

CoDoL: Conditional Domain Prompt Learning for Out-of-Distribution Generalization

SafetyDGX agent

arXiv:2509.15330v2 Announce Type: replace Abstract: Recent advances in pre-training vision-language models (VLMs), e.g., contrastive language-image pre-training (CLIP) methods, have shown great potent

Cohort-attention Evaluation Metrics for Tied Data

SafetyDGX agent

arXiv:2503.12755v3 Announce Type: replace Abstract: Artificial intelligence (AI) has significantly improved medical screening accuracy, particularly in cancer detection and risk assessment. However, t

CONFLUX: A Latent Diusion Model for 3D Chest-CT Synthesis with RL Post-Training

SafetyDGX agent

arXiv:2607.02998v1 Announce Type: cross Abstract: Controllable generative models of 3D medical images can synthesize volumes with specified clinical attributes, but this demands samples that are simul

CoRE-VLA: Towards Scalable and Robust Vision-Language-Action Modeling via Conditional Routing of Experts

SafetyDGX agent

arXiv:2607.03693v1 Announce Type: new Abstract: Vision-language-action (VLA) models have advanced generalist robotic manipulation, yet real-world deployment reveals a fundamental challenge: robots are

Counterfactual Methods for Detecting Unfairness in Anti-Money Laundering Algorithms

SafetyDGX agent

arXiv:2607.05101v1 Announce Type: new Abstract: The application of machine learning-based predictive algorithms to Anti-Money Laundering (AML) has grown rapidly, driven by the vast volume of financial

Court filing: Meta says four US states seek 1.4T over claims it designed Facebook and Instagram to addict youth and misled the public; its market cap is ~1.5T (Diana Novak Jones/Reuters)

SafetyDGX agent

Diana Novak Jones / Reuters: Court filing: Meta says four US states seek 1.4T over claims it designed Facebook and Instagram to addict youth and misled the public; its market cap is ~1.5T — Meta Platf

Coverage-Controlled Preference Mining from Noisy Claim Verification for Evidence-Grounded Generation

SafetyDGX agent

arXiv:2603.10494v2 Announce Type: replace Abstract: Evidence-grounded generation produces summaries whose claims should be supported by supplied evidence, but claim-level verifiers provide noisy feedb

Covert Trait Propagation Is Representation Alignment: Mechanistic Evidence from Hidden-Channel Distillation

SafetyDGX agent

arXiv:2607.04432v1 Announce Type: cross Abstract: A student model trained on pure uniform noise can still inherit its teacher's digit-classification ability, provided the two share initialization. Pre

CRODA-ST: Single-Target Cross-Receiver Open-Set Radio Fingerprint Recognition

SafetyDGX agent

arXiv:2607.02567v1 Announce Type: cross Abstract: Radio frequency fingerprint identification (RFFI) provides a physical-layer credential for Internet of Things devices, but open-set decisions become f

Cross-Embodiment Robot Manipulation via a Unified Hand Action Space

SafetyDGX agent

arXiv:2607.03570v1 Announce Type: new Abstract: Robot manipulation policies are typically tied to specific robotic hand embodiments, limiting the transfer of learned behaviors across platforms with di

CrossHallu: Do Hallucination Signals Generalize Across Languages and Domains in Large Language Model's Internals?

SafetyDGX agent

arXiv:2607.04029v1 Announce Type: new Abstract: Recent hallucination detection techniques in large language models (LLMs) focus on directly extracting features from a model's internal representations

Current as Touch: Proprioceptive Contact Feedback for Compliant Dexterous Manipulation

SafetyDGX agent

arXiv:2607.03529v1 Announce Type: new Abstract: Compliance is essential for dexterous manipulation, yet existing solutions often rely on external tactile or force sensors that are costly, fragile, and

CV-DCLR: Causal-Visual Dynamic Label Refinement for Robust Zero-Shot Learning

SafetyDGX agent

arXiv:2607.02601v1 Announce Type: new Abstract: Zero-Shot Learning (ZSL) facilitates knowledge transfer via shared semantic spaces. However, a critical bottleneck in this paradigm is Semantic Entangle

Decentralized Aggregation of LLM Predictions via Wagering Mechanisms

SafetyDGX agent

arXiv:2607.04389v1 Announce Type: new Abstract: It is increasingly common to aggregate predictions from multiple LLMs, each with domain expertise or access to private tools and data, to improve collec

Decomposed Prompting Does Not Fix Knowledge Gaps, But Helps Models Say 'I Don't Know'

SafetyDGX agent

arXiv:2602.04853v2 Announce Type: replace Abstract: Large language models often struggle to recognize their knowledge limits in closed-book question answering, leading to confident hallucinations. Whi

Deep Learning for Dynamic Programming with Recursive Utility

SafetyDGX agent

arXiv:2607.04278v1 Announce Type: cross Abstract: We propose the first deep learning algorithm, the Certainty Equivalent Learning (CEL) algorithm, for solving high-dimensional discrete-time dynamic pr

DeGenseGS: Geometrically and Semantically Decoupled Surgical Scene Understanding in 4D Gaussian Splatting

SafetyDGX agent

arXiv:2607.04761v1 Announce Type: new Abstract: Real-time, text-promptable 4D reconstruction is indispensable for autonomous surgical interaction. Severe misalignment between semantic meaning and phys

Demonstrating Generalization Failures via Mixtures of Conditional Policies

SafetyDGX agent

arXiv:2607.03478v1 Announce Type: new Abstract: Post-training of frontier language models is conducted on curated task suites, and inevitably leaves a distribution shift between training and deploymen

Denoised Conformal Alignment for Reliable Selection of Conditional Average Treatment Effect Predictions

SafetyDGX agent

arXiv:2607.03161v1 Announce Type: cross Abstract: In selective deployment, practitioners act only on a model-chosen subset of individuals based on predicted conditional average treatment effects, but

Deterministic Envelopes for Tamed SGLD: Decoupling Stochastic Gradient Noise and Localizing Taming

SafetyDGX agent

arXiv:2606.05242v2 Announce Type: replace-cross Abstract: Stochastic gradient Langevin algorithms often use tamed denominators to stabilize superlinear drifts. This paper shows that when the denominat

DiCE-CIR: Direct Composition Learning for Efficient Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2607.04665v1 Announce Type: new Abstract: Zero-shot composed image retrieval (ZS-CIR) aims to retrieve a target image from a multimodal query consisting of a reference image and an edit text des

DICT: Data Injection and Contrastive Trajectory Refinement for Conditional Image Generation with Diffusion Models

SafetyDGX agent

arXiv:2607.03899v1 Announce Type: new Abstract: Diffusion models have become a dominant paradigm for conditional image generation, yet existing approaches generally follow two directions: task-specifi

Diffusion-Guided Uncertainty-Aware Delayed Policy Optimization

SafetyDGX agent

arXiv:2607.05064v1 Announce Type: new Abstract: Reinforcement learning in real world environments often suffers from severe performance degradation due to delayed feedback. Existing approaches typical

Displacement Preserving Relational Distillation for Robust Medical Segmentation

SafetyDGX agent

arXiv:2607.04599v1 Announce Type: new Abstract: Accurate 3D medical segmentation is limited by anatomical variability and high computational costs. While knowledge distillation (KD) offers a route for

Distribution-free Deviation Bounds and The Role of Domain Knowledge in Learning via Model Selection with Cross-validation Risk Estimation

SafetyDGX agent

arXiv:2303.08777v3 Announce Type: replace-cross Abstract: Cross-validation is one of the most widely used tools for risk estimation and model selection in statistics and machine learning, yet its theo

Diverse Normal Prototypes-Guided Contrastive Reconstruction for Medical Anomaly Detection

SafetyDGX agent

arXiv:2508.19573v2 Announce Type: replace Abstract: Anomaly detection in medical images is challenging due to limited annotations and the domain gap. Existing reconstruction-based methods often rely o

Do ECG Foundation Models Transfer to Rare Cardiac Diseases? Evidence from Brugada Syndrome Detection

SafetyDGX agent

arXiv:2607.03009v1 Announce Type: new Abstract: Background: Foundation models (FMs) trained on large-scale unlabeled physiological data have emerged as a promising paradigm for medical artificial inte

Do Vision-Language-Action Models Mean What They Say? On the Role of Faithfulness in Embodied Reasoning

SafetyDGX agent

arXiv:2607.04681v1 Announce Type: cross Abstract: Embodied Chain-of-Thought has emerged as a promising mechanism to enhance robot decision-making and interpretability in black-box Vision-Language Acti

dOPSD: On-Policy Self-Distillation for Diffusion Language Models

SafetyDGX agent

arXiv:2607.04428v1 Announce Type: cross Abstract: Diffusion large language models (dLLMs) generate text by iteratively denoising a masked sequence, offering a parallel alternative to autoregressive mo

Double-Helix Active Geometry: LiDAR-Anchored Multi-View Depth with Selective Abstention

SafetyDGX agent

arXiv:2607.02561v1 Announce Type: cross Abstract: Consumer depth sensors such as the LiDAR scanner on recent iPhones provide metric range, but their useful range is short and their returns are sparse.

DSWAM: A Dual-System World Action Foundation Model for Fine-Grained Robot Manipulation

SafetyDGX agent

arXiv:2607.04927v1 Announce Type: cross Abstract: World Action Models (WAMs) provide a promising alternative to Vision-Language-Action (VLA) policies by using video-based world modeling as dense super

Efficient bias mitigation in T2I diffusion models using Concept Graphs

SafetyDGX agent

arXiv:2607.03397v1 Announce Type: new Abstract: Text-to-Image diffusion models often propagate harmful bias inherited from the training data. Existing bias mitigation techniques typically intervene on

EGRA:Toward Enhanced Behavior Graphs and Representation Alignment for Multimodal Recommendation

SafetyDGX agent

arXiv:2508.16170v2 Announce Type: replace-cross Abstract: MultiModal Recommendation (MMR) systems have emerged as a promising solution for improving recommendation quality by leveraging rich item-side

← Previous
1…8586878889…242
Next →