AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
10 Jul 2026

NFTR: From Provable Mode-Averaging to Geodesic Subgoal Selection in Offline Goal-Conditioned RL

SafetyDGX agent

arXiv:2607.07855v1 Announce Type: new Abstract: Hierarchical Implicit Q-Learning (HIQL), an offline goal-conditioned RL method, selects subgoals by value-function advantages alone. This rule has two c

One of the biggest threats to US innovation isn’t just research funding - it’s losing the next generation of scientists. Phd admissions at l…

SafetyDGX agent

One of the biggest threats to US innovation isn’t just research funding - it’s losing the next generation of scientists. Phd admissions at leading research universities fell 15% this year. Scientific

Open-ended Multi-agent Autocurricula via Visual Inspection of Policies with Multi-modal LLMs

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2607.08193v1 Announce Type: cross Abstract: Open-ended curricula in Reinforcement Learning (RL) aim to train generally-capable agents by identifying tasks that facilitate learning increasingly c

OpenAI's head of safety, Johannes Heidecke, is leaving as OpenAI integrates its research and safety teams; Mia Glaese will become VP of research and safety (Maxwell Zeff/Wired)

SafetyDGX agent

Maxwell Zeff / Wired: OpenAI's head of safety, Johannes Heidecke, is leaving as OpenAI integrates its research and safety teams; Mia Glaese will become VP of research and safety — Johannes Heidecke's

OPSD-V: On-Policy Self-Distillation for Post-Training Few-Step Autoregressive Video Generators

SafetyDGX agent

arXiv:2607.08766v1 Announce Type: new Abstract: We propose OPSD-V, an on-policy self-distillation paradigm for post-training few-step autoregressive (AR) video diffusion models. Existing few-step AR v

Persona Cartography: Charting Language Model Personality Traits in Weight Space

SafetyDGX agent

arXiv:2607.07916v1 Announce Type: new Abstract: Large language models exhibit recurring behavioural patterns -- personas -- that shape generalisation and safety, but we lack reliable tools for decompo

PhyMAGIC: Physical Motion-Aware Generative Inference with Confidence-guided LLM

SafetyDGX agent

arXiv:2505.16456v3 Announce Type: replace Abstract: Recent advances in 3D content generation have amplified demand for dynamic models that are both visually realistic and physically consistent. Howeve

Physics-Guided Biomechanical Gait Adaptation for Humanoid Locomotion on Extreme Sloped Terrains

SafetyDGX agent

arXiv:2607.07830v1 Announce Type: new Abstract: Model-free reinforcement learning has enabled impressive humanoid locomotion; however, control on steep slopes remains largely unexplored. Unlike flat o

PLURAL: A Global Dataset for Value Alignment

SafetyDGX agent

arXiv:2607.08034v1 Announce Type: cross Abstract: Large language models (LLMs) are used worldwide, yet disproportionately reflect Western values, limiting their ability to represent diverse value syst

Post-Training in End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2607.08072v1 Announce Type: new Abstract: End-to-end models that map multimodal inputs directly to future trajectories/maneuvers have emerged as an increasingly prominent research paradigm in au

Prismata: Confining Cross-Site Prompt Injection in Web Agents

SafetyDGX agent

arXiv:2607.08147v1 Announce Type: cross Abstract: Autonomous web agents promise to automate everyday browsing tasks, but inherit one of the web's oldest attack surfaces. Cross-Site Scripting proved th

reminds me of the time Greg Brockman personally downloaded YouTube videos for training purposes, as reported in NYT.

SafetyDGX agent

reminds me of the time Greg Brockman personally downloaded YouTube videos for training purposes, as reported in NYT. 🦔Apple is suing OpenAI for systematic trade secret theft weeks before the IPO. The

Safe Flow Q-Learning: Offline Safe Reinforcement Learning with Reachability-Based Flow Policies

SafetyDGX agent

arXiv:2603.15136v2 Announce Type: replace-cross Abstract: Offline safe reinforcement learning (RL) seeks reward-maximizing policies from static datasets under strict safety constraints. Existing metho

SASGeo: Stability-Aware Semantic Map Localization for GNSS-Denied UAVs -- A Framework and Synthetic Proof of Concept

SafetyDGX agent

arXiv:2607.07737v1 Announce Type: cross Abstract: GNSS-denied unmanned aerial vehicles require occasional absolute position fixes to bound the drift of visual-inertial odometry. Cross-view image retri

Scalable and Culturally Specific Stereotype Dataset Construction via Human-LLM Collaboration

SafetyDGX agent

arXiv:2607.07895v1 Announce Type: new Abstract: Research on stereotypes in large language models (LLMs) has largely focused on English-speaking contexts, due to the lack of datasets in other languages

Search-based Testing of Vision Language Models for In-Car Scene Understanding

SafetyDGX agent

arXiv:2607.02300v2 Announce Type: replace Abstract: In the automotive domain, in-car scene understanding (ISU) enables the detection of safety-critical events, such as driver distraction, and supports

Securing Autonomous Vehicle Systems via Twin-Aware Federated Reinforcement Learning

SafetyDGX agent

arXiv:2607.08137v1 Announce Type: cross Abstract: Federated reinforcement learning (FRL) is crucial for enabling collaborative learning across multiple agents without sharing raw data, thereby enhanci

SeFA-Policy: Fast and Accurate Visuomotor Policy Learning with Selective Flow Alignment

SafetyDGX agent

arXiv:2511.08583v2 Announce Type: replace-cross Abstract: Developing efficient and accurate visuomotor policies poses a central challenge in robotic imitation learning. While recent rectified flow app

SkillPlug: Unsupervised Skill Mining for Few-Shot Adaptation in Robotic Manipulation

SafetyDGX agent

arXiv:2607.08354v1 Announce Type: new Abstract: Learning transferable visuomotor imitation policies that generalize across diverse manipulation tasks and adapt rapidly to new tasks from only a handful

Soft Robotic Exogloves for Dexterous Mobility -- Towards Personalized Rehabilitation

SafetyDGX agent

arXiv:2607.07968v1 Announce Type: new Abstract: Soft robotic exogloves can provide hand rehabilitation and assistance. Fitting these gloves often relies on standardized measurements not tailored to th

Spatio-Temporal Scheduling Prediction Under Backhaul Delay for Resilient Coordinated Beamforming

SafetyDGX agent

arXiv:2607.08454v1 Announce Type: cross Abstract: Coordinated beamforming in distributed 5G networks relies on the timely exchange of inter-cell scheduling information, but backhaul latency makes this

SPHINX: First Explain, Then Explore

SafetyDGX agent

arXiv:2606.17482v2 Announce Type: replace Abstract: Generating adversarial driving scenarios is critical for evaluating and improving autonomous vehicle decision-making systems in simulation. Recent a

Statistical Efficiency and Inference of Quantile Distributional Reinforcement Learning

SafetyDGX agent

arXiv:2607.08444v1 Announce Type: cross Abstract: In this paper, we study quantile-based distributional reinforcement learning from the perspective of statistical efficiency. We focus on distributiona

Texture Representations in Deep Vision Models: Comparing CNNs, Vision Transformers, and Human Perception

SafetyDGX agent

arXiv:2607.08321v1 Announce Type: new Abstract: In computational vision science, Convolutional Neural Networks (CNNs) have emerged as a popular model of biological vision because of the alignment they

The Contribution of XAI for the Safe Development and Certification of AI: An Expert-Based Analysis

SafetyDGX agent

arXiv:2408.02379v2 Announce Type: replace-cross Abstract: Developing and certifying safe - or so-called trustworthy - AI has become an increasingly salient issue, especially in light of upcoming regul

TNODEV: Toolbox for Neural ODE Verification

SafetyDGX agent

arXiv:2606.16567v2 Announce Type: replace Abstract: Neural ordinary differential equations (neural ODE) gained attention in safety critical settings such as continuous-time controllers for cyber-physi

Trustworthy Machine Learning through the Lens of Combinatorial Optimization: Survey and Research Perspectives

SafetyDGX agent

arXiv:2607.07762v1 Announce Type: new Abstract: Modern machine learning (ML) increasingly relies on complex models whose behavior is difficult to characterize beyond empirical performance metrics. Acr

Two Axes of LLM Abstention: Answer Correctness and Question Answerability

SafetyDGX agent

arXiv:2607.08456v1 Announce Type: cross Abstract: A model should refuse two different things: answers it would get wrong, and questions it should not answer at all, such as unanswerable ones or ones r

UltraX: Refining Pre-Training Data at Scale with Adaptive Programmatic Editing

SafetyDGX agent

arXiv:2607.08646v1 Announce Type: cross Abstract: As available training data approaches its physical limit, gains from Scaling Laws have begun to diminish. Consequently, improving Large Language Model

Unified Face Attack Detection via Fine-Grained Semantic Guidance

SafetyDGX agent

arXiv:2607.08156v1 Announce Type: new Abstract: The growing applications of facial recognition systems are accompanied by increasingly diverse security threats. Existing datasets lack detailed textual

Validating LLMs in social science: Epistemic threats and emerging norms

SafetyDGX agent

arXiv:2607.07915v1 Announce Type: cross Abstract: Large language models (LLMs) are reshaping social science methodology. Researchers increasingly prompt language models to generate quantitative measur

When Structured Sparse Autoencoders Learn Consistent Concepts Across Modalities

SafetyDGX agent

arXiv:2607.08605v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) have emerged as a promising technique for mechanistic interpretability by learning a set of sparse latent features in large

When Synthetic Speech Is All You Have: Better Call GRPO

SafetyDGX agent

arXiv:2607.08409v1 Announce Type: cross Abstract: LLM-based ASR adapted to regulated domains such as banking is bottlenecked by privacy: real speech is costly and legally constrained to collect, makin

Who Gets Missed in the Tail? Thresholded Subgroup Underdiagnosis in Long-Tailed Chest X-ray Classification

SafetyDGX agent

arXiv:2607.07717v1 Announce Type: cross Abstract: In chest X-ray (CXR) classification, acceptable ranking performance can still leave rare-positive patients below threshold, especially within subgroup

Workflow as Knowledge: Semantic Persistence for LLM-Mediated Workflows

SafetyDGX agent

arXiv:2607.08740v1 Announce Type: new Abstract: Large language model (LLM) applications increasingly use explicit workflows for tool use, retrieval, branching, checkpointing, and human approval. Exist

XALPHA: A Memory-Driven AI Quant Researcher for Hypothesis-to-Code Alpha Discovery

SafetyDGX agent

arXiv:2607.08332v1 Announce Type: new Abstract: Financial markets are noisy, non-stationary, and high-dimensional, making it difficult to discover predictive and robust trading signals. Alpha discover

XFACTORS: Disentangled Information Bottleneck via Contrastive Supervision

SafetyDGX agent

arXiv:2601.21688v2 Announce Type: replace-cross Abstract: Disentangled representation learning aims to map independent factors of variation to independent representation components. On one hand, purel

Zoom-IQA: Image Quality Assessment with Reliable Region-Aware Reasoning

SafetyDGX agent

arXiv:2601.02918v3 Announce Type: replace Abstract: Image Quality Assessment (IQA) is a long-standing problem in computer vision. Previous methods typically focus on predicting numerical scores withou

9 Jul 2026

A Distributionally Robust Optimisation Approach to Fair Credit Scoring

SafetyDGX agent

arXiv:2402.01811v2 Announce Type: replace Abstract: Credit scoring has been catalogued by the European Commission and the Executive Office of the US President as a high-risk classification task, in li

Ad Headline Generation using Self-Critical Masked Language Model

SafetyDGX agent

arXiv:2607.06818v1 Announce Type: cross Abstract: For any E-commerce website it is a nontrivial problem to build enduring advertisements that attract shoppers. It is hard to pass the creative quality

AGAPI-Agents: An Open-Access Agentic AI Platform for Accelerated Materials Design on AtomGPT.org

SafetyDGX agent

arXiv:2512.11935v2 Announce Type: replace Abstract: Agentic AI systems increasingly connect large language models (LLMs) to external scientific tools, yet whether and when tool access improves predict

Agentic Data Environments

SafetyDGX agent

arXiv:2607.07397v1 Announce Type: new Abstract: Autonomous agents promise substantial gains in speed, scale, and labor efficiency, but their failures can impose abrupt and often irreversible costs. Th

AI for Cultural Heritage Textiles: Fine-Tuned Latent Diffusion for Novel Ulos Motif Synthesis

SafetyDGX agent

arXiv:2607.06590v1 Announce Type: new Abstract: Preserving and revitalising traditional textiles such as Ulos, a cultural heritage of the Batak ethnic group in North Sumatra, Indonesia, requires balan

`Attention-Guided Cross-Temporal Clustering for Self-Supervised Video Object Segmentation

SafetyDGX agent

arXiv:2607.07230v1 Announce Type: new Abstract: Video object segmentation (VOS) is a fundamental task in video understanding, requiring accurate delineation and consistent tracking of objects across f

Automatic Echocardiography Segmentation via Transition Probability Correlation for Stable Semantic Extraction

SafetyDGX agent

arXiv:2607.07580v1 Announce Type: new Abstract: While echocardiography is essential for cardiovascular diagnosis, inherent speckle noise and low signal-to-noise ratio often lead to ambiguous semantic

Avoiding unsafe sets when training with Langevin Dynamics

SafetyDGX agent

arXiv:2607.07538v1 Announce Type: new Abstract: Training a model with noisy gradient descent can be idealized as overdamped Langevin dynamics on the loss landscape, and a natural safety question is to

Behavior Foundations for Quadruped Robots: ABot-C0 Technical Report

SafetyDGX agent

arXiv:2607.07370v1 Announce Type: cross Abstract: In embodied intelligence systems, the motion controller serves as the critical bridge between semantic reasoning and physical execution. Humanoid cont

Behavior Leverage Imbalance in Multi-Teacher On-Policy Distillation

SafetyDGX agent

arXiv:2607.07050v1 Announce Type: new Abstract: Agentic language models must learn when to call tools, when to consume tool responses, and when to answer directly. This makes multi-teacher on-policy d

CARLA-GS: Decoupling Representation, Reasoning, and Physics Simulation for Autonomous Driving Corner-Case Synthesis

SafetyDGX agent

arXiv:2607.07601v1 Announce Type: cross Abstract: Safety evaluation for autonomous driving is dominated by rare, safety-critical interactions, motivating simulators that can deliberately synthesize co

ContrastiveCFG: Guiding Diffusion Sampling by Contrasting Positive and Negative Concepts

SafetyDGX agent

arXiv:2411.17077v2 Announce Type: replace-cross Abstract: As Classifier-Free Guidance (CFG) has proven effective in conditional diffusion model sampling for improved condition alignment, many applicat

D2PO: Optimizing Diffusion Samplers via Dynamic Preference

SafetyDGX agent

arXiv:2607.06609v1 Announce Type: cross Abstract: We propose D2PO (Dynamic Direct Preference Optimization), a principled framework for optimizing diffusion sampling policies with respect to timestep s

DASH: Dynamic Audio-Driven Semantic Chunking for Efficient Omnimodal Token Compression

SafetyDGX agent

arXiv:2603.15685v2 Announce Type: replace-cross Abstract: Omnimodal large language models (OmniLLMs) jointly process audio and visual streams, but the resulting long multimodal token sequences make in

Deep Reinforcement Learning for Reliability Based Bi-Objective Portfolio Optimization

SafetyDGX agent

arXiv:2607.06610v1 Announce Type: cross Abstract: Portfolio optimization under uncertainty is inherently a multi-objective decision problem involving complex interactions among return, risk, market dy

DiaLLM: An Investigation into the Robustness-Generation Gap in English Dialect Adaptation

SafetyDGX agent

arXiv:2607.07669v1 Announce Type: cross Abstract: Large language models increasingly understand dialectal English, yet still produce only standard, US-leaning English, leaving dialectal generation, th

Discovering Geometric Biases in 3D Face Reconstruction: A Curvature-Aware Spectral Framework for Fairness Evaluation

SafetyDGX agent

arXiv:2607.07486v1 Announce Type: new Abstract: 3D Morphable Models (3DMMs) remain the standard parametric shape priors for many state-of-the-art 3D face reconstruction algorithms. However, as these m

Do Counterfactually Fair Image Classifiers Satisfy Group Fairness? -- A Theoretical and Empirical Study

SafetyDGX agent

arXiv:2607.06603v1 Announce Type: cross Abstract: The notion of algorithmic fairness has been actively explored from various aspects of fairness, such as counterfactual fairness (CF) and group fairnes

Dual Latent Memory in Vision-Language-Action Models for Robotic Manipulation

SafetyDGX agent

arXiv:2607.07608v1 Announce Type: cross Abstract: Mainstream Vision-Language-Action (VLA) models predict actions primarily from the current observation under a Markovian assumption, thus struggling wi

EmbodiedGen V2: An Agentic, Simulation-Ready 3D World Engine for Embodied AI

SafetyDGX agent

arXiv:2607.07459v1 Announce Type: cross Abstract: We present EmbodiedGen V2, a generative 3D world engine for building executable sim-ready environments for embodied intelligence. Sim-ready 3D asset g

EMO-R3: Reflective Reinforcement Learning for Emotional Reasoning in Multimodal Large Language Models

SafetyDGX agent

arXiv:2602.23802v2 Announce Type: replace Abstract: Multimodal Large Language Models (MLLMs) have shown remarkable progress in visual reasoning and understanding tasks but still struggle to capture th

Entropy Pacing Policy Optimization for Multi-Task Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2607.07178v1 Announce Type: cross Abstract: Recent breakthroughs of Reinforcement Learning (RL) have highlighted its potential for complex agentic Large Language Model (LLM) tasks. However, exis

← Previous
1…3738394041…212
Next →