AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
3 Jul 2026

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

SafetyDGX agent

arXiv:2607.02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference

Quantifying the Uncertainty of Blindly Estimated Room Embeddings Using a Dispersion-Calibrated Score

SafetyDGX agent

arXiv:2607.01527v1 Announce Type: cross Abstract: Room embeddings derived from reverberant speech are often unreliable: speech content and recording degradation can alter the representation even when

Quantum-Inspired Vision: Leveraging Wave-Particle Duality for Low-Illumination Enhancement

SafetyDGX agent

arXiv:2607.01731v1 Announce Type: cross Abstract: This study provides a theoretical expansion of the recent Data Relativistic Uncertainty (DRU) framework by formalizing a physics-to-AI paradigm for im

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Rank-Then-Act: Reward-Free Control from Frame-Order Progress

SafetyDGX agent

arXiv:2607.01897v1 Announce Type: cross Abstract: We introduce Rank-Then-Act (RTA), a framework for learning control policies from expert video demonstrations without environment rewards. RTA trains a

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

SafetyDGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

SafetyDGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

SABER: A Semantic-Aligned Brain Network Analysis Framework via Multi-scale Hypergraphs

SafetyDGX agent

arXiv:2607.01901v1 Announce Type: cross Abstract: Effective brain disease diagnosis requires the synergy of brain connectivity patterns and high-level semantic knowledge. Existing methods, however, la

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

SafetyDGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

Safeguarding LLM Agents from Misalignment through Provenance Analysis

SafetyDGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

Sim2Real-AD: A Modular Sim-to-Real Framework for Deploying VLM-Guided Reinforcement Learning in Real-World Autonomous Driving

SafetyDGX agent

arXiv:2604.03497v2 Announce Type: replace-cross Abstract: Vision-language-model (VLM)-guided reinforcement learning (RL) has recently attracted significant attention for it, replacing brittle hand-cra

SPLC: Social Preference Learning for Crowd Robot Navigation

SafetyDGX agent

arXiv:2607.01925v1 Announce Type: new Abstract: Offline reinforcement learning (RL) holds significant potential for crowd robot navigation in human-robot coexistence applications. However, the inheren

Structuring the Space of Sociotechnical Alignment

SafetyDGX agent

arXiv:2607.01250v1 Announce Type: cross Abstract: Sociotechnical alignment concerns the social desirability of AI behavior and is thus inherently normative, not merely technical. While NLP research in

The Rising Unsustainability of AI Graphics Cards Production

SafetyDGX agent

arXiv:2607.01258v1 Announce Type: cross Abstract: The rapid advancement of Artificial Intelligence (AI) has been accompanied by significant increases in computational and environmental costs, driven b

Tight Lower Bounds for the Multi-Secretary Problem via Bellman Certificates

SafetyDGX agent

arXiv:2607.02150v1 Announce Type: cross Abstract: This paper studies additive regret in the multi-secretary problem, defined as the gap between the expected offline prophet reward and the reward of th

Towards Learning Representations of Policies in Two-Player Zero-Sum Imperfect-Information Games

SafetyDGX agent

arXiv:2607.01498v1 Announce Type: new Abstract: We investigate the problem of learning useful policy representations (embeddings) in two-player zero-sum imperfect-information games. We make three cont

Transformer Geometry Observatory TGO-II: Representational Similarity Observatory

SafetyDGX agent

arXiv:2607.02386v1 Announce Type: cross Abstract: While Vision Transformers have achieved remarkable success across computer vision and language applications, the geometric evolution of their internal

Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models

SafetyDGX agent

arXiv:2512.01715v2 Announce Type: replace Abstract: Vision-language-action (VLA) models that generate continuous action chunks via flow matching lack an internal signal for judging whether a given pre

VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment

SafetyDGX agent

arXiv:2607.01586v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have recently advanced robotic manipulation, yet the effects of different robot-data pre-training paradigms remai

WaveLander: A Generalizable Hierarchical Control Framework for UAV Landing on Wave-Disturbed Platforms via Reinforcement Learning

SafetyDGX agent

arXiv:2607.01281v1 Announce Type: new Abstract: Autonomous landing of unmanned aerial vehicles (UAVs) on wave-disturbed marine platforms remains challenging due to stochastic platform motion, time-var

WBMM: Windowed Batch Matrix Multiplication for Efficient Large Receptive Field Convolution

SafetyDGX agent

arXiv:2607.02097v1 Announce Type: cross Abstract: Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory ac

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates

SafetyDGX agent

arXiv:2607.02507v1 Announce Type: new Abstract: LLM agents will increasingly act in socially structured settings where role, audience, and relational context can shape what is advantageous or costly t

When Sample Selection Bias Precipitates Model Collapse

SafetyDGX agent

arXiv:2606.13732v2 Announce Type: replace Abstract: The proliferation of recursive training on synthetic data can alleviate data scarcity but risks model collapse, where repeated training erodes distr

When Should Service Agents Reconsider? Difficulty-Routed Control in Customer-Service Operations

SafetyDGX agent

arXiv:2607.01426v1 Announce Type: new Abstract: Autonomous customer-service agents are shifting from conversational interfaces toward operational execution roles: they retrieve firm records, apply ser

Wind-Aware Reinforcement Learning Control of a Small Quadrotor Using Learned Onboard Wind Estimation in Simulated Atmospheric Turbulence

SafetyDGX agent

arXiv:2607.01528v1 Announce Type: new Abstract: Small multirotor aircraft are increasingly tasked with operations in the atmospheric boundary layer, where turbulent winds comparable to the vehicle's a

WorldSample: Closed-loop Real-robot RL with World Modelling

SafetyDGX agent

arXiv:2607.02431v1 Announce Type: cross Abstract: Reinforcement learning (RL) can overcome the demonstration-coverage limitation of imitation learning (IL) by allowing robots to improve through trial-

Wow, even I was surprised how high AI ranked! Very good to see!

SafetyDGX agent

Wow, even I was surprised how high AI ranked! Very good to see! Interesting poll of Hill staffers from @PunchbowlNews. 250 years is a long time! But interesting to see that 'losing control of AI' is t

2 Jul 2026

A Category Theory Account of AI Identity

SafetyDGX agent

arXiv:2607.00220v1 Announce Type: cross Abstract: Artificial intelligence (AI) systems are routinely modified after deployment through retraining and changes in their environments. These transformatio

A Filtered Mixture-of-Generators for Fully Synthetic Survival Training

SafetyDGX agent

arXiv:2607.00127v1 Announce Type: new Abstract: Survival analysis models time-to-event data, but in clinical settings training data are costly and scarce: events accrue over years of follow-up, cohort

A Mechanism-Driven Theory of Phase Transitions in Active Learning

SafetyDGX agent

arXiv:2607.00144v1 Announce Type: cross Abstract: Active learning (AL) performance is known to be budget-dependent, yet regimes are typically defined by heuristic label counts that fail to generalize

A Multi-Resolution Finite-Volume Inspired Deep Learning Framework for Spatiotemporal Dynamics Prediction

SafetyDGX agent

arXiv:2607.00460v1 Announce Type: cross Abstract: Predicting complex spatiotemporal dynamics in physical processes often demands computationally expensive numerical methods or data-driven neural netwo

A small tax on every token produced could be transformative, without putting the government into bed with a specific company. And because ev…

SafetyDGX agent

A small tax on every token produced could be transformative, without putting the government into bed with a specific company. And because every token draws on uncompensated contributions from multiple

A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors

Local AiDGX agent

arXiv:2603.21048v2 Announce Type: replace-cross Abstract: The identification of hazardous driving behaviors from in-cabin video streams is essential for enhancing road safety and supporting the detect

Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization

SafetyDGX agent

arXiv:2607.00531v1 Announce Type: cross Abstract: Scientific reasoning is an increasingly important capability of large language models, yet improving the robustness and efficiency of training such re

Active Spatial Guidance: Eliminating Injected Positional Mechanisms in Vision Transformers

SafetyDGX agent

arXiv:2607.00580v1 Announce Type: new Abstract: Vision Transformers (ViTs) commonly rely on injected positional mechanisms to address self-attention's permutation invariance. Motivated by the spatial

Aligning Sentence Embeddings to Human Concepts via Sparse Autoencoders

SafetyDGX agent

arXiv:2607.00023v1 Announce Type: cross Abstract: Dense sentence embeddings are fundamental to modern Retrieval-Augmented Generation (RAG) systems but suffer from a lack of interpretability due to fea

ASPIRE: Agentic /Skills Discovery for Robotics

SafetyDGX agent

arXiv:2607.00272v1 Announce Type: cross Abstract: Traditional robot programming is challenging: it requires orchestrating multimodal perception, managing physical contact dynamics, and handling divers

Attribute-Prompted Kernel Hashing for Unsupervised Data-Efficient Cross-Modal Retrieval

SafetyDGX agent

arXiv:2607.00379v1 Announce Type: cross Abstract: Unsupervised cross-modal hashing enables efficient retrieval of semantically related instances across different modalities without requiring manual se

AutoSpeed: Annotation-Free Stage-Adaptive Motion Speed Learning for Robot Manipulation

SafetyDGX agent

arXiv:2607.01051v1 Announce Type: new Abstract: Different stages of manipulation tasks exhibit varying levels of difficulty, suggesting stage-dependent motion speeds and temporal prediction horizons.

Bounded Morality: Defining the Space of Moral Computation

SafetyDGX agent

arXiv:2607.00002v1 Announce Type: new Abstract: Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as stati

BrainFIBRE: A Foundation Model via Information Decomposition for Brain Microstructure

SafetyDGX agent

arXiv:2607.00573v1 Announce Type: new Abstract: Diffusion MRI probes brain microstructure with particular sensitivity to early cerebrovascular and neurodegenerative changes. Neurite Orientation Disper

Caption Bottleneck Models

SafetyDGX agent

arXiv:2607.00578v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide interpretability by routing predictions through a layer of human-understandable concepts. However, defining an

ClinRAG-GRAPH: Clinical-prior Retrieval-Augmented Graph Model with Domain Adversarial Learning for Breast pCR Prediction

SafetyDGX agent

arXiv:2607.00798v1 Announce Type: new Abstract: Neoadjuvant chemotherapy (NAC) response prediction is clinically important for treatment stratification in breast cancer. However, robust pre-treatment

congrats @yudapearl!

SafetyDGX agent

congrats @yudapearl! Judea Pearl Named AI Pioneer by Boston Global Forum in Honor of America’s 250th Anniversary https://samueli.ucla.edu/judea-pearl-named-ai-pioneer-by-boston-global-forum-in-honor-o

'consensus' can't just built by big tech! let's all be careful of regulatory capture and extreme concentration of power!

SafetyDGX agent

'consensus' can't just built by big tech! let's all be careful of regulatory capture and extreme concentration of power! NEW: Anthropic announces it is drafting a consensus framework with Amazon, Micr

Constructive Alignment: Governing Preference Dynamics in Human-AI Interaction

SafetyDGX agent

arXiv:2607.00001v1 Announce Type: new Abstract: Most approaches to AI alignment treat human preferences as fixed targets to be inferred and optimized. This assumption conflicts with extensive empirica

Continuous Speculative Decoding for Autoregressive Image Generation

SafetyDGX agent

arXiv:2411.11925v3 Announce Type: replace Abstract: Continuous visual autoregressive (AR) models have demonstrated promising performance in image generation, but their inherently sequential nature res

Dataset Biases and Shortcut Learning in Motion-Based AI-Generated Video Detection

SafetyDGX agent

arXiv:2607.00948v1 Announce Type: new Abstract: The visual quality of AI-generated videos has improved drastically in recent years, making it increasingly difficult for humans to distinguish between r

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

SafetyDGX agent

arXiv:2607.01170v1 Announce Type: cross Abstract: Generative reasoning re-rankers achieve strong recommendation accuracy by emitting a chain-of-thought before re-ordering a candidate list, but they ar

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation

SafetyDGX agent

arXiv:2607.01208v1 Announce Type: cross Abstract: Language models deployed in high-stakes roles can potentially favor certain entities, brands, or viewpoints, steering user decisions at scale. Such pr

Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts

SafetyDGX agent

arXiv:2607.00666v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models often fail to perform the same learned tasks under environmental shifts, such as changes in camera pose and shifts

Dual-Informed Vertical Expansion for Multi-Objective Node Selection in Anytime Conflict-Based Search

SafetyDGX agent

arXiv:2607.00156v1 Announce Type: new Abstract: Conflict-Based Search (CBS) is a leading exact algorithm for Multi-Agent Path Finding (MAPF), but its high-level node-selection rule is usually treated

ELMP: Efficient Learning for Motion Planning via Analytical Policy Gradients

SafetyDGX agent

arXiv:2607.00215v1 Announce Type: new Abstract: Neural Motion Planners (NMPs) enable fast reactive motion generation, but adapting them to new environments typically requires recollecting large expert

Enhancing Hardware Fault Tolerance in Machines with Reinforcement Learning Policy Gradient Algorithms

SafetyDGX agent

arXiv:2407.15283v2 Announce Type: replace-cross Abstract: Industry is moving toward autonomous, network-connected machines that detect and adapt to changing conditions, including hardware faults. Conv

EPO: Boosting 3D Foundation Models with Edge-based Pose Optimization

SafetyDGX agent

arXiv:2607.00579v1 Announce Type: new Abstract: We introduce extbf{Edge-based Pose Optimization (EPO)}, a trackless geometric optimization framework specifically designed to boost the Structure-from-M

EquiSteer: Cross-Attention Steering Towards a Fairer Text-Guided Image Generation

SafetyDGX agent

arXiv:2607.01147v1 Announce Type: new Abstract: Text-to-image diffusion models power everyday creative tasks, but they still reproduce the demographic biases in their training data. On common prompts

FAR: Failure-Aware Retry for Test-Time Recovery and Continual Policy Improvement

SafetyDGX agent

arXiv:2607.01111v1 Announce Type: cross Abstract: Robot policies inevitably encounter failures when deployed in real environments. Naive retries often repeat the same mistakes, while many existing rec

ForAug: Mitigating Biases in Image Classification via Controlled Image Compositions

SafetyDGX agent

arXiv:2503.09399v4 Announce Type: replace-cross Abstract: Large-scale image classification datasets exhibit strong compositional biases: objects tend to be centered, appear at characteristic scales, a

FrameONE: Hierarchical Motion Modeling for Universal Multi-View Echocardiographic Keyframe Detection

SafetyDGX agent

arXiv:2607.00748v1 Announce Type: new Abstract: Accurate detection of end-systole (ES) and end-diastole (ED) frames is fundamental to echocardiographic assessment. Existing methods are typically devel

From Pixels to Temporal Correlations: Learning Informative Representations for Reinforcement Learning Pre-training

SafetyDGX agent

arXiv:2607.00811v1 Announce Type: new Abstract: Unsupervised pre-training on large-scale datasets has demonstrated significant potential for improving the sample efficiency and performance of Reinforc

From Prior to Pro: Efficient Skill Mastery via Distribution Contractive RL Finetuning

SafetyDGX agent

arXiv:2603.10263v2 Announce Type: replace-cross Abstract: We introduce Distribution Contractive Reinforcement Learning (DICE-RL), a framework that uses reinforcement learning (RL) as a 'distribution c

← Previous
1…9091929394…242
Next →