AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
1 Jul 2026

Ask the World Before Acting: Budgeted Environment Probing for World-Model Calibration

SafetyDGX agent

arXiv:2606.31422v1 Announce Type: new Abstract: Long-horizon language agents do not only choose actions; they carry a private model of the world from one decision to the next. When that model drifts,

Automating Cause-Effect Specification with Knowledge Graphs and Large Language Models

SafetyDGX agent

arXiv:2606.31614v1 Announce Type: cross Abstract: Engineering specifications such as interlocks, alarm rationalization tables, and cause-and-effect (C&E) matrices remain central to process control and

Autonomous UAV Navigation for Individual Wildlife Re-Identification

SafetyDGX agent

arXiv:2606.31772v1 Announce Type: new Abstract: Reliable individual re-identification (re-ID) of wildlife is essential for population monitoring, behavioral tracking, and conservation policy evaluatio


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

AVTok: 1D Unified Tokenization for Holistic Audio-Video Generation

SafetyDGX agent

arXiv:2606.30811v1 Announce Type: new Abstract: Audio-video generation has recently gained unprecedented research attention, aiming to synthesize high-quality sounding video content with fine-grained

Behavior Cloning is Not All You Need: The Optimality of On-Policy Distillation for Noisy Expert Feedback

SafetyDGX agent

arXiv:2606.30923v1 Announce Type: cross Abstract: Imitation Learning is a natural framework for learning in sequential decision-making systems and has emerged as the dominant paradigm through which we

Beyond the Expressivity-Trainability Paradox: A Dynamical Lie Algebra Perspective on Navigating Barren Plateaus in Quantum Machine Learning

SafetyDGX agent

arXiv:2606.31536v1 Announce Type: new Abstract: As Quantum Machine Learning (QML) transitions toward practical implementation, the field faces a critical architectural bottleneck that challenges the f

BlockPilot: Instance-Adaptive Policy Learning for Diffusion-based Speculative Decoding

SafetyDGX agent

arXiv:2606.31315v1 Announce Type: new Abstract: Speculative decoding accelerates inference by using a lightweight draft model to generate candidate tokens in parallel, and are then verified by the tar

BP-TTA: Balanced and Prototype-Guided Test-Time Adaptation in Dynamic Scenarios

SafetyDGX agent

arXiv:2606.31420v1 Announce Type: new Abstract: Test-Time Adaptation (TTA) enables models trained on a source domain to adapt online to unlabeled test data under distribution shifts. While recent TTA

Breaking Failure Cascades: Step-Aware Reinforcement Learning for Medical Multimodal Reasoning

SafetyDGX agent

arXiv:2606.31825v1 Announce Type: cross Abstract: Recent multimodal large language models have shown great promise in clinical image reasoning, but existing post-training pipelines remain predominantl

Bridging Information Asymmetry: A Hierarchical Framework for Deterministic Blind Face Restoration

SafetyDGX agent

arXiv:2601.19506v4 Announce Type: replace Abstract: Blind face restoration remains a persistent challenge due to the inherent ill-posedness of reconstructing holistic structures from severely constrai

Can Tabular In-Context Learners Generalize to Biomolecular Property Prediction?

SafetyDGX agent

arXiv:2606.31126v1 Announce Type: new Abstract: Predicting biomolecular properties from limited labeled data is a central bottleneck in protein engineering and small-molecule design. As strong pretrai

CasaMaestro: Multi-View Panoramas for House-Scale 3D Reconstruction

SafetyDGX agent

arXiv:2606.31086v1 Announce Type: new Abstract: The rise of home-deployed embodied AI systems is driving a growing need for fast, metric 3D reconstruction of residential spaces to support navigation,

Certified Speculative Execution for Untrusted AI Agents

SafetyDGX agent

arXiv:2606.31023v1 Announce Type: cross Abstract: Hard-constrained sequential decision systems have no certified way to spend the test-time compute of modern AI: executing the multi-step drafts of a l

ChronoFlow-Policy: Unifying Past-Current-Future Interaction Flow in Visuomotor Policy Learning

SafetyDGX agent

arXiv:2606.31493v1 Announce Type: new Abstract: Visual signals play a crucial role in policy learning by enabling models to capture object motion and interaction dynamics. Just as humans reason about

CLExEval: A Human-in-the-Loop Framework for Qualitative Evaluation of LLM Clinical Reasoning

SafetyDGX agent

arXiv:2606.31608v1 Announce Type: new Abstract: Large Language Models (LLMs) achieve strong results on many medical benchmarks, but their clinical reasoning remains difficult to evaluate reliably. A c

CLOUDADV: Decision-Aligned Instance Sizing with Zero-Shot Foundation Models under Drift

SafetyDGX agent

arXiv:2606.31470v1 Announce Type: new Abstract: Cloud virtual machines are often overprovisioned, creating avoidable cost and operational inefficiency. We present CLOUDADV, an interactive engineer-fac

Corruption Robust Offline Reinforcement Learning with Human Feedback

SafetyDGX agent

arXiv:2402.06734v2 Announce Type: replace-cross Abstract: We study data corruption robustness for reinforcement learning with human feedback (RLHF) in an offline setting. Given an offline dataset of p

Cross-Space Distillation: Teaching One-Step Students with Modern Diffusion Teachers

SafetyDGX agent

arXiv:2606.32020v1 Announce Type: new Abstract: Modern one-step diffusion models achieve impressive quality through distribution-based timestep distillation. Yet, they rely on a critical assumption: T

DataEvolver: Self-Evolving Multi-Agent Data Construction for Text-Rich Image Generation

SafetyDGX agent

arXiv:2606.31537v1 Announce Type: new Abstract: Text-rich image generation is one of the most challenging settings in image generation, since models must simultaneously produce visually realistic imag

DDIAgents: Mechanism-Conditioned Context Flow for Drug-Drug Interaction Prediction

SafetyDGX agent

arXiv:2606.31085v1 Announce Type: new Abstract: Drug-drug interaction (DDI) prediction is essential for medication safety, yet it requires reasoning over heterogeneous biomedical evidence whose releva

Deep Reinforcement Learning for Spacecraft Attitude Control During Atmospheric Re-Entry

SafetyDGX agent

arXiv:2606.31291v1 Announce Type: new Abstract: Deep reinforcement learning has the potential to solve attitude control problems more adaptively, precisely, and robustly by handling nonlinear dynamics

Designing Privacy-Preserving Visual Perception for Robot Navigation Based on User Privacy Preferences

SafetyDGX agent

arXiv:2604.06382v2 Announce Type: replace Abstract: Visual navigation is a fundamental capability of mobile service robots, yet the onboard cameras required for such navigation can capture privacy-sen

DISCOVER: A Solver for Distributional Counterfactual Explanations

SafetyDGX agent

arXiv:2603.16436v2 Announce Type: replace Abstract: Counterfactual explanations (CE) explain model decisions by identifying input modifications that lead to different predictions. Most existing method

ECHO: Prune to act, trace to learn with selective turn memory in agentic RL

SafetyDGX agent

arXiv:2606.31650v1 Announce Type: cross Abstract: Long-horizon language agents must repeatedly interact with tools, accumulate evidence, and make decisions under bounded context windows. Existing cont

Efficient Sim-to-Real Transfer of World-Action Models from Synthetic Priors

SafetyDGX agent

arXiv:2606.31101v1 Announce Type: new Abstract: Bridging the sim-to-real gap is a core challenge in deploying learned manipulation policies. Sim-to-real learning is attractive because it can replace e

EgoVITA: Learning to Plan and Verify for Egocentric Video Reasoning

SafetyDGX agent

arXiv:2511.18242v3 Announce Type: replace Abstract: Egocentric video understanding requires procedural reasoning under partial observability and continuously shifting viewpoints. Current multimodal la

ELASTIC: Efficiently Learning to Adaptively Scale Test-Time Compute for Generative Control Policies

SafetyDGX agent

arXiv:2606.31132v1 Announce Type: new Abstract: Generative control policies (GCPs), such as diffusion policies and flow-based vision-language-action models, enable test-time scaling in robot control.

End-to-End Efficient RL for Linear Bellman Complete MDPs with Deterministic Transitions

SafetyDGX agent

arXiv:2603.23461v2 Announce Type: replace Abstract: We study reinforcement learning (RL) with linear function approximation in Markov Decision Processes (MDPs) satisfying linear Bellman completeness -

ERA: Entropy-Guided Visual Token Pruning with Rectified Attention for Efficient MLLMs

SafetyDGX agent

arXiv:2606.31982v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) incur prohibitive inference costs due to long visual token sequences. Training-free visual token reduction prov

Evil Spectra: How Optimisers can Amplify or Suppress Emergent Misalignment

SafetyDGX agent

arXiv:2606.31591v1 Announce Type: cross Abstract: Emergent misalignment (EM) is a recently discovered phenomenon in LLMs where fine-tuning on a narrow misaligned task, such as writing insecure code, l

Evo-PI: Aligning Medical Reasoning via Evolving Principle-Guided Supervision

SafetyDGX agent

arXiv:2606.31800v1 Announce Type: new Abstract: Despite recent progress, the reasoning capabilities of large multimodal language models (MLLMs) remain fundamentally constrained by static supervision,

Explainable Artificial Intelligence For The Detection and Characterisation of Stage B Heart Failure

SafetyDGX agent

arXiv:2606.30665v1 Announce Type: cross Abstract: Stage B heart failure is characterized by asymptomatic structural or functional cardiac abnormalities. Identifying individuals at this stage is clinic

ExPLoRe: Expert Patch-Level Loss Routing for Multi-Objective Masked Image Modeling

SafetyDGX agent

arXiv:2606.31201v1 Announce Type: new Abstract: Multi-objective masked image modeling (MIM) combines complementary learning signals (token distillation, CLS alignment, and pixel reconstruction) but ex

FastDSAC: Enhancing Policy Plasticity via Constrained Exploration for Scalable Humanoid Locomotion

SafetyDGX agent

arXiv:2606.31691v1 Announce Type: new Abstract: Scalable reinforcement learning has popularized high-throughput sampling architectures, which significantly compresses the training time for off-policy

Fine-grained Motion Retrieval via Joint-Angle Motion Images and Token-Patch Late Interaction

SafetyDGX agent

arXiv:2603.09930v2 Announce Type: replace Abstract: Text-motion retrieval aims to learn a semantically aligned latent space between natural language descriptions and 3D human motion skeleton sequences

Finite Difference Flow Optimization for RL Post-Training of Text-to-Image Models

SafetyDGX agent

arXiv:2603.12893v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become a standard technique for post-training diffusion-based image synthesis models, as it enables learning f

FLARE-AI: Flaw Reporting for AI

SafetyDGX agent

arXiv:2606.31567v1 Announce Type: cross Abstract: Flaw reporting for deployed AI systems is fundamental to identifying system failures and improving AI safety. Yet the AI reporting ecosystem is fragme

Freeform Preference Learning for Robotic Manipulation

SafetyDGX agent

arXiv:2606.32027v1 Announce Type: cross Abstract: Reward design remains a central bottleneck for autonomous robot policy improvement, especially in long-horizon manipulation tasks where sparse success

From Failure to Alignment: A Requirements Engineering Framework for Machine Learning Systems

SafetyDGX agent

arXiv:2606.31589v1 Announce Type: cross Abstract: Organisations designing, developing, and deploying machine learning systems (MLS) need to be able to check that these systems are trustworthy, and com

From Propositional to Perceptual Asymmetry: Extending Frictive Policy Optimization to Asymmetric Partial Information Dialogue

SafetyDGX agent

arXiv:2606.30973v1 Announce Type: new Abstract: Frictive Policy Optimization (FPO; Pustejovsky et al., 2025) treats friction in collaborative dialogue -- misalignment, misunderstanding, repair -- as a

G2P: Gaussian-to-Point Attribute Alignment for Boundary-Aware 3D Segmentation

SafetyDGX agent

arXiv:2601.03510v3 Announce Type: replace Abstract: Point cloud segmentation is critical for 3D scene understanding. However, sparse and irregular point distributions provide limited appearance eviden

GEAR: Guided End-to-End AutoRegression for Image Synthesis

SafetyDGX agent

arXiv:2606.32039v1 Announce Type: new Abstract: Visual generative models are typically trained in two stages. A tokenizer is first trained for reconstruction and then frozen, after which a generator i

GR2 Technical Report

SafetyDGX agent

arXiv:2606.31984v1 Announce Type: cross Abstract: Industrial recommendation systems serve billions of users through a multi-stage funnel -- retrieval, early-stage ranking, and re-ranking -- where the

GRAPE: Graph-Augmented Prototype Explanations for Interactive Medical Image Diagnosis

SafetyDGX agent

arXiv:2606.30901v1 Announce Type: new Abstract: Prototype-based medical image classifiers present three clinical limitations: they treat findings as independent, silently amplify unsafe physician feed

Great new episode with @So8res on @PeterMcCormack's podcast, give it a watch! https://www.youtube.com/watch?v=1laTnwAdaLA

SafetyDGX agent

Connor Leahy promoted a podcast episode featuring So8res (a researcher/AI safety figure) on Peter McCormack's podcast, sharing a YouTube link to the full video on X/Twitter. The post encourages viewer

GUIDE: Resolving Domain Bias in GUI Agents through Real-Time Web Video Retrieval and Plug-and-Play Annotation

SafetyDGX agent

arXiv:2603.26266v3 Announce Type: replace Abstract: Large vision-language models have endowed GUI agents with strong general capabilities for interface understanding and interaction. However, due to i

Harnessing Textual Refusal Directions for Multimodal Safety

SafetyDGX agent

arXiv:2606.31876v1 Announce Type: new Abstract: To improve safety in Large Language Models (LLMs) we can either perform post-training alignment or exploit refusal directions in the activation space. B

HistoriQA-ThirdRepublic: Multi-Hop Question Answering Corpus for Historical Research, Parliamentary Debates from the French Third Republic (1870-1940)

SafetyDGX agent

arXiv:2606.31325v1 Announce Type: new Abstract: We present HistoriQA-ThirdRepublic: a French-language dataset of multi-hop historical questions derived from parliamentary debates and newspapers of the

How Human Feedback Shapes AI-generated Community Notes

SafetyDGX agent

arXiv:2606.30905v1 Announce Type: cross Abstract: Community Notes, a bridging-based crowd-sourced fact-checking system, has emerged as a new mechanism for moderating misleading information on social m

Incentive Aware AI Regulations: A Credal Characterisation

SafetyDGX agent

arXiv:2603.05175v2 Announce Type: replace Abstract: The rapid proliferation of AI applications has intensified debate on effective regulation of these black-box services. Effective regulation must bal

InfiniVerse: Occupancy Guided Unbounded Scene Generation for Autonomous Driving

SafetyDGX agent

arXiv:2606.31109v1 Announce Type: new Abstract: Generating realistic, controllable, and temporally coherent urban environments is a critical yet unresolved challenge in the autonomous driving communit

IterCAD: An Iterative Multimodal Agent for Visually-Grounded CAD Generation and Editing

SafetyDGX agent

arXiv:2606.13368v2 Announce Type: replace Abstract: Computer-Aided Design is pivotal in modern manufacturing, yet existing automated methods predominantly rely on open-loop, one-shot generation, creat

Keep Policy Gradient in Charge: Sibling-Guided Credit Distillation for Long-Horizon Tool-Use Agents

SafetyDGX agent

arXiv:2606.12634v2 Announce Type: replace-cross Abstract: Long-horizon tool-use reinforcement learning learns from outcome verification, but trajectory-level advantages are broadcast over reasoning, A

LabGuard: Grounding Natural-Language Laboratory Rules into Runtime Guards for Embodied Laboratory Agents

SafetyDGX agent

arXiv:2606.31045v1 Announce Type: new Abstract: Scientific embodied agents are increasingly capable of carrying out laboratory procedures, but executing these procedures safely in dynamic laboratory e

LaMP: Learning Vision-Language-Action Policy with 3D Scene Flow as Latent Motion Prior

SafetyDGX agent

arXiv:2603.25399v2 Announce Type: replace Abstract: We introduce extbf{LaMP}, a dual-expert Vision-Language-Action framework that embeds dense 3D scene flow as a latent motion prior for robotic manipu

Language-Assisted Super-Resolution from Real-World Low-Resolution Patches

SafetyDGX agent

arXiv:2606.31363v1 Announce Type: new Abstract: Single image super-resolution aims to reconstruct high-resolution (HR) images from low-resolution (LR) inputs. Training SR models typically requires pai

Layout-Conditioned Autoregressive Text-to-Image Generation via Structured Masking

SafetyDGX agent

arXiv:2509.12046v2 Announce Type: replace-cross Abstract: Although autoregressive (AR) models have demonstrated remarkable success in image generation, extending these models to layout-conditioned gen

Learning Dexterous Grasping from Sparse Taxonomy Guidance

SafetyDGX agent

arXiv:2604.04138v2 Announce Type: replace-cross Abstract: Dexterous manipulation requires planning a grasp configuration suited to the object and task, which is then executed through coordinated multi

Learning Structurally Consistent Representations for Multi-View Radar Semantic Segmentation

SafetyDGX agent

arXiv:2606.31609v1 Announce Type: cross Abstract: Radar sensors provide reliable perception under adverse weather and lighting conditions, but their sparse, noisy, and weakly semantic measurements mak

Learning Where to Look: A Reinforcement Learning Framework for Robust Micro-Ultrasound Prostate Cancer Detection

SafetyDGX agent

arXiv:2606.30951v1 Announce Type: cross Abstract: Micro-ultrasound (muUS) is a new, emerging, and promising imaging modality for prostate cancer (PCa) detection, but accurate identification of suspici

← Previous
1…5051525354…212
Next →