AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,708 results
12 May 2026

DAPE: Dynamic Non-uniform Alignment and Progressive Detail Enhancement Techniques for Improving the Performance of Efficient Visual Language Models

SafetyDGX agent

arXiv:2605.08902v1 Announce Type: cross Abstract: In recent years, pre-trained visual-linguistic models have demonstrated tremendous potential, becoming a crucial foundational framework for numerous d

DARE: Difficulty-Adaptive Reinforcement Learning with Co-Evolved Difficulty Estimation

SafetyDGX agent

arXiv:2605.09188v1 Announce Type: cross Abstract: Reinforcement learning improves the reasoning ability of large language models but remains costly and sample-inefficient, as many rollouts provide wea

Data-driven transport modelling without overfit

SafetyDGX agent

arXiv:2605.08801v1 Announce Type: new Abstract: Macroscopic transport modelling aims to predict traffic flows after proposed public policy interventions, such as a new road or railway section or a tem


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Debugging the Debuggers: Failure-Anchored Structured Recovery for Software Engineering Agents

SafetyDGX agent

arXiv:2605.08717v1 Announce Type: cross Abstract: Software engineering agents are increasingly deployed in evaluable engineering environments, yet post-failure recovery remains costly, manual, and ad

Decoupling Endpoint and Semantic Transition Learning for Zero-Shot Composed Image Retrieval

SafetyDGX agent

arXiv:2605.08389v1 Announce Type: cross Abstract: Zero-shot composed image retrieval (ZS-CIR) retrieves a target image from a reference image and a text modification without human-annotated CIR triple

Dependency-Aware Discrete Diffusion for Scene Graph Generation

SafetyDGX agent

arXiv:2605.09065v1 Announce Type: new Abstract: Scene graphs (SGs) represent objects and their relationships as structured graphs, enabling applications in image generation, robotics, and 3D understan

DexWrist: A Robotic Wrist for Constrained and Dynamic Manipulation

SafetyDGX agent

arXiv:2507.01008v3 Announce Type: replace Abstract: Development of dexterous manipulation hardware has primarily focused on hands and grippers. However, these end-effectors are often paired with bulky

dFlowGRPO: Rate-Aware Policy Optimization for Discrete Flow Models

SafetyDGX agent

arXiv:2605.09291v1 Announce Type: new Abstract: Discrete flow models (DFMs) are a class of flexible generative models for generating discrete data, and diffusion large language models (dLLMs) can be v

DGPO: Beyond Pairwise Preferences with Directional Consistent Groupwise Optimization

SafetyDGX agent

arXiv:2605.10863v1 Announce Type: new Abstract: Although Large Language Models (LLMs) have made remarkable progress, current preference optimization methods still struggle to align directional consist

Disentangled Representation Learning via Flow Matching

SafetyDGX agent

arXiv:2602.05214v2 Announce Type: replace Abstract: Disentangled representation learning aims to capture the underlying explanatory factors of observed data, enabling a principled understanding of the

Do Linear Probes Generalize Better in Persona Coordinates?

SafetyDGX agent

arXiv:2605.09391v1 Announce Type: new Abstract: It is becoming increasingly necessary to have monitors check for harmful behaviors during language model interactions, but text-only monitoring has not

Drum Synthesis from Expressive Drum Grids via Neural Audio Codecs

SafetyDGX agent

arXiv:2605.10281v1 Announce Type: cross Abstract: Generating realistic drum audio directly from symbolic representations is a challenging task at the intersection of music perception and machine learn

DuetFair: Coupling Inter- and Intra-Subgroup Robustness for Fair Medical Image Segmentation

SafetyDGX agent

arXiv:2605.10521v1 Announce Type: cross Abstract: Medical image segmentation models can perform unevenly across subgroups. Most existing fairness methods focus on improving average subgroup performanc

Dynamic Skill Lifecycle Management for Agentic Reinforcement Learning

SafetyDGX agent

arXiv:2605.10923v1 Announce Type: cross Abstract: Large language model agents increasingly rely on external skills to solve complex tasks, where skills act as modular units that extend their capabilit

DynaMiCS: Fine-tuning LLMs with Performance Constraints using Dynamic Mixtures

SafetyDGX agent

arXiv:2605.10770v1 Announce Type: new Abstract: Multi-domain fine-tuning of large language models requires improving performance on target domains while preserving performance on constrained domains,

E-TCAV: Formalizing Penultimate Proxies for Efficient Concept Based Interpretability

SafetyDGX agent

arXiv:2605.10261v1 Announce Type: new Abstract: TCAV (Testing with Concept Activation Vectors) is an interpretability method that assesses the alignment between the internal representations of a train

Effective Explanations Support Planning Under Uncertainty

SafetyDGX agent

arXiv:2605.08406v1 Announce Type: cross Abstract: Explaining how to get from A to B can be challenging. It requires mentally simulating what the listener will do based on what they are told. To captur

EGL-SCA: Structural Credit Assignment for Co-Evolving Instructions and Tools in Graph Reasoning Agents

SafetyDGX agent

arXiv:2605.10366v1 Announce Type: new Abstract: Graph reasoning agents operating from natural-language inputs must solve a coupled problem: they must reconstruct a structured graph instance from text,

ElasticFlow: One-Step Physics-Consistent Policy with Elastic Time Horizons for Language-Guided Manipulation

SafetyDGX agent

arXiv:2605.08799v1 Announce Type: new Abstract: Diffusion policies have demonstrated exceptional performance in embodied AI. However, their iterative denoising process results in high latency, and exi

Embodied AI in Action: Insights from SAE World Congress 2026 on Safety, Trust, Robotics, and Real-World Deployment

SafetyDGX agent

arXiv:2605.10653v1 Announce Type: new Abstract: Embodied artificial intelligence is rapidly moving from research into real-world systems such as autonomous vehicles, mobile robots, and industrial mach

Emergence of Physical Intelligence via Controllable Information Production

SafetyDGX agent

arXiv:2601.22449v2 Announce Type: replace Abstract: Intrinsic Motivation (IM) aims to train agents without external rewards, enabling useful behavior to emerge from the agent's interaction with its en

Empowering VLMs for Few-Shot Multimodal Time Series Classification via Tailored Agentic Reasoning

SafetyDGX agent

arXiv:2605.09395v1 Announce Type: new Abstract: In this paper, we propose the first VLnderline{extbf{M}} nderline{extbf{a}}gentic nderline{extbf{r}}easoning framework for few-nderline{extbf{s}}hot mul

Equivariant Reinforcement Learning for Clifford Quantum Circuit Synthesis

SafetyDGX agent

arXiv:2605.10910v1 Announce Type: cross Abstract: We consider the problem of synthesizing Clifford quantum circuits for devices with all-to-all qubit connectivity. We approach this task as a reinforce

EROAS: 3D Efficient Reactive Obstacle Avoidance System for Autonomous Underwater Vehicles using 2.5D Forward-Looking Sonar

SafetyDGX agent

arXiv:2411.05516v3 Announce Type: replace Abstract: Autonomous Underwater Vehicles (AUVs) have advanced significantly in obstacle detection and path planning through sonar, cameras, and learning-based

ETS: Energy-Guided Test-Time Scaling for Training-Free RL Alignment

SafetyDGX agent

arXiv:2601.21484v2 Announce Type: replace Abstract: Reinforcement Learning (RL) post-training alignment for language models is effective, but also costly and unstable in practice, owing to its complic

EvoMAS: Learning Execution-Time Workflows for Multi-Agent Systems

SafetyDGX agent

arXiv:2605.08769v1 Announce Type: new Abstract: Large language model (LLM)-based multi-agent systems have shown strong potential on complex tasks through agent specialization, tool use, and collaborat

EvoPref: Multi-Objective Evolutionary Optimization Discovers Diverse LLM Alignments Beyond Gradient Descent

SafetyDGX agent

arXiv:2605.09777v1 Announce Type: cross Abstract: Gradient-based preference optimization methods for large language model (LLM) alignment suffer from preference collapse, converging to narrow behavior

EvoStreaming: Your Offline Video Model Is a Natively Streaming Assistant

SafetyDGX agent

arXiv:2605.10343v1 Announce Type: cross Abstract: Streaming video understanding demands more than watching longer videos: assistants must decide when to speak in real time, balancing responsiveness ag

Execution Envelopes: A Shared Admission Contract for Backend AI Execution Requests

SafetyDGX agent

arXiv:2605.08267v1 Announce Type: cross Abstract: Enterprise AI backends increasingly admit heterogeneous execution requests across model deployment, inference, evaluation, data movement, and agentic

Expert Evaluation and the Limits of Human Feedback in Mental Health AI Safety Testing

SafetyDGX agent

arXiv:2601.18061v3 Announce Type: replace Abstract: Learning from human feedback~(LHF) assumes that expert judgments, appropriately aggregated, yield valid ground truth for training and evaluating AI

Explanation-Aware Learning for Enhanced Interpretability in Biomedical Imaging

SafetyDGX agent

arXiv:2605.10054v1 Announce Type: new Abstract: Deep neural networks for medical image diagnosis often achieve high predictive accuracy while relying on spurious or clinically irrelevant visual cues,

Explicit Stair Geometry Conditioning for Robust Humanoid Locomotion

SafetyDGX agent

arXiv:2605.09944v1 Announce Type: new Abstract: Robust humanoid stair climbing remains challenging due to geometric discontinuities, sensitivity to step height variations, and perception uncertainty i

Exploration-Driven Optimization for Test-Time Large Language Model Reasoning

SafetyDGX agent

arXiv:2605.09853v1 Announce Type: new Abstract: Post-training techniques combined with inference-time scaling significantly enhance the reasoning and alignment capabilities of large language models (L

Extended Wasserstein-GAN Approach to Causal Distribution Learning: Density-Free Estimation and Minimax Optimality

SafetyDGX agent

arXiv:2605.10206v1 Announce Type: cross Abstract: Distributional causal inference requires estimating not only average treatment effects but also interventional outcome distributions, including quanti

f-GRPO and Beyond: Divergence-Based Reinforcement Learning Algorithms for General LLM Alignment

SafetyDGX agent

arXiv:2602.05946v3 Announce Type: replace Abstract: Recent work shows that preference alignment objectives can be interpreted as divergence estimators between aligned (preferred) & unaligned (less-pre

FairHealth: An Open-Source Python Library for Trustworthy Healthcare AI in Low-Resource Settings

SafetyDGX agent

arXiv:2605.08198v1 Announce Type: cross Abstract: We present FairHealth, an open-source Python library that provides a unified, modular framework for trustworthy machine learning in healthcare applica

Fairness of Explanations in Artificial Intelligence (AI): A Unifying Framework, Axioms, and Future Direction toward Responsible AI

SafetyDGX agent

arXiv:2605.09852v1 Announce Type: new Abstract: Machine learning algorithms are being used in high-stakes decisions, including those in criminal justice, healthcare, credit, and employment. The resear

Fairness vs Performance: Characterizing the Pareto Frontier of Algorithmic Decision Systems

SafetyDGX agent

arXiv:2605.10604v1 Announce Type: cross Abstract: Designing fair algorithmic decision systems requires balancing model performance with fairness toward affected individuals: More fairness might requir

Fast Rates for Offline Contextual Bandits with Forward-KL Regularization under Single-Policy Concentrability

SafetyDGX agent

arXiv:2605.09214v1 Announce Type: cross Abstract: Kullback-Leibler (KL) regularization is ubiquitous in reinforcement learning algorithms in the form of reverse or forward KL. Recent studies have demo

Few-Click-Driven Interactive 3D Segmentation with Semantic Embedding

SafetyDGX agent

arXiv:2605.08925v1 Announce Type: new Abstract: Interactive segmentation allows efficient label generation by leveraging user-provided clicks to progressively refine predictions, which is critical whe

Flag Varieties: A Geometric Framework for Deep Network Alignment

SafetyDGX agent

arXiv:2605.09861v1 Announce Type: cross Abstract: Alignment, the tendency of adjacent weight matrices in deep networks to develop compatible subspace orientations, underlies gradient flow, Neural Coll

FlashAR: Efficient Post-Training Acceleration for Autoregressive Image Generation

SafetyDGX agent

arXiv:2605.09430v1 Announce Type: new Abstract: Large-scale autoregressive models have demonstrated remarkable capabilities in image generation. However, their sequential raster-scan decoding relies o

Force Policy: Learning Hybrid Force-Position Control Policy under Interaction Frame for Contact-Rich Manipulation

SafetyDGX agent

arXiv:2602.22088v2 Announce Type: replace Abstract: Contact-rich manipulation demands human-like integration of perception and force feedback: vision should guide task progress, while high-frequency i

Formal Policy Enforcement for Real-World Agentic Systems

SafetyDGX agent

arXiv:2602.16708v3 Announce Type: replace-cross Abstract: Security policy enforcement in contemporary agentic systems predominantly consists of embedding natural-language policies within an agent's sy

FrameTwin: Curve-Anchored Gaussian Alignment from Sparse Views for Adaptive Wireframe 3D Printing

SafetyDGX agent

arXiv:2605.09362v1 Announce Type: cross Abstract: We present FrameTwin, a curve-anchored Gaussian alignment framework that uses sparse-view images to close the control loop for adaptive wireframe 3D p

Frequency Adapter with SAM for Generalized Medical Image Segmentation

SafetyDGX agent

arXiv:2605.09925v1 Announce Type: new Abstract: Medical image segmentation is a critical task in computer-aided diagnosis and treatment planning. However, deep learning models often struggle to genera

Frictional Q-Learning

SafetyDGX agent

arXiv:2509.19771v5 Announce Type: replace-cross Abstract: Off-policy reinforcement learning suffers from extrapolation errors when a learned policy selects actions that are weakly supported in the rep

From Holo Pockets to Electron Density: GPT-style Drug Design with Density

SafetyDGX agent

arXiv:2605.08767v1 Announce Type: new Abstract: Recent advances in generative modeling have enabled significant progress in structure-based drug design (SBDD). Existing methods typically condition mol

From Passive Reuse to Active Reasoning: Grounding Large Language Models for Neuro-Symbolic Experience Replay

SafetyDGX agent

arXiv:2605.09419v1 Announce Type: new Abstract: While experience replay is essential for data efficiency in reinforcement learning (RL), standard methods treat the replay buffer as a passive memory sy

Functional Graphs for Predicting and Explaining Goal Failure in Sparse Goal-Conditioned RL

SafetyDGX agent

arXiv:2605.09335v1 Announce Type: new Abstract: Sparse goal-conditioned reinforcement learning can produce policies whose failures are hidden by aggregate success rates. We analyze trained goal-condit

GemDepth: Geometry-Embedded Features for 3D-Consistent Video Depth

SafetyDGX agent

arXiv:2605.10525v1 Announce Type: new Abstract: Video depth estimation extends monocular prediction into the temporal domain to ensure coherence. However, existing methods often suffer from spatial bl

Gender Fairness in Audio Deepfake Detection: Performance and Disparity Analysis

SafetyDGX agent

arXiv:2603.09007v2 Announce Type: replace-cross Abstract: Audio deepfake detection aims to detect real human voices from those generated by Artificial Intelligence (AI) and has emerged as a significan

Generalized Category Discovery in Federated Graph Learning

SafetyDGX agent

arXiv:2605.08178v1 Announce Type: cross Abstract: Federated Graph Learning (FGL) enables collaborative learning over distributed graph data, yet existing approaches largely rely on a closed-world assu

Generative Adversarial Post-Training Mitigates Reward Hacking in Live Human-AI Music Interaction

SafetyDGX agent

arXiv:2511.17879v4 Announce Type: replace Abstract: Most applications of generative AI involve a sequential interaction in which a person inputs a prompt and waits for a response, and where reaction t

GNN for Structural Displacement Prediction

SafetyDGX agent

arXiv:2605.08303v1 Announce Type: cross Abstract: Accurate prediction of structural displacements under external loading is fundamental to structural health monitoring and seismic safety assessment. A

Good science starts with intellectual honesty, and I am not seeing it. I 100% believe the quote below and zero percent believe the seemingly…

SafetyDGX agent

Good science starts with intellectual honesty, and I am not seeing it. I 100% believe the quote below and zero percent believe the seemingly similar– but actually very different—quote that @geoffreyhi

Governed Metaprogramming for Intelligent Systems: Reclassifying Eval as a Governed Effect

SafetyDGX agent

arXiv:2605.05248v2 Announce Type: replace-cross Abstract: AI systems increasingly synthesize executable structure at runtime: LLMs generate programs, agents construct workflows,self-improving systems

Governing AI-Assisted Security Operations: A Design Science Framework for Operational Decision Support

SafetyDGX agent

arXiv:2605.09534v1 Announce Type: cross Abstract: Engineering managers increasingly must decide how to introduce generative artificial intelligence (AI), retrieval-augmented generation, and coding age

GuardAD: Safeguarding Autonomous Driving MLLMs via Markovian Safety Logic

SafetyDGX agent

arXiv:2605.10386v1 Announce Type: new Abstract: Multimodal large language models (MLLMs) are increasingly integrated into autonomous driving (AD) systems; however, they remain vulnerable to diverse sa

Guided Streaming Stochastic Interpolant Policy

SafetyDGX agent

arXiv:2605.10051v1 Announce Type: cross Abstract: Inference-time guidance is essential for steering generative robot policies toward dynamic objectives without retraining, yet existing methods are lar

← Previous
1…148149150151152…212
Next →