AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,704 results
28 Apr 2026

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems

SafetyDGX agent

arXiv:2604.23022v1 Announce Type: cross Abstract: Two-stage recommender systems first choose a candidate generator and then rank items within the generated set. Because the generator decides which ite

Certified geometric robustness -- Super-DeepG

SafetyDGX agent

arXiv:2604.24379v1 Announce Type: new Abstract: Safety-critical applications are required to perform as expected in normal operations. Image processing functions are often required to be insensitive t

CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning

SafetyDGX agent

arXiv:2604.23308v1 Announce Type: new Abstract: Offline multi-agent reinforcement learning (MARL) enables policy learning from fixed datasets, but is prone to coordination failure: agents trained on s


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs

SafetyDGX agent

arXiv:2604.22785v1 Announce Type: new Abstract: Large language model (LLM) deployments increasingly rely on multi-agent architectures in which multiple models either compete through routing mechanisms

CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule Generation

SafetyDGX agent

arXiv:2604.23307v1 Announce Type: cross Abstract: Dual-target molecule generation, which focuses on discovering compounds capable of interacting with two target proteins, has garnered significant atte

COMO: Closed-Loop Optical Molecule Recognition with Minimum Risk Training

SafetyDGX agent

arXiv:2604.23546v1 Announce Type: cross Abstract: Optical chemical structure recognition (OCSR) translates molecular images into machine-readable representations like SMILES strings or molecular graph

Complex SGD and Directional Bias in Reproducing Kernel Hilbert Spaces

SafetyDGX agent

arXiv:2604.23017v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is a known stochastic iterative method popular for large-scale convex optimization problems due to its simple implemen

Computer Vision-Based Early Detection of Container Loss at Sea

SafetyDGX agent

arXiv:2604.24193v1 Announce Type: new Abstract: Containerised shipping underpins global trade, yet container loss at sea remains a persistent safety, environmental, and economic challenge. Despite com

Conditional Imputation for Within-Modality Missingness in Multi-Modal Federated Learning

SafetyDGX agent

arXiv:2604.23112v1 Announce Type: new Abstract: Multimodal Federated Learning (MMFL) enables privacy-preserving collaborative training, but real-world clinical applications often suffer from within-mo

Conflict-Aware Harmonized Rotational Gradient for Multiscale Kinetic Regimes

SafetyDGX agent

arXiv:2604.24745v1 Announce Type: new Abstract: In this paper, we propose a harmonized rotational gradient method, termed HRGrad, for simultaneously tackling multiscale time-dependent kinetic problems

ConsDreamer: Advancing Multi-View Consistency for Zero-Shot Text-to-3D Generation

SafetyDGX agent

arXiv:2504.02316v4 Announce Type: replace-cross Abstract: Recent advances in zero-shot text-to-3D generation have revolutionized 3D content creation by enabling direct synthesis from textual descripti

Context-Aware Hospitalization Forecasting Evaluations for Decision Support using LLMs

SafetyDGX agent

arXiv:2604.23949v1 Announce Type: new Abstract: Medical and public health experts must make real-time resource decisions, such as expanding hospital bed capacity, based on projected hospitalization tr

Control Barrier Functions Solved with Hierarchical Quadratic Programming for Safe Physical Human-Robot Interaction

SafetyDGX agent

arXiv:2604.23039v1 Announce Type: new Abstract: Physical human-robot interaction offers the potential to leverage human intelligence and robot physical capabilities to enable a range of exciting appli

Cooperative Informative Sensing for Monitoring Dynamic Indoor Environments via Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2604.23179v1 Announce Type: cross Abstract: Monitoring human activity in indoor environments is important for applications such as facility management, safety assessment, and space utilization a

Cooptimizing Safety and Performance Using Safety Value-Constrained Model Predictive Control

SafetyDGX agent

arXiv:2604.23863v1 Announce Type: new Abstract: Autonomous systems are increasingly deployed in real-world environments, where they must achieve high performance while maintaining safety under state a

CT-Guided Spatially-varying Regularization for Voxel-Wise Deformable Whole-Body PET Registration

SafetyDGX agent

arXiv:2604.22905v1 Announce Type: cross Abstract: Whole-body Positron Emission Tomography (PET) registration is essential for multi-parametric tumor characterization and assessment of metastatic disea

CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning

SafetyDGX agent

arXiv:2601.13262v2 Announce Type: replace Abstract: While large language models (LLMs) have shown to perform well on monolingual mathematical and commonsense reasoning, they remain unreliable for mult

Data-efficient Targeted Token-level Preference Optimization for LLM-based Text-to-Speech

SafetyDGX agent

arXiv:2510.05799v2 Announce Type: replace-cross Abstract: Aligning text-to-speech (TTS) system outputs with human feedback through preference optimization has been shown to effectively improve the rob

Designing escalation criteria for international AI incident response: criteria, triggers, and thresholds

SafetyDGX agent

arXiv:2604.23183v1 Announce Type: cross Abstract: AI incident reporting requirements are emerging in regulation and policy, yet no operational criteria exist for determining when a detected AI inciden

Designing Instance-Level Sampling Schedules via REINFORCE with James-Stein Shrinkage

SafetyDGX agent

arXiv:2511.22177v2 Announce Type: replace-cross Abstract: Most post-training methods for text-to-image samplers focus on model weights: either fine-tuning the backbone for alignment or distilling it f

DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning

SafetyDGX agent

arXiv:2601.16046v2 Announce Type: replace-cross Abstract: Language-driven dexterous grasp generation requires the models to understand task semantics, 3D geometry, and complex hand-object interactions

Discovering Agentic Safety Specifications from 1-Bit Danger Signals

SafetyDGX agent

arXiv:2604.23210v1 Announce Type: new Abstract: Can large language model agents discover hidden safety objectives through experience alone? We introduce EPO-Safe (Experiential Prompt Optimization for

Discovering Failure Modes in Vision-Language Models using RL

SafetyDGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making

SafetyDGX agent

arXiv:2604.23557v1 Announce Type: cross Abstract: Building scalable and reusable multi-agent decision policies from offline datasets remains a challenge in offline multi-agent reinforcement learning (

Do Synthetic Trajectories Reflect Real Reward Hacking? A Systematic Study on Monitoring In-the-Wild Hacking in Code Generation

SafetyDGX agent

arXiv:2604.23488v1 Announce Type: new Abstract: Reward hacking in code generation, where models exploit evaluation loopholes to obtain full reward without correctly solving the tasks, poses a critical

Do Transaction-Level and Actor-Level AML Queues Agree? An Empirical Evaluation of Granularity Effects on the Elliptic++ Graph

SafetyDGX agent

arXiv:2604.23494v1 Announce Type: new Abstract: Graph-based anti-money laundering (AML) systems on blockchain networks can score suspicious activity at two granularity levels -- transactions or actor

Does Machine Unlearning Preserve Clinical Safety? A Risk Analysis for Medical Image Classification

SafetyDGX agent

arXiv:2604.23854v1 Announce Type: new Abstract: The application of Deep Learning in medical diagnosis must balance patient safety with compliance with data protection regulations. Machine Unlearning e

DPEPO: Diverse Parallel Exploration Policy Optimization for LLM-based Agents

SafetyDGX agent

arXiv:2604.24320v1 Announce Type: new Abstract: Large language model (LLM) agents that follow the sequential 'reason-then-act' paradigm have achieved superior performance in many complex tasks.However

DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models

SafetyDGX agent

arXiv:2604.24357v1 Announce Type: cross Abstract: Diffusion language models generate without a fixed left-to-right order, making token ordering a central algorithmic choice: which tokens should be rev

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

SafetyDGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence

SafetyDGX agent

arXiv:2604.23325v1 Announce Type: cross Abstract: Emotionally talking head video generation aims to generate expressive portrait videos with accurate lip synchronization and emotional facial expressio

Early Warning of Intraoperative Adverse Events via Transformer-Driven Multi-Label Learning

SafetyDGX agent

arXiv:2603.05212v2 Announce Type: replace-cross Abstract: Early warning of intraoperative adverse events plays a vital role in reducing surgical risk and improving patient safety. While deep learning

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

SafetyDGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

SafetyDGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

EU countries and lawmakers reach an impasse on a deal watering down the EU's AI Act due to some parties seeking exemptions for already regulated industries (Foo Yun Chee/Reuters)

SafetyDGX agent

Foo Yun Chee / Reuters: EU countries and lawmakers reach an impasse on a deal watering down the EU's AI Act due to some parties seeking exemptions for already regulated industries — EU countries and E

Evaluating Language Models' Evaluations of Games

SafetyDGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

Explanation Quality Assessment as Ranking with Listwise Rewards

SafetyDGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

Extending Precipitation Nowcasting Horizons via Spectral Fusion of Radar Observations and Foundation Model Priors

SafetyDGX agent

arXiv:2603.21768v3 Announce Type: replace-cross Abstract: Precipitation nowcasting is critical for disaster mitigation and aviation safety. However, radar-only models frequently suffer from a lack of

Extreme bandits

SafetyDGX agent

arXiv:2604.24545v1 Announce Type: cross Abstract: In many areas of medicine, security, and life sciences, we want to allocate limited resources to different sources in order to detect extreme values.

Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents

SafetyDGX agent

arXiv:2604.23990v1 Announce Type: new Abstract: This paper presents PSA-Eval, a failure-centered runtime evaluation framework for deployed trilingual public-space agents. The central claim is that, wh

FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data

SafetyDGX agent

arXiv:2604.24572v1 Announce Type: new Abstract: The Observational Medical Outcomes Partnership Common Data Model (OMOP CDM), maintained by the Observational Health Data Sciences and Informatics (OHDSI

Federated Cross-Modal Retrieval with Missing Modalities via Semantic Routing and Adapter Personalization

SafetyDGX agent

arXiv:2604.22885v1 Announce Type: cross Abstract: Federated cross-modal retrieval faces severe challenges from heterogeneous client data, particularly non-IID semantic distributions and missing modali

Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning

SafetyDGX agent

arXiv:2602.07605v3 Announce Type: replace-cross Abstract: Any entity in the visual world can be hierarchically grouped based on shared characteristics and mapped to fine-grained sub-categories. While

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification

SafetyDGX agent

arXiv:2604.23588v1 Announce Type: new Abstract: Financial AI systems must produce answers grounded in specific regulatory filings, yet current LLMs fabricate metrics, invent citations, and miscalculat

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

SafetyDGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

Fragmented AI policy threatens US leadership as government scrambles to keep pace

SafetyDGX agent

AI policy fragmentation is emerging as a critical risk for Washington, and without a federal standard, a patchwork of conflicting state-level rules threatens to undermine American competitiveness. Tha

From Optimization to Prediction: Transformer-Based Path-Flow Estimation to the Traffic Assignment Problem

SafetyDGX agent

arXiv:2510.19889v2 Announce Type: replace-cross Abstract: The traffic assignment problem is essential for traffic flow analysis, traditionally solved using mathematical programs under the Equilibrium

From Stateless Queries to Autonomous Actions: A Layered Security Framework for Agentic AI Systems

SafetyDGX agent

arXiv:2604.23338v1 Announce Type: cross Abstract: Agentic AI systems face security challenges that stateless large language models do not. They plan across extended horizons, maintain persistent memor

From what I can tell, Elon Musk just stepped on his own toes at the trial, making it all about him instead of the promises Altman and Brockm…

SafetyDGX agent

From what I can tell, Elon Musk just stepped on his own toes at the trial, making it all about him instead of the promises Altman and Brockman broke. OpenAI will slay his ego on cross. He should have

GAMED.AI: A Hierarchical Multi-Agent Framework for Automated Educational Game Generation

SafetyDGX agent

arXiv:2604.23947v1 Announce Type: new Abstract: We introduce GameDAI, a hierarchical multi-agent framework that transforms instructor-provided questions into fully playable, pedagogically grounded edu

Generative Control as Optimization: Time Unconditional Flow Matching for Adaptive and Robust Robotic Control

SafetyDGX agent

arXiv:2603.17834v2 Announce Type: replace-cross Abstract: Diffusion models and flow matching have become a cornerstone of robotic imitation learning, yet they suffer from a structural inefficiency whe

Google, 2001: Don’t Be Evil Google, 2026: How much does mass surveillance pay?

SafetyDGX agent

This post by AI researcher Gary Marcus contrasts Google's original 'Don't Be Evil' corporate motto from 2001 with a critical question about the company's 2026 practices, likely arguing that Google has

Governing What You Cannot Observe: Adaptive Runtime Governance for Autonomous AI Agents

SafetyDGX agent

arXiv:2604.24686v1 Announce Type: new Abstract: Autonomous AI agents can remain fully authorized and still become unsafe as behavior drifts, adversaries adapt, and decision patterns shift without any

Grammar-Constrained Refinement of Safety Operational Rules Using Language in the Loop: What Could Go Wrong

SafetyDGX agent

arXiv:2604.23523v1 Announce Type: cross Abstract: Safety specifications in cyber-physical systems (CPS) capture the operational conditions the system must satisfy to operate safely within its intended

Guided Speculative Inference for Efficient Test-Time Alignment of LLMs

SafetyDGX agent

arXiv:2506.04118v3 Announce Type: replace Abstract: We propose Guided Speculative Inference (GSI), a novel algorithm for efficient reward-guided decoding in large language models. GSI combines soft be

Hierarchical Prototype-based Domain Priors for Multiple Instance Learning in Multimodal Histopathology Analysis

SafetyDGX agent

arXiv:2604.23982v1 Announce Type: new Abstract: Digital pathology has fundamentally altered diagnostic workflows by enabling the computational analysis of gigapixel Whole Slide Images (WSIs), yet effe

Hierarchical RL-MPC Control for Dynamic Wake Steering in Wind Farms

SafetyDGX agent

arXiv:2604.22797v1 Announce Type: cross Abstract: Wind farm wake steering optimization is challenging due to complex flow physics and changing conditions. This paper presents a hierarchical framework

Hindsight Preference Optimization for Financial Time Series Advisory

SafetyDGX agent

arXiv:2604.23988v1 Announce Type: cross Abstract: Time series models predict numbers; decision-makers need advisory -- directional signals with reasoning, actionable suggestions, and risk management.

honestly, one of the worst things this nation has ever done to itself.

SafetyDGX agent

honestly, one of the worst things this nation has ever done to itself. No amount of words can convey how far behind this country will be in every avenue for a number of years because of this administr

Humanoid Whole-Body Badminton via Multi-Stage Reinforcement Learning

SafetyDGX agent

arXiv:2511.11218v3 Announce Type: replace Abstract: Humanoid robots have demonstrated strong capabilities for interacting with static scenes across locomotion and manipulation, yet dynamic real-world

← Previous
1…175176177178179…212
Next →