AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

Can Compact Language Models Search Like Agents? Distillation-Guided Policy Optimization for Preserving Agentic RAG Capabilities

DGX agent

arXiv:2508.20324v4 Announce Type: replace Abstract: Reinforcement Learning has emerged as a dominant post-training approach to elicit agentic RAG behaviors such as search and planning from language mo

safetyarxiv-cs-cl
28 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

CAP-CoT: Cycle Adversarial Prompt for Improving Chain of Thoughts in LLM Reasoning

DGX agent

arXiv:2604.23270v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has emerged as a simple and effective way to elicit step-by-step solutions from large language models (LLMs). However,

safetyarxiv-cs-ai
28 Apr 2026
Safety

CAPSULE: Control-Theoretic Action Perturbations for Safe Uncertainty-Aware Reinforcement Learning

DGX agent

arXiv:2604.23576v1 Announce Type: cross Abstract: Ensuring safe exploration in high-dimensional systems with unknown dynamics remains a significant challenge. Existing safe reinforcement learning meth

safetyarxiv-cs-ai
28 Apr 2026
Safety

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems

DGX agent

arXiv:2604.23022v1 Announce Type: cross Abstract: Two-stage recommender systems first choose a candidate generator and then rank items within the generated set. Because the generator decides which ite

safetyarxiv-cs-lg
28 Apr 2026
Safety

Certified geometric robustness -- Super-DeepG

DGX agent

arXiv:2604.24379v1 Announce Type: new Abstract: Safety-critical applications are required to perform as expected in normal operations. Image processing functions are often required to be insensitive t

safetyarxiv-cs-ai
28 Apr 2026
Safety

CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning

DGX agent

arXiv:2604.23308v1 Announce Type: new Abstract: Offline multi-agent reinforcement learning (MARL) enables policy learning from fixed datasets, but is prone to coordination failure: agents trained on s

safetyarxiv-cs-lg
28 Apr 2026
Safety

CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs

DGX agent

arXiv:2604.22785v1 Announce Type: new Abstract: Large language model (LLM) deployments increasingly rely on multi-agent architectures in which multiple models either compete through routing mechanisms

safetyarxiv-cs-lg
28 Apr 2026
Safety

CombiMOTS: Combinatorial Multi-Objective Tree Search for Dual-Target Molecule Generation

DGX agent

arXiv:2604.23307v1 Announce Type: cross Abstract: Dual-target molecule generation, which focuses on discovering compounds capable of interacting with two target proteins, has garnered significant atte

safetyarxiv-cs-ai
28 Apr 2026
Safety

COMO: Closed-Loop Optical Molecule Recognition with Minimum Risk Training

DGX agent

arXiv:2604.23546v1 Announce Type: cross Abstract: Optical chemical structure recognition (OCSR) translates molecular images into machine-readable representations like SMILES strings or molecular graph

safetyarxiv-cs-ai
28 Apr 2026
Safety

Complex SGD and Directional Bias in Reproducing Kernel Hilbert Spaces

DGX agent

arXiv:2604.23017v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is a known stochastic iterative method popular for large-scale convex optimization problems due to its simple implemen

safetyarxiv-cs-lg
28 Apr 2026
Safety

Computer Vision-Based Early Detection of Container Loss at Sea

DGX agent

arXiv:2604.24193v1 Announce Type: new Abstract: Containerised shipping underpins global trade, yet container loss at sea remains a persistent safety, environmental, and economic challenge. Despite com

safetyarxiv-cs-cv
28 Apr 2026
Safety

Conditional Imputation for Within-Modality Missingness in Multi-Modal Federated Learning

DGX agent

arXiv:2604.23112v1 Announce Type: new Abstract: Multimodal Federated Learning (MMFL) enables privacy-preserving collaborative training, but real-world clinical applications often suffer from within-mo

safetyarxiv-cs-lg
28 Apr 2026
Safety

Conflict-Aware Harmonized Rotational Gradient for Multiscale Kinetic Regimes

DGX agent

arXiv:2604.24745v1 Announce Type: new Abstract: In this paper, we propose a harmonized rotational gradient method, termed HRGrad, for simultaneously tackling multiscale time-dependent kinetic problems

safetyarxiv-cs-lg
28 Apr 2026
Safety

ConsDreamer: Advancing Multi-View Consistency for Zero-Shot Text-to-3D Generation

DGX agent

arXiv:2504.02316v4 Announce Type: replace-cross Abstract: Recent advances in zero-shot text-to-3D generation have revolutionized 3D content creation by enabling direct synthesis from textual descripti

safetyarxiv-cs-ai
28 Apr 2026
Safety

Context-Aware Hospitalization Forecasting Evaluations for Decision Support using LLMs

DGX agent

arXiv:2604.23949v1 Announce Type: new Abstract: Medical and public health experts must make real-time resource decisions, such as expanding hospital bed capacity, based on projected hospitalization tr

safetyarxiv-cs-ai
28 Apr 2026
Safety

Control Barrier Functions Solved with Hierarchical Quadratic Programming for Safe Physical Human-Robot Interaction

DGX agent

arXiv:2604.23039v1 Announce Type: new Abstract: Physical human-robot interaction offers the potential to leverage human intelligence and robot physical capabilities to enable a range of exciting appli

safetyarxiv-cs-ro
28 Apr 2026
Safety

Cooperative Informative Sensing for Monitoring Dynamic Indoor Environments via Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.23179v1 Announce Type: cross Abstract: Monitoring human activity in indoor environments is important for applications such as facility management, safety assessment, and space utilization a

safetyarxiv-cs-ai
28 Apr 2026
Safety

Cooptimizing Safety and Performance Using Safety Value-Constrained Model Predictive Control

DGX agent

arXiv:2604.23863v1 Announce Type: new Abstract: Autonomous systems are increasingly deployed in real-world environments, where they must achieve high performance while maintaining safety under state a

safetyarxiv-cs-ro
28 Apr 2026
Safety

CT-Guided Spatially-varying Regularization for Voxel-Wise Deformable Whole-Body PET Registration

DGX agent

arXiv:2604.22905v1 Announce Type: cross Abstract: Whole-body Positron Emission Tomography (PET) registration is essential for multi-parametric tumor characterization and assessment of metastatic disea

safetyarxiv-cs-ai
28 Apr 2026
Safety

CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning

DGX agent

arXiv:2601.13262v2 Announce Type: replace Abstract: While large language models (LLMs) have shown to perform well on monolingual mathematical and commonsense reasoning, they remain unreliable for mult

safetyarxiv-cs-ai
28 Apr 2026
Safety

Data-efficient Targeted Token-level Preference Optimization for LLM-based Text-to-Speech

DGX agent

arXiv:2510.05799v2 Announce Type: replace-cross Abstract: Aligning text-to-speech (TTS) system outputs with human feedback through preference optimization has been shown to effectively improve the rob

safetyarxiv-cs-ai
28 Apr 2026
Safety

Designing escalation criteria for international AI incident response: criteria, triggers, and thresholds

DGX agent

arXiv:2604.23183v1 Announce Type: cross Abstract: AI incident reporting requirements are emerging in regulation and policy, yet no operational criteria exist for determining when a detected AI inciden

safetyarxiv-cs-ai
28 Apr 2026
Safety

Designing Instance-Level Sampling Schedules via REINFORCE with James-Stein Shrinkage

DGX agent

arXiv:2511.22177v2 Announce Type: replace-cross Abstract: Most post-training methods for text-to-image samplers focus on model weights: either fine-tuning the backbone for alignment or distilling it f

safetyarxiv-cs-cv
28 Apr 2026
Safety

DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning

DGX agent

arXiv:2601.16046v2 Announce Type: replace-cross Abstract: Language-driven dexterous grasp generation requires the models to understand task semantics, 3D geometry, and complex hand-object interactions

safetyarxiv-cs-cv
28 Apr 2026
Safety

Discovering Agentic Safety Specifications from 1-Bit Danger Signals

DGX agent

arXiv:2604.23210v1 Announce Type: new Abstract: Can large language model agents discover hidden safety objectives through experience alone? We introduce EPO-Safe (Experiential Prompt Optimization for

safetyarxiv-cs-ai
28 Apr 2026
Safety

Discovering Failure Modes in Vision-Language Models using RL

DGX agent

arXiv:2604.04733v2 Announce Type: replace-cross Abstract: Vision-language Models (VLMs), despite achieving strong performance on multimodal benchmarks, often misinterpret straightforward visual concep

safetyarxiv-cs-ai
28 Apr 2026
Safety

DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making

DGX agent

arXiv:2604.23557v1 Announce Type: cross Abstract: Building scalable and reusable multi-agent decision policies from offline datasets remains a challenge in offline multi-agent reinforcement learning (

safetyarxiv-cs-ai
28 Apr 2026
Safety

Do Synthetic Trajectories Reflect Real Reward Hacking? A Systematic Study on Monitoring In-the-Wild Hacking in Code Generation

DGX agent

arXiv:2604.23488v1 Announce Type: new Abstract: Reward hacking in code generation, where models exploit evaluation loopholes to obtain full reward without correctly solving the tasks, poses a critical

safetyarxiv-cs-lg
28 Apr 2026
Safety

Do Transaction-Level and Actor-Level AML Queues Agree? An Empirical Evaluation of Granularity Effects on the Elliptic++ Graph

DGX agent

arXiv:2604.23494v1 Announce Type: new Abstract: Graph-based anti-money laundering (AML) systems on blockchain networks can score suspicious activity at two granularity levels -- transactions or actor

safetyarxiv-cs-ai
28 Apr 2026
Safety

Does Machine Unlearning Preserve Clinical Safety? A Risk Analysis for Medical Image Classification

DGX agent

arXiv:2604.23854v1 Announce Type: new Abstract: The application of Deep Learning in medical diagnosis must balance patient safety with compliance with data protection regulations. Machine Unlearning e

safetyarxiv-cs-ai
28 Apr 2026
Safety

DPEPO: Diverse Parallel Exploration Policy Optimization for LLM-based Agents

DGX agent

arXiv:2604.24320v1 Announce Type: new Abstract: Large language model (LLM) agents that follow the sequential 'reason-then-act' paradigm have achieved superior performance in many complex tasks.However

safetyarxiv-cs-cl
28 Apr 2026
Safety

DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models

DGX agent

arXiv:2604.24357v1 Announce Type: cross Abstract: Diffusion language models generate without a fixed left-to-right order, making token ordering a central algorithmic choice: which tokens should be rev

safetyarxiv-cs-ai
28 Apr 2026
Safety

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

DGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

safetyarxiv-cs-ai
28 Apr 2026
Safety

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence

DGX agent

arXiv:2604.23325v1 Announce Type: cross Abstract: Emotionally talking head video generation aims to generate expressive portrait videos with accurate lip synchronization and emotional facial expressio

safetyarxiv-cs-ai
28 Apr 2026
Safety

Early Warning of Intraoperative Adverse Events via Transformer-Driven Multi-Label Learning

DGX agent

arXiv:2603.05212v2 Announce Type: replace-cross Abstract: Early warning of intraoperative adverse events plays a vital role in reducing surgical risk and improving patient safety. While deep learning

safetyarxiv-cs-ai
28 Apr 2026
Safety

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

DGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

safetyarxiv-cs-cl
28 Apr 2026
Safety

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

DGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

safetyarxiv-cs-ai
28 Apr 2026
Safety

EU countries and lawmakers reach an impasse on a deal watering down the EU's AI Act due to some parties seeking exemptions for already regulated industries (Foo Yun Chee/Reuters)

DGX agent

Foo Yun Chee / Reuters: EU countries and lawmakers reach an impasse on a deal watering down the EU's AI Act due to some parties seeking exemptions for already regulated industries — EU countries and E

safetytechmeme
28 Apr 2026
Safety

Evaluating Language Models' Evaluations of Games

DGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

safetyarxiv-cs-ai
28 Apr 2026
Safety

Explanation Quality Assessment as Ranking with Listwise Rewards

DGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

safetyarxiv-cs-ai
28 Apr 2026
Safety

Extending Precipitation Nowcasting Horizons via Spectral Fusion of Radar Observations and Foundation Model Priors

DGX agent

arXiv:2603.21768v3 Announce Type: replace-cross Abstract: Precipitation nowcasting is critical for disaster mitigation and aviation safety. However, radar-only models frequently suffer from a lack of

safetyarxiv-cs-ai
28 Apr 2026
Safety

Extreme bandits

DGX agent

arXiv:2604.24545v1 Announce Type: cross Abstract: In many areas of medicine, security, and life sciences, we want to allocate limited resources to different sources in order to detect extreme values.

safetyarxiv-cs-lg
28 Apr 2026
Safety

Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents

DGX agent

arXiv:2604.23990v1 Announce Type: new Abstract: This paper presents PSA-Eval, a failure-centered runtime evaluation framework for deployed trilingual public-space agents. The central claim is that, wh

safetyarxiv-cs-ai
28 Apr 2026
Safety

FastOMOP: A Foundational Architecture for Reliable Agentic Real-World Evidence Generation on OMOP CDM data

DGX agent

arXiv:2604.24572v1 Announce Type: new Abstract: The Observational Medical Outcomes Partnership Common Data Model (OMOP CDM), maintained by the Observational Health Data Sciences and Informatics (OHDSI

safetyarxiv-cs-ai
28 Apr 2026
Safety

Federated Cross-Modal Retrieval with Missing Modalities via Semantic Routing and Adapter Personalization

DGX agent

arXiv:2604.22885v1 Announce Type: cross Abstract: Federated cross-modal retrieval faces severe challenges from heterogeneous client data, particularly non-IID semantic distributions and missing modali

safetyarxiv-cs-ai
28 Apr 2026
Safety

Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning

DGX agent

arXiv:2602.07605v3 Announce Type: replace-cross Abstract: Any entity in the visual world can be hierarchically grouped based on shared characteristics and mapped to fine-grained sub-categories. While

safetyarxiv-cs-ai
28 Apr 2026
Safety

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification

DGX agent

arXiv:2604.23588v1 Announce Type: new Abstract: Financial AI systems must produce answers grounded in specific regulatory filings, yet current LLMs fabricate metrics, invent citations, and miscalculat

safetyarxiv-cs-ai
28 Apr 2026
Safety

Fix Initial Codes and Iteratively Refine Textual Directions Toward Safe Multi-Turn Code Correction

DGX agent

arXiv:2604.23989v1 Announce Type: cross Abstract: Recent work on large language models (LLMs) has emphasized the importance of scaling inference compute. From this perspective, the state-of-the-art me

safetyarxiv-cs-ai
28 Apr 2026
← Previous
1…219220221222223…265
Next →