AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Additive Control Variates Dominate Self-Normalisation in Off-Policy Evaluation

DGX agent

arXiv:2602.14914v2 Announce Type: replace Abstract: Off-policy evaluation (OPE) is essential for assessing ranking and recommendation systems without costly online interventions. Self-Normalised Inver

safetyarxiv-cs-lg
28 Apr 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Adversary-Free Counterfactual Prediction via Information-Regularized Representations

DGX agent

arXiv:2510.15479v2 Announce Type: replace Abstract: We study counterfactual prediction under assignment bias and propose a mathematically grounded, information-theoretic approach that removes treatmen

safetyarxiv-cs-lg
28 Apr 2026
Safety

Algorithmic Administration and the EU AI Act: Legal Principles for Public Sector Use of AI

DGX agent

arXiv:2604.22765v1 Announce Type: cross Abstract: The increasing use of artificial intelligence (AI) by public authorities introduces both opportunities for innovation and significant challenges for t

safetyarxiv-cs-ai
28 Apr 2026
Safety

Aligning with Your Own Voice: Self-Corrected Preference Learning for Hallucination Mitigation in LVLMs

DGX agent

arXiv:2604.24395v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) frequently suffer from hallucinations. Existing preference learning-based approaches largely rely on proprietary mo

safetyarxiv-cs-ai
28 Apr 2026
Safety

ANCHOR: LLM-driven Subject Conditioning for Text-to-Image Synthesis

DGX agent

arXiv:2404.10141v2 Announce Type: replace-cross Abstract: Text-to-image (T2I) models have achieved remarkable progress in high-quality image synthesis, yet most benchmarks rely on simple, self-contain

safetyarxiv-cs-cl
28 Apr 2026
Safety

AnemiaVision: Non-Invasive Anemia Detection via Smartphone Imagery Using EfficientNet-B3 with TrivialAugmentWide, Mixup Augmentation, and Persistent Patient History Management

DGX agent

arXiv:2604.22964v1 Announce Type: new Abstract: Anemia affects over one billion people globally and remains severely under-diagnosed in low-resource regions where laboratory blood tests are inaccessib

safetyarxiv-cs-cv
28 Apr 2026
Safety

Animalbooth: multimodal feature enhancement for animal subject personalization

DGX agent

arXiv:2509.16702v2 Announce Type: replace Abstract: Personalized animal image generation is challenging due to rich appearance cues and large morphological variability. Existing approaches often exhib

safetyarxiv-cs-cv
28 Apr 2026
Safety

Bellman Residual Minimization for Control: Geometry, Stationarity, and Convergence

DGX agent

arXiv:2601.18840v3 Announce Type: replace Abstract: Markov decision problems are most commonly solved via dynamic programming. Another approach is Bellman residual minimization, which directly minimiz

safetyarxiv-cs-lg
28 Apr 2026
Safety

Beyond Match Maximization and Fairness: Retention-Optimized Two-Sided Matching

DGX agent

arXiv:2602.15752v2 Announce Type: replace Abstract: On two-sided matching platforms such as online dating and recruiting, recommendation algorithms often aim to maximize the total number of matches. H

safetyarxiv-cs-lg
28 Apr 2026
Safety

Breaking Lock-In: Preserving Steerability under Low-Data VLA Post-Training

DGX agent

arXiv:2604.23121v1 Announce Type: cross Abstract: Have you ever post-trained a generalist vision-language-action (VLA) policy on a small demonstration dataset, only to find that it stops responding to

safetyarxiv-cs-cv
28 Apr 2026
Safety

Bridging Reasoning and Action: Hybrid LLM-RL Framework for Efficient Cross-Domain Task-Oriented Dialogue

DGX agent

arXiv:2604.23345v1 Announce Type: new Abstract: Cross-domain task-oriented dialogue requires reasoning over implicit and explicit feasibility constraints while planning long-horizon, multi-turn action

safetyarxiv-cs-cl
28 Apr 2026
Safety

BVI-Mamba: Video Enhancement Using a Visual State-Space Model for Low-Light and Underwater Environments

DGX agent

arXiv:2604.23655v1 Announce Type: new Abstract: Videos captured in low-light and underwater conditions often suffer from distortions such as noise, low contrast, color imbalance, and blur. These issue

safetyarxiv-cs-cv
28 Apr 2026
Safety

CA-IDD: Cross-Attention Guided Identity-Conditional Diffusion for Identity-Consistent Face Swapping

DGX agent

arXiv:2604.24493v1 Announce Type: new Abstract: Face swapping aims to optimize realistic facial image generation by leveraging the identity of a source face onto a target face while preserving pose, e

safetyarxiv-cs-cv
28 Apr 2026
Safety

Can Compact Language Models Search Like Agents? Distillation-Guided Policy Optimization for Preserving Agentic RAG Capabilities

DGX agent

arXiv:2508.20324v4 Announce Type: replace Abstract: Reinforcement Learning has emerged as a dominant post-training approach to elicit agentic RAG behaviors such as search and planning from language mo

safetyarxiv-cs-cl
28 Apr 2026
Safety

CASP: Support-Aware Offline Policy Selection for Two-Stage Recommender Systems

DGX agent

arXiv:2604.23022v1 Announce Type: cross Abstract: Two-stage recommender systems first choose a candidate generator and then rank items within the generated set. Because the generator decides which ite

safetyarxiv-cs-lg
28 Apr 2026
Safety

CODA: Coordination via On-Policy Diffusion for Multi-Agent Offline Reinforcement Learning

DGX agent

arXiv:2604.23308v1 Announce Type: new Abstract: Offline multi-agent reinforcement learning (MARL) enables policy learning from fixed datasets, but is prone to coordination failure: agents trained on s

safetyarxiv-cs-lg
28 Apr 2026
Safety

CoFi-PGMA: Counterfactual Policy Gradients under Filtered Feedback for Multi-Agent LLMs

DGX agent

arXiv:2604.22785v1 Announce Type: new Abstract: Large language model (LLM) deployments increasingly rely on multi-agent architectures in which multiple models either compete through routing mechanisms

safetyarxiv-cs-lg
28 Apr 2026
Safety

COMO: Closed-Loop Optical Molecule Recognition with Minimum Risk Training

DGX agent

arXiv:2604.23546v1 Announce Type: cross Abstract: Optical chemical structure recognition (OCSR) translates molecular images into machine-readable representations like SMILES strings or molecular graph

safetyarxiv-cs-ai
28 Apr 2026
Safety

Complex SGD and Directional Bias in Reproducing Kernel Hilbert Spaces

DGX agent

arXiv:2604.23017v1 Announce Type: new Abstract: Stochastic Gradient Descent (SGD) is a known stochastic iterative method popular for large-scale convex optimization problems due to its simple implemen

safetyarxiv-cs-lg
28 Apr 2026
Safety

Conditional Imputation for Within-Modality Missingness in Multi-Modal Federated Learning

DGX agent

arXiv:2604.23112v1 Announce Type: new Abstract: Multimodal Federated Learning (MMFL) enables privacy-preserving collaborative training, but real-world clinical applications often suffer from within-mo

safetyarxiv-cs-lg
28 Apr 2026
Safety

Conflict-Aware Harmonized Rotational Gradient for Multiscale Kinetic Regimes

DGX agent

arXiv:2604.24745v1 Announce Type: new Abstract: In this paper, we propose a harmonized rotational gradient method, termed HRGrad, for simultaneously tackling multiscale time-dependent kinetic problems

safetyarxiv-cs-lg
28 Apr 2026
Safety

ConsDreamer: Advancing Multi-View Consistency for Zero-Shot Text-to-3D Generation

DGX agent

arXiv:2504.02316v4 Announce Type: replace-cross Abstract: Recent advances in zero-shot text-to-3D generation have revolutionized 3D content creation by enabling direct synthesis from textual descripti

safetyarxiv-cs-ai
28 Apr 2026
Safety

CT-Guided Spatially-varying Regularization for Voxel-Wise Deformable Whole-Body PET Registration

DGX agent

arXiv:2604.22905v1 Announce Type: cross Abstract: Whole-body Positron Emission Tomography (PET) registration is essential for multi-parametric tumor characterization and assessment of metastatic disea

safetyarxiv-cs-ai
28 Apr 2026
Safety

CURE-Med: Curriculum-Informed Reinforcement Learning for Multilingual Medical Reasoning

DGX agent

arXiv:2601.13262v2 Announce Type: replace Abstract: While large language models (LLMs) have shown to perform well on monolingual mathematical and commonsense reasoning, they remain unreliable for mult

safetyarxiv-cs-ai
28 Apr 2026
Safety

Data-efficient Targeted Token-level Preference Optimization for LLM-based Text-to-Speech

DGX agent

arXiv:2510.05799v2 Announce Type: replace-cross Abstract: Aligning text-to-speech (TTS) system outputs with human feedback through preference optimization has been shown to effectively improve the rob

safetyarxiv-cs-ai
28 Apr 2026
Safety

DextER: Language-driven Dexterous Grasp Generation with Embodied Reasoning

DGX agent

arXiv:2601.16046v2 Announce Type: replace-cross Abstract: Language-driven dexterous grasp generation requires the models to understand task semantics, 3D geometry, and complex hand-object interactions

safetyarxiv-cs-cv
28 Apr 2026
Safety

DLM: Unified Decision Language Models for Offline Multi-Agent Sequential Decision Making

DGX agent

arXiv:2604.23557v1 Announce Type: cross Abstract: Building scalable and reusable multi-agent decision policies from offline datasets remains a challenge in offline multi-agent reinforcement learning (

safetyarxiv-cs-ai
28 Apr 2026
Safety

Do Synthetic Trajectories Reflect Real Reward Hacking? A Systematic Study on Monitoring In-the-Wild Hacking in Code Generation

DGX agent

arXiv:2604.23488v1 Announce Type: new Abstract: Reward hacking in code generation, where models exploit evaluation loopholes to obtain full reward without correctly solving the tasks, poses a critical

safetyarxiv-cs-lg
28 Apr 2026
Safety

Do Transaction-Level and Actor-Level AML Queues Agree? An Empirical Evaluation of Granularity Effects on the Elliptic++ Graph

DGX agent

arXiv:2604.23494v1 Announce Type: new Abstract: Graph-based anti-money laundering (AML) systems on blockchain networks can score suspicious activity at two granularity levels -- transactions or actor

safetyarxiv-cs-ai
28 Apr 2026
Safety

DPEPO: Diverse Parallel Exploration Policy Optimization for LLM-based Agents

DGX agent

arXiv:2604.24320v1 Announce Type: new Abstract: Large language model (LLM) agents that follow the sequential 'reason-then-act' paradigm have achieved superior performance in many complex tasks.However

safetyarxiv-cs-cl
28 Apr 2026
Safety

DPRM: A Plug-in Doob h transform-induced Token-Ordering Module for Diffusion Language Models

DGX agent

arXiv:2604.24357v1 Announce Type: cross Abstract: Diffusion language models generate without a fixed left-to-right order, making token ordering a central algorithmic choice: which tokens should be rev

safetyarxiv-cs-ai
28 Apr 2026
Safety

DVPO: Distributional Value Modeling-based Policy Optimization for LLM Post-Training

DGX agent

arXiv:2512.03847v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has shown strong performance in LLM post-training, but real-world deployment often involves noisy or incomplete su

safetyarxiv-cs-ai
28 Apr 2026
Safety

EAD-Net: Emotion-Aware Talking Head Generation with Spatial Refinement and Temporal Coherence

DGX agent

arXiv:2604.23325v1 Announce Type: cross Abstract: Emotionally talking head video generation aims to generate expressive portrait videos with accurate lip synchronization and emotional facial expressio

safetyarxiv-cs-ai
28 Apr 2026
Safety

Efficient Rationale-based Retrieval: On-policy Distillation from Generative Rerankers based on JEPA

DGX agent

arXiv:2604.23336v1 Announce Type: cross Abstract: Unlike traditional fact-based retrieval, rationale-based retrieval typically necessitates cross-encoding of query-document pairs using large language

safetyarxiv-cs-cl
28 Apr 2026
Safety

EL3DD: Extended Latent 3D Diffusion for Language Conditioned Multitask Manipulation

DGX agent

arXiv:2511.13312v2 Announce Type: replace-cross Abstract: Acting in human environments is a crucial capability for general-purpose robots, necessitating a robust understanding of natural language and

safetyarxiv-cs-ai
28 Apr 2026
Safety

Evaluating Language Models' Evaluations of Games

DGX agent

arXiv:2510.10930v2 Announce Type: replace-cross Abstract: Reasoning is not just about solving problems -- it is also about evaluating which problems are worth solving at all. Evaluations of artificial

safetyarxiv-cs-ai
28 Apr 2026
Safety

Explanation Quality Assessment as Ranking with Listwise Rewards

DGX agent

arXiv:2604.24176v1 Announce Type: new Abstract: We reformulate explanation quality assessment as a ranking problem rather than a generation problem. Instead of optimizing models to produce a single 'b

safetyarxiv-cs-ai
28 Apr 2026
Safety

Extreme bandits

DGX agent

arXiv:2604.24545v1 Announce Type: cross Abstract: In many areas of medicine, security, and life sciences, we want to allocate limited resources to different sources in order to detect extreme values.

safetyarxiv-cs-lg
28 Apr 2026
Safety

Failure-Centered Runtime Evaluation for Deployed Trilingual Public-Space Agents

DGX agent

arXiv:2604.23990v1 Announce Type: new Abstract: This paper presents PSA-Eval, a failure-centered runtime evaluation framework for deployed trilingual public-space agents. The central claim is that, wh

safetyarxiv-cs-ai
28 Apr 2026
Safety

Federated Cross-Modal Retrieval with Missing Modalities via Semantic Routing and Adapter Personalization

DGX agent

arXiv:2604.22885v1 Announce Type: cross Abstract: Federated cross-modal retrieval faces severe challenges from heterogeneous client data, particularly non-IID semantic distributions and missing modali

safetyarxiv-cs-ai
28 Apr 2026
Safety

Fine-R1: Make Multi-modal LLMs Excel in Fine-Grained Visual Recognition by Chain-of-Thought Reasoning

DGX agent

arXiv:2602.07605v3 Announce Type: replace-cross Abstract: Any entity in the visual world can be hierarchically grouped based on shared characteristics and mapped to fine-grained sub-categories. While

safetyarxiv-cs-ai
28 Apr 2026
Safety

FinGround: Detecting and Grounding Financial Hallucinations via Atomic Claim Verification

DGX agent

arXiv:2604.23588v1 Announce Type: new Abstract: Financial AI systems must produce answers grounded in specific regulatory filings, yet current LLMs fabricate metrics, invent citations, and miscalculat

safetyarxiv-cs-ai
28 Apr 2026
Safety

GAMED.AI: A Hierarchical Multi-Agent Framework for Automated Educational Game Generation

DGX agent

arXiv:2604.23947v1 Announce Type: new Abstract: We introduce GameDAI, a hierarchical multi-agent framework that transforms instructor-provided questions into fully playable, pedagogically grounded edu

safetyarxiv-cs-ai
28 Apr 2026
Safety

Hierarchical Prototype-based Domain Priors for Multiple Instance Learning in Multimodal Histopathology Analysis

DGX agent

arXiv:2604.23982v1 Announce Type: new Abstract: Digital pathology has fundamentally altered diagnostic workflows by enabling the computational analysis of gigapixel Whole Slide Images (WSIs), yet effe

safetyarxiv-cs-cv
28 Apr 2026
Safety

Hindsight Preference Optimization for Financial Time Series Advisory

DGX agent

arXiv:2604.23988v1 Announce Type: cross Abstract: Time series models predict numbers; decision-makers need advisory -- directional signals with reasoning, actionable suggestions, and risk management.

safetyarxiv-cs-ai
28 Apr 2026
Safety

Humanoid Whole-Body Badminton via Multi-Stage Reinforcement Learning

DGX agent

arXiv:2511.11218v3 Announce Type: replace Abstract: Humanoid robots have demonstrated strong capabilities for interacting with static scenes across locomotion and manipulation, yet dynamic real-world

safetyarxiv-cs-ro
28 Apr 2026
Safety

In-Sync: Adaptation of Speech Aware Large Language Models for ASR with Word Level Timestamp Predictions

DGX agent

arXiv:2604.22817v1 Announce Type: cross Abstract: Recent advances in speech-aware language models have coupled strong acoustic encoders with large language models, enabling systems that move beyond tr

safetyarxiv-cs-cl
28 Apr 2026
Safety

InCoM: Intent-Driven Perception and Structured Coordination for Mobile Manipulation

DGX agent

arXiv:2602.23024v2 Announce Type: replace Abstract: Mobile manipulation is a fundamental capability for general-purpose robotic agents, requiring both coordinated control of the mobile base and manipu

safetyarxiv-cs-ro
28 Apr 2026
← Previous
1…208209210211212…257
Next →