AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Motion-Guided Semantic Alignment with Negative Prompts for Zero-Shot Video Action Recognition

DGX agent

arXiv:2604.17062v1 Announce Type: new Abstract: Zero-shot action recognition is challenging due to the semantic gap between seen and unseen classes. We present a novel framework that enhances CLIP wit

safetyarxiv-cs-cv
21 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Navigating Distribution Shifts in Medical Image Analysis: A Survey

DGX agent

arXiv:2411.05824v3 Announce Type: replace-cross Abstract: Medical Image Analysis (MedIA) has become indispensable in modern healthcare, enhancing clinical diagnostics and personalized treatment. Despi

safetyarxiv-cs-cv
21 Apr 2026
Safety

Negative Advantage Is a Double-Edged Sword: Calibrating Advantage in GRPO for Deep Search

DGX agent

arXiv:2604.18235v1 Announce Type: new Abstract: Deep search agents can autonomously initiate multi-turn interactions with search engines, thereby exhibiting strong question-answering capabilities. Suc

safetyarxiv-cs-cl
21 Apr 2026
Safety

OmniVLA-RL: A Vision-Language-Action Model with Spatial Understanding and Online RL

DGX agent

arXiv:2604.17706v1 Announce Type: new Abstract: Visual-Language-Action (VLA) models represent a paradigm shift in embodied AI, yet existing frameworks often struggle with imprecise spatial perception,

safetyarxiv-cs-ro
21 Apr 2026
Safety

On the Convergence and Size Transferability of Continuous-depth Graph Neural Networks

DGX agent

arXiv:2510.03923v2 Announce Type: replace Abstract: Continuous-depth graph neural networks, also known as Graph Neural Differential Equations (GNDEs), combine the structural inductive bias of Graph Ne

safetyarxiv-cs-lg
21 Apr 2026
Safety

On the Importance of Tactile Sensing for Imitation Learning: A Case Study on Robotic Match Lighting

DGX agent

arXiv:2504.13618v4 Announce Type: replace Abstract: The field of robotic manipulation has advanced significantly in recent years. At the sensing level, several novel tactile sensors have been develope

safetyarxiv-cs-ro
21 Apr 2026
Safety

On the Shelf Life of Fine-Tuned LLM-Judges: Future-Proofing, Backward-Compatibility, and Question Generalization

DGX agent

arXiv:2509.23542v2 Announce Type: replace Abstract: The LLM-as-a-judge paradigm is widely used in both evaluating free-text model responses and reward modeling for model alignment and fine-tuning. Rec

safetyarxiv-cs-cl
21 Apr 2026
Safety

Operationalizing Fairness in Text-to-Image Models: A Survey of Bias, Fairness Audits and Mitigation Strategies

DGX agent

arXiv:2604.16516v1 Announce Type: new Abstract: Text-to-Image (T2I) generation models have been widely adopted across various industries, yet are criticized for frequently exhibiting societal stereoty

safetyarxiv-cs-cv
21 Apr 2026
Safety

OPSDL: On-Policy Self-Distillation for Long-Context Language Models

DGX agent

arXiv:2604.17535v1 Announce Type: new Abstract: Extending the effective context length of large language models (LLMs) remains a central challenge for real-world applications. While recent post-traini

safetyarxiv-cs-cl
21 Apr 2026
Safety

OVOD-Agent: A Markov-Bandit Framework for Proactive Visual Reasoning and Self-Evolving Detection

DGX agent

arXiv:2511.21064v2 Announce Type: replace-cross Abstract: Open-Vocabulary Object Detection (OVOD) aims to enable detectors to generalize across categories by leveraging semantic information. Although

safetyarxiv-cs-cv
21 Apr 2026
Safety

PaTaRM: Bridging Pairwise and Pointwise Signals via Preference-Aware Task-Adaptive Reward Modeling

DGX agent

arXiv:2510.24235v3 Announce Type: replace Abstract: Reward models (RMs) are central to reinforcement learning from human feedback (RLHF), providing the critical supervision signals that align large la

safetyarxiv-cs-lg
21 Apr 2026
Safety

Peerispect: Claim Verification in Scientific Peer Reviews

DGX agent

arXiv:2604.17667v1 Announce Type: new Abstract: Peer review is central to scientific publishing, yet reviewers frequently include claims that are subjective, rhetorical, or misaligned with the submitt

safetyarxiv-cs-cl
21 Apr 2026
Safety

PEPR: Privileged Event-based Predictive Regularization for Domain Generalization

DGX agent

arXiv:2602.04583v2 Announce Type: replace Abstract: Deep neural networks for visual perception are highly susceptible to domain shift, which poses a critical challenge for real-world deployment under

safetyarxiv-cs-cv
21 Apr 2026
Safety

Plasticity Loss in Deep Reinforcement Learning: A Survey

DGX agent

arXiv:2411.04832v3 Announce Type: replace-cross Abstract: Plasticity refers to a network's ability to adapt to changing data distributions, which is crucial for the successful training of deep reinfor

safetyarxiv-cs-lg
21 Apr 2026
Safety

Policy Testing in Markov Decision Processes

DGX agent

arXiv:2505.15342v2 Announce Type: replace-cross Abstract: We study the policy testing problem in discounted Markov decision processes (MDPs) in the fixed-confidence setting under a generative model wi

safetyarxiv-cs-lg
21 Apr 2026
Safety

PrinciplismQA: A Philosophy-Grounded Approach to Assessing LLM-Human Clinical Medical Ethics Alignment

DGX agent

arXiv:2508.05132v2 Announce Type: replace Abstract: As medical LLMs transition to clinical deployment, assessing their ethical reasoning capability becomes critical. While achieving high accuracy on k

safetyarxiv-cs-cl
21 Apr 2026
Safety

ProtoCLIP: Prototype-Aligned Latent Refinement for Robust Zero-Shot Chest X-Ray Classification

DGX agent

arXiv:2604.18444v1 Announce Type: cross Abstract: Zero-shot vision-language models (VLMs) have shown promise for chest radiograph classification, but their performance is often limited by confounding

safetyarxiv-cs-cv
21 Apr 2026
Safety

Q-SINDy: Quantum-Kernel Sparse Identification of Nonlinear Dynamics with Provable Coefficient Debiasing

DGX agent

arXiv:2604.16779v1 Announce Type: cross Abstract: Quantum feature maps offer expressive embeddings for classical learning tasks, and augmenting sparse identification of nonlinear dynamics (SINDy) with

safetyarxiv-cs-lg
21 Apr 2026
Safety

RAYEN: Imposition of Hard Convex Constraints on Neural Networks

DGX agent

arXiv:2307.08336v2 Announce Type: replace Abstract: Despite the numerous applications of convex constraints in Robotics, enforcing them within learning-based frameworks remains an open challenge. Exis

safetyarxiv-cs-lg
21 Apr 2026
Safety

Reasoning on the Manifold: Bidirectional Consistency for Self-Verification in Diffusion Language Models

DGX agent

arXiv:2604.16565v1 Announce Type: new Abstract: While Diffusion Large Language Models (dLLMs) offer structural advantages for global planning, efficiently verifying that they arrive at correct answers

safetyarxiv-cs-lg
21 Apr 2026
Safety

RemoteShield: Enable Robust Multimodal Large Language Models for Earth Observation

DGX agent

arXiv:2604.17243v1 Announce Type: new Abstract: A robust Multimodal Large Language Model (MLLM) for Earth Observation should maintain consistent interpretation and reasoning under realistic input vari

safetyarxiv-cs-cv
21 Apr 2026
Safety

Render-of-Thought: Rendering Textual Chain-of-Thought as Images for Visual Latent Reasoning

DGX agent

arXiv:2601.14750v3 Announce Type: replace Abstract: Chain-of-Thought (CoT) prompting has achieved remarkable success in unlocking the reasoning capabilities of Large Language Models (LLMs). Although C

safetyarxiv-cs-cl
21 Apr 2026
Safety

Rethinking the Comparison Unit in Sequence-Level Reinforcement Learning: An Equal-Length Paired Training Framework from Loss Correction to Sample Construction

DGX agent

arXiv:2604.17328v1 Announce Type: new Abstract: This paper investigates the length problem in sequence-level relative reinforcement learning. We observe that, although existing methods partially allev

safetyarxiv-cs-lg
21 Apr 2026
Safety

Retrieval-Augmented Multimodal Model for Fake News Detection

DGX agent

arXiv:2604.18112v1 Announce Type: new Abstract: In recent years, multimodal multidomain fake news detection has garnered increasing attention. Nevertheless, this direction presents two significant cha

safetyarxiv-cs-cl
21 Apr 2026
Safety

Reward Score Matching: Unifying Reward-based Fine-tuning for Flow and Diffusion Models

DGX agent

arXiv:2604.17415v1 Announce Type: cross Abstract: Reward-based fine-tuning aims to steer a pretrained diffusion or flow-based generative model toward higher-reward samples while remaining close to the

safetyarxiv-cs-cv
21 Apr 2026
Safety

REZE: Representation Regularization for Domain-adaptive Text Embedding Pre-finetuning

DGX agent

arXiv:2604.17257v1 Announce Type: new Abstract: Recent text embedding models are often adapted to specialized domains via contrastive pre-finetuning (PFT) on a naive collection of scattered, heterogen

safetyarxiv-cs-cl
21 Apr 2026
Safety

Robust Tool Use via Fission-GRPO: Learning to Recover from Execution Errors

DGX agent

arXiv:2601.15625v2 Announce Type: replace Abstract: Large language models (LLMs) can call tools effectively, yet they remain brittle in multi-turn execution: after a tool-call error, smaller models of

safetyarxiv-cs-lg
21 Apr 2026
Safety

S-GRPO: Unified Post-Training for Large Vision-Language Models

DGX agent

arXiv:2604.16557v1 Announce Type: cross Abstract: Current post-training methodologies for adapting Large Vision-Language Models (LVLMs) generally fall into two paradigms: Supervised Fine-Tuning (SFT)

safetyarxiv-cs-cl
21 Apr 2026
Safety

Scalable Neighborhood-Based Multi-Agent Actor-Critic

DGX agent

arXiv:2604.18190v1 Announce Type: new Abstract: We propose MADDPG-K, a scalable extension to Multi-Agent Deep Deterministic Policy Gradient (MADDPG) that addresses the computational limitations of cen

safetyarxiv-cs-lg
21 Apr 2026
Safety

Scalable Physics-Informed Neural Differential Equations and Data-Driven Algorithms for HVAC Systems

DGX agent

arXiv:2604.18438v1 Announce Type: new Abstract: We present a scalable, data-driven simulation framework for large-scale heating, ventilation, and air conditioning (HVAC) systems that couples physics-i

safetyarxiv-cs-lg
21 Apr 2026
Safety

See Through the Noise: Improving Domain Generalization in Gaze Estimation

DGX agent

arXiv:2604.16562v1 Announce Type: new Abstract: Generalizable gaze estimation methods have garnered increasing attention due to their critical importance in real-world applications and have achieved s

safetyarxiv-cs-cv
21 Apr 2026
Safety

Sharpening Lightweight Models for Generalized Polyp Segmentation: A Boundary Guided Distillation from Foundation Models

DGX agent

arXiv:2604.17865v1 Announce Type: new Abstract: Automated polyp segmentation is critical for early colorectal cancer detection and its prevention, yet remains challenging due to weak boundaries, large

safetyarxiv-cs-cv
21 Apr 2026
Safety

Soft Label Pruning and Quantization for Large-Scale Dataset Distillation

DGX agent

arXiv:2604.18135v1 Announce Type: new Abstract: Large-scale dataset distillation requires storing auxiliary soft labels that can be 30-40x larger on ImageNet-1K and 200x larger on ImageNet-21K than th

safetyarxiv-cs-cv
21 Apr 2026
Safety

Source-Free Domain Adaptation with Vision-Language Prior

DGX agent

arXiv:2604.17748v1 Announce Type: new Abstract: Source-Free Domain Adaptation (SFDA) seeks to adapt a source model, which is pre-trained on a supervised source domain, for a target domain, with only a

safetyarxiv-cs-cv
21 Apr 2026
Safety

(Sparse) Attention to the Details: Preserving Spectral Fidelity in ML-based Weather Forecasting Models

DGX agent

arXiv:2604.16429v1 Announce Type: cross Abstract: We introduce Mosaic, a probabilistic weather forecasting model that addresses two principal sources of spectral degradation in ML-based weather predic

safetyarxiv-cs-cv
21 Apr 2026
Safety

Spectral bandits for smooth graph functions

DGX agent

arXiv:2604.18420v1 Announce Type: cross Abstract: Smooth functions on graphs have wide applications in manifold and semi-supervised learning. In this paper, we study a bandit problem where the payoffs

safetyarxiv-cs-lg
21 Apr 2026
Safety

Speculative Verification: Exploiting Information Gain to Refine Speculative Decoding

DGX agent

arXiv:2509.24328v2 Announce Type: replace Abstract: LLMs have low GPU efficiency and high latency due to autoregressive decoding. Speculative decoding (SD) mitigates this using a small draft model to

safetyarxiv-cs-cl
21 Apr 2026
Safety

SpiralThinker: Latent Reasoning through an Iterative Process with Text-Latent Interleaving

DGX agent

arXiv:2511.08983v2 Announce Type: replace Abstract: Recent advances in large reasoning models have been driven by reinforcement learning and test-time scaling, accompanied by growing interest in laten

safetyarxiv-cs-cl
21 Apr 2026
Safety

SPS: Steering Probability Squeezing for Better Exploration in Reinforcement Learning for Large Language Models

DGX agent

arXiv:2604.16995v1 Announce Type: new Abstract: Reinforcement learning (RL) has emerged as a promising paradigm for training reasoning-oriented models by leveraging rule-based reward signals. However,

safetyarxiv-cs-cl
21 Apr 2026
Safety

Sub-metre Lunar DEM Generation and Validation from Chandrayaan-2 OHRC Multi-View Imagery Using an Open-Source Pipeline

DGX agent

arXiv:2604.01032v3 Announce Type: replace Abstract: High-resolution digital elevation models (DEMs) of the lunar surface are essential for surface mobility planning, landing site characterization, and

safetyarxiv-cs-cv
21 Apr 2026
Safety

Support Sufficiency as Consequence-Sensitive Compression in Belief Arbitration

DGX agent

arXiv:2604.16434v1 Announce Type: cross Abstract: When a system commits to a hypothesis, much of the evidential structure behind that commitment is lost to compression. Standard accounts assume that s

safetyarxiv-cs-lg
21 Apr 2026
Safety

SynAgent: Generalizable Cooperative Humanoid Manipulation via Solo-to-Cooperative Agent Synergy

DGX agent

arXiv:2604.18557v1 Announce Type: new Abstract: Controllable cooperative humanoid manipulation is a fundamental yet challenging problem for embodied intelligence, due to severe data scarcity, complexi

safetyarxiv-cs-cv
21 Apr 2026
Safety

SynopticBench: Evaluating Vision-Language Models on Generating Weather Forecast Discussions of the Future

DGX agent

arXiv:2604.16451v1 Announce Type: new Abstract: Recent advances in visual-language models (VLMs) have led to significant improvements in a plethora of complex multimodal tasks like image captioning, r

safetyarxiv-cs-cl
21 Apr 2026
Safety

Synthia: Scalable Grounded Persona Generation from Social Media Data

DGX agent

arXiv:2507.14922v2 Announce Type: replace Abstract: Persona-driven simulations are increasingly used in computational social science, yet their validity critically depends on the fidelity of the under

safetyarxiv-cs-cl
21 Apr 2026
Safety

Task Matters: Knowledge Requirements Shape LLM Responses to Context-Memory Conflict

DGX agent

arXiv:2506.06485v4 Announce Type: replace Abstract: Large language models (LLMs) draw on both contextual information and parametric memory, yet these sources can conflict. Prior studies have largely e

safetyarxiv-cs-cl
21 Apr 2026
Safety

TeMuDance: Contrastive Alignment-Based Textual Control for Music-Driven Dance Generation

DGX agent

arXiv:2604.17005v1 Announce Type: new Abstract: Existing music-driven dance generation approaches have achieved strong realism and effective audio-motion alignment. However, they generally lack semant

safetyarxiv-cs-cv
21 Apr 2026
Safety

The Illusion of Certainty: Decoupling Capability and Calibration in On-Policy Distillation

DGX agent

arXiv:2604.16830v1 Announce Type: new Abstract: On-policy distillation (OPD) is an increasingly important paradigm for post-training language models. However, we identify a pervasive Scaling Law of Mi

safetyarxiv-cs-lg
21 Apr 2026
Safety

The Less You Depend, The More You Learn: Synthesizing Novel Views from Sparse, Unposed Images with Minimal 3D Knowledge

DGX agent

arXiv:2506.09885v2 Announce Type: replace Abstract: Recent advances in feed-forward Novel View Synthesis (NVS) have led to a divergence between two design philosophies: bias-driven methods, which rely

safetyarxiv-cs-cv
21 Apr 2026
← Previous
1…219220221222223…257
Next →