AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,745
  • Agents7,195
  • Applications5,151
  • Concepts5
  • Hardware1,740
  • Industry6,080
  • Local Ai4,671
  • Model Releases22,272
  • Research19,012
  • Safety12,702
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,745Total entries
1Added by human
83,744Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Drift-Based Policy Optimization: Native One-Step Policy Learning for Online Robot Control

DGX agent

arXiv:2604.03540v2 Announce Type: replace Abstract: Although multi-step generative policies achieve strong performance in robotic manipulation by modeling multimodal action distributions, they require

safetyarxiv-cs-ro
10 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

EchoTorrent: Towards Swift, Sustained, and Streaming Multi-Modal Video Generation

DGX agent

arXiv:2602.13669v4 Announce Type: replace Abstract: Recent multi-modal video generation models have achieved high visual quality, but their prohibitive latency and limited temporal stability hinder re

safetyarxiv-cs-cv
10 Apr 2026
Safety

EgoVerse: An Egocentric Human Dataset for Robot Learning from Around the World

DGX agent

arXiv:2604.07607v1 Announce Type: cross Abstract: Robot learning increasingly depends on large and diverse data, yet robot data collection remains expensive and difficult to scale. Egocentric human da

safetyarxiv-cs-cv
10 Apr 2026
Safety

Equivariant Multi-agent Reinforcement Learning for Multimodal Vehicle-to-Infrastructure Systems

DGX agent

arXiv:2604.06914v1 Announce Type: new Abstract: In this paper, we study a vehicle-to-infrastructure (V2I) system where distributed base stations (BSs) acting as road-side units (RSUs) collect multimod

safetyarxiv-cs-lg
10 Apr 2026
Safety

Faithful GRPO: Improving Visual Spatial Reasoning in Multimodal Language Models via Constrained Policy Optimization

DGX agent

arXiv:2604.08476v1 Announce Type: new Abstract: Multimodal reasoning models (MRMs) trained with reinforcement learning with verifiable rewards (RLVR) show improved accuracy on visual reasoning benchma

safetyarxiv-cs-cv
10 Apr 2026
Safety

Force-Aware Residual DAgger via Trajectory Editing for Precision Insertion with Impedance Control

DGX agent

arXiv:2603.04038v2 Announce Type: replace Abstract: Imitation learning (IL) has shown strong potential for contact-rich precision insertion tasks. However, its practical deployment is often hindered b

safetyarxiv-cs-ro
10 Apr 2026
Safety

FP4 Explore, BF16 Train: Diffusion Reinforcement Learning via Efficient Rollout Scaling

DGX agent

arXiv:2604.06916v1 Announce Type: cross Abstract: Reinforcement-Learning-based post-training has recently emerged as a promising paradigm for aligning text-to-image diffusion models with human prefere

safetyarxiv-cs-ai
10 Apr 2026
Safety

From Ground Truth to Measurement: A Statistical Framework for Human Labeling

DGX agent

arXiv:2604.07591v1 Announce Type: cross Abstract: Supervised machine learning assumes that labeled data provide accurate measurements of the concepts models are meant to learn. Yet in practice, human

safetyarxiv-cs-cl
10 Apr 2026
Safety

FVD: Inference-Time Alignment of Diffusion Models via Fleming-Viot Resampling

DGX agent

arXiv:2604.06779v1 Announce Type: new Abstract: We introduce Fleming-Viot Diffusion (FVD), an inference-time alignment method that resolves the diversity collapse commonly observed in Sequential Monte

safetyarxiv-cs-ai
10 Apr 2026
Safety

Geometric Properties of the Voronoi Tessellation in Latent Semantic Manifolds of Large Language Models

DGX agent

arXiv:2604.06767v1 Announce Type: new Abstract: Language models operate on discrete tokens but compute in continuous vector spaces, inducing a Voronoi tessellation over the representation manifold. We

safetyarxiv-cs-lg
10 Apr 2026
Safety

GIFT: Group-Relative Implicit Fine-Tuning Integrates GRPO with DPO and UNA

DGX agent

arXiv:2510.23868v4 Announce Type: replace Abstract: This paper proposes extit{Group-relative Implicit Fine-Tuning (GIFT)}, a reinforcement learning framework for aligning large language models (LLMs

safetyarxiv-cs-lg
10 Apr 2026
Safety

Governed Capability Evolution for Embodied Agents: Safe Upgrade, Compatibility Checking, and Runtime Rollback for Embodied Capability Modules

DGX agent

arXiv:2604.08059v1 Announce Type: new Abstract: Embodied agents are increasingly expected to improve over time by updating their executable capabilities rather than rewriting the agent itself. Prior w

safetyarxiv-cs-ro
10 Apr 2026
Safety

Grasp as You Dream: Imitating Functional Grasping from Generated Human Demonstrations

DGX agent

arXiv:2604.07517v1 Announce Type: new Abstract: Building generalist robots capable of performing functional grasping in everyday, open-world environments remains a significant challenge due to the vas

safetyarxiv-cs-ro
10 Apr 2026
Safety

Guiding a Diffusion Model by Swapping Its Tokens

DGX agent

arXiv:2604.08048v1 Announce Type: new Abstract: Classifier-Free Guidance (CFG) is a widely used inference-time technique to boost the image quality of diffusion models. Yet, its reliance on text condi

safetyarxiv-cs-cv
10 Apr 2026
Safety

How Independent are Large Language Models? A Statistical Framework for Auditing Behavioral Entanglement and Reweighting Verifier Ensembles

DGX agent

arXiv:2604.07650v1 Announce Type: cross Abstract: The rapid growth of the large language model (LLM) ecosystem raises a critical question: are seemingly diverse models truly independent? Shared pretra

safetyarxiv-cs-cl
10 Apr 2026
Safety

How Psychological Learning Paradigms Shaped and Constrained Artificial Intelligence

DGX agent

arXiv:2603.18203v3 Announce Type: replace Abstract: Current artificial intelligence systems struggle with systematic compositional reasoning: the capacity to recombine known components in novel config

safetyarxiv-cs-cl
10 Apr 2026
Safety

How to Evaluate Speech Translation with Source-Aware Neural MT Metrics

DGX agent

arXiv:2511.03295v3 Announce Type: replace-cross Abstract: Automatic evaluation of ST systems is typically performed by comparing translation hypotheses with one or more reference translations. While e

safetyarxiv-cs-ai
10 Apr 2026
Safety

HumanLLM: Benchmarking and Improving LLM Anthropomorphism via Human Cognitive Patterns

DGX agent

arXiv:2601.10198v3 Announce Type: replace Abstract: Large Language Models (LLMs) have demonstrated remarkable capabilities in reasoning and generation, serving as the foundation for advanced persona s

safetyarxiv-cs-cl
10 Apr 2026
Safety

Improving Semantic Uncertainty Quantification in Language Model Question-Answering via Token-Level Temperature Scaling

DGX agent

arXiv:2604.07172v1 Announce Type: new Abstract: Calibration is central to reliable semantic uncertainty quantification, yet prior work has largely focused on discrimination, neglecting calibration. As

safetyarxiv-cs-lg
10 Apr 2026
Safety

Incremental Residual Reinforcement Learning Toward Real-World Learning for Social Navigation

DGX agent

arXiv:2604.07945v1 Announce Type: new Abstract: As the demand for mobile robots continues to increase, social navigation has emerged as a critical task, driving active research into deep reinforcement

safetyarxiv-cs-ro
10 Apr 2026
Safety

Inside-Out: Measuring Generalization in Vision Transformers Through Inner Workings

DGX agent

arXiv:2604.08192v1 Announce Type: cross Abstract: Reliable generalization metrics are fundamental to the evaluation of machine learning models. Especially in high-stakes applications where labeled tar

safetyarxiv-cs-cv
10 Apr 2026
Safety

Karma Mechanisms for Decentralised, Cooperative Multi Agent Path Finding

DGX agent

arXiv:2604.07970v1 Announce Type: cross Abstract: Multi-Agent Path Finding (MAPF) is a fundamental coordination problem in large-scale robotic and cyber-physical systems, where multiple agents must co

safetyarxiv-cs-ro
10 Apr 2026
Safety

KD-MARL: Resource-Aware Knowledge Distillation in Multi-Agent Reinforcement Learning

DGX agent

arXiv:2604.06691v1 Announce Type: new Abstract: Real world deployment of multi agent reinforcement learning MARL systems is fundamentally constrained by limited compute memory and inference time. Whil

safetyarxiv-cs-ai
10 Apr 2026
Safety

LangDriveCTRL: Natural Language Controllable Driving Scene Editing with Multi-modal Agents

DGX agent

arXiv:2512.17445v2 Announce Type: replace Abstract: LangDriveCTRL is a natural-language-controllable framework for editing real-world driving videos to synthesize diverse traffic scenarios. It represe

safetyarxiv-cs-cv
10 Apr 2026
Safety

Large Language Model Post-Training: A Unified View of Off-Policy and On-Policy Learning

DGX agent

arXiv:2604.07941v1 Announce Type: new Abstract: Post-training has become central to turning pretrained large language models (LLMs) into aligned and deployable systems. Recent progress spans supervise

safetyarxiv-cs-cl
10 Apr 2026
Safety

Learning to Negotiate: Multi-Agent Deliberation for Collective Value Alignment in LLMs

DGX agent

arXiv:2603.10476v2 Announce Type: replace Abstract: LLM alignment has progressed in single-agent settings through paradigms such as RL with human feedback (RLHF), while recent work explores scalable a

safetyarxiv-cs-cl
10 Apr 2026
Safety

Limits of Difficulty Scaling: Hard Samples Yield Diminishing Returns in GRPO-Tuned SLMs

DGX agent

arXiv:2604.06298v1 Announce Type: new Abstract: Recent alignment work on Large Language Models (LLMs) suggests preference optimization can improve reasoning by shifting probability mass toward better

safetyarxiv-cs-lg
10 Apr 2026
Safety

MCLR: Improving Conditional Modeling via Inter-Class Likelihood-Ratio Maximization and Unifying Classifier-Free Guidance with Alignment Objectives

DGX agent

arXiv:2603.22364v2 Announce Type: replace-cross Abstract: Diffusion models have achieved state-of-the-art performance in generative modeling, but their success often relies heavily on classifier-free

safetyarxiv-cs-cv
10 Apr 2026
Safety

MDP modeling for multi-stage stochastic programs

DGX agent

arXiv:2509.22981v2 Announce Type: replace Abstract: We study a class of multi-stage stochastic programs, which incorporate modeling features from Markov decision processes (MDPs). This class includes

safetyarxiv-cs-lg
10 Apr 2026
Safety

MemReader: From Passive to Active Extraction for Long-Term Agent Memory

DGX agent

arXiv:2604.07877v1 Announce Type: new Abstract: Long-term memory is fundamental for personalized and autonomous agents, yet populating it remains a bottleneck. Existing systems treat memory extraction

safetyarxiv-cs-cl
10 Apr 2026
Safety

Mixture Proportion Estimation and Weakly-supervised Kernel Test for Conditional Independence

DGX agent

arXiv:2604.07191v1 Announce Type: cross Abstract: Mixture proportion estimation (MPE) aims to estimate class priors from unlabeled data. This task is a critical component in weakly supervised learning

safetyarxiv-cs-ai
10 Apr 2026
Safety

MO-RiskVAE: A Multi-Omics Variational Autoencoder for Survival Risk Modeling in Multiple MyelomaMO-RiskVAE

DGX agent

arXiv:2604.06267v1 Announce Type: cross Abstract: Multimodal variational autoencoders (VAEs) have emerged as a powerful framework for survival risk modeling in multiple myeloma by integrating heteroge

safetyarxiv-cs-ai
10 Apr 2026
Safety

MonoUNet: A Robust Tiny Neural Network for Automated Knee Cartilage Segmentation on Point-of-Care Ultrasound Devices

DGX agent

arXiv:2604.07780v1 Announce Type: cross Abstract: Objective: To develop a robust and compact deep learning model for automated knee cartilage segmentation on point-of-care ultrasound (POCUS) devices.

safetyarxiv-cs-cv
10 Apr 2026
Safety

MorphDistill: Distilling Unified Morphological Knowledge from Pathology Foundation Models for Colorectal Cancer Survival Prediction

DGX agent

arXiv:2604.06390v1 Announce Type: cross Abstract: Background: Colorectal cancer (CRC) remains a leading cause of cancer-related mortality worldwide. Accurate survival prediction is essential for treat

safetyarxiv-cs-ai
10 Apr 2026
Safety

MotionScape: A Large-Scale Real-World Highly Dynamic UAV Video Dataset for World Models

DGX agent

arXiv:2604.07991v1 Announce Type: new Abstract: Recent advances in world models have demonstrated strong capabilities in simulating physical reality, making them an increasingly important foundation f

safetyarxiv-cs-cv
10 Apr 2026
Safety

MSCT: Differential Cross-Modal Attention for Deepfake Detection

DGX agent

arXiv:2604.07741v1 Announce Type: new Abstract: Audio-visual deepfake detection typically employs a complementary multi-modal model to check the forgery traces in the video. These methods primarily ex

safetyarxiv-cs-cv
10 Apr 2026
Safety

Multi-agent Reach-avoid MDP via Potential Games and Low-rank Policy Structure

DGX agent

arXiv:2410.17690v2 Announce Type: replace-cross Abstract: We optimize finite horizon multi-agent reach-avoid Markov decision process (MDP) via local feedback policies. The global feedback polic

safetyarxiv-cs-ro
10 Apr 2026
Safety

Multi-Faceted Self-Consistent Preference Alignment for Query Rewriting in Conversational Search

DGX agent

arXiv:2604.06771v1 Announce Type: cross Abstract: Conversational Query Rewriting (CQR) aims to rewrite ambiguous queries to achieve more efficient conversational search. Early studies have predominant

safetyarxiv-cs-ai
10 Apr 2026
Safety

Multi-Turn Reasoning LLMs for Task Offloading in Mobile Edge Computing

DGX agent

arXiv:2604.07148v1 Announce Type: new Abstract: Emerging computation-intensive applications impose stringent latency requirements on resource-constrained mobile devices. Mobile Edge Computing (MEC) ad

safetyarxiv-cs-lg
10 Apr 2026
Safety

Neural Computers

DGX agent

arXiv:2604.06425v1 Announce Type: cross Abstract: We propose a new frontier: Neural Computers (NCs) -- an emerging machine form that unifies computation, memory, and I/O in a learned runtime state. Un

safetyarxiv-cs-ai
10 Apr 2026
Safety

On the Global Photometric Alignment for Low-Level Vision

DGX agent

arXiv:2604.08172v1 Announce Type: new Abstract: Supervised low-level vision models rely on pixel-wise losses against paired references, yet paired training sets exhibit per-pair photometric inconsiste

safetyarxiv-cs-cv
10 Apr 2026
Safety

OpenVLThinkerV2: A Generalist Multimodal Reasoning Model for Multi-domain Visual Tasks

DGX agent

arXiv:2604.08539v1 Announce Type: cross Abstract: Group Relative Policy Optimization (GRPO) has emerged as the de facto Reinforcement Learning (RL) objective driving recent advancements in Multimodal

safetyarxiv-cs-cl
10 Apr 2026
Safety

OVS-DINO: Open-Vocabulary Segmentation via Structure-Aligned SAM-DINO with Language Guidance

DGX agent

arXiv:2604.08461v1 Announce Type: new Abstract: Open-Vocabulary Segmentation (OVS) aims to segment image regions beyond predefined category sets by leveraging semantic descriptions. While CLIP based a

safetyarxiv-cs-cv
10 Apr 2026
Safety

OxEnsemble: Fair Ensembles for Low-Data Classification

DGX agent

arXiv:2512.09665v2 Announce Type: replace Abstract: We address the problem of fair classification in settings where data is scarce and unbalanced across demographic groups. Such low-data regimes are c

safetyarxiv-cs-cv
10 Apr 2026
Safety

Part^{2}GS: Part-aware Modeling of Articulated Objects using 3D Gaussian Splatting

DGX agent

arXiv:2506.17212v2 Announce Type: replace Abstract: Articulated objects are common in the real world, yet modeling their structure and motion remains a challenging task for 3D reconstruction methods.

safetyarxiv-cs-cv
10 Apr 2026
Safety

Path Regularization: A Near-Complete and Optimal Nonasymptotic Generalization Theory for Multilayer Neural Networks and Double Descent Phenomenon

DGX agent

arXiv:2503.02129v2 Announce Type: replace-cross Abstract: Path regularization has shown to be a very effective regularization to train neural networks, leading to a better generalization property than

safetyarxiv-cs-ai
10 Apr 2026
Safety

PEER: Unified Process-Outcome Reinforcement Learning for Structured Empathetic Reasoning

DGX agent

arXiv:2508.09521v2 Announce Type: replace Abstract: Emotional support conversations require more than fluent responses. Supporters need to understand the seeker's situation and emotions, adopt an appr

safetyarxiv-cs-cl
10 Apr 2026
Safety

Personalizing Text-to-Image Generation to Individual Taste

DGX agent

arXiv:2604.07427v1 Announce Type: new Abstract: Modern text-to-image (T2I) models generate high-fidelity visuals but remain indifferent to individual user preferences. While existing reward models opt

safetyarxiv-cs-cv
10 Apr 2026
← Previous
1…234235236237238…257
Next →