AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

Position: Good Embodied Reward Models Need Bad Behavior Data

DGX agent

arXiv:2606.01036v1 Announce Type: new Abstract: This position paper argues that to obtain reliable embodied reward models, the community must invest in ``bad'' robot data: failed, suboptimal, error-pr

safetyarxiv-cs-ro
2 Jun 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Position: Stop Preaching and Start Practising Data Frugality for Responsible Development of AI

DGX agent

arXiv:2602.19789v2 Announce Type: replace Abstract: This position paper argues that the machine learning community must move from preaching to practising data frugality for responsible artificial inte

safetyarxiv-cs-lg
2 Jun 2026
Safety

Post-Deterministic Distributed Systems: A New Foundation for Trustworthy Autonomous Infrastructure

DGX agent

arXiv:2606.01722v1 Announce Type: cross Abstract: For decades, distributed systems have typically assumed that correct participants execute protocol-specified behavior with stable, externally defined,

safetyarxiv-cs-ai
2 Jun 2026
Safety

PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning

DGX agent

arXiv:2606.00395v1 Announce Type: cross Abstract: Mixture of Experts (MoE) Large Language Models (LLMs) achieve strong performance at scale. However, reinforcement learning (RL) on MoE-based LLMs ofte

safetyarxiv-cs-ai
2 Jun 2026
Safety

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

DGX agent

arXiv:2602.07298v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of

safetyarxiv-cs-ai
2 Jun 2026
Safety

Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning

DGX agent

arXiv:2602.02098v2 Announce Type: replace-cross Abstract: Multi-task reinforcement learning trains generalist policies that can execute multiple tasks. While recent years have seen significant progres

safetyarxiv-cs-ai
2 Jun 2026
Safety

Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States

DGX agent

arXiv:2606.00970v1 Announce Type: new Abstract: We study risk-neutral control in Markov decision processes with an absorbing catastrophic state. Even though rewards are linear and the agent has no uti

safetyarxiv-cs-ai
2 Jun 2026
Safety

Quantifying the Salience of Geo-Cultural Values for Pluralistic Safety Alignment

DGX agent

arXiv:2606.00369v1 Announce Type: cross Abstract: Safe global deployment of AI models requires alignment with human values that vary across cultures. Yet rater pools in safety evaluation datasets rema

safetyarxiv-cs-lg
2 Jun 2026
Safety

Quantitative Movement Testing: Measuring Patient Movements from a Single Smartphone Video

DGX agent

arXiv:2606.02301v1 Announce Type: cross Abstract: Chronic pain diminishes quality of life by decreasing functional ability, yet objectively measuring this functional impact remains challenging in real

safetyarxiv-cs-ai
2 Jun 2026
Safety

RADE: Random Add-Drop Edge as a Regularizer

DGX agent

arXiv:2606.00757v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) suffer from overfitting and over-squashing of long-range information. Stochastic graph augmentations (e.g., edge deletion)

safetyarxiv-cs-lg
2 Jun 2026
Safety

RAFT: Data Refinement and Adaptive Distillation for Domain Fine-Tuning with Alleviated Forgetting

DGX agent

arXiv:2606.00147v1 Announce Type: cross Abstract: Domain-specific supervised fine-tuning (SFT) often improves in-domain performance at the cost of degrading a model's general capabilities. We view thi

safetyarxiv-cs-ai
2 Jun 2026
Safety

RAIGen: Rare Attribute Identification in Text-to-Image Generative Models

DGX agent

arXiv:2602.06806v2 Announce Type: replace Abstract: Text-to-image diffusion models achieve impressive generation quality but inherit and amplify training-data biases, skewing coverage of semantic attr

safetyarxiv-cs-cv
2 Jun 2026
Safety

RankByGene: Gene-Guided Histopathology Representation Learning Through Cross-Modal Ranking Consistency

DGX agent

arXiv:2411.15076v3 Announce Type: replace-cross Abstract: Spatial transcriptomics (ST) provides essential spatial context by mapping gene expression within tissue, enabling detailed study of cellular

safetyarxiv-cs-cv
2 Jun 2026
Safety

REAL: Resolving Knowledge Conflicts in Knowledge-Intensive Visual Question Answering via Reasoning-Pivot Alignment

DGX agent

arXiv:2602.14065v2 Announce Type: replace Abstract: Knowledge-intensive Visual Question Answering (KI-VQA) frequently suffers from severe knowledge conflicts caused by the inherent limitations of open

safetyarxiv-cs-ai
2 Jun 2026
Safety

Real-Time Sensing of Inaccessible Physical Fields via an Edge-Deployable Hardware-Portable Graph Neural Operator

DGX agent

arXiv:2604.01802v2 Announce Type: replace Abstract: Real-time inference of inaccessible interior physical fields from sparse boundary observations is a fundamental but unresolved problem in scientific

safetyarxiv-cs-lg
2 Jun 2026
Safety

REBot: From RAG to CatRAG with Semantic Enrichment and Graph Routing

DGX agent

arXiv:2510.01800v3 Announce Type: replace Abstract: Academic regulation advising is essential for helping students interpret and comply with institutional policies, yet building effective systems requ

safetyarxiv-cs-ai
2 Jun 2026
Safety

Reconsidering Positional Supervision in Masked Diffusion Language Model Training

DGX agent

arXiv:2601.22947v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by unmasking tokens in parallel and have recently emerged as alternatives to autoregressive l

safetyarxiv-cs-cl
2 Jun 2026
Safety

RedDebate: Safer Responses Through Multi-Agent Red Teaming Debates

DGX agent

arXiv:2506.11083v3 Announce Type: replace Abstract: We introduce RedDebate, a novel multi-agent debate framework that provides the foundation for Large Language Models (LLMs) to identify and mitigate

safetyarxiv-cs-cl
2 Jun 2026
Safety

ReFLEX: Length-Generalizable CSI Denoising for MIMO-OFDM via Relative-Frequency Bias

DGX agent

arXiv:2606.00263v1 Announce Type: cross Abstract: This letter studies CSI denoising for MIMO--OFDM with variable NR resource block (RB) allocations. ReFLEX is a length-generalizable Transformer whose

safetyarxiv-cs-lg
2 Jun 2026
Safety

Regime-Adaptive Continual Learning for Portfolio Management

DGX agent

arXiv:2606.00143v1 Announce Type: cross Abstract: Financial markets are inherently non-stationary, exhibiting frequent regime shifts and structural changes that render traditional Portfolio Management

safetyarxiv-cs-ai
2 Jun 2026
Safety

Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief

DGX agent

arXiv:2606.00680v1 Announce Type: new Abstract: Offline reinforcement learning (RL) aims to optimize policies from pre-collected datasets. A bottleneck of this paradigm is managing epistemic uncertain

safetyarxiv-cs-ai
2 Jun 2026
Safety

Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)

DGX agent

arXiv:2512.18333v2 Announce Type: replace-cross Abstract: This paper proposes a new Reinforcement Learning (RL) based control architecture for quadrotors. With the literature focusing on controlling t

safetyarxiv-cs-ai
2 Jun 2026
Safety

Reinforcement Learning with Pairwise Preferences in Long-Term Decision Problems

DGX agent

arXiv:2606.00367v1 Announce Type: cross Abstract: Reinforcement learning problems typically define the goal as maximizing the expected value of a scalar reward function. But, pairwise preferences are

safetyarxiv-cs-ai
2 Jun 2026
Safety

Repurposing Adversarial Perturbations for Continual Learning: From Defense to Active Alignment

DGX agent

arXiv:2606.02322v1 Announce Type: cross Abstract: In dynamic environments, large language models need to keep adapting to new tasks, but continual learning often suffers from forgetting, limited trans

safetyarxiv-cs-ai
2 Jun 2026
Safety

RESBev: Making BEV Perception More Robust

DGX agent

arXiv:2603.09529v2 Announce Type: replace Abstract: Bird's-eye-view (BEV) perception has emerged as a cornerstone of autonomous driving systems, providing a structured, ego-centric representation crit

safetyarxiv-cs-cv
2 Jun 2026
Safety

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

DGX agent

arXiv:2606.01619v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) enables LLM agents to improve continuously from environment rewards, yet the resulting policies do not systematicall

safetyarxiv-cs-ai
2 Jun 2026
Safety

Restoring Initial Noise Sensitivity in Text-to-Image Distillation via Geometric Alignment

DGX agent

arXiv:2606.01651v1 Announce Type: new Abstract: Generative distillation significantly accelerates text-to-image (T2I) generation by compressing multi-step trajectories into few-step student models whi

safetyarxiv-cs-cv
2 Jun 2026
Safety

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation

DGX agent

arXiv:2507.02792v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models have shown remarkable success in generating high-quality images from text prompts. Recent efforts extend these

safetyarxiv-cs-cv
2 Jun 2026
Safety

RL-ACRGNet: Reinforcement Learning-Based Chest Radiology Report Generation Network

DGX agent

arXiv:2606.02035v1 Announce Type: new Abstract: Medical imaging interpretation is a foundational pillar of modern clinical diagnostics, yet the manual generation of radiology reports remains a time-co

safetyarxiv-cs-ai
2 Jun 2026
Safety

RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.01281v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for enhancing the reasoning capabilities of large language mo

safetyarxiv-cs-ai
2 Jun 2026
Safety

RoboDream: Compositional World Models for Scalable Robot Data Synthesis

DGX agent

arXiv:2606.02577v1 Announce Type: cross Abstract: Scaling robot learning requires large-scale, diverse demonstrations, yet real-world data collection via teleoperation remains prohibitively expensive

safetyarxiv-cs-cv
2 Jun 2026
Safety

Robust Integrated Planning and Control for Quadrotors in Dynamic Environments via NMPC with CBF Penalties

DGX agent

arXiv:2606.01038v1 Announce Type: new Abstract: This paper presents a new robust integrated planning and control (IPC) strategy for multirotor uncrewed aerial vehicles. We propose a nonlinear model pr

safetyarxiv-cs-ro
2 Jun 2026
Safety

Robust Shielding for Safe Reinforcement Learning

DGX agent

arXiv:2606.00270v1 Announce Type: new Abstract: Shielding is an effective approach to formally guarantee the safety of reinforcement learning agents in Markov decision processes (MDPs). However, exist

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safe2Drive: Evaluating Safe Driving Behaviors of E2E Autonomous Driving Models

DGX agent

arXiv:2606.00191v1 Announce Type: cross Abstract: Recent end-to-end (E2E) autonomous driving policies achieve high driving scores in closed-loop simulations. Yet it remains unclear whether these polic

safetyarxiv-cs-cv
2 Jun 2026
Safety

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

DGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

safetyarxiv-cs-ai
2 Jun 2026
Safety

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment

DGX agent

arXiv:2606.02530v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) with human values often degrades their general capabilities, termed the alignment tax. Existing methods mitigate t

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safety Alignment of LMs via Non-cooperative Games

DGX agent

arXiv:2512.20806v3 Announce Type: replace Abstract: Ensuring the safety of language models (LMs) while maintaining their usefulness remains a critical challenge in AI alignment. Current approaches rel

safetyarxiv-cs-ai
2 Jun 2026
Safety

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

DGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

safetyarxiv-cs-lg
2 Jun 2026
Safety

Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Mitigated by Machine Unlearning

DGX agent

arXiv:2503.11832v5 Announce Type: replace Abstract: Recent vision language models (VLMs) have made remarkable strides in generative modeling with multimodal inputs, particularly text and images. Howev

safetyarxiv-cs-ai
2 Jun 2026
Safety

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

DGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

safetyarxiv-cs-lg
2 Jun 2026
Safety

Scalar-Measurement Attitude Estimation on mathbf{SO}(3) with Bias Compensation

DGX agent

arXiv:2603.02478v2 Announce Type: replace-cross Abstract: Attitude estimation methods typically rely on full vector measurements from inertial sensors such as accelerometers and magnetometers. This pa

safetyarxiv-cs-ro
2 Jun 2026
Safety

SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation

DGX agent

arXiv:2606.01940v1 Announce Type: new Abstract: Existing methods for category-level object articulation from a single 3D observation often rely on dense supervision, multi-frame inputs, or CAD templat

safetyarxiv-cs-cv
2 Jun 2026
Safety

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

DGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

safetyarxiv-cs-ai
2 Jun 2026
Safety

SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search

DGX agent

arXiv:2606.01325v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit cha

safetyarxiv-cs-lg
2 Jun 2026
Safety

SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems

DGX agent

arXiv:2606.01741v1 Announce Type: cross Abstract: Distributed event-based systems have become a common substrate for Internet-scale publish/subscribe services, IoT telemetry, cloud-native microservice

safetyarxiv-cs-ai
2 Jun 2026
Safety

Semantic Retrieval for Product Search in E-Commerce

DGX agent

arXiv:2606.01504v1 Announce Type: cross Abstract: Semantic retrieval in e-commerce must handle short, noisy, and colloquial queries over large product catalogs with fine-grained attribute distinctions

safetyarxiv-cs-lg
2 Jun 2026
Safety

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

DGX agent

arXiv:2602.16220v2 Announce Type: replace Abstract: Modeling multiscale patterns is crucial for long-term time series forecasting (TSF). However, redundancy and noise in time series, together with sem

safetyarxiv-cs-lg
2 Jun 2026
Safety

SentimentLens: Reconciling Sentiment and Ratings via Dual-Modality in the Hospitality Sector

DGX agent

arXiv:2606.00084v1 Announce Type: cross Abstract: Online travel platforms generate vast volumes of user-generated hotel reviews, offering rich opportunities to understand traveler experiences at scale

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…124125126127128…267
Next →