AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
2 Jun 2026

PR2: Predictive Routing Replay for MoE-Based LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.00395v1 Announce Type: cross Abstract: Mixture of Experts (MoE) Large Language Models (LLMs) achieve strong performance at scale. However, reinforcement learning (RL) on MoE-based LLMs ofte

Principled Synthetic Data Enables the First Scaling Laws for LLMs in Recommendation

SafetyDGX agent

arXiv:2602.07298v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) represent a promising frontier for recommender systems, yet their development has been impeded by the absence of

Probabilistic Performance Guarantees for Multi-Task Reinforcement Learning

SafetyDGX agent

arXiv:2602.02098v2 Announce Type: replace-cross Abstract: Multi-task reinforcement learning trains generalist policies that can execute multiple tasks. While recent years have seen significant progres


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Prospect-Theory Behavior from Bellman Optimality in MDPs with Catastrophic States

SafetyDGX agent

arXiv:2606.00970v1 Announce Type: new Abstract: We study risk-neutral control in Markov decision processes with an absorbing catastrophic state. Even though rewards are linear and the agent has no uti

Quantifying the Salience of Geo-Cultural Values for Pluralistic Safety Alignment

SafetyDGX agent

arXiv:2606.00369v1 Announce Type: cross Abstract: Safe global deployment of AI models requires alignment with human values that vary across cultures. Yet rater pools in safety evaluation datasets rema

Quantitative Movement Testing: Measuring Patient Movements from a Single Smartphone Video

SafetyDGX agent

arXiv:2606.02301v1 Announce Type: cross Abstract: Chronic pain diminishes quality of life by decreasing functional ability, yet objectively measuring this functional impact remains challenging in real

RADE: Random Add-Drop Edge as a Regularizer

SafetyDGX agent

arXiv:2606.00757v1 Announce Type: new Abstract: Graph Neural Networks (GNNs) suffer from overfitting and over-squashing of long-range information. Stochastic graph augmentations (e.g., edge deletion)

RAFT: Data Refinement and Adaptive Distillation for Domain Fine-Tuning with Alleviated Forgetting

SafetyDGX agent

arXiv:2606.00147v1 Announce Type: cross Abstract: Domain-specific supervised fine-tuning (SFT) often improves in-domain performance at the cost of degrading a model's general capabilities. We view thi

RAIGen: Rare Attribute Identification in Text-to-Image Generative Models

SafetyDGX agent

arXiv:2602.06806v2 Announce Type: replace Abstract: Text-to-image diffusion models achieve impressive generation quality but inherit and amplify training-data biases, skewing coverage of semantic attr

RankByGene: Gene-Guided Histopathology Representation Learning Through Cross-Modal Ranking Consistency

SafetyDGX agent

arXiv:2411.15076v3 Announce Type: replace-cross Abstract: Spatial transcriptomics (ST) provides essential spatial context by mapping gene expression within tissue, enabling detailed study of cellular

REAL: Resolving Knowledge Conflicts in Knowledge-Intensive Visual Question Answering via Reasoning-Pivot Alignment

SafetyDGX agent

arXiv:2602.14065v2 Announce Type: replace Abstract: Knowledge-intensive Visual Question Answering (KI-VQA) frequently suffers from severe knowledge conflicts caused by the inherent limitations of open

Real-Time Sensing of Inaccessible Physical Fields via an Edge-Deployable Hardware-Portable Graph Neural Operator

SafetyDGX agent

arXiv:2604.01802v2 Announce Type: replace Abstract: Real-time inference of inaccessible interior physical fields from sparse boundary observations is a fundamental but unresolved problem in scientific

REBot: From RAG to CatRAG with Semantic Enrichment and Graph Routing

SafetyDGX agent

arXiv:2510.01800v3 Announce Type: replace Abstract: Academic regulation advising is essential for helping students interpret and comply with institutional policies, yet building effective systems requ

Reconsidering Positional Supervision in Masked Diffusion Language Model Training

SafetyDGX agent

arXiv:2601.22947v2 Announce Type: replace Abstract: Masked diffusion language models (MDLMs) generate text by unmasking tokens in parallel and have recently emerged as alternatives to autoregressive l

RedDebate: Safer Responses Through Multi-Agent Red Teaming Debates

SafetyDGX agent

arXiv:2506.11083v3 Announce Type: replace Abstract: We introduce RedDebate, a novel multi-agent debate framework that provides the foundation for Large Language Models (LLMs) to identify and mitigate

ReFLEX: Length-Generalizable CSI Denoising for MIMO-OFDM via Relative-Frequency Bias

SafetyDGX agent

arXiv:2606.00263v1 Announce Type: cross Abstract: This letter studies CSI denoising for MIMO--OFDM with variable NR resource block (RB) allocations. ReFLEX is a length-generalizable Transformer whose

Regime-Adaptive Continual Learning for Portfolio Management

SafetyDGX agent

arXiv:2606.00143v1 Announce Type: cross Abstract: Financial markets are inherently non-stationary, exhibiting frequent regime shifts and structural changes that render traditional Portfolio Management

Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief

SafetyDGX agent

arXiv:2606.00680v1 Announce Type: new Abstract: Offline reinforcement learning (RL) aims to optimize policies from pre-collected datasets. A bottleneck of this paradigm is managing epistemic uncertain

Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)

SafetyDGX agent

arXiv:2512.18333v2 Announce Type: replace-cross Abstract: This paper proposes a new Reinforcement Learning (RL) based control architecture for quadrotors. With the literature focusing on controlling t

Reinforcement Learning with Pairwise Preferences in Long-Term Decision Problems

SafetyDGX agent

arXiv:2606.00367v1 Announce Type: cross Abstract: Reinforcement learning problems typically define the goal as maximizing the expected value of a scalar reward function. But, pairwise preferences are

Repurposing Adversarial Perturbations for Continual Learning: From Defense to Active Alignment

SafetyDGX agent

arXiv:2606.02322v1 Announce Type: cross Abstract: In dynamic environments, large language models need to keep adapting to new tasks, but continual learning often suffers from forgetting, limited trans

RESBev: Making BEV Perception More Robust

SafetyDGX agent

arXiv:2603.09529v2 Announce Type: replace Abstract: Bird's-eye-view (BEV) perception has emerged as a cornerstone of autonomous driving systems, providing a structured, ego-centric representation crit

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

SafetyDGX agent

arXiv:2606.01619v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) enables LLM agents to improve continuously from environment rewards, yet the resulting policies do not systematicall

Restoring Initial Noise Sensitivity in Text-to-Image Distillation via Geometric Alignment

SafetyDGX agent

arXiv:2606.01651v1 Announce Type: new Abstract: Generative distillation significantly accelerates text-to-image (T2I) generation by compressing multi-step trajectories into few-step student models whi

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation

SafetyDGX agent

arXiv:2507.02792v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models have shown remarkable success in generating high-quality images from text prompts. Recent efforts extend these

RL-ACRGNet: Reinforcement Learning-Based Chest Radiology Report Generation Network

SafetyDGX agent

arXiv:2606.02035v1 Announce Type: new Abstract: Medical imaging interpretation is a foundational pillar of modern clinical diagnostics, yet the manual generation of radiology reports remains a time-co

RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.01281v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for enhancing the reasoning capabilities of large language mo

RoboDream: Compositional World Models for Scalable Robot Data Synthesis

SafetyDGX agent

arXiv:2606.02577v1 Announce Type: cross Abstract: Scaling robot learning requires large-scale, diverse demonstrations, yet real-world data collection via teleoperation remains prohibitively expensive

Robust Integrated Planning and Control for Quadrotors in Dynamic Environments via NMPC with CBF Penalties

SafetyDGX agent

arXiv:2606.01038v1 Announce Type: new Abstract: This paper presents a new robust integrated planning and control (IPC) strategy for multirotor uncrewed aerial vehicles. We propose a nonlinear model pr

Robust Shielding for Safe Reinforcement Learning

SafetyDGX agent

arXiv:2606.00270v1 Announce Type: new Abstract: Shielding is an effective approach to formally guarantee the safety of reinforcement learning agents in Markov decision processes (MDPs). However, exist

Safe2Drive: Evaluating Safe Driving Behaviors of E2E Autonomous Driving Models

SafetyDGX agent

arXiv:2606.00191v1 Announce Type: cross Abstract: Recent end-to-end (E2E) autonomous driving policies achieve high driving scores in closed-loop simulations. Yet it remains unclear whether these polic

SafeMCP: Proactive Power Regulation for LLM Agent Defense via Environment-Grounded Look-Ahead Reasoning

SafetyDGX agent

arXiv:2606.01991v1 Announce Type: new Abstract: As Large Language Model (LLM) agents increasingly leverage the Model Context Protocol (MCP) to operate in complex environments, the expansion of their a

SafeSteer: Localized On-Policy Distillation for Efficient Safety Alignment

SafetyDGX agent

arXiv:2606.02530v1 Announce Type: new Abstract: Aligning Large Language Models (LLMs) with human values often degrades their general capabilities, termed the alignment tax. Existing methods mitigate t

Safety Alignment of LMs via Non-cooperative Games

SafetyDGX agent

arXiv:2512.20806v3 Announce Type: replace Abstract: Ensuring the safety of language models (LMs) while maintaining their usefulness remains a critical challenge in AI alignment. Current approaches rel

Safety Game: Inference-Time Alignment of Black-Box LLMs via Constrained Optimization

SafetyDGX agent

arXiv:2510.09330v3 Announce Type: replace Abstract: Ensuring that large language models (LLMs) comply with safety requirements is a central challenge in AI deployment. Existing alignment approaches pr

Safety Mirage: How Spurious Correlations Undermine VLM Safety Fine-Tuning and Can Be Mitigated by Machine Unlearning

SafetyDGX agent

arXiv:2503.11832v5 Announce Type: replace Abstract: Recent vision language models (VLMs) have made remarkable strides in generative modeling with multimodal inputs, particularly text and images. Howev

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

SafetyDGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

Scalar-Measurement Attitude Estimation on mathbf{SO}(3) with Bias Compensation

SafetyDGX agent

arXiv:2603.02478v2 Announce Type: replace-cross Abstract: Attitude estimation methods typically rely on full vector measurements from inertial sensors such as accelerometers and magnetometers. This pa

SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation

SafetyDGX agent

arXiv:2606.01940v1 Announce Type: new Abstract: Existing methods for category-level object articulation from a single 3D observation often rely on dense supervision, multi-frame inputs, or CAD templat

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

SafetyDGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search

SafetyDGX agent

arXiv:2606.01325v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit cha

SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems

SafetyDGX agent

arXiv:2606.01741v1 Announce Type: cross Abstract: Distributed event-based systems have become a common substrate for Internet-scale publish/subscribe services, IoT telemetry, cloud-native microservice

Semantic Retrieval for Product Search in E-Commerce

SafetyDGX agent

arXiv:2606.01504v1 Announce Type: cross Abstract: Semantic retrieval in e-commerce must handle short, noisy, and colloquial queries over large product catalogs with fine-grained attribute distinctions

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

SafetyDGX agent

arXiv:2602.16220v2 Announce Type: replace Abstract: Modeling multiscale patterns is crucial for long-term time series forecasting (TSF). However, redundancy and noise in time series, together with sem

SentimentLens: Reconciling Sentiment and Ratings via Dual-Modality in the Hospitality Sector

SafetyDGX agent

arXiv:2606.00084v1 Announce Type: cross Abstract: Online travel platforms generate vast volumes of user-generated hotel reviews, offering rich opportunities to understand traveler experiences at scale

Set-Supervised Diffusion Policy: Learning Action-Chunking Diffusion through Corrections

SafetyDGX agent

arXiv:2606.01865v1 Announce Type: new Abstract: Diffusion policies have recently emerged as a powerful framework for robotic manipulation. However, like other behavior cloning methods, they remain vul

Shape-Prior-Based Point Cloud Completion for Single-Stage Fully Sparse 3D Object Detection

SafetyDGX agent

arXiv:2606.00688v1 Announce Type: new Abstract: Single-stage fully sparse 3D object detectors rely on point clouds data to detect objects in autonomous driving scenarios. However, the sparsity and inc

Shape Your Body: Value Gradients for Multi-Embodiment Robot Design

SafetyDGX agent

arXiv:2606.00702v1 Announce Type: cross Abstract: We propose to turn generalist multi-embodiment value functions into reusable models for robot design. Instead of running a new reinforcement learning

SHERLOCK: Towards Dynamic Knowledge Adaptation in LLM-enhanced E-commerce Risk Management

SafetyDGX agent

arXiv:2510.08948v4 Announce Type: replace-cross Abstract: Effective e-commerce risk management requires in-depth case investigations to identify emerging fraud patterns in highly adversarial environme

Silent Failures in Federated Personalization of Foundation Models

SafetyDGX agent

arXiv:2606.00947v1 Announce Type: cross Abstract: Foundation models are increasingly personalized on decentralized private data through federated learning and are now deployed at scale under growing r

Silent Failures in Physical AI: A Literature Review of Runtime Action Authorization for Autonomous Systems

SafetyDGX agent

arXiv:2606.00090v1 Announce Type: cross Abstract: Physical AI systems increasingly map multimodal observations, language instructions, and learned world representations into physically consequential a

SilentDrift: Exploiting Action Chunking for Stealthy Backdoor Attacks on Vision-Language-Action Models

SafetyDGX agent

arXiv:2601.14323v2 Announce Type: replace-cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed in safety-critical robotic applications, yet their security vulnerabilities rema

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards

SafetyDGX agent

arXiv:2510.01167v2 Announce Type: replace-cross Abstract: Aligning large language models to human preferences is inherently multidimensional, yet most pipelines collapse heterogeneous signals into a s

SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training

SafetyDGX agent

arXiv:2606.02355v1 Announce Type: new Abstract: Long-horizon LLM agents can benefit from reusable skills, yet existing skill-based methods often rely on external skill generators during training or pe

Skill or Skip? Learning Selective Skill Invocation in Agentic Tasks via Dual-Granularity Preference Learning

SafetyDGX agent

arXiv:2606.00510v1 Announce Type: cross Abstract: Agent skills are callable procedural modules that provide reusable knowledge and execution policies for complex agentic tasks. However, existing metho

SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models

SafetyDGX agent

arXiv:2606.00664v1 Announce Type: cross Abstract: Embodied world models have emerged as a promising paradigm in robotics by predicting how robot actions affect the surrounding scene. However, the roll

SORA: Free Second-Order Attacks in Fast Adversarial Training

SafetyDGX agent

arXiv:2606.00738v1 Announce Type: cross Abstract: Adversarial Training (AT) is a leading defense against adversarial examples but often suffers from Catastrophic Overfitting (CO) in efficient single-s

source: https://easternherald.com/2026/06/01/ai-trillion-dollar-frenzy-venture-capital-reality-check/

SafetyDGX agent

Gary Marcus discusses the disconnect between venture capital hype surrounding AI's trillion-dollar potential and the practical realities of AI development and market viability. The article likely exam

SpaceX has joined three other industry groups in calling on the FCC to adopt a nationwide policy to automatically unlock phones tied to a ca…

SafetyDGX agent

SpaceX has joined three other industry groups in calling on the FCC to adopt a nationwide policy to automatically unlock phones tied to a carrier’s network 180 days after activation. “Automatic mobile

SPADER: Step-wise Peer Advantage with Diversity-Aware Exploration Rewards for Multi-Answer Question Answering

SafetyDGX agent

arXiv:2606.00593v1 Announce Type: cross Abstract: Large language models are increasingly deployed as tool-augmented agents to acquire information beyond parametric knowledge. While recent work has imp

← Previous
1…99100101102103…214
Next →