AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

ReFLEX: Length-Generalizable CSI Denoising for MIMO-OFDM via Relative-Frequency Bias

DGX agent

arXiv:2606.00263v1 Announce Type: cross Abstract: This letter studies CSI denoising for MIMO--OFDM with variable NR resource block (RB) allocations. ReFLEX is a length-generalizable Transformer whose

safetyarxiv-cs-lg
2 Jun 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Regime-Adaptive Continual Learning for Portfolio Management

DGX agent

arXiv:2606.00143v1 Announce Type: cross Abstract: Financial markets are inherently non-stationary, exhibiting frequent regime shifts and structural changes that render traditional Portfolio Management

safetyarxiv-cs-ai
2 Jun 2026
Safety

Regularized Offline Policy Optimization with Posterior Hybrid Bayesian Belief

DGX agent

arXiv:2606.00680v1 Announce Type: new Abstract: Offline reinforcement learning (RL) aims to optimize policies from pre-collected datasets. A bottleneck of this paradigm is managing epistemic uncertain

safetyarxiv-cs-ai
2 Jun 2026
Safety

Reinforcement Learning Position Control of a Quadrotor Using Soft Actor-Critic (SAC)

DGX agent

arXiv:2512.18333v2 Announce Type: replace-cross Abstract: This paper proposes a new Reinforcement Learning (RL) based control architecture for quadrotors. With the literature focusing on controlling t

safetyarxiv-cs-ai
2 Jun 2026
Safety

Reinforcement Learning with Pairwise Preferences in Long-Term Decision Problems

DGX agent

arXiv:2606.00367v1 Announce Type: cross Abstract: Reinforcement learning problems typically define the goal as maximizing the expected value of a scalar reward function. But, pairwise preferences are

safetyarxiv-cs-ai
2 Jun 2026
Model Releases

Relative Energy Learning for LiDAR Out-of-Distribution Detection

DGX agent

arXiv:2511.06720v3 Announce Type: replace Abstract: Out-of-distribution (OOD) detection is a critical requirement for reliable autonomous driving, where safety depends on recognizing road obstacles an

model-releasesarxiv-cs-cv
2 Jun 2026
Safety

Repurposing Adversarial Perturbations for Continual Learning: From Defense to Active Alignment

DGX agent

arXiv:2606.02322v1 Announce Type: cross Abstract: In dynamic environments, large language models need to keep adapting to new tasks, but continual learning often suffers from forgetting, limited trans

safetyarxiv-cs-ai
2 Jun 2026
Safety

ReSkill: Reconciling Skill Creation with Policy Optimization in Agentic RL

DGX agent

arXiv:2606.01619v1 Announce Type: new Abstract: Agentic reinforcement learning (RL) enables LLM agents to improve continuously from environment rewards, yet the resulting policies do not systematicall

safetyarxiv-cs-ai
2 Jun 2026
Safety

Restoring Initial Noise Sensitivity in Text-to-Image Distillation via Geometric Alignment

DGX agent

arXiv:2606.01651v1 Announce Type: new Abstract: Generative distillation significantly accelerates text-to-image (T2I) generation by compressing multi-step trajectories into few-step student models whi

safetyarxiv-cs-cv
2 Jun 2026
Safety

RichControl: Structure- and Appearance-Rich Training-Free Spatial Control for Text-to-Image Generation

DGX agent

arXiv:2507.02792v5 Announce Type: replace Abstract: Text-to-image (T2I) diffusion models have shown remarkable success in generating high-quality images from text prompts. Recent efforts extend these

safetyarxiv-cs-cv
2 Jun 2026
Safety

RL-ACRGNet: Reinforcement Learning-Based Chest Radiology Report Generation Network

DGX agent

arXiv:2606.02035v1 Announce Type: new Abstract: Medical imaging interpretation is a foundational pillar of modern clinical diagnostics, yet the manual generation of radiology reports remains a time-co

safetyarxiv-cs-ai
2 Jun 2026
Safety

RLVR without Ineffective Samples: Group Prioritized Off-Policy Optimization for LLM Reasoning

DGX agent

arXiv:2606.01281v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for enhancing the reasoning capabilities of large language mo

safetyarxiv-cs-ai
2 Jun 2026
Safety

RoboDream: Compositional World Models for Scalable Robot Data Synthesis

DGX agent

arXiv:2606.02577v1 Announce Type: cross Abstract: Scaling robot learning requires large-scale, diverse demonstrations, yet real-world data collection via teleoperation remains prohibitively expensive

safetyarxiv-cs-cv
2 Jun 2026
Safety

Scalable Ride-Sourcing Vehicle Rebalancing with Service Accessibility Guarantee: A Constrained Mean-Field Reinforcement Learning Approach

DGX agent

arXiv:2503.24183v3 Announce Type: replace Abstract: The expansion of ride-sourcing services such as Uber and Lyft has reshaped urban transportation by offering flexible, on-demand mobility via mobile

safetyarxiv-cs-lg
2 Jun 2026
Safety

Scalar-Measurement Attitude Estimation on mathbf{SO}(3) with Bias Compensation

DGX agent

arXiv:2603.02478v2 Announce Type: replace-cross Abstract: Attitude estimation methods typically rely on full vector measurements from inertial sensors such as accelerometers and magnetometers. This pa

safetyarxiv-cs-ro
2 Jun 2026
Safety

SCAPO: Self-Supervised Category-Level Articulated Pose Estimation from a Single 3D Observation

DGX agent

arXiv:2606.01940v1 Announce Type: new Abstract: Existing methods for category-level object articulation from a single 3D observation often rely on dense supervision, multi-frame inputs, or CAD templat

safetyarxiv-cs-cv
2 Jun 2026
Safety

SceneSmith: Agentic Generation of Simulation-Ready Indoor Scenes

DGX agent

arXiv:2602.09153v2 Announce Type: replace-cross Abstract: Simulation has become a key tool for training and evaluating home robots at scale, yet existing environments fail to capture the diversity and

safetyarxiv-cs-ai
2 Jun 2026
Safety

SEArch: Optimistic Policy Selection Between Scene Noise and Drift for UAV Radar Search

DGX agent

arXiv:2606.01325v1 Announce Type: cross Abstract: Unmanned Aerial Vehicles (UAVs) equipped with radar sensors are deployed for target search missions in diverse environments, where targets exhibit cha

safetyarxiv-cs-lg
2 Jun 2026
Safety

SECUREVENT: Hybrid AI/ML Security Monitoring for Distributed Event-Based Systems

DGX agent

arXiv:2606.01741v1 Announce Type: cross Abstract: Distributed event-based systems have become a common substrate for Internet-scale publish/subscribe services, IoT telemetry, cloud-native microservice

safetyarxiv-cs-ai
2 Jun 2026
Safety

Semantic Retrieval for Product Search in E-Commerce

DGX agent

arXiv:2606.01504v1 Announce Type: cross Abstract: Semantic retrieval in e-commerce must handle short, noisy, and colloquial queries over large product catalogs with fine-grained attribute distinctions

safetyarxiv-cs-lg
2 Jun 2026
Safety

SEMixer: Semantics Enhanced MLP-Mixer for Multiscale Mixing and Long-term Time Series Forecasting

DGX agent

arXiv:2602.16220v2 Announce Type: replace Abstract: Modeling multiscale patterns is crucial for long-term time series forecasting (TSF). However, redundancy and noise in time series, together with sem

safetyarxiv-cs-lg
2 Jun 2026
Safety

SentimentLens: Reconciling Sentiment and Ratings via Dual-Modality in the Hospitality Sector

DGX agent

arXiv:2606.00084v1 Announce Type: cross Abstract: Online travel platforms generate vast volumes of user-generated hotel reviews, offering rich opportunities to understand traveler experiences at scale

safetyarxiv-cs-ai
2 Jun 2026
Safety

Set-Supervised Diffusion Policy: Learning Action-Chunking Diffusion through Corrections

DGX agent

arXiv:2606.01865v1 Announce Type: new Abstract: Diffusion policies have recently emerged as a powerful framework for robotic manipulation. However, like other behavior cloning methods, they remain vul

safetyarxiv-cs-ro
2 Jun 2026
Safety

Shape-Prior-Based Point Cloud Completion for Single-Stage Fully Sparse 3D Object Detection

DGX agent

arXiv:2606.00688v1 Announce Type: new Abstract: Single-stage fully sparse 3D object detectors rely on point clouds data to detect objects in autonomous driving scenarios. However, the sparsity and inc

safetyarxiv-cs-cv
2 Jun 2026
Safety

Shape Your Body: Value Gradients for Multi-Embodiment Robot Design

DGX agent

arXiv:2606.00702v1 Announce Type: cross Abstract: We propose to turn generalist multi-embodiment value functions into reusable models for robot design. Instead of running a new reinforcement learning

safetyarxiv-cs-ai
2 Jun 2026
Safety

SHERLOCK: Towards Dynamic Knowledge Adaptation in LLM-enhanced E-commerce Risk Management

DGX agent

arXiv:2510.08948v4 Announce Type: replace-cross Abstract: Effective e-commerce risk management requires in-depth case investigations to identify emerging fraud patterns in highly adversarial environme

safetyarxiv-cs-ai
2 Jun 2026
Safety

Silent Failures in Federated Personalization of Foundation Models

DGX agent

arXiv:2606.00947v1 Announce Type: cross Abstract: Foundation models are increasingly personalized on decentralized private data through federated learning and are now deployed at scale under growing r

safetyarxiv-cs-ai
2 Jun 2026
Safety

Simultaneous Multi-objective Alignment Across Verifiable and Non-verifiable Rewards

DGX agent

arXiv:2510.01167v2 Announce Type: replace-cross Abstract: Aligning large language models to human preferences is inherently multidimensional, yet most pipelines collapse heterogeneous signals into a s

safetyarxiv-cs-ai
2 Jun 2026
Safety

SIRI: Self-Internalizing Reinforcement Learning with Intrinsic Skills for LLM Agent Training

DGX agent

arXiv:2606.02355v1 Announce Type: new Abstract: Long-horizon LLM agents can benefit from reusable skills, yet existing skill-based methods often rely on external skill generators during training or pe

safetyarxiv-cs-ai
2 Jun 2026
Safety

Skill or Skip? Learning Selective Skill Invocation in Agentic Tasks via Dual-Granularity Preference Learning

DGX agent

arXiv:2606.00510v1 Announce Type: cross Abstract: Agent skills are callable procedural modules that provide reusable knowledge and execution policies for complex agentic tasks. However, existing metho

safetyarxiv-cs-ai
2 Jun 2026
Safety

SKIP: Sparse Keyframe Interpolation Paradigm for Efficient Embodied World Models

DGX agent

arXiv:2606.00664v1 Announce Type: cross Abstract: Embodied world models have emerged as a promising paradigm in robotics by predicting how robot actions affect the surrounding scene. However, the roll

safetyarxiv-cs-cv
2 Jun 2026
Safety

SORA: Free Second-Order Attacks in Fast Adversarial Training

DGX agent

arXiv:2606.00738v1 Announce Type: cross Abstract: Adversarial Training (AT) is a leading defense against adversarial examples but often suffers from Catastrophic Overfitting (CO) in efficient single-s

safetyarxiv-cs-ai
2 Jun 2026
Safety

source: https://easternherald.com/2026/06/01/ai-trillion-dollar-frenzy-venture-capital-reality-check/

DGX agent

Gary Marcus discusses the disconnect between venture capital hype surrounding AI's trillion-dollar potential and the practical realities of AI development and market viability. The article likely exam

safetygary-marcus--x
2 Jun 2026
Safety

SpaceX has joined three other industry groups in calling on the FCC to adopt a nationwide policy to automatically unlock phones tied to a ca…

DGX agent

SpaceX has joined three other industry groups in calling on the FCC to adopt a nationwide policy to automatically unlock phones tied to a carrier’s network 180 days after activation. “Automatic mobile

safetyelon-musk--x
2 Jun 2026
Safety

SPADER: Step-wise Peer Advantage with Diversity-Aware Exploration Rewards for Multi-Answer Question Answering

DGX agent

arXiv:2606.00593v1 Announce Type: cross Abstract: Large language models are increasingly deployed as tool-augmented agents to acquire information beyond parametric knowledge. While recent work has imp

safetyarxiv-cs-ai
2 Jun 2026
Safety

Spatial-Temporal Decoupled Reference Conditioning for Identity-Preserving Text-to-Video Generation

DGX agent

arXiv:2606.02441v1 Announce Type: new Abstract: Identity-preserving video generation (IPVG) aims to synthesize high-fidelity videos that follow text prompts while faithfully preserving a reference ide

safetyarxiv-cs-cv
2 Jun 2026
Safety

Spatiotemporal Multi-Task Graph Transformer for Trip-Level Transit Prediction

DGX agent

arXiv:2606.00572v1 Announce Type: new Abstract: Passenger count data from public transit systems reveals urban mobility patterns and is essential for planning, operation, and optimisation. However, no

safetyarxiv-cs-lg
2 Jun 2026
Safety

SpeedAug: Policy Acceleration via Tempo-Enriched Policy and RL Fine-Tuning

DGX agent

arXiv:2512.00062v2 Announce Type: replace-cross Abstract: Robotic policy learning for complex real-world manipulation tasks has seen rapid recent progress, enabled in large part by the ability to coll

safetyarxiv-cs-ai
2 Jun 2026
Safety

SS-ZKR: Spatial-Semantic Zero-Knowledge Routing for Privacy-Preserving Multi-Agent Collaboration

DGX agent

arXiv:2606.00962v1 Announce Type: cross Abstract: Foundational agent interoperability standards, notably the Agent-to-Agent (A2A) protocol and the Model Context Protocol (MCP), have advanced multi-age

safetyarxiv-cs-ai
2 Jun 2026
Safety

Stabilizing Policy Optimization via Logits Convexity

DGX agent

arXiv:2603.00963v2 Announce Type: replace-cross Abstract: While reinforcement learning (RL) has been central to the recent success of large language models (LLMs), RL optimization is notoriously unsta

safetyarxiv-cs-cl
2 Jun 2026
Safety

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that th…

DGX agent

// State-Externalizing Harnesses // A new paradigm is emerging on how to effectively build agents and harnesses. If there is a state that the environment can maintain reliably, it probably doesn't bel

safetydair-ai--x
2 Jun 2026
Safety

StressDream: Steering Video World Models for Robust Policy Evaluation and Improvement

DGX agent

arXiv:2606.00267v1 Announce Type: cross Abstract: Video world models (WMs) have shown promise for policy evaluation and improvement by imagining realistic future observations conditioned on ego-robot

safetyarxiv-cs-ai
2 Jun 2026
Safety

SWARD: Stochastic Window-Attention-Based Relational Distillation for Cross-Architectural Semantic Segmentation

DGX agent

arXiv:2606.00999v1 Announce Type: new Abstract: Large-scale vision foundation models have driven substantial gains on dense prediction tasks such as semantic segmentation, but their size makes deploym

safetyarxiv-cs-cv
2 Jun 2026
Safety

Sympatheia: Emotionally Adaptive Voice Assistant with Continuous Affect Conditioning

DGX agent

arXiv:2606.00851v1 Announce Type: cross Abstract: Empathetic spoken dialogue systems must infer a user's emotional state to respond appropriately, yet everyday speech often carries weak, neutral, or a

safetyarxiv-cs-cl
2 Jun 2026
Safety

T-POP: Test-Time Personalization with Online Preference Feedback

DGX agent

arXiv:2509.24696v2 Announce Type: replace-cross Abstract: Personalizing large language models (LLMs) to individual user preferences is a critical step beyond generating generically helpful responses.

safetyarxiv-cs-ai
2 Jun 2026
Safety

Task-Induced Representational Invariances Depend on Learning Objective in Deep RL

DGX agent

arXiv:2606.01868v1 Announce Type: new Abstract: Reinforcement Learning (RL) has long served as a model for goal-directed animal behavior in neuroscience. Modern deep RL has shown remarkable success ac

safetyarxiv-cs-lg
2 Jun 2026
Safety

Thanks everyone for reading! Longer version some additional points on crowd psychology and links to a pair of great new podcasts with the le…

DGX agent

Thanks everyone for reading! Longer version some additional points on crowd psychology and links to a pair of great new podcasts with the legendary investors @realsteveeisman and @gnoble79 and converg

safetygary-marcus--x
2 Jun 2026
Safety

The Paradox of Outcome Optimization: A Causal Information-Theoretic Bound on Reasoning Shortcuts in LLMs

DGX agent

arXiv:2606.00674v1 Announce Type: cross Abstract: Large Language Models (LLMs) aligned via outcome-based Reinforcement Learning (RL) frequently exhibit a critical failure mode: they achieve high perfo

safetyarxiv-cs-ai
2 Jun 2026
← Previous
1…165166167168169…302
Next →