AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Verifiable Agentic Infrastructure: Proof-Derived Authorization for Sovereign AI Systems

DGX agent

arXiv:2605.15228v1 Announce Type: new Abstract: Modern cloud and enterprise systems rely on identity-centric authorization, assuming that callers possessing valid credentials are safe to execute comma

safetyarxiv-cs-ai
18 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Video Models Can Reason with Verifiable Rewards

DGX agent

arXiv:2605.15458v1 Announce Type: new Abstract: Video diffusion models have made rapid progress in perceptual realism and temporal coherence, but they remain primarily optimized for plausible generati

safetyarxiv-cs-cv
18 May 2026
Safety

VSPO: Vector-Steered Policy Optimization for Behavioral Control

DGX agent

arXiv:2605.15604v1 Announce Type: cross Abstract: Modern language models often need to optimize a primary accuracy objective while also accommodating secondary behavioral preferences, such as verbosit

safetyarxiv-cs-cl
18 May 2026
Safety

What Is Preference Optimization Doing, and Why?

DGX agent

arXiv:2512.00778v2 Announce Type: replace Abstract: Preference optimization (PO) is indispensable for large language models (LLMs), with methods such as direct preference optimization (DPO) and proxim

safetyarxiv-cs-lg
18 May 2026
Safety

When and Why Adversarial Training Improves PINNs: A Neural Tangent Kernel Perspective

DGX agent

arXiv:2605.15959v1 Announce Type: cross Abstract: Physics-informed neural networks (PINNs) are powerful surrogates for differential equations but are notoriously difficult to train due to spectral bia

safetyarxiv-cs-ai
18 May 2026
Safety

When Importance Sampling Misallocates Credit: Asymmetric Ratios for Outcome-Supervised RL

DGX agent

arXiv:2510.06062v2 Announce Type: replace Abstract: Reinforcement learning (RL) has shown great promise in large language models (LLMs) post-training, which typically rely on token-level clipping to m

safetyarxiv-cs-cl
18 May 2026
Safety

When Latent Geometry Is Not Enough: Draft-Conditioned Latent Refinement for Non-Autoregressive Text Generation

DGX agent

arXiv:2605.15557v1 Announce Type: new Abstract: Continuous diffusion and flow models are attractive for non-autoregressive text generation because they can update all positions in parallel. A major di

safetyarxiv-cs-cl
18 May 2026
Safety

A cross-species neural foundation model for end-to-end speech decoding

DGX agent

arXiv:2511.21740v5 Announce Type: replace-cross Abstract: Speech brain-computer interfaces (BCIs) aim to restore communication for people with paralysis by translating neural activity into text. Most

safetyarxiv-cs-ai
15 May 2026
Safety

A Security Analysis of the OpenClaw AI Agent Framework

DGX agent

arXiv:2603.27517v3 Announce Type: replace-cross Abstract: AI agent frameworks connecting large language model (LLM) reasoning to host execution surfaces -- shell, filesystem, containers, and messaging

safetyarxiv-cs-ai
15 May 2026
Safety

Achieving Approximate Symmetry Is Exponentially Easier than Exact Symmetry

DGX agent

arXiv:2512.11855v2 Announce Type: replace-cross Abstract: Enforcing exact symmetry in machine learning models often yields significant gains in scientific applications, serving as a powerful inductive

safetyarxiv-cs-ai
15 May 2026
Safety

Active Learners as Efficient PRP Rerankers

DGX agent

arXiv:2605.14236v1 Announce Type: cross Abstract: Pairwise Ranking Prompting (PRP) elicits pairwise preference judgments from an LLM, which are then aggregated into a ranking, usually via classical so

safetyarxiv-cs-ai
15 May 2026
Safety

ActivePusher: Active Learning and Planning with Residual Physics for Nonprehensile Manipulation

DGX agent

arXiv:2506.04646v4 Announce Type: replace-cross Abstract: Planning with learned dynamics models offers a promising approach toward versatile real-world manipulation, particularly in nonprehensile sett

safetyarxiv-cs-lg
15 May 2026
Safety

AIS: Adaptive Importance Sampling for Quantized RL

DGX agent

arXiv:2605.13907v1 Announce Type: cross Abstract: Reinforcement learning (RL) for large language models (LLMs) is dominated by the cost of rollout generation, which has motivated the use of low-precis

safetyarxiv-cs-ai
15 May 2026
Safety

Aligning Latent Geometry for Spherical Flow Matching in Image Generation

DGX agent

arXiv:2605.15193v1 Announce Type: new Abstract: Latent flow matching for image generation usually transports Gaussian noise to variational autoencoder latents along linear paths. Both endpoints, howev

safetyarxiv-cs-cv
15 May 2026
Safety

AMiD: Knowledge Distillation for LLMs with alpha-mixture Assistant Distribution

DGX agent

arXiv:2510.15982v3 Announce Type: replace-cross Abstract: Autoregressive large language models (LLMs) have achieved remarkable improvement across many tasks but incur high computational and memory cos

safetyarxiv-cs-ai
15 May 2026
Safety

Anti-Length Shift: Dynamic Outlier Truncation for Training Efficient Reasoning Models

DGX agent

arXiv:2601.03969v2 Announce Type: replace Abstract: Large reasoning models enhanced by reinforcement learning with verifiable rewards have achieved significant performance gains by extending their cha

safetyarxiv-cs-ai
15 May 2026
Safety

AQKA: Active Quantum Kernel Acquisition Under a Shot Budget

DGX agent

arXiv:2605.14672v1 Announce Type: new Abstract: Estimating an N imes N quantum kernel from circuit fidelities requires Theta(N^2 S) measurement shots, the dominant bottleneck for deployment on near-te

safetyarxiv-cs-lg
15 May 2026
Safety

ASH: Agents that Self-Hone via Embodied Learning

DGX agent

arXiv:2605.14211v1 Announce Type: new Abstract: Long-horizon embodied tasks remain a fundamental challenge in AI, as current methods rely on hand-engineered rewards or action-labeled demonstrations, n

safetyarxiv-cs-ai
15 May 2026
Safety

AutoMoT: A Unified Vision-Language-Action Model with Asynchronous Mixture-of-Transformers for End-to-End Autonomous Driving

DGX agent

arXiv:2603.14851v3 Announce Type: replace Abstract: Integrating vision-language models (VLMs) into end-to-end (E2E) autonomous driving (AD) systems has shown promise in improving scene understanding.

safetyarxiv-cs-cv
15 May 2026
Safety

Before the Body Moves: Learning Anticipatory Joint Intent for Language-Conditioned Humanoid Control

DGX agent

arXiv:2605.14417v1 Announce Type: cross Abstract: Natural language is an intuitive interface for humanoid robots, yet streaming whole-body control requires control representations that are executable

safetyarxiv-cs-cv
15 May 2026
Safety

Beyond Instance-Level Self-Supervision in 3D Multi-Modal Medical Imaging

DGX agent

arXiv:2605.14654v1 Announce Type: new Abstract: Self-supervised pre-training methods in medical imaging typically treat each individual as an isolated instance, learning representations through augmen

safetyarxiv-cs-cv
15 May 2026
Safety

BiSpikCLM: A Spiking Language Model integrating Softmax-Free Spiking Attention and Spike-Aware Alignment Distillation

DGX agent

arXiv:2605.13859v1 Announce Type: cross Abstract: Spiking Neural Networks (SNNs) offer promising energy-efficient alternatives to large language models (LLMs) due to their event-driven nature and ultr

safetyarxiv-cs-ai
15 May 2026
Safety

Boosting Reinforcement Learning with Verifiable Rewards via Randomly Selected Few-Shot Guidance

DGX agent

arXiv:2605.15012v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has achieved great success in developing Large Language Models (LLMs) with chain-of-thought roll

safetyarxiv-cs-ai
15 May 2026
Safety

Break-the-Beat! Controllable MIDI-to-Drum Audio Synthesis

DGX agent

arXiv:2605.14555v1 Announce Type: cross Abstract: Current methods for creating drum loop audio in digital music production, such as using one-shot samples or resampling, often demand non-trivial effor

safetyarxiv-cs-ai
15 May 2026
Safety

CHASM: Cross-frequency Harmonized Axis-Separable Mixing for Spectral Token Operators

DGX agent

arXiv:2605.14727v1 Announce Type: new Abstract: Spectral token mixers based on Fourier transforms provide an efficient way to model global interactions in visual feature maps. Existing designs often e

safetyarxiv-cs-cv
15 May 2026
Safety

COAL: Counterfactual and Observation-Enhanced Alignment Learning for Discriminative Referring Multi-Object Tracking

DGX agent

arXiv:2605.14795v1 Announce Type: new Abstract: Referring Multi-Object Tracking (RMOT) faces a fundamental structural contradiction between the high-discriminability demand and the sparse semantic sup

safetyarxiv-cs-cv
15 May 2026
Safety

Collaborative Yet Personalized Policy Training: Single-Timescale Federated Actor-Critic

DGX agent

arXiv:2605.14423v1 Announce Type: cross Abstract: Despite the popularity of the actor-critic method and the practical needs of collaborative policy training, existing works typically either overlook e

safetyarxiv-cs-ai
15 May 2026
Safety

Comparing Developer and LLM Biases in Code Evaluation

DGX agent

arXiv:2603.24586v2 Announce Type: replace-cross Abstract: As LLMs are increasingly used as judges in code applications, they should be evaluated in realistic interactive settings that capture partial

safetyarxiv-cs-cl
15 May 2026
Safety

Complacent, Not Sycophantic: Reframing Large Language Models and Designing AI Literacy for Complacent Machines

DGX agent

arXiv:2605.14544v1 Announce Type: new Abstract: Large language models are often described as sycophantic, in the sense that they appear to flatter users or mirror their beliefs. We argue that this lab

safetyarxiv-cs-ai
15 May 2026
Safety

Compositional Sparsity as an Inductive Bias for Neural Architecture Design

DGX agent

arXiv:2605.14764v1 Announce Type: cross Abstract: Identifying the structural priors that enable Deep Neural Networks (DNNs) to overcome the curse of dimensionality is a fundamental challenge in machin

safetyarxiv-cs-ai
15 May 2026
Safety

CrystalReasoner: Reasoning and RL for Property-Conditioned Crystal Structure Generation

DGX agent

arXiv:2605.14344v1 Announce Type: new Abstract: Generative modeling has emerged as a promising approach for crystal structure discovery. However, existing LLM-based generative models struggle with low

safetyarxiv-cs-ai
15 May 2026
Safety

CuSearch: Curriculum Rollout Sampling via Search Depth for Agentic RAG

DGX agent

arXiv:2605.11611v2 Announce Type: replace Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for training agentic retrieval-augmented generation (RAG)

safetyarxiv-cs-ai
15 May 2026
Safety

DAPL: Integration of Positive and Negative Descriptions in Text-Based Person Search

DGX agent

arXiv:2405.07459v3 Announce Type: replace Abstract: Text-based person search (TBPS) aims to retrieve specific images of individuals from large datasets using textual descriptions. Existing TBPS method

safetyarxiv-cs-cv
15 May 2026
Safety

Deciphering Neural Reparameterized Full-Waveform Inversion with Neural Sensitivity Kernel and Wave Tangent Kernel

DGX agent

arXiv:2605.14370v1 Announce Type: cross Abstract: Full-waveform inversion (FWI) estimates unknown parameters in the wave equation from limited boundary measurements. Recent advances in neural reparame

safetyarxiv-cs-ai
15 May 2026
Safety

Deepchecks: Evaluating Retrieval-Augmented Generation (RAG)

DGX agent

arXiv:2605.14488v1 Announce Type: new Abstract: Large Language Models (LLMs) augmented with Retrieval-Augmented Generation (RAG) techniques are revolutionizing applications across multiple domains, su

safetyarxiv-cs-ai
15 May 2026
Safety

Delta Forcing: Trust Region Steering for Interactive Autoregressive Video Generation

DGX agent

arXiv:2605.14382v1 Announce Type: new Abstract: Interactive real-time autoregressive video generation is essential for applications such as content creation and world modeling, where visual content mu

safetyarxiv-cs-cv
15 May 2026
Safety

Diagnosing Training Inference Mismatch in LLM Reinforcement Learning

DGX agent

arXiv:2605.14220v1 Announce Type: cross Abstract: Modern LLM RL systems separate rollout generation from policy optimization. These two stages are expected to produce token probabilities that match ex

safetyarxiv-cs-ai
15 May 2026
Safety

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models

DGX agent

arXiv:2605.15055v1 Announce Type: cross Abstract: Reinforcement learning has emerged as a powerful tool for improving diffusion-based text-to-image models, but existing methods are largely limited to

safetyarxiv-cs-cv
15 May 2026
Safety

Dimension-Level Intent Fidelity Evaluation for Large Language Models: Evidence from Structured Prompt Ablation

DGX agent

arXiv:2605.14517v1 Announce Type: cross Abstract: Holistic evaluation scores capture overall output quality but do not distinguish whether a model reproduced the structural form of a user's request fr

safetyarxiv-cs-ai
15 May 2026
Safety

Distance-Matrix Wasserstein Statistics for Scalable Gromov--Wasserstein Learning

DGX agent

arXiv:2605.14981v1 Announce Type: new Abstract: Gromov--Wasserstein (GW) distances compare graphs, shapes, and point clouds through internal distances, without requiring a common coordinate system. Th

safetyarxiv-cs-lg
15 May 2026
Safety

Distribution Corrected Offline Data Distillation for Large Language Models

DGX agent

arXiv:2605.14071v1 Announce Type: new Abstract: Distilling reasoning traces from strong large language models into smaller ones is a promising route to improve intelligence in resource-constrained set

safetyarxiv-cs-cl
15 May 2026
Safety

Distributions as Actions: A Unified Framework for Diverse Action Spaces

DGX agent

arXiv:2506.16608v3 Announce Type: replace-cross Abstract: We introduce a novel reinforcement learning (RL) framework that treats parameterized action distributions as actions, redefining the boundary

safetyarxiv-cs-ai
15 May 2026
Safety

Do Language Models Align with Brains? Prediction Scores Are Not Enough

DGX agent

arXiv:2605.14025v1 Announce Type: cross Abstract: Brain-language model comparisons often interpret neural prediction scores as evidence that model representations capture brain-relevant language compu

safetyarxiv-cs-ai
15 May 2026
Safety

DSSP: Diffusion State Space Policy with Full-History Encoding

DGX agent

arXiv:2605.14598v1 Announce Type: new Abstract: Diffusion-based imitation learning has shown strong promise for robot manipulation. However, most existing policies condition only on the current observ

safetyarxiv-cs-ro
15 May 2026
Safety

Dual Hierarchical Dialogue Policy Learning for Legal Inquisitive Conversational Agents

DGX agent

arXiv:2605.14057v1 Announce Type: new Abstract: Most existing dialogue systems are user-driven, primarily designed to fulfill user requests. However, in many critical real-world scenarios, a conversat

safetyarxiv-cs-cl
15 May 2026
Safety

Dynamic Mixed-Precision Routing for Efficient Multi-step LLM Interaction

DGX agent

arXiv:2602.02711v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance in long-horizon decision-making tasks through multi-step interaction and reasoning at test t

safetyarxiv-cs-ai
15 May 2026
Safety

Efficient Generative Retrieval for E-commerce Search with Semantic Cluster IDs and Expert-Guided RL

DGX agent

arXiv:2605.14434v1 Announce Type: cross Abstract: Generative retrieval offers a promising alternative by unifying the fragmented multi-stage retrieval process into a single end-to-end model. However,

safetyarxiv-cs-ai
15 May 2026
Safety

EponaV2: Driving World Model with Comprehensive Future Reasoning

DGX agent

arXiv:2605.14696v1 Announce Type: new Abstract: Data scaling plays a pivotal role in the pursuit of general intelligence. However, the prevailing perception-planning paradigm in autonomous driving rel

safetyarxiv-cs-cv
15 May 2026
← Previous
1…179180181182183…260
Next →