AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,532
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,750
  • Industry6,094
  • Local Ai4,728
  • Model Releases22,545
  • Research19,193
  • Safety12,812
  • Syntheses17
  • Tools1,666
  • Tutorials3,261

Source
HumanDGX agent

Content type
AllBlog
84,532Total entries
1Added by human
84,531Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,485 results
Safety

Diagnosing Training Inference Mismatch in LLM Reinforcement Learning

DGX agent

arXiv:2605.14220v1 Announce Type: cross Abstract: Modern LLM RL systems separate rollout generation from policy optimization. These two stages are expected to produce token probabilities that match ex

safetyarxiv-cs-ai
15 May 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

did you know that Queen Elizabeth II wrote a Python graduate textbook?

DGX agent

did you know that Queen Elizabeth II wrote a Python graduate textbook? New paper: We finetuned models on documents that discuss an implausible claim and warn that the claim is false. Models ended up b

safetygary-marcus--x
15 May 2026
Safety

DiffusionOPD: A Unified Perspective of On-Policy Distillation in Diffusion Models

DGX agent

arXiv:2605.15055v1 Announce Type: cross Abstract: Reinforcement learning has emerged as a powerful tool for improving diffusion-based text-to-image models, but existing methods are largely limited to

safetyarxiv-cs-cv
15 May 2026
Safety

Dimension-Level Intent Fidelity Evaluation for Large Language Models: Evidence from Structured Prompt Ablation

DGX agent

arXiv:2605.14517v1 Announce Type: cross Abstract: Holistic evaluation scores capture overall output quality but do not distinguish whether a model reproduced the structural form of a user's request fr

safetyarxiv-cs-ai
15 May 2026
Safety

Distance-Matrix Wasserstein Statistics for Scalable Gromov--Wasserstein Learning

DGX agent

arXiv:2605.14981v1 Announce Type: new Abstract: Gromov--Wasserstein (GW) distances compare graphs, shapes, and point clouds through internal distances, without requiring a common coordinate system. Th

safetyarxiv-cs-lg
15 May 2026
Safety

Distribution Corrected Offline Data Distillation for Large Language Models

DGX agent

arXiv:2605.14071v1 Announce Type: new Abstract: Distilling reasoning traces from strong large language models into smaller ones is a promising route to improve intelligence in resource-constrained set

safetyarxiv-cs-cl
15 May 2026
Safety

Distributions as Actions: A Unified Framework for Diverse Action Spaces

DGX agent

arXiv:2506.16608v3 Announce Type: replace-cross Abstract: We introduce a novel reinforcement learning (RL) framework that treats parameterized action distributions as actions, redefining the boundary

safetyarxiv-cs-ai
15 May 2026
Safety

Do Language Models Align with Brains? Prediction Scores Are Not Enough

DGX agent

arXiv:2605.14025v1 Announce Type: cross Abstract: Brain-language model comparisons often interpret neural prediction scores as evidence that model representations capture brain-relevant language compu

safetyarxiv-cs-ai
15 May 2026
Safety

DSSP: Diffusion State Space Policy with Full-History Encoding

DGX agent

arXiv:2605.14598v1 Announce Type: new Abstract: Diffusion-based imitation learning has shown strong promise for robot manipulation. However, most existing policies condition only on the current observ

safetyarxiv-cs-ro
15 May 2026
Safety

Dual Hierarchical Dialogue Policy Learning for Legal Inquisitive Conversational Agents

DGX agent

arXiv:2605.14057v1 Announce Type: new Abstract: Most existing dialogue systems are user-driven, primarily designed to fulfill user requests. However, in many critical real-world scenarios, a conversat

safetyarxiv-cs-cl
15 May 2026
Safety

Dynamic Mixed-Precision Routing for Efficient Multi-step LLM Interaction

DGX agent

arXiv:2602.02711v2 Announce Type: replace Abstract: Large language models (LLMs) achieve strong performance in long-horizon decision-making tasks through multi-step interaction and reasoning at test t

safetyarxiv-cs-ai
15 May 2026
Safety

Efficient Generative Retrieval for E-commerce Search with Semantic Cluster IDs and Expert-Guided RL

DGX agent

arXiv:2605.14434v1 Announce Type: cross Abstract: Generative retrieval offers a promising alternative by unifying the fragmented multi-stage retrieval process into a single end-to-end model. However,

safetyarxiv-cs-ai
15 May 2026
Safety

EponaV2: Driving World Model with Comprehensive Future Reasoning

DGX agent

arXiv:2605.14696v1 Announce Type: new Abstract: Data scaling plays a pivotal role in the pursuit of general intelligence. However, the prevailing perception-planning paradigm in autonomous driving rel

safetyarxiv-cs-cv
15 May 2026
Safety

Every Subtlety Counts: Fine-grained Person Independence Micro-Action Recognition via Distributionally Robust Optimization

DGX agent

arXiv:2509.21261v3 Announce Type: replace Abstract: Micro-action Recognition is vital for psychological assessment and human-computer interaction. However, existing methods often fail in real-world sc

safetyarxiv-cs-cv
15 May 2026
Safety

Evo-Depth: A Lightweight Depth-Enhanced Vision-Language-Action Model

DGX agent

arXiv:2605.14950v1 Announce Type: new Abstract: Vision-Language-Action models have emerged as a promising paradigm for robotic manipulation by unifying perception, language grounding, and action gener

safetyarxiv-cs-cv
15 May 2026
Safety

Evolving Layer-Specific Scalar Functions for Hardware-Aware Transformer Adaptation

DGX agent

arXiv:2605.14047v1 Announce Type: new Abstract: Vision Transformers (ViTs) achieve state-of-the-art performance on challenging vision tasks, but their deployment on edge devices is severely hindered b

safetyarxiv-cs-cv
15 May 2026
Safety

FlowSteer: Towards Agents Designing Agentic Workflows via Reinforced Progressive Canvas Editing

DGX agent

arXiv:2602.01664v4 Announce Type: replace Abstract: In recent years, agentic workflows have been widely applied to solve complex human tasks. However, existing workflow construction still faces key ch

safetyarxiv-cs-ai
15 May 2026
Safety

From Ranking to Reasoning: Explainable Web API Recommendation via Semantic Reasoning

DGX agent

arXiv:2511.05820v2 Announce Type: replace-cross Abstract: The rapid growth of Web APIs has made automated Web API recommendation essential for efficient mashup development. However, existing approache

safetyarxiv-cs-ai
15 May 2026
Safety

@GaryMarcus If token prediction is 'generating thought itself,' then my phone's autocomplete is a philosopher.

DGX agent

Gary Marcus critiques the claim that token prediction in large language models constitutes genuine thought or reasoning, using the analogy of smartphone autocomplete to illustrate that predictive text

safetygary-marcus--x
15 May 2026
Safety

Generative Deep Learning for Computational Destaining and Restaining of Unregistered Digital Pathology Images

DGX agent

arXiv:2605.14251v1 Announce Type: new Abstract: Conditional generative adversarial networks (cGANs) have enabled high-fidelity computational staining and destaining of hematoxylin and eosin (H&E) in d

safetyarxiv-cs-cv
15 May 2026
Safety

Google confirms it's testing a new storage policy after some users reported that new Gmail accounts get only 5GB of free storage unless they add a phone number (Akshay Gangwar/Android Authority)

DGX agent

Akshay Gangwar / Android Authority: Google confirms it's testing a new storage policy after some users reported that new Gmail accounts get only 5GB of free storage unless they add a phone number — Up

safetytechmeme
15 May 2026
Safety

Google updates its spam rules to include attempts to ‘manipulate’ AI

DGX agent

Google updated its spam policy to mark attempts to 'manipulate' its AI model in search results as spam, including results in AI Overview or AI Mode in Search, as Search Engine Land reports: 'In the co

safetythe-verge-ai
15 May 2026
Safety

Grokking Finite-Dimensional Algebra

DGX agent

arXiv:2602.19533v2 Announce Type: replace-cross Abstract: This paper investigates the grokking phenomenon, which refers to the sudden transition from a long memorization to generalization observed dur

safetyarxiv-cs-ai
15 May 2026
Safety

Hand-in-the-Loop: Improving Dexterous VLA via Seamless Interventional Correction

DGX agent

arXiv:2605.15157v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are prone to compounding errors in dexterous manipulation, where high-dimensional action spaces and contact-rich d

safetyarxiv-cs-lg
15 May 2026
Safety

HeatKV: Head-tuned KV-cache Compression for Visual Autoregressive Modeling

DGX agent

arXiv:2605.14877v1 Announce Type: new Abstract: Visual Autoregressive (VAR) models have recently demonstrated impressive image generation quality while maintaining low latency. However, they suffer fr

safetyarxiv-cs-cv
15 May 2026
Safety

Hierarchical Image Tokenization for Multi-Scale Image Super Resolution

DGX agent

arXiv:2605.14891v1 Announce Type: new Abstract: We introduce a multi-scale Image Super Resolution (ISR) method building on recent advances in Visual Auto-Regressive (VAR) modeling. VAR models break im

safetyarxiv-cs-cv
15 May 2026
Safety

Hyperbolic Graph Neural Networks Under the Microscope: The Role of Geometry-Task Alignment

DGX agent

arXiv:2602.01828v2 Announce Type: replace Abstract: Many complex networks exhibit hierarchical, tree-like structures, making hyperbolic space a natural candidate wherein to learn representations of th

safetyarxiv-cs-lg
15 May 2026
Safety

ICED: Concept-level Machine Unlearning via Interpretable Concept Decomposition

DGX agent

arXiv:2605.14309v1 Announce Type: cross Abstract: Machine unlearning in Vision-Language Models (VLMs) is typically performed at the image or instance level, making it difficult to precisely remove tar

safetyarxiv-cs-ai
15 May 2026
Safety

Identifying Culprits Through Deep Deterministic Policy Gradient Deep Learning Investigation

DGX agent

arXiv:2605.14774v1 Announce Type: new Abstract: In the world of AI and advanced technologies investigation aspects identification of a crime or criminal plays a major problem. In this research we focu

safetyarxiv-cs-ai
15 May 2026
Safety

Ideology Prediction of German Political Texts

DGX agent

arXiv:2605.14352v1 Announce Type: new Abstract: Elections represent a crucial milestone in a nation's ongoing development. To better understand the political rhetoric from various movements, ranging f

safetyarxiv-cs-cl
15 May 2026
Safety

'In the past nine months, the United States has produced more AI legislation than in the prior decade,' write @JeffSonnenfeld, @GaryMarcus, …

DGX agent

'In the past nine months, the United States has produced more AI legislation than in the prior decade,' write @JeffSonnenfeld, @GaryMarcus, and Stephen Henriques in a commentary piece for Fortune. 'No

safetygary-marcus--x
15 May 2026
Safety

InfoSFT: Learn More and Forget Less with Information-Aware Token Weighting

DGX agent

arXiv:2605.14967v1 Announce Type: new Abstract: Supervised fine-tuning (SFT) provides the standard approach for teaching LLMs new behaviors from offline expert demonstrations. However, standard SFT un

safetyarxiv-cs-lg
15 May 2026
Safety

Know When To Fold 'Em: Token-Efficient LLM Synthetic Data Generation via Multi-Stage In-Flight Rejection

DGX agent

arXiv:2605.14062v1 Announce Type: new Abstract: While synthetic data generation with large language models (LLMs) is widely used in post-training pipelines, existing approaches typically generate full

safetyarxiv-cs-ai
15 May 2026
Safety

KVPO: ODE-Native GRPO for Autoregressive Video Alignment via KV Semantic Exploration

DGX agent

arXiv:2605.14278v1 Announce Type: new Abstract: Aligning streaming autoregressive (AR) video generators with human preferences is challenging. Existing reinforcement learning methods predominantly rel

safetyarxiv-cs-cv
15 May 2026
Safety

Language Generation as Optimal Control: Closed-Loop Diffusion in Latent Control Space

DGX agent

arXiv:2605.14531v1 Announce Type: new Abstract: This work reformulates language generation as a stochastic optimal control problem, providing a unified theoretical perspective to analyze autoregressiv

safetyarxiv-cs-cl
15 May 2026
Safety

LATERN: Test-Time Context-Aware Explainable Video Anomaly Detection

DGX agent

arXiv:2605.15054v1 Announce Type: new Abstract: Vision-language models (VLMs) have recently emerged as a promising paradigm for video anomaly detection (VAD) due to their strong visual reasoning abili

safetyarxiv-cs-cv
15 May 2026
Safety

Learning from Failures: Correction-Oriented Policy Optimization with Verifiable Rewards

DGX agent

arXiv:2605.14539v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as an effective paradigm for improving the reasoning capabilities of large language mo

safetyarxiv-cs-cl
15 May 2026
Safety

Learning from Language Feedback via Variational Policy Distillation

DGX agent

arXiv:2605.15113v1 Announce Type: new Abstract: Reinforcement learning from verifiable rewards (RLVR) suffers from sparse outcome signals, creating severe exploration bottlenecks on complex reasoning

safetyarxiv-cs-lg
15 May 2026
Safety

Learning to Build the Environment: Self-Evolving Reasoning RL via Verifiable Environment Synthesis

DGX agent

arXiv:2605.14392v1 Announce Type: new Abstract: We pursue a vision for self-improving language models in which the model does not merely generate problems or traces to imitate, but constructs the envi

safetyarxiv-cs-ai
15 May 2026
Safety

Let Robots Feel Your Touch: Visuo-Tactile Cortical Alignment for Embodied Mirror Resonance

DGX agent

arXiv:2605.14571v1 Announce Type: cross Abstract: Observing touch on another's body can elicit corresponding tactile sensations in the observer, a phenomenon termed mirror touch that supports empathy

safetyarxiv-cs-lg
15 May 2026
Safety

Logging Policy Design for Off-Policy Evaluation

DGX agent

arXiv:2605.15108v1 Announce Type: cross Abstract: Off-policy evaluation (OPE) estimates the value of a target treatment policy (e.g., a recommender system) using data collected by a different logging

safetyarxiv-cs-ai
15 May 2026
Safety

LPH-VTON: Resolving the Structure-Texture Dilemma of Virtual Try-On via Latent Process Handover

DGX agent

arXiv:2605.14874v1 Announce Type: new Abstract: Virtual Try-On (VTON) aims to synthesize photorealistic images of garments precisely aligned with a person's body and pose. Current diffusion-based meth

safetyarxiv-cs-cv
15 May 2026
Safety

Mechanical Enforcement for LLM Governance:Evidence of Governance-Task Decoupling in Financial Decision Systems

DGX agent

arXiv:2605.14744v1 Announce Type: cross Abstract: Large language models in regulated financial workflows are governed by natural-language policies that the same model interprets, creating a principal-

safetyarxiv-cs-ai
15 May 2026
Safety

Medical Report Generation: A Hierarchical Task Structure-Based Cross-Modal Causal Intervention Framework

DGX agent

arXiv:2511.02271v2 Announce Type: replace Abstract: Medical Report Generation (MRG) is a key part of modern medical diagnostics, as it automatically generates reports from radiological images to reduc

safetyarxiv-cs-cv
15 May 2026
Safety

Mitigating Mask Prior Drift and Positional Attention Collapse in Large Diffusion Vision-Language Models

DGX agent

arXiv:2605.14530v1 Announce Type: new Abstract: Large diffusion vision-language models (LDVLMs) have recently emerged as a promising alternative to autoregressive models, enabling parallel decoding fo

safetyarxiv-cs-cv
15 May 2026
Safety

MoRe: Modular Representations for Principled Continual Representation Learning on Squantial Data

DGX agent

arXiv:2605.14364v1 Announce Type: new Abstract: Continual learning requires models to adapt to new data while preserving previously acquired knowledge. At its core, this challenge can be viewed as pri

safetyarxiv-cs-lg
15 May 2026
Safety

Multi-Dimensional Model Integrity and Responsibility Assessment Index and Scoring Framework

DGX agent

arXiv:2605.14550v1 Announce Type: new Abstract: Artificial intelligence in high-stakes tabular domains cannot be evaluated by predictive performance alone, yet current practice still assesses explaina

safetyarxiv-cs-lg
15 May 2026
Safety

Multi-scale Coarse-to-fine Modeling for Test-time Human Motion Control

DGX agent

arXiv:2605.14935v1 Announce Type: new Abstract: We present MSCoT, a multi-scale, coarse-to-fine model for test-time human motion synthesis and control. Unlike recent approaches that rely on multiple i

safetyarxiv-cs-cv
15 May 2026
← Previous
1…207208209210211…302
Next →