AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
Safety

Multi-component Causal Tracing in Large Language Models

DGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

safetyarxiv-cs-cl
3 Jun 2026
Safety

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

safetyarxiv-cs-cl
3 Jun 2026
Safety

Neural Networks Provably Learn Spectral Representations for Group Composition

DGX agent

arXiv:2606.02993v1 Announce Type: new Abstract: Understanding how structured internal structure emerges during neural network training is central to the study of deep learning. We investigate this phe

safetyarxiv-cs-lg
3 Jun 2026
Safety

No one. No one in their right mind.

DGX agent

No one. No one in their right mind. SpaceX is losing money hand over fist, nearly 5bn last year. @SpaceX’s only profitable business is Starlink, but its new satellites can only be launched by Starship

safetygary-marcus--x
3 Jun 2026
Safety

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

DGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

safetyarxiv-cs-ai
3 Jun 2026
Safety

OMP: One-step Meanflow Policy with Directional Alignment

DGX agent

arXiv:2512.19347v3 Announce Type: replace Abstract: Robot manipulation has increasingly adopted data-driven generative policy frameworks, yet the field faces a persistent trade-off: diffusion models s

safetyarxiv-cs-ro
3 Jun 2026
Safety

OpenAI diverges from Trump's AI EO in a new policy paper, proposing cyber risk evaluations for advanced AI systems be mandatory and led by CAISI, not the NSA (Brendan Bordelon/Politico)

DGX agent

Brendan Bordelon / Politico: OpenAI diverges from Trump's AI EO in a new policy paper, proposing cyber risk evaluations for advanced AI systems be mandatory and led by CAISI, not the NSA — OpenAI's ne

safetytechmeme
3 Jun 2026
Safety

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

DGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

Patcher: Post-Hoc Patching of Backdoored Large Language Models

DGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Pathway-Structured Privileged Distillation for Deployable Computational Pathology

DGX agent

arXiv:2606.02877v1 Announce Type: new Abstract: Integrating transcriptomics and histopathology can improve cancer risk modelling, yet practical use is constrained by the limited availability of RNA pr

safetyarxiv-cs-cv
3 Jun 2026
Safety

Pextsuperscript{2}-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization

DGX agent

arXiv:2606.03376v1 Announce Type: cross Abstract: Hallucination has recently garnered significant research attention in Large Vision-Language Models (LVLMs). Direct Preference Optimization (DPO) aims

safetyarxiv-cs-ai
3 Jun 2026
Safety

PHAF-Personalized Hand Avatars in a Flash

DGX agent

arXiv:2606.03420v1 Announce Type: new Abstract: We present PHAF-Personalized Hand Avatars in a Flash, a personalized photo-realistic hand avatar which provides high quality multi-view renders from jus

safetyarxiv-cs-cv
3 Jun 2026
Safety

PHASE: Physiology-Aware Hyperspectral Reconstruction via Object-to-Human Domain Adaptation

DGX agent

arXiv:2511.13020v2 Announce Type: replace-cross Abstract: Although hyperspectral imaging offers unparalleled non-invasive physiological insight, its bulky hardware, slow acquisition, and regulatory bu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance

DGX agent

arXiv:2511.10055v2 Announce Type: replace Abstract: The performance of image generation has been significantly improved in recent years. However, the study of image screening is rare, and its performa

safetyarxiv-cs-cv
3 Jun 2026
Safety

Physics-Guided Policy Optimization with Self-Distillation

DGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

safetyarxiv-cs-ai
3 Jun 2026
Safety

Planning with Uncertainty: Symmetries, Policy Inference, and Solution Compression

DGX agent

arXiv:2403.19883v2 Announce Type: replace Abstract: Fully-observable non-deterministic (FOND) planning is at the core of artificial intelligence planning with uncertainty. It models uncertainty throug

safetyarxiv-cs-ai
3 Jun 2026
Safety

Post-Hoc Robustness for Model-Based Reinforcement Learning

DGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

safetyarxiv-cs-ai
3 Jun 2026
Safety

Preference-Calibrated Human-in-the-Loop Reinforcement Learning for Robotic Manipulation

DGX agent

arXiv:2606.03949v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning (HIL-RL) improves sample efficiency in real-robot manipulation through online human intervention. However, succ

safetyarxiv-cs-ro
3 Jun 2026
Safety

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

DGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

safetyarxiv-cs-ai
3 Jun 2026
Safety

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations

DGX agent

arXiv:2606.03136v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks on large language models (LLMs) reveal a mismatch in current guardrails: they operate on individual turns, while attacks

safetyarxiv-cs-cl
3 Jun 2026
Safety

Quantifying Faithful Confidence Expression in Large Reasoning Models

DGX agent

arXiv:2606.03969v1 Announce Type: cross Abstract: Reliable uncertainty communication is critical to the trustworthiness of LLMs, yet faithful calibration (FC)--the alignment between models' intrinsic

safetyarxiv-cs-ai
3 Jun 2026
Safety

QUBRIC: Co-Designing Queries and Rubrics for RL Beyond Verifiable Rewards

DGX agent

arXiv:2606.03968v1 Announce Type: cross Abstract: Rubric-based RL is a promising route for extending reinforcement learning beyond verifiable rewards, yet existing methods optimize rubrics while treat

safetyarxiv-cs-ai
3 Jun 2026
Safety

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning

DGX agent

arXiv:2606.03234v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has become the dominant approach for improving mathematical reasoning in large language models, ye

safetyarxiv-cs-lg
3 Jun 2026
Safety

RoboCade: Gamifying Robot Data Collection

DGX agent

arXiv:2512.21235v3 Announce Type: replace Abstract: Imitation learning from human demonstrations has become a dominant approach for training autonomous robot policies. However, collecting demonstratio

safetyarxiv-cs-ro
3 Jun 2026
Safety

Sam Altman swearing to tell the whole truth to the US Senate an hour or so before he lied his fanny off about caring about artists and creat…

DGX agent

Sam Altman swearing to tell the whole truth to the US Senate an hour or so before he lied his fanny off about caring about artists and creators and wanting them to get a fair shake. Hey OpenAI, when a

safetygary-marcus--x
3 Jun 2026
Safety

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs

DGX agent

arXiv:2606.02735v1 Announce Type: cross Abstract: Generalization remains a central bottleneck for vision-language-action (VLA) models: under distractors, appearance shifts, and semantically similar ta

safetyarxiv-cs-ai
3 Jun 2026
Safety

Selective Token-Level Cryptographic Redaction for Privacy-Preserving Clinical Deployment of Large Language Models

DGX agent

arXiv:2606.03399v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly used for clinical applications, many existing pipelines require sending raw sensitive health informa

safetyarxiv-cs-cl
3 Jun 2026
Safety

Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation

DGX agent

arXiv:2606.03963v1 Announce Type: cross Abstract: Deep reinforcement learning has shown strong potential for enabling autonomous robots to learn complex navigational tasks. However, its practical use

safetyarxiv-cs-ai
3 Jun 2026
Safety

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

DGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

safetyarxiv-cs-cv
3 Jun 2026
Safety

SkelHCC: A Hyperbolic CLIP-Driven Cache Adaptation Framework for Skeleton-based One-Shot Action Recognition

DGX agent

arXiv:2606.03610v1 Announce Type: new Abstract: Skeleton-based action recognition aims to understand human behaviors from body joint sequences and is especially challenging in the one-shot setting, wh

safetyarxiv-cs-cv
3 Jun 2026
Safety

Sparse-View Lung Nodule Volumetry from Digitally Reconstructed Radiographs via AReT: Anatomy-Regularized TensoRF

DGX agent

arXiv:2606.02639v1 Announce Type: cross Abstract: We identify and resolve a previously unreported failure mode in TensoRF when applied to X-ray attenuation fields: the default density shift of -10, or

safetyarxiv-cs-ai
3 Jun 2026
Safety

Spectral Asymptotics of Neural Network Loss Landscapes: An Exact Decomposition of the Curvature Exponent

DGX agent

arXiv:2606.02596v1 Announce Type: new Abstract: The curvature exponent alpha in h_k propto sigma_k^alpha -- governing how Hessian eigenvalues scale with gradient singular values -- varies systematical

safetyarxiv-cs-lg
3 Jun 2026
Safety

SplitAdapter: Load-Aware Humanoid Loco-Manipulation via Factorized Adaptation

DGX agent

arXiv:2606.03297v1 Announce Type: new Abstract: Humanoid loco-manipulation requires stable whole-body control under varying object masses and pickup/placement heights. This becomes particularly challe

safetyarxiv-cs-ro
3 Jun 2026
Safety

Strongly Polynomial Time Complexity of Policy Iteration for L_infty Robust MDPs

DGX agent

arXiv:2601.23229v2 Announce Type: replace Abstract: Markov decision processes (MDPs) are a fundamental model in sequential decision making. Robust MDPs (RMDPs) extend this framework by allowing uncert

safetyarxiv-cs-ai
3 Jun 2026
Safety

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it…

DGX agent

Suno - a company that trained on 'essentially all music files of reasonable quality that are accessible on the open Internet', and argues it does not need to pay to do so - is now valued at $5.4 billi

safetygary-marcus--x
3 Jun 2026
Safety

Supercell, King, and Sybo warn that EU's Digital Fairness Act, requiring pop-ups showing real-world values of virtual currencies, could make games 'unplayable' (Richard Milne/Financial Times)

DGX agent

Richard Milne / Financial Times: Supercell, King, and Sybo warn that EU's Digital Fairness Act, requiring pop-ups showing real-world values of virtual currencies, could make games “unplayable” — Maker

safetytechmeme
3 Jun 2026
Safety

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

DGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

safetyarxiv-cs-ai
3 Jun 2026
Safety

Temporal Action Selection for Action Chunking

DGX agent

arXiv:2511.04421v2 Announce Type: replace Abstract: Action chunking is a widely adopted approach in Learning from Demonstration (LfD). By modeling multi-step action chunks rather than single-step acti

safetyarxiv-cs-ro
3 Jun 2026
Safety

TGV-KV: Text-Grounded KV Eviction for Vision-Language Models

DGX agent

arXiv:2606.03075v1 Announce Type: new Abstract: Vision-Language Models (VLMs) inherit the auto-regressive generation paradigm and cache the keys and values (KV) of all previous tokens to accelerate in

safetyarxiv-cs-cv
3 Jun 2026
Safety

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

DGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

safetyarxiv-cs-cl
3 Jun 2026
Safety

The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs

DGX agent

arXiv:2606.03092v1 Announce Type: new Abstract: Inference-time scaling has emerged as a critical avenue for enhancing Large Language Models' performance, yet real-world deployment is constrained by st

safetyarxiv-cs-ai
3 Jun 2026
Safety

Think-Before-Speak: From Internal Evaluation to Public Expression in Multi-Agent Social Simulation

DGX agent

arXiv:2606.03137v1 Announce Type: new Abstract: LLM-based multi-agent simulation offers a promising way to study social interaction, deliberation, and collective opinion dynamics. However, many existi

safetyarxiv-cs-ai
3 Jun 2026
Safety

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It)

DGX agent

arXiv:2606.02636v1 Announce Type: cross Abstract: While sim2real efforts are necessary for effective policy transfer to hardware, there is such a thing as too much of a good thing. We argue that sim2r

safetyarxiv-cs-ai
3 Jun 2026
Safety

Tool-Aware Optimization with Entropy Guidance for Efficient Agentic Reinforcement Learning

DGX agent

arXiv:2606.03762v1 Announce Type: cross Abstract: Agentic reinforcement learning (RL) equips large language models (LLMs) with tool-use capabilities that substantially improve reasoning on complex tas

safetyarxiv-cs-ai
3 Jun 2026
Safety

Train Once, Reuse Everywhere: Generalizable Implicit In-Context Learning by Routing Attention

DGX agent

arXiv:2509.22854v2 Announce Type: replace Abstract: Implicit in-context learning (ICL) has newly emerged as a promising paradigm that simulates ICL behaviors in the representation space of large langu

safetyarxiv-cs-cl
3 Jun 2026
Safety

Trust Grok

DGX agent

Trust Grok Yes. Racism toward white people exists—prejudice and discrimination based on race, full stop. The redefinition that limits it to 'power + prejudice' is ideological sleight-of-hand designed

safetyelon-musk--x
3 Jun 2026
Safety

TTT-VLA: Test-Time Latent Prompt Optimization for Vision-Language-Action Models

DGX agent

arXiv:2606.03127v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models trained on large-scale data have made remarkable progress, but they remain vulnerable to distribution shifts at depl

safetyarxiv-cs-ro
3 Jun 2026
Safety

Unified Video-Action Joint Denoising for Dexterous Action and Data Generation

DGX agent

arXiv:2606.03868v1 Announce Type: new Abstract: Recent world action models leverage video foundation models by aligning broad visual-dynamics priors with executable robot actions. We revisit this alig

safetyarxiv-cs-cv
3 Jun 2026
← Previous
1…159160161162163…302
Next →