AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Letting Tutor Personas Speak Up for LLMs: Learning Steering Vectors from Dialogue via Preference Optimization

DGX agent

arXiv:2602.07639v2 Announce Type: replace Abstract: With the emergence of large language models (LLMs) as a powerful class of generative artificial intelligence (AI), their use in tutoring has become

safetyarxiv-cs-cl
3 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

DGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

safetyarxiv-cs-ai
3 Jun 2026
Safety

Libra: Efficient Resource Management for Agentic RL Post-Training

DGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

safetyarxiv-cs-ai
3 Jun 2026
Safety

LoCAtion: Long-time Collaborative Attention Framework for High Dynamic Range Video Reconstruction

DGX agent

arXiv:2603.14377v2 Announce Type: replace Abstract: Prevailing High Dynamic Range (HDR) video reconstruction methods are fundamentally trapped in a fragile alignment-and-fusion paradigm. While explici

safetyarxiv-cs-cv
3 Jun 2026
Safety

Margin Play: A Multi-Agent System For Public Policy Analysis In The Brazilian Equatorial Margin

DGX agent

arXiv:2606.02614v1 Announce Type: cross Abstract: The Brazilian Equatorial Margin (BEM) is Brazil's next offshore oil frontier, with operations expected to begin in 2026 in the Foz do Amazonas basin.

safetyarxiv-cs-ai
3 Jun 2026
Safety

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

DGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

safetyarxiv-cs-ai
3 Jun 2026
Safety

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

DGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

safetyarxiv-cs-ai
3 Jun 2026
Safety

Mitigating False Credit Propagation: Probabilistic Graphical Reward Aggregation for Rubric-Based Reinforcement Learning

DGX agent

arXiv:2606.03361v1 Announce Type: new Abstract: Rubric-based rewards are increasingly used for open-ended language model post-training, but criterion-level scores are often aggregated as independent u

safetyarxiv-cs-lg
3 Jun 2026
Safety

Multi-component Causal Tracing in Large Language Models

DGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

safetyarxiv-cs-cl
3 Jun 2026
Safety

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

DGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

safetyarxiv-cs-cl
3 Jun 2026
Safety

Neural Networks Provably Learn Spectral Representations for Group Composition

DGX agent

arXiv:2606.02993v1 Announce Type: new Abstract: Understanding how structured internal structure emerges during neural network training is central to the study of deep learning. We investigate this phe

safetyarxiv-cs-lg
3 Jun 2026
Safety

NVIDIA OmniDreams: Real-Time Generative World Model for Closed-Loop Autonomous Vehicle Simulation

DGX agent

arXiv:2606.03159v1 Announce Type: cross Abstract: As autonomous vehicle capabilities advance, the safe evaluation of driving policies in long-tail scenarios remains a critical bottleneck. In closed-lo

safetyarxiv-cs-ai
3 Jun 2026
Safety

OMP: One-step Meanflow Policy with Directional Alignment

DGX agent

arXiv:2512.19347v3 Announce Type: replace Abstract: Robot manipulation has increasingly adopted data-driven generative policy frameworks, yet the field faces a persistent trade-off: diffusion models s

safetyarxiv-cs-ro
3 Jun 2026
Safety

PAND: Prompt-Aware Neighborhood Distillation for Lightweight Fine-Grained Visual Classification

DGX agent

arXiv:2602.07768v3 Announce Type: replace-cross Abstract: Distilling knowledge from large Vision-Language Models (VLMs) into lightweight networks is crucial yet challenging in Fine-Grained Visual Clas

safetyarxiv-cs-ai
3 Jun 2026
Local Ai

Patcher: Post-Hoc Patching of Backdoored Large Language Models

DGX agent

arXiv:2606.02995v1 Announce Type: cross Abstract: Large language models remain vulnerable to jailbreak backdoor attacks, where adversaries poison safety alignment data to embed hidden triggers that by

local-aiarxiv-cs-ai
3 Jun 2026
Safety

Pathway-Structured Privileged Distillation for Deployable Computational Pathology

DGX agent

arXiv:2606.02877v1 Announce Type: new Abstract: Integrating transcriptomics and histopathology can improve cancer risk modelling, yet practical use is constrained by the limited availability of RNA pr

safetyarxiv-cs-cv
3 Jun 2026
Safety

Pextsuperscript{2}-DPO: Grounding Hallucination in Perceptual Processing via Calibration Direct Preference Optimization

DGX agent

arXiv:2606.03376v1 Announce Type: cross Abstract: Hallucination has recently garnered significant research attention in Large Vision-Language Models (LVLMs). Direct Preference Optimization (DPO) aims

safetyarxiv-cs-ai
3 Jun 2026
Safety

PHAF-Personalized Hand Avatars in a Flash

DGX agent

arXiv:2606.03420v1 Announce Type: new Abstract: We present PHAF-Personalized Hand Avatars in a Flash, a personalized photo-realistic hand avatar which provides high quality multi-view renders from jus

safetyarxiv-cs-cv
3 Jun 2026
Safety

PHASE: Physiology-Aware Hyperspectral Reconstruction via Object-to-Human Domain Adaptation

DGX agent

arXiv:2511.13020v2 Announce Type: replace-cross Abstract: Although hyperspectral imaging offers unparalleled non-invasive physiological insight, its bulky hardware, slow acquisition, and regulatory bu

safetyarxiv-cs-ai
3 Jun 2026
Safety

Physical Plausibility Reasoning via HCM-GRPO: Empowering Compact Model for Superior Performance

DGX agent

arXiv:2511.10055v2 Announce Type: replace Abstract: The performance of image generation has been significantly improved in recent years. However, the study of image screening is rare, and its performa

safetyarxiv-cs-cv
3 Jun 2026
Safety

Physics-Guided Policy Optimization with Self-Distillation

DGX agent

arXiv:2606.03620v1 Announce Type: cross Abstract: Self-distilled policy optimization (SDPO) has become a popular paradigm for LLM post-training, where a model learns from its own predictions condition

safetyarxiv-cs-ai
3 Jun 2026
Safety

Planning with Uncertainty: Symmetries, Policy Inference, and Solution Compression

DGX agent

arXiv:2403.19883v2 Announce Type: replace Abstract: Fully-observable non-deterministic (FOND) planning is at the core of artificial intelligence planning with uncertainty. It models uncertainty throug

safetyarxiv-cs-ai
3 Jun 2026
Safety

Post-Hoc Robustness for Model-Based Reinforcement Learning

DGX agent

arXiv:2606.03521v1 Announce Type: cross Abstract: To improve the real-world applicability of reinforcement learning (RL), the field of adversarially robust RL studies how to train agents under adversa

safetyarxiv-cs-ai
3 Jun 2026
Safety

Preference-Calibrated Human-in-the-Loop Reinforcement Learning for Robotic Manipulation

DGX agent

arXiv:2606.03949v1 Announce Type: new Abstract: Human-in-the-loop reinforcement learning (HIL-RL) improves sample efficiency in real-robot manipulation through online human intervention. However, succ

safetyarxiv-cs-ro
3 Jun 2026
Safety

PrimeSVT: An Automated Memory-aware Pruning Framework with Prioritized Compression Policy for Spiking Vision Transformers

DGX agent

arXiv:2606.03428v1 Announce Type: cross Abstract: The large sizes of Spiking Vision Transformers (SViTs) still hinder their embedded implementation, highlighting the need for model compression. State-

safetyarxiv-cs-ai
3 Jun 2026
Safety

PsychoPass: Geometric Profiling of Multi-Turn Adversarial LLM Conversations

DGX agent

arXiv:2606.03136v1 Announce Type: cross Abstract: Multi-turn jailbreak attacks on large language models (LLMs) reveal a mismatch in current guardrails: they operate on individual turns, while attacks

safetyarxiv-cs-cl
3 Jun 2026
Safety

Quantifying Faithful Confidence Expression in Large Reasoning Models

DGX agent

arXiv:2606.03969v1 Announce Type: cross Abstract: Reliable uncertainty communication is critical to the trustworthiness of LLMs, yet faithful calibration (FC)--the alignment between models' intrinsic

safetyarxiv-cs-ai
3 Jun 2026
Safety

QUBRIC: Co-Designing Queries and Rubrics for RL Beyond Verifiable Rewards

DGX agent

arXiv:2606.03968v1 Announce Type: cross Abstract: Rubric-based RL is a promising route for extending reinforcement learning beyond verifiable rewards, yet existing methods optimize rubrics while treat

safetyarxiv-cs-ai
3 Jun 2026
Safety

Right Makes Might: Aligning Verified Hidden States Empowers RL Reasoning

DGX agent

arXiv:2606.03234v1 Announce Type: new Abstract: Reinforcement Learning from Verifiable Rewards (RLVR) has become the dominant approach for improving mathematical reasoning in large language models, ye

safetyarxiv-cs-lg
3 Jun 2026
Safety

RoboCade: Gamifying Robot Data Collection

DGX agent

arXiv:2512.21235v3 Announce Type: replace Abstract: Imitation learning from human demonstrations has become a dominant approach for training autonomous robot policies. However, collecting demonstratio

safetyarxiv-cs-ro
3 Jun 2026
Safety

See Less, Specify More: Visual Evidence Budgets for Generalizable VLAs

DGX agent

arXiv:2606.02735v1 Announce Type: cross Abstract: Generalization remains a central bottleneck for vision-language-action (VLA) models: under distractors, appearance shifts, and semantically similar ta

safetyarxiv-cs-ai
3 Jun 2026
Safety

Selective Token-Level Cryptographic Redaction for Privacy-Preserving Clinical Deployment of Large Language Models

DGX agent

arXiv:2606.03399v1 Announce Type: new Abstract: While large language models (LLMs) are increasingly used for clinical applications, many existing pipelines require sending raw sensitive health informa

safetyarxiv-cs-cl
3 Jun 2026
Safety

Self-Refining Agentic Reinforcement Learning for Vision-Conditioned UAV Navigation

DGX agent

arXiv:2606.03963v1 Announce Type: cross Abstract: Deep reinforcement learning has shown strong potential for enabling autonomous robots to learn complex navigational tasks. However, its practical use

safetyarxiv-cs-ai
3 Jun 2026
Safety

SimuScene: Simulation-Ready Compositional 3D Scene Reconstruction from a Single Image

DGX agent

arXiv:2606.03994v1 Announce Type: new Abstract: Reconstructing interactive, simulation-ready 3D scenes from a single image is a critical bottleneck for robotic manipulation. While recent single-image

safetyarxiv-cs-cv
3 Jun 2026
Safety

SkelHCC: A Hyperbolic CLIP-Driven Cache Adaptation Framework for Skeleton-based One-Shot Action Recognition

DGX agent

arXiv:2606.03610v1 Announce Type: new Abstract: Skeleton-based action recognition aims to understand human behaviors from body joint sequences and is especially challenging in the one-shot setting, wh

safetyarxiv-cs-cv
3 Jun 2026
Safety

Sparse-View Lung Nodule Volumetry from Digitally Reconstructed Radiographs via AReT: Anatomy-Regularized TensoRF

DGX agent

arXiv:2606.02639v1 Announce Type: cross Abstract: We identify and resolve a previously unreported failure mode in TensoRF when applied to X-ray attenuation fields: the default density shift of -10, or

safetyarxiv-cs-ai
3 Jun 2026
Safety

Spectral Asymptotics of Neural Network Loss Landscapes: An Exact Decomposition of the Curvature Exponent

DGX agent

arXiv:2606.02596v1 Announce Type: new Abstract: The curvature exponent alpha in h_k propto sigma_k^alpha -- governing how Hessian eigenvalues scale with gradient singular values -- varies systematical

safetyarxiv-cs-lg
3 Jun 2026
Safety

SplitAdapter: Load-Aware Humanoid Loco-Manipulation via Factorized Adaptation

DGX agent

arXiv:2606.03297v1 Announce Type: new Abstract: Humanoid loco-manipulation requires stable whole-body control under varying object masses and pickup/placement heights. This becomes particularly challe

safetyarxiv-cs-ro
3 Jun 2026
Safety

Strongly Polynomial Time Complexity of Policy Iteration for L_infty Robust MDPs

DGX agent

arXiv:2601.23229v2 Announce Type: replace Abstract: Markov decision processes (MDPs) are a fundamental model in sequential decision making. Robust MDPs (RMDPs) extend this framework by allowing uncert

safetyarxiv-cs-ai
3 Jun 2026
Safety

Taiji: Pareto Optimal Policy Optimization with Semantics-IDs Trade-off for Industrial LLM-Enhanced Recommendation

DGX agent

arXiv:2606.03866v1 Announce Type: cross Abstract: Scaling recommender systems via large language models (LLMs) has become a prominent trend in the industry. However, aligning the LLM's semantic space

safetyarxiv-cs-ai
3 Jun 2026
Safety

Temporal Action Selection for Action Chunking

DGX agent

arXiv:2511.04421v2 Announce Type: replace Abstract: Action chunking is a widely adopted approach in Learning from Demonstration (LfD). By modeling multi-step action chunks rather than single-step acti

safetyarxiv-cs-ro
3 Jun 2026
Safety

TGV-KV: Text-Grounded KV Eviction for Vision-Language Models

DGX agent

arXiv:2606.03075v1 Announce Type: new Abstract: Vision-Language Models (VLMs) inherit the auto-regressive generation paradigm and cache the keys and values (KV) of all previous tokens to accelerate in

safetyarxiv-cs-cv
3 Jun 2026
Safety

The Ghost Annotator: a Framework to Explore Human Label Variation in Content Moderation through Conformal Prediction

DGX agent

arXiv:2606.02911v1 Announce Type: new Abstract: Current research primarily focuses on model performance, while comparatively less attention has been devoted to uncertainty estimation, particularly in

safetyarxiv-cs-cl
3 Jun 2026
Safety

The Shadow Price of Reasoning: Economic Perspective on Optimal Budget Allocation for LLMs

DGX agent

arXiv:2606.03092v1 Announce Type: new Abstract: Inference-time scaling has emerged as a critical avenue for enhancing Large Language Models' performance, yet real-world deployment is constrained by st

safetyarxiv-cs-ai
3 Jun 2026
Safety

Think-Before-Speak: From Internal Evaluation to Public Expression in Multi-Agent Social Simulation

DGX agent

arXiv:2606.03137v1 Announce Type: new Abstract: LLM-based multi-agent simulation offers a promising way to study social interaction, deliberation, and collective opinion dynamics. However, many existi

safetyarxiv-cs-ai
3 Jun 2026
Safety

Too Much of a Good Thing: When sim2real Efforts Impede Policy Learning (And What to Do About It)

DGX agent

arXiv:2606.02636v1 Announce Type: cross Abstract: While sim2real efforts are necessary for effective policy transfer to hardware, there is such a thing as too much of a good thing. We argue that sim2r

safetyarxiv-cs-ai
3 Jun 2026
Safety

Tool-Aware Optimization with Entropy Guidance for Efficient Agentic Reinforcement Learning

DGX agent

arXiv:2606.03762v1 Announce Type: cross Abstract: Agentic reinforcement learning (RL) equips large language models (LLMs) with tool-use capabilities that substantially improve reasoning on complex tas

safetyarxiv-cs-ai
3 Jun 2026
Safety

Train Once, Reuse Everywhere: Generalizable Implicit In-Context Learning by Routing Attention

DGX agent

arXiv:2509.22854v2 Announce Type: replace Abstract: Implicit in-context learning (ICL) has newly emerged as a promising paradigm that simulates ICL behaviors in the representation space of large langu

safetyarxiv-cs-cl
3 Jun 2026
← Previous
1…139140141142143…260
Next →