AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

Position: Anthropomorphic Misalignment Research Needs Stronger Evidence

DGX agent

arXiv:2606.07612v1 Announce Type: cross Abstract: We argue that many Anthropomorphic Misalignment Research (AMR) studies need stronger evidence to ensure that they can provide a robust foundation for

safetyarxiv-cs-ai
9 Jun 2026
Safety

PriFT: Prior-Support Guided Supervised Fine-Tuning

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

arXiv:2606.09396v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is an efficient approach for downstream task adaptation and often serves as the initialization stage for reinforcement le

safetyarxiv-cs-lg
9 Jun 2026
Safety

Prisma-World: Camera-Controllable Multi-Agent Video World Model

DGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

safetyarxiv-cs-cv
9 Jun 2026
Safety

PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models

DGX agent

arXiv:2606.08926v1 Announce Type: new Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring

safetyarxiv-cs-lg
9 Jun 2026
Safety

Property-Informed Diffusion-Based Text-to-Microstructure Generation

DGX agent

arXiv:2606.08150v1 Announce Type: new Abstract: Designing 3D metamaterial microstructures that meet the intended functions remains a major challenge, as it typically requires domain expertise, iterati

safetyarxiv-cs-cv
9 Jun 2026
Safety

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization

DGX agent

arXiv:2606.09711v1 Announce Type: new Abstract: Reward hacking is usually studied after it becomes visible, once a model earns high proxy reward while failing the intended task. We instead study what

safetyarxiv-cs-ai
9 Jun 2026
Safety

PRPO: Perception-Reinforced Policy Optimization via Token-Level Dynamic Advantage Reshaping

DGX agent

arXiv:2606.08708v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective paradigm for improving the reasoning capability of Large Vision-Language M

safetyarxiv-cs-cv
9 Jun 2026
Safety

PTDL:Multi-Terrain Fall Recovery via Phase-Terrain Decoupled Learning

DGX agent

arXiv:2606.08922v1 Announce Type: new Abstract: Humanoid robots can fall on slopes, gravel, and uneven ground in unstructured environments. We target integrated fall recovery and locomotion: rebuildin

safetyarxiv-cs-ro
9 Jun 2026
Safety

Q-VGM: Q-Guided Value-Gradient Matching for Flow-Matching VLA Policies

DGX agent

arXiv:2606.08015v1 Announce Type: new Abstract: We propose Q-Guided Value-Gradient Matching (Q-VGM), an off-policy reinforcement learning (RL) method that tackles a long-standing challenge in fine-tun

safetyarxiv-cs-ro
9 Jun 2026
Safety

QDS-SNN: Energy-efficient Quantum Deeply-Supervised Spiking Neural Network Algorithm for Traffic Sign Recognition

DGX agent

arXiv:2606.07657v1 Announce Type: cross Abstract: Traffic sign recognition is crucial for intelligent transportation and autonomous driving, as it can improve driving efficiency and ensure road safety

safetyarxiv-cs-lg
9 Jun 2026
Safety

QnRL: Quantum-Native Reinforcement Learning

DGX agent

arXiv:2606.08276v1 Announce Type: cross Abstract: Quantum reinforcement learning (QRL) is a promising approach to learn effective decision strategies across several applications with stochastic enviro

safetyarxiv-cs-lg
9 Jun 2026
Safety

Quantifying Uncertainty in Space Debris Capture with Active Tether-Net Systems Caused by Noisy Observations

DGX agent

arXiv:2606.07580v1 Announce Type: cross Abstract: As Low Earth Orbit has grown more crowded with space debris, the need for reliable and efficient debris removal solutions becomes more urgent. An acti

safetyarxiv-cs-lg
9 Jun 2026
Safety

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

DGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

safetyarxiv-cs-ai
9 Jun 2026
Safety

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

DGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

safetyarxiv-cs-ai
9 Jun 2026
Safety

ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

DGX agent

arXiv:2606.09381v1 Announce Type: new Abstract: Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the

safetyarxiv-cs-ro
9 Jun 2026
Safety

Region-Wise Correspondence Prediction between Manga Line Art Images

DGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reinforcement Learning for Flow-Matching Policies with Density Transport

DGX agent

arXiv:2606.08602v1 Announce Type: cross Abstract: We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is t

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reinforcement learning in linear embedding space unlocks generalizable control across soft robot configurations

DGX agent

arXiv:2606.08104v1 Announce Type: new Abstract: Soft-bodied organisms such as octopuses and elephant trunks exhibit remarkable morphological adaptability, dynamically reconfiguring body shape and stif

safetyarxiv-cs-ro
9 Jun 2026
Safety

Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning

DGX agent

arXiv:2606.08436v1 Announce Type: new Abstract: The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language querie

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reliable to Expressive: A Curriculum for Rubric-Following Safety Judges

DGX agent

arXiv:2606.09165v1 Announce Type: new Abstract: Safety judges are increasingly deployed to evaluate model outputs against evolving criteria, yet recent meta-evaluation work shows they remain brittle u

safetyarxiv-cs-ai
9 Jun 2026
Safety

Repair Before Veto, When Repair Is Hidden: Quantum-Accessible Features for Repair-Augmented Constraint Learning

DGX agent

arXiv:2606.08020v1 Announce Type: cross Abstract: Hard-constraint decision systems usually veto infeasible candidates. This is too rigid when the system can act: if a known affordable repair would mak

safetyarxiv-cs-ai
9 Jun 2026
Safety

Rethinking the Divergence Regularization in LLM RL

DGX agent

arXiv:2606.09821v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a key component of post-training large language models (LLMs). In practice, LLM RL is often off-policy because of

safetyarxiv-cs-lg
9 Jun 2026
Safety

Revisiting Articulated Parts Perception in Robot Manipulation

DGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

safetyarxiv-cs-cv
9 Jun 2026
Safety

Revisiting the shutdown problem

DGX agent

arXiv:2606.08296v1 Announce Type: new Abstract: A key premise in leading arguments for existential risk from artificial intelligence is that malfunctioning artificial agents could not be easily shut d

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Safety

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have beh…

DGX agent

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have behaved? The OAI / Anthropic values difference is deeply misund

safetygary-marcus--x
9 Jun 2026
Safety

RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments

DGX agent

arXiv:2511.07317v2 Announce Type: replace-cross Abstract: We introduce Reinforcement Learning (RL) with Adaptive Verifiable Environments (RLVE), an approach using verifiable environments that procedur

safetyarxiv-cs-lg
9 Jun 2026
Safety

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

DGX agent

arXiv:2602.11934v2 Announce Type: replace Abstract: Robot manipulation often fails in the final millimeters: a policy may recognize the right object yet miss the pose offsets, boundaries, or pre-conta

safetyarxiv-cs-ro
9 Jun 2026
Safety

RPO-PDT: Demonstrating Role-Play-Based Knowledge Adaptation for Student Support Dialogue (Demonstration System)

DGX agent

arXiv:2606.09255v1 Announce Type: new Abstract: We present RPO-PDT: a retrieval-grounded, role-play-based dialogue system for adaptive student support in higher education. RPO-PDT is: (1) able to prov

safetyarxiv-cs-ro
9 Jun 2026
Safety

SAD-Flower: Flow Matching for Safe, Admissible, and Dynamically Consistent Planning

DGX agent

arXiv:2511.05355v3 Announce Type: replace Abstract: Flow matching (FM) has shown promising results in data-driven planning. However, it inherently lacks formal guarantees for ensuring state and action

safetyarxiv-cs-lg
9 Jun 2026
Safety

Safe, Fluent and Acceptable Motion Generation and Execution for Human--Robot Interaction in Manufacturing Environments

DGX agent

arXiv:2606.08741v1 Announce Type: new Abstract: Robots operating in human environments must not only ensure physical safety but also exhibit behaviors that are understandable, fluent, and acceptable t

safetyarxiv-cs-ro
9 Jun 2026
Safety

Safety is Contextual, LLM-Judges Are Not: Navigating the Rigid Priors of Evaluators

DGX agent

arXiv:2606.07874v1 Announce Type: new Abstract: LLMs-as-judges are the only way to evaluate safety at scale. Despite their importance, LLM-judges themselves are rarely evaluated beyond human agreement

safetyarxiv-cs-ai
9 Jun 2026
Safety

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

DGX agent

arXiv:2606.07602v1 Announce Type: cross Abstract: LLM-based LEGO assembly generation requires both semantic grounding and physical feasibility. We identify a data-induced failure mode, PhysHack, in wh

safetyarxiv-cs-ai
9 Jun 2026
Safety

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

DGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

safetyarxiv-cs-ai
9 Jun 2026
Safety

Scaling Neural Network Verification with Tensor Parallelism and Fully Sharded Data Parallelism

DGX agent

arXiv:2606.09377v1 Announce Type: cross Abstract: Formal neural network verification -- proving that a network satisfies safety properties for all inputs in a specified domain -- is bounded in practic

safetyarxiv-cs-ai
9 Jun 2026
Safety

SciTrace: Trajectory-Aware Safety Reasoning for Scientific Discovery Agents

DGX agent

arXiv:2606.08234v1 Announce Type: new Abstract: LLM-based scientific agents have shown strong capacity for autonomous research, yet their safety layers remain structurally divorced from core reasoning

safetyarxiv-cs-ai
9 Jun 2026
Safety

SecureClaw: Clawing Back Control of LLM Agents

DGX agent

arXiv:2606.09549v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents face two distinct security failures: unauthorized external actions and exposure of sensitive plaintext in

safetyarxiv-cs-ai
9 Jun 2026
Safety

See More, Match Better: Multi-Source Feature Fusion for Two-View Correspondence Learning

DGX agent

arXiv:2606.09262v1 Announce Type: new Abstract: Two-view correspondence learning aims to distinguish true correspondences (inliers) from false ones (outliers) in image pairs by leveraging their underl

safetyarxiv-cs-cv
9 Jun 2026
Safety

SEF-CLGC at SemEval-2026 Task 11: Logical Notation Impact on Language Model Performance

DGX agent

arXiv:2606.09157v1 Announce Type: cross Abstract: This paper revisits our pipeline called Syllogistic Evaluation Framework-Common Logic Grammar Construction (SEF-CLGC). We combine formal logical notat

safetyarxiv-cs-ai
9 Jun 2026
Safety

Self-Evolving Scientific Agent Discovers Generalizable Physically-Reasoned Fluid Control

DGX agent

arXiv:2606.08405v1 Announce Type: new Abstract: While data-intensive deep reinforcement learning can optimize complex control policies, scientific discovery in physical systems fundamentally requires

safetyarxiv-cs-ai
9 Jun 2026
Safety

Self-Supervised Learning with a Multi-Task Latent Space Objective

DGX agent

arXiv:2602.05845v2 Announce Type: replace Abstract: We propose a multi-task formulation of self-predictive Siamese SSL in which each spatial transformation defines a distinct latent-space alignment ta

safetyarxiv-cs-cv
9 Jun 2026
Safety

Semantic Quorum Assurance: Collective Certification for Non-Deterministic AI Infrastructure

DGX agent

arXiv:2606.08021v1 Announce Type: cross Abstract: As large language model (LLM) agents are integrated into autonomous cloud operations, distributed systems face a semantic reliability problem: propose

safetyarxiv-cs-ai
9 Jun 2026
Safety

SemDINO: A DINOv3-Driven Network for Cross-Temporal Semantic Alignment in Change Detection

DGX agent

arXiv:2606.09772v1 Announce Type: new Abstract: Semantic change detection (SCD) aims to simultaneously locate land-cover changes and identify semantic categories before and after transition. However,

safetyarxiv-cs-cv
9 Jun 2026
Safety

Sequential statistical inference for Large Language Models: Representation, validity, and monitoring

DGX agent

arXiv:2606.07624v1 Announce Type: new Abstract: This discussion argues that sequential statistical inference can naturally contribute to LLM trustworthiness. In deployment, LLM systems are queried rep

safetyarxiv-cs-lg
9 Jun 2026
Safety

SG-OPD: Sign-Gated On-Policy Distillation via Sign-Consistency Gating and Phased Teacher Sampling

DGX agent

arXiv:2606.09304v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories with dense per-token supervision from a stronger teacher, and often outperforms

safetyarxiv-cs-lg
9 Jun 2026
Safety

sGPO: Trading Inference FLOPs for Training Efficiency in RLVR

DGX agent

arXiv:2606.08854v1 Announce Type: cross Abstract: Standard Reinforcement Learning with Verifiable Rewards (RLVR) training allocates a fixed rollout budget to every query, without regard for what each

safetyarxiv-cs-ai
9 Jun 2026
Safety

SMI: Efficient Self-Supervised Learning via Mutual-Information-Inspired Dependency Optimization

DGX agent

arXiv:2606.08332v1 Announce Type: new Abstract: Self-supervised learning (SSL) has achieved remarkable representation learning performance, but many existing methods rely on large batch sizes, memory

safetyarxiv-cs-cv
9 Jun 2026
Safety

'So There's a Catch-22 Here': How Early Adopters Who Build Multi-Agent LLM Systems Conceptualize Transparency

DGX agent

arXiv:2606.08323v1 Announce Type: cross Abstract: Multi-agent large language model (LLM) systems are rapidly emerging, yet transparency, a cornerstone of responsible AI, remains under-defined in these

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…102103104105106…267
Next →