AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,606
  • Agents7,269
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,099
  • Local Ai4,731
  • Model Releases22,585
  • Research19,194
  • Safety12,820
  • Syntheses17
  • Tools1,668
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,606Total entries
1Added by human
84,605Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,494 results
Safety

Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity Constraints

DGX agent

arXiv:2601.23221v2 Announce Type: replace Abstract: As acquiring reliable ground-truth labels is usually costly, or infeasible, crowdsourcing and aggregation of noisy human annotations is the typical

safetyarxiv-cs-lg
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

OrderDP: A Theoretically Guaranteed Lossless Dynamic Data Pruning Framework

DGX agent

arXiv:2606.08574v1 Announce Type: cross Abstract: Data pruning (DP), as an oft-stated strategy to alleviate heavy training burdens, reduces the volume of training samples according to a well-defined p

safetyarxiv-cs-cv
9 Jun 2026
Safety

Outage Detection in Self-Healing Smart Grids Using Reinforcement Learning with Spectral Graph Neural Networks

DGX agent

arXiv:2606.07583v1 Announce Type: cross Abstract: Self-healing smart grids can quickly adjust their network configuration during outages to minimize power disruptions. During an outage, several action

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

DGX agent

arXiv:2606.08543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, wher

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

DGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

safetyarxiv-cs-ai
9 Jun 2026
Safety

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization

DGX agent

arXiv:2605.06582v2 Announce Type: replace Abstract: Many operations on sensory data -- comparison, memory, retrieval, and reasoning -- are naturally expressed over discrete symbolic structures. In lan

safetyarxiv-cs-lg
9 Jun 2026
Safety

PairWise Image Finder: An Open-source Tool for Finding Visually Aligned Street-Level Image Pairs for Urban Perception Studies

DGX agent

arXiv:2606.08795v1 Announce Type: new Abstract: Change detection and scene recognition techniques have been widely applied to Street View Imagery (SVI) to understand changes in scenes across the years

safetyarxiv-cs-cv
9 Jun 2026
Safety

Path Planning Using Deep Deterministic Policy Gradient: A Reinforcement Learning Approach

DGX agent

arXiv:2606.07855v1 Announce Type: new Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge because the problem is nonlinear and nonconvex even in sim

safetyarxiv-cs-ro
9 Jun 2026
Safety

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

DGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

safetyarxiv-cs-ai
9 Jun 2026
Safety

Payoff scaling shapes cooperation in LLM agents across languages

DGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

safetyarxiv-cs-ai
9 Jun 2026
Safety

PBSD: Privileged Bayesian Self-Distillation for Long-Horizon Credit Assignment

DGX agent

arXiv:2606.09348v1 Announce Type: new Abstract: Long-horizon agentic tasks pose a fundamental credit assignment challenge for outcome-base reinforcement learning: trajectory-level rewards verify final

safetyarxiv-cs-lg
9 Jun 2026
Safety

Petri Net Modeling and Deadlock-Free Scheduling of Attachable Heterogeneous AGV Systems

DGX agent

arXiv:2508.00724v2 Announce Type: replace-cross Abstract: The increasing demand for flexible automation has accelerated the adoption of heterogeneous automated guided vehicles (AGVs). This work invest

safetyarxiv-cs-ro
9 Jun 2026
Safety

Physically Consistent Null Space Alignment for Detection of Low-Magnitude False Data Injection Attacks

DGX agent

arXiv:2606.08473v1 Announce Type: new Abstract: False data injection attacks (FDIAs) introducing small measurement perturbations can still cause large deviations in power system state estimation when

safetyarxiv-cs-lg
9 Jun 2026
Safety

Polaffini: A feature-based approach for robust affine and polyaffine image registration

DGX agent

arXiv:2602.17337v2 Announce Type: replace Abstract: In this work we present Polaffini, a robust and versatile framework for anatomically grounded registration. Medical image registration is dominated

safetyarxiv-cs-cv
9 Jun 2026
Safety

PriFT: Prior-Support Guided Supervised Fine-Tuning

DGX agent

arXiv:2606.09396v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is an efficient approach for downstream task adaptation and often serves as the initialization stage for reinforcement le

safetyarxiv-cs-lg
9 Jun 2026
Safety

Prisma-World: Camera-Controllable Multi-Agent Video World Model

DGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

safetyarxiv-cs-cv
9 Jun 2026
Safety

PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models

DGX agent

arXiv:2606.08926v1 Announce Type: new Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring

safetyarxiv-cs-lg
9 Jun 2026
Safety

Property-Informed Diffusion-Based Text-to-Microstructure Generation

DGX agent

arXiv:2606.08150v1 Announce Type: new Abstract: Designing 3D metamaterial microstructures that meet the intended functions remains a major challenge, as it typically requires domain expertise, iterati

safetyarxiv-cs-cv
9 Jun 2026
Safety

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization

DGX agent

arXiv:2606.09711v1 Announce Type: new Abstract: Reward hacking is usually studied after it becomes visible, once a model earns high proxy reward while failing the intended task. We instead study what

safetyarxiv-cs-ai
9 Jun 2026
Safety

PRPO: Perception-Reinforced Policy Optimization via Token-Level Dynamic Advantage Reshaping

DGX agent

arXiv:2606.08708v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective paradigm for improving the reasoning capability of Large Vision-Language M

safetyarxiv-cs-cv
9 Jun 2026
Safety

PTDL:Multi-Terrain Fall Recovery via Phase-Terrain Decoupled Learning

DGX agent

arXiv:2606.08922v1 Announce Type: new Abstract: Humanoid robots can fall on slopes, gravel, and uneven ground in unstructured environments. We target integrated fall recovery and locomotion: rebuildin

safetyarxiv-cs-ro
9 Jun 2026
Safety

Q-VGM: Q-Guided Value-Gradient Matching for Flow-Matching VLA Policies

DGX agent

arXiv:2606.08015v1 Announce Type: new Abstract: We propose Q-Guided Value-Gradient Matching (Q-VGM), an off-policy reinforcement learning (RL) method that tackles a long-standing challenge in fine-tun

safetyarxiv-cs-ro
9 Jun 2026
Safety

QnRL: Quantum-Native Reinforcement Learning

DGX agent

arXiv:2606.08276v1 Announce Type: cross Abstract: Quantum reinforcement learning (QRL) is a promising approach to learn effective decision strategies across several applications with stochastic enviro

safetyarxiv-cs-lg
9 Jun 2026
Safety

Quantifying Uncertainty in Space Debris Capture with Active Tether-Net Systems Caused by Noisy Observations

DGX agent

arXiv:2606.07580v1 Announce Type: cross Abstract: As Low Earth Orbit has grown more crowded with space debris, the need for reliable and efficient debris removal solutions becomes more urgent. An acti

safetyarxiv-cs-lg
9 Jun 2026
Safety

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

DGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

safetyarxiv-cs-ai
9 Jun 2026
Safety

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

DGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

safetyarxiv-cs-ai
9 Jun 2026
Safety

ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

DGX agent

arXiv:2606.09381v1 Announce Type: new Abstract: Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the

safetyarxiv-cs-ro
9 Jun 2026
Safety

Region-Wise Correspondence Prediction between Manga Line Art Images

DGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reinforcement Learning for Flow-Matching Policies with Density Transport

DGX agent

arXiv:2606.08602v1 Announce Type: cross Abstract: We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is t

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reinforcement learning in linear embedding space unlocks generalizable control across soft robot configurations

DGX agent

arXiv:2606.08104v1 Announce Type: new Abstract: Soft-bodied organisms such as octopuses and elephant trunks exhibit remarkable morphological adaptability, dynamically reconfiguring body shape and stif

safetyarxiv-cs-ro
9 Jun 2026
Safety

Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning

DGX agent

arXiv:2606.08436v1 Announce Type: new Abstract: The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language querie

safetyarxiv-cs-cv
9 Jun 2026
Safety

Repair Before Veto, When Repair Is Hidden: Quantum-Accessible Features for Repair-Augmented Constraint Learning

DGX agent

arXiv:2606.08020v1 Announce Type: cross Abstract: Hard-constraint decision systems usually veto infeasible candidates. This is too rigid when the system can act: if a known affordable repair would mak

safetyarxiv-cs-ai
9 Jun 2026
Safety

Rethinking the Divergence Regularization in LLM RL

DGX agent

arXiv:2606.09821v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a key component of post-training large language models (LLMs). In practice, LLM RL is often off-policy because of

safetyarxiv-cs-lg
9 Jun 2026
Safety

Revisiting Articulated Parts Perception in Robot Manipulation

DGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Safety

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have beh…

DGX agent

ridiculous oversimplification, from a prominent OpenAI employee no less. does this capture your impression of how the two companies have behaved? The OAI / Anthropic values difference is deeply misund

safetygary-marcus--x
9 Jun 2026
Safety

RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments

DGX agent

arXiv:2511.07317v2 Announce Type: replace-cross Abstract: We introduce Reinforcement Learning (RL) with Adaptive Verifiable Environments (RLVE), an approach using verifiable environments that procedur

safetyarxiv-cs-lg
9 Jun 2026
Safety

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

DGX agent

arXiv:2602.11934v2 Announce Type: replace Abstract: Robot manipulation often fails in the final millimeters: a policy may recognize the right object yet miss the pose offsets, boundaries, or pre-conta

safetyarxiv-cs-ro
9 Jun 2026
Safety

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

DGX agent

arXiv:2606.07602v1 Announce Type: cross Abstract: LLM-based LEGO assembly generation requires both semantic grounding and physical feasibility. We identify a data-induced failure mode, PhysHack, in wh

safetyarxiv-cs-ai
9 Jun 2026
Safety

SAW: Stage-Aware Dynamic Weighting for Multi-Objective Reinforcement Learning in Large Language Models

DGX agent

arXiv:2606.07705v1 Announce Type: cross Abstract: Although multi-objective reinforcement learning (MORL) is central to aligning large language models with complex human preferences, the prevailing pra

safetyarxiv-cs-ai
9 Jun 2026
Safety

SecureClaw: Clawing Back Control of LLM Agents

DGX agent

arXiv:2606.09549v1 Announce Type: cross Abstract: Tool-using large language model (LLM) agents face two distinct security failures: unauthorized external actions and exposure of sensitive plaintext in

safetyarxiv-cs-ai
9 Jun 2026
Safety

See More, Match Better: Multi-Source Feature Fusion for Two-View Correspondence Learning

DGX agent

arXiv:2606.09262v1 Announce Type: new Abstract: Two-view correspondence learning aims to distinguish true correspondences (inliers) from false ones (outliers) in image pairs by leveraging their underl

safetyarxiv-cs-cv
9 Jun 2026
Safety

SEF-CLGC at SemEval-2026 Task 11: Logical Notation Impact on Language Model Performance

DGX agent

arXiv:2606.09157v1 Announce Type: cross Abstract: This paper revisits our pipeline called Syllogistic Evaluation Framework-Common Logic Grammar Construction (SEF-CLGC). We combine formal logical notat

safetyarxiv-cs-ai
9 Jun 2026
Safety

Self-Evolving Scientific Agent Discovers Generalizable Physically-Reasoned Fluid Control

DGX agent

arXiv:2606.08405v1 Announce Type: new Abstract: While data-intensive deep reinforcement learning can optimize complex control policies, scientific discovery in physical systems fundamentally requires

safetyarxiv-cs-ai
9 Jun 2026
Safety

Self-Supervised Learning with a Multi-Task Latent Space Objective

DGX agent

arXiv:2602.05845v2 Announce Type: replace Abstract: We propose a multi-task formulation of self-predictive Siamese SSL in which each spatial transformation defines a distinct latent-space alignment ta

safetyarxiv-cs-cv
9 Jun 2026
Safety

SemDINO: A DINOv3-Driven Network for Cross-Temporal Semantic Alignment in Change Detection

DGX agent

arXiv:2606.09772v1 Announce Type: new Abstract: Semantic change detection (SCD) aims to simultaneously locate land-cover changes and identify semantic categories before and after transition. However,

safetyarxiv-cs-cv
9 Jun 2026
Safety

Sequential statistical inference for Large Language Models: Representation, validity, and monitoring

DGX agent

arXiv:2606.07624v1 Announce Type: new Abstract: This discussion argues that sequential statistical inference can naturally contribute to LLM trustworthiness. In deployment, LLM systems are queried rep

safetyarxiv-cs-lg
9 Jun 2026
Safety

SG-OPD: Sign-Gated On-Policy Distillation via Sign-Consistency Gating and Phased Teacher Sampling

DGX agent

arXiv:2606.09304v1 Announce Type: cross Abstract: On-policy distillation (OPD) trains a student on its own trajectories with dense per-token supervision from a stronger teacher, and often outperforms

safetyarxiv-cs-lg
9 Jun 2026
← Previous
1…146147148149150…302
Next →