AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,588
  • Agents7,266
  • Applications5,200
  • Concepts5
  • Hardware1,756
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,577
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,588Total entries
1Added by human
84,587Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

DGX agent

arXiv:2601.12263v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval

safetyarxiv-cs-ai
9 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without Training

DGX agent

arXiv:2601.03256v2 Announce Type: replace Abstract: We present Muses, the first training-free method for fantastic 3D creature generation in a feed-forward paradigm. Previous methods, which rely on pa

safetyarxiv-cs-cv
9 Jun 2026
Safety

NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI Analysis

DGX agent

arXiv:2606.07635v1 Announce Type: cross Abstract: Multimodal neuroimaging fusion of functional MRI (fMRI) and diffusion tensor imaging (DTI) provides complementary information for cognitive impairment

safetyarxiv-cs-ai
9 Jun 2026
Safety

Neutrality Bites: Gender Representation in AI-Generated Animal Stories

DGX agent

arXiv:2606.07969v1 Announce Type: cross Abstract: Gender bias in AI-generated stories is a well-documented problem. While much attention has been paid to reducing or mitigating this bias, it is not al

safetyarxiv-cs-ai
9 Jun 2026
Safety

No Modality Left Behind: Adapting to Missing Modalities via Knowledge Distillation for Brain Tumor Segmentation

DGX agent

arXiv:2509.15017v2 Announce Type: replace Abstract: Accurate brain tumor segmentation is essential for preoperative evaluation and personalized treatment. Multi-modal MRI is widely used due to its abi

safetyarxiv-cs-cv
9 Jun 2026
Safety

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

DGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

safetyarxiv-cs-ai
9 Jun 2026
Safety

Normality Calibration in Semi-supervised Graph Anomaly Detection

DGX agent

arXiv:2510.02014v3 Announce Type: replace Abstract: Graph anomaly detection (GAD) has attracted growing interest for its crucial ability to uncover irregular patterns in broad applications. Semi-super

safetyarxiv-cs-lg
9 Jun 2026
Safety

OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation

DGX agent

arXiv:2606.08548v1 Announce Type: new Abstract: Recent progress in robot manipulation has been largely driven by learning from large-scale demonstrations. For humanoid robot loco-manipulation tasks, h

safetyarxiv-cs-ro
9 Jun 2026
Safety

omega-EVA: Envision, Verify, and Act with Latent Interactive World Models

DGX agent

arXiv:2606.09457v1 Announce Type: new Abstract: Embodied policies typically map current observations directly to actions, leaving candidate-action consequences implicit. World models provide predictiv

safetyarxiv-cs-ro
9 Jun 2026
Safety

Operationalising the Superficial Alignment Hypothesis via Task Complexity

DGX agent

arXiv:2602.15829v2 Announce Type: replace Abstract: The superficial alignment hypothesis (SAH) posits that large language models learn most of their knowledge during pre-training, and that post-traini

safetyarxiv-cs-lg
9 Jun 2026
Safety

Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity Constraints

DGX agent

arXiv:2601.23221v2 Announce Type: replace Abstract: As acquiring reliable ground-truth labels is usually costly, or infeasible, crowdsourcing and aggregation of noisy human annotations is the typical

safetyarxiv-cs-lg
9 Jun 2026
Safety

OrderDP: A Theoretically Guaranteed Lossless Dynamic Data Pruning Framework

DGX agent

arXiv:2606.08574v1 Announce Type: cross Abstract: Data pruning (DP), as an oft-stated strategy to alleviate heavy training burdens, reduces the volume of training samples according to a well-defined p

safetyarxiv-cs-cv
9 Jun 2026
Safety

Outage Detection in Self-Healing Smart Grids Using Reinforcement Learning with Spectral Graph Neural Networks

DGX agent

arXiv:2606.07583v1 Announce Type: cross Abstract: Self-healing smart grids can quickly adjust their network configuration during outages to minimize power disruptions. During an outage, several action

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

DGX agent

arXiv:2606.08543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, wher

safetyarxiv-cs-ai
9 Jun 2026
Safety

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

DGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

safetyarxiv-cs-ai
9 Jun 2026
Safety

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization

DGX agent

arXiv:2605.06582v2 Announce Type: replace Abstract: Many operations on sensory data -- comparison, memory, retrieval, and reasoning -- are naturally expressed over discrete symbolic structures. In lan

safetyarxiv-cs-lg
9 Jun 2026
Safety

PairWise Image Finder: An Open-source Tool for Finding Visually Aligned Street-Level Image Pairs for Urban Perception Studies

DGX agent

arXiv:2606.08795v1 Announce Type: new Abstract: Change detection and scene recognition techniques have been widely applied to Street View Imagery (SVI) to understand changes in scenes across the years

safetyarxiv-cs-cv
9 Jun 2026
Safety

Path Planning Using Deep Deterministic Policy Gradient: A Reinforcement Learning Approach

DGX agent

arXiv:2606.07855v1 Announce Type: new Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge because the problem is nonlinear and nonconvex even in sim

safetyarxiv-cs-ro
9 Jun 2026
Safety

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

DGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

safetyarxiv-cs-ai
9 Jun 2026
Safety

Payoff scaling shapes cooperation in LLM agents across languages

DGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

safetyarxiv-cs-ai
9 Jun 2026
Safety

PBSD: Privileged Bayesian Self-Distillation for Long-Horizon Credit Assignment

DGX agent

arXiv:2606.09348v1 Announce Type: new Abstract: Long-horizon agentic tasks pose a fundamental credit assignment challenge for outcome-base reinforcement learning: trajectory-level rewards verify final

safetyarxiv-cs-lg
9 Jun 2026
Safety

Petri Net Modeling and Deadlock-Free Scheduling of Attachable Heterogeneous AGV Systems

DGX agent

arXiv:2508.00724v2 Announce Type: replace-cross Abstract: The increasing demand for flexible automation has accelerated the adoption of heterogeneous automated guided vehicles (AGVs). This work invest

safetyarxiv-cs-ro
9 Jun 2026
Safety

Physically Consistent Null Space Alignment for Detection of Low-Magnitude False Data Injection Attacks

DGX agent

arXiv:2606.08473v1 Announce Type: new Abstract: False data injection attacks (FDIAs) introducing small measurement perturbations can still cause large deviations in power system state estimation when

safetyarxiv-cs-lg
9 Jun 2026
Safety

Polaffini: A feature-based approach for robust affine and polyaffine image registration

DGX agent

arXiv:2602.17337v2 Announce Type: replace Abstract: In this work we present Polaffini, a robust and versatile framework for anatomically grounded registration. Medical image registration is dominated

safetyarxiv-cs-cv
9 Jun 2026
Safety

PriFT: Prior-Support Guided Supervised Fine-Tuning

DGX agent

arXiv:2606.09396v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is an efficient approach for downstream task adaptation and often serves as the initialization stage for reinforcement le

safetyarxiv-cs-lg
9 Jun 2026
Safety

Prisma-World: Camera-Controllable Multi-Agent Video World Model

DGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

safetyarxiv-cs-cv
9 Jun 2026
Safety

PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models

DGX agent

arXiv:2606.08926v1 Announce Type: new Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring

safetyarxiv-cs-lg
9 Jun 2026
Safety

Property-Informed Diffusion-Based Text-to-Microstructure Generation

DGX agent

arXiv:2606.08150v1 Announce Type: new Abstract: Designing 3D metamaterial microstructures that meet the intended functions remains a major challenge, as it typically requires domain expertise, iterati

safetyarxiv-cs-cv
9 Jun 2026
Safety

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization

DGX agent

arXiv:2606.09711v1 Announce Type: new Abstract: Reward hacking is usually studied after it becomes visible, once a model earns high proxy reward while failing the intended task. We instead study what

safetyarxiv-cs-ai
9 Jun 2026
Safety

PRPO: Perception-Reinforced Policy Optimization via Token-Level Dynamic Advantage Reshaping

DGX agent

arXiv:2606.08708v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective paradigm for improving the reasoning capability of Large Vision-Language M

safetyarxiv-cs-cv
9 Jun 2026
Safety

PTDL:Multi-Terrain Fall Recovery via Phase-Terrain Decoupled Learning

DGX agent

arXiv:2606.08922v1 Announce Type: new Abstract: Humanoid robots can fall on slopes, gravel, and uneven ground in unstructured environments. We target integrated fall recovery and locomotion: rebuildin

safetyarxiv-cs-ro
9 Jun 2026
Safety

Q-VGM: Q-Guided Value-Gradient Matching for Flow-Matching VLA Policies

DGX agent

arXiv:2606.08015v1 Announce Type: new Abstract: We propose Q-Guided Value-Gradient Matching (Q-VGM), an off-policy reinforcement learning (RL) method that tackles a long-standing challenge in fine-tun

safetyarxiv-cs-ro
9 Jun 2026
Safety

QnRL: Quantum-Native Reinforcement Learning

DGX agent

arXiv:2606.08276v1 Announce Type: cross Abstract: Quantum reinforcement learning (QRL) is a promising approach to learn effective decision strategies across several applications with stochastic enviro

safetyarxiv-cs-lg
9 Jun 2026
Safety

Quantifying Uncertainty in Space Debris Capture with Active Tether-Net Systems Caused by Noisy Observations

DGX agent

arXiv:2606.07580v1 Announce Type: cross Abstract: As Low Earth Orbit has grown more crowded with space debris, the need for reliable and efficient debris removal solutions becomes more urgent. An acti

safetyarxiv-cs-lg
9 Jun 2026
Safety

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

DGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

safetyarxiv-cs-ai
9 Jun 2026
Safety

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

DGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

safetyarxiv-cs-ai
9 Jun 2026
Safety

ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

DGX agent

arXiv:2606.09381v1 Announce Type: new Abstract: Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the

safetyarxiv-cs-ro
9 Jun 2026
Safety

Region-Wise Correspondence Prediction between Manga Line Art Images

DGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reinforcement Learning for Flow-Matching Policies with Density Transport

DGX agent

arXiv:2606.08602v1 Announce Type: cross Abstract: We present an online reinforcement learning (RL) algorithm for fine-tuning flow-matching policies in continuous-control problems. Our key insight is t

safetyarxiv-cs-ai
9 Jun 2026
Safety

Reinforcement learning in linear embedding space unlocks generalizable control across soft robot configurations

DGX agent

arXiv:2606.08104v1 Announce Type: new Abstract: Soft-bodied organisms such as octopuses and elephant trunks exhibit remarkable morphological adaptability, dynamically reconfiguring body shape and stif

safetyarxiv-cs-ro
9 Jun 2026
Safety

Reinforcing Temporal Answer Grounding in Instructional Video via Candidate-Aware Causal Reasoning

DGX agent

arXiv:2606.08436v1 Announce Type: new Abstract: The task of temporal answer grounding in instructional video (TAGV), which aims to locate precise video segments that respond to natural language querie

safetyarxiv-cs-cv
9 Jun 2026
Safety

Repair Before Veto, When Repair Is Hidden: Quantum-Accessible Features for Repair-Augmented Constraint Learning

DGX agent

arXiv:2606.08020v1 Announce Type: cross Abstract: Hard-constraint decision systems usually veto infeasible candidates. This is too rigid when the system can act: if a known affordable repair would mak

safetyarxiv-cs-ai
9 Jun 2026
Safety

Rethinking the Divergence Regularization in LLM RL

DGX agent

arXiv:2606.09821v1 Announce Type: new Abstract: Reinforcement learning (RL) has become a key component of post-training large language models (LLMs). In practice, LLM RL is often off-policy because of

safetyarxiv-cs-lg
9 Jun 2026
Safety

Revisiting Articulated Parts Perception in Robot Manipulation

DGX agent

arXiv:2606.08103v1 Announce Type: cross Abstract: We are surrounded by various objects with movable, articulated parts, e.g., box, handle, door. An accurate and generalizable perception of articulated

safetyarxiv-cs-cv
9 Jun 2026
Safety

Reward Shaping for (Inference-Time) Alignment: A Stackelberg Game Perspective

DGX agent

arXiv:2602.02572v2 Announce Type: replace-cross Abstract: Existing alignment methods directly use the reward model learned from user preference data to optimize an LLM policy, subject to KL regulariza

safetyarxiv-cs-ai
9 Jun 2026
Safety

RLVE: Scaling Up Reinforcement Learning for Language Models with Adaptive Verifiable Environments

DGX agent

arXiv:2511.07317v2 Announce Type: replace-cross Abstract: We introduce Reinforcement Learning (RL) with Adaptive Verifiable Environments (RLVE), an approach using verifiable environments that procedur

safetyarxiv-cs-lg
9 Jun 2026
Safety

Robot-DIFT: Correspondence-Sensitive Diffusion Features for Contact-Rich Robot Manipulation

DGX agent

arXiv:2602.11934v2 Announce Type: replace Abstract: Robot manipulation often fails in the final millimeters: a policy may recognize the right object yet miss the pose offsets, boundaries, or pre-conta

safetyarxiv-cs-ro
9 Jun 2026
Safety

Sample-Efficient Post-Training for LEGO Spatial-Physics Reasoning

DGX agent

arXiv:2606.07602v1 Announce Type: cross Abstract: LLM-based LEGO assembly generation requires both semantic grounding and physical feasibility. We identify a data-induced failure mode, PhysHack, in wh

safetyarxiv-cs-ai
9 Jun 2026
← Previous
1…129130131132133…260
Next →