AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Generalizable VLA Finetuning via Representation Anchoring and Language-Action Alignment

DGX agent

arXiv:2607.13429v1 Announce Type: cross Abstract: Finetuning a pretrained vision-language model (VLM) on robot demonstrations via behavior cloning (BC) has become the standard recipe for vision-langua

safetyarxiv-cs-cv
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

GNN-DIP: Neural Corridor Selection for Decomposition-Based Motion Planning

DGX agent

arXiv:2603.12361v2 Announce Type: replace Abstract: Motion planning through narrow passages remains a core challenge: sampling-based planners rarely place samples inside these narrow but critical regi

safetyarxiv-cs-ro
16 Jul 2026
Safety

Groc-PO: Grounded Context Preference Optimization for Truthful Multimodal LLMs

DGX agent

arXiv:2607.13712v1 Announce Type: cross Abstract: Despite the rapid progress of Multimodal Large Language Models (MLLMs), they still suffer from untruthfulness issues, such as visual hallucinations, c

safetyarxiv-cs-ai
16 Jul 2026
Safety

HELP: Human-Efficient Large-Scale Robot Post-Training with Rollout Segmentation

DGX agent

arXiv:2607.09776v2 Announce Type: replace-cross Abstract: When adapting Vision Language Action (VLA) models to downstream tasks, multiple rounds of post-training are often required to progressively ad

safetyarxiv-cs-ai
16 Jul 2026
Safety

Human4K: A Large-Scale 4K Multi-View Mocap Dataset for Whole-Body 3D Human Reconstruction

DGX agent

arXiv:2607.13646v1 Announce Type: cross Abstract: Recent advances in 3D human reconstruction have improved overall performance, yet current models still fail in the most challenging real-world scenari

safetyarxiv-cs-ai
16 Jul 2026
Safety

Introducing Human-Centeredness in AI-Assisted Lexicography

DGX agent

arXiv:2607.11808v2 Announce Type: replace-cross Abstract: This paper proposes a human-centered artificial intelligence (HCAI) framework for AI-assisted lexicography. While generative AI offers signifi

safetyarxiv-cs-ai
16 Jul 2026
Safety

Inverse-LLaVA: Rethinking Multimodal Alignment via Text-to-Vision Mapping

DGX agent

arXiv:2508.12466v2 Announce Type: replace-cross Abstract: Traditional multimodal learning approaches rely on alignment pre-training to bridge vision and language modalities, typically by projecting vi

safetyarxiv-cs-ai
16 Jul 2026
Safety

Joint On-and-Off Policy Learning for Vision-and-Language Navigation

DGX agent

arXiv:2607.13461v1 Announce Type: new Abstract: Vision-and-Language Navigation (VLN) necessitates an embodied agent to navigate in the physical world by adhering to natural language instructions. Rece

safetyarxiv-cs-ro
16 Jul 2026
Safety

LAPO: Leave-One-Turn Attribution for Self-Generated Process Rewards in Multi-Turn Search Reasoning

DGX agent

arXiv:2607.13501v1 Announce Type: new Abstract: Reinforcement learning for multi-turn search reasoning typically relies on terminal outcome rewards, which cannot distinguish useful, redundant, and har

safetyarxiv-cs-ai
16 Jul 2026
Safety

Learning Red Agent Policy from Observations for Neurosymbolic Autonomous Cyber Agents

DGX agent

arXiv:2606.18223v2 Announce Type: replace-cross Abstract: With sophisticated cyber-attacks becoming increasingly prevalent, modern networks require intelligent autonomous cyber-defense agents trained

safetyarxiv-cs-ai
16 Jul 2026
Safety

Learning to Learn-at-Test-Time: Language Agents with Learnable Adaptation Policies

DGX agent

arXiv:2604.00830v3 Announce Type: replace-cross Abstract: Test-Time Learning (TTL) enables language agents to iteratively refine their performance through repeated interactions with the environment at

safetyarxiv-cs-ai
16 Jul 2026
Safety

Marker-free deformable registration and fusion for augmented reality-guided positive margin localization during tumor resection surgery

DGX agent

arXiv:2607.13343v1 Announce Type: new Abstract: Positive margins in head and neck oncologic surgery require mapping specimen-side pathology findings to the patient resection bed. This is challenging b

safetyarxiv-cs-cv
16 Jul 2026
Safety

MASPRM: Multi-Agent System Process Reward Model

DGX agent

arXiv:2510.24803v3 Announce Type: replace-cross Abstract: Inference-time search over multi-agent systems (MAS) wastes compute when it cannot identify which agent's intermediate message advanced progre

safetyarxiv-cs-ai
16 Jul 2026
Safety

Memory as a Controlled Process: Learned Adaptive Memory Management for LLM Agents

DGX agent

arXiv:2607.13591v1 Announce Type: cross Abstract: Large Language Model (LLM) agents increasingly rely on external memory systems to accumulate experience across tasks. Yet nearly all existing approach

safetyarxiv-cs-ai
16 Jul 2026
Safety

Mind the Gap: Action Rebinding Attacks against Android GUI Agents

DGX agent

arXiv:2601.12349v3 Announce Type: replace-cross Abstract: Large multimodal model powered GUI agents are emerging as high-privilege operators on mobile platforms, entrusted to perceive screen content a

safetyarxiv-cs-ai
16 Jul 2026
Safety

Music-to-Dance Generation via Atomic Movements

DGX agent

arXiv:2607.13978v1 Announce Type: cross Abstract: Music-driven dance generation aims to produce human motion that is both rhythmically synchronized and semantically consistent with music. While recent

safetyarxiv-cs-ai
16 Jul 2026
Safety

Non-Expansive Two-Time-Scale Stochastic Approximation: A Fixed-Schedule One-Quarter Barrier and Bias-Corrected Acceleration

DGX agent

arXiv:2607.13414v1 Announce Type: cross Abstract: Non-expansive two-time-scale stochastic approximation is governed by a slow stochastic Krasnoselskii--Mann fixed-point iteration rather than by contra

safetyarxiv-cs-lg
16 Jul 2026
Safety

NSNQuant: A Double Normalization Approach for Calibration-Free Low-Bit Vector Quantization of KV Cache

DGX agent

arXiv:2505.18231v3 Announce Type: replace-cross Abstract: Large Language Model (LLM) inference is typically memory-intensive, especially when processing large batch sizes and long sequences, due to th

safetyarxiv-cs-ai
16 Jul 2026
Safety

On the Sublinear Regret of Continuous K-Max Bandits

DGX agent

arXiv:2502.13467v2 Announce Type: replace Abstract: The K-Max combinatorial multi-armed bandit problem arises in applications such as recommendation and distributed decision making, where the reward i

safetyarxiv-cs-lg
16 Jul 2026
Safety

Optimal and Efficient Contextual Combinatorial Semi-bandits with General Function Approximation

DGX agent

arXiv:2607.13686v1 Announce Type: new Abstract: We study the contextual combinatorial semi-bandit (CCSB) problem with general reward function approximation. At each round, the learner observes a conte

safetyarxiv-cs-lg
16 Jul 2026
Safety

OrDA: Orthogonal Disentanglement of Access Habits Framework for Homepage Marketing Block Recommendations

DGX agent

arXiv:2607.13420v1 Announce Type: new Abstract: Clicks on homepage marketing blocks are driven by a dual-mechanism of content interest and access habits. However, habitual clicks often create Pseudo-P

safetyarxiv-cs-lg
16 Jul 2026
Safety

PhysClaw-0: A Symbiotic Agentic System for Robot Autonomy via Language Corrections

DGX agent

arXiv:2607.14047v1 Announce Type: new Abstract: Autonomous data collection governs the volume and quality of real-world trajectories for manipulation policy learning. Existing pipelines reduce human e

safetyarxiv-cs-ro
16 Jul 2026
Safety

Pretraining in Actor-Critic Reinforcement Learning for Locomotion

DGX agent

arXiv:2510.12363v4 Announce Type: replace-cross Abstract: The pretraining-finetuning paradigm has facilitated numerous transformative advancements in artificial intelligence research in recent years.

safetyarxiv-cs-lg
16 Jul 2026
Safety

Price of Fairness in Bandits: A Tight Minimax Characterization

DGX agent

arXiv:2607.13402v1 Announce Type: cross Abstract: In bandit problems, standard regret-minimizing algorithms treat exploration as an amortized cost, which can expose early participants to unfair ex-ant

safetyarxiv-cs-ai
16 Jul 2026
Safety

Privacy Preserving Recommender Systems Balancing Personalization with Privacy

DGX agent

arXiv:2607.13328v1 Announce Type: cross Abstract: Personalized recommendation systems are central to modern e-commerce and retail platforms, but they typically rely on centralized storage of detailed

safetyarxiv-cs-ai
16 Jul 2026
Safety

PUe: Biased Positive-Unlabeled Learning Enhancement by Causal Inference

DGX agent

arXiv:2607.13428v1 Announce Type: new Abstract: Positive-Unlabeled (PU) learning aims to achieve high-accuracy binary classification with limited labeled positive examples and numerous unlabeled ones.

safetyarxiv-cs-lg
16 Jul 2026
Safety

RADAR: Closed-Loop Robotic Data Generation via Semantic Planning and Autonomous Causal Environment Reset

DGX agent

arXiv:2603.11811v2 Announce Type: replace-cross Abstract: The acquisition of large-scale physical interaction data, a critical prerequisite for modern robot learning, is severely bottlenecked by the p

safetyarxiv-cs-ai
16 Jul 2026
Safety

Reverse to Advance: Teleoperation-Cost Effective Hard Policy Learning from Reversed Easy Tasks

DGX agent

arXiv:2607.13455v1 Announce Type: new Abstract: High-quality teleoperation datasets are costly to collect, particularly for hard tasks. We observe that many tasks exhibit directional asymmetry: comple

safetyarxiv-cs-ro
16 Jul 2026
Safety

SARFA: Segment Anything with Radiomic Feature Alignment

DGX agent

arXiv:2607.13323v1 Announce Type: new Abstract: The Segment Anything Model (SAM) has demonstrated strong generalizability across a variety of segmentation tasks. However, SAM often struggles in situat

safetyarxiv-cs-cv
16 Jul 2026
Safety

SCOPE-RL: Optimizing Reasoning Paths Before and After Success

DGX agent

arXiv:2607.11506v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) optimizes LLMs using sparse verifiable final-answer rewards. This sparse anchor reliably verif

safetyarxiv-cs-lg
16 Jul 2026
Safety

ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation

DGX agent

arXiv:2607.13124v1 Announce Type: cross Abstract: Structured pruning is a hardware-friendly way to compress LLMs, but it is mostly validated on multiple-choice recognition tasks, while the same compre

safetyarxiv-cs-ai
16 Jul 2026
Safety

SIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement Learning

DGX agent

arXiv:2607.13931v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) drives multimodal reasoning, but answer-level correctness does not guarantee that a vision-languag

safetyarxiv-cs-cv
16 Jul 2026
Safety

Structured Reinforcement Learning for Bayesian Persuasion : Application to Intelligent Interactive Driving

DGX agent

arXiv:2607.13576v1 Announce Type: new Abstract: Interactive driving, wherein an intelligent lead vehicle equipped with real-time traffic data coordinates route choices of connected vehicles, offers a

safetyarxiv-cs-lg
16 Jul 2026
Safety

The Cafe in Amsterdam: When the Incumbent Becomes the Oracle

DGX agent

arXiv:2607.13393v1 Announce Type: cross Abstract: A field can reformulate its computations freely exactly where its demand is stated independently of any incumbent implementation, and finds itself una

safetyarxiv-cs-ai
16 Jul 2026
Safety

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

DGX agent

arXiv:2607.13612v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performa

safetyarxiv-cs-ai
16 Jul 2026
Safety

ThinkBLOX: 3D Indoor Scene Generation with Progressive Reasoning

DGX agent

arXiv:2607.13539v1 Announce Type: new Abstract: While traditional graphics methods often synthesize 3D indoor scenes autoregressively or hierarchically, recent vision-language model (VLM)-based genera

safetyarxiv-cs-cv
16 Jul 2026
Safety

Track and Caption Any Motion: Open-Vocabulary Spatiotemporal Captioning via Trajectory-Conditioned Generation

DGX agent

arXiv:2512.10607v2 Announce Type: replace Abstract: We present TCAM (Track and Caption Any Motion), a generative framework that watches a video and with no text query and no region prompt decides what

safetyarxiv-cs-cv
16 Jul 2026
Safety

Vision-Based Obstacle Separation for Strawberry Harvesting in Clusters Using Hierarchical Reinforcement Learning

DGX agent

arXiv:2607.13799v1 Announce Type: new Abstract: Selective harvesting in clustered strawberry environments is challenging because ripe fruits are often occluded by surrounding unripe fruits, making dir

safetyarxiv-cs-ro
16 Jul 2026
Safety

Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback

DGX agent

arXiv:2607.13389v1 Announce Type: new Abstract: Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pi

safetyarxiv-cs-lg
16 Jul 2026
Safety

AAAI-26 Dual Submissions: Novel Challenges

DGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

safetyarxiv-cs-ai
15 Jul 2026
Safety

Accuracy and Normalized Accuracy under Length Bias: Analysis, Guidelines, and a Bayesian Alternative

DGX agent

arXiv:2607.12767v1 Announce Type: new Abstract: Multiple-choice benchmarks that rank candidate completions by conditional log-probability suffer from a length bias: because log-probabilities sum over

safetyarxiv-cs-ai
15 Jul 2026
Safety

Auditable Context-Aware HFMD Forecasting with Structured LLM Agents

DGX agent

arXiv:2511.23276v2 Announce Type: replace Abstract: Effective HFMD surveillance requires forecasts capturing both time-series patterns and contextual drivers such as school calendars, weather, and pol

safetyarxiv-cs-lg
15 Jul 2026
Safety

Beyond Coordinate Gauge: An Audited Protocol for Detecting Donor-Specific Functional Fingerprints after Neural Collapse

DGX agent

arXiv:2607.11967v1 Announce Type: cross Abstract: Independently trained neural networks have no shared neuron-index reference frame, so comparing them requires accounting for coordinate freedom. Neura

safetyarxiv-cs-ai
15 Jul 2026
Safety

Beyond Parallel Tracking: Interactive Multi-Feature Fusion Drives Semantic Reconstruction from Non-invasive Brain Recordings

DGX agent

arXiv:2607.12071v1 Announce Type: new Abstract: Continuous semantic reconstruction from non-invasive neural recordings remains limited by the representational mismatch between semantic feature spaces

safetyarxiv-cs-cl
15 Jul 2026
Safety

Calculating Mutual Information between a Reward Maximizer and its Environment

DGX agent

arXiv:2602.12963v2 Announce Type: replace Abstract: An important question in the field of AI is the extent to which successful behaviour requires an internal representation of the world. In this work,

safetyarxiv-cs-ai
15 Jul 2026
Safety

Calibrated Selective Prediction Using Deep Ensembles for ROI-Based Thyroid Nodule Ultrasound Classification Under Dataset Shift: A Retrospective Evaluation

DGX agent

arXiv:2607.12075v1 Announce Type: cross Abstract: Background: Deep learning models can classify thyroid nodules on ultrasound, but reliable clinical decision support also requires calibrated probabili

safetyarxiv-cs-ai
15 Jul 2026
Safety

Calibration-First Reward-Component Auditing for Reinforcement Learning Control in Smart Greenhouses

DGX agent

arXiv:2607.11959v1 Announce Type: new Abstract: Greenhouse reinforcement learning can test climate-control ideas at a speed and scale that is difficult to achieve with crop experiments alone. For smar

safetyarxiv-cs-ai
15 Jul 2026
Safety

Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction

DGX agent

arXiv:2607.12835v1 Announce Type: new Abstract: Rubric-based evaluation is a promising approach for assessing open-ended outputs from LLM-based research agents, particularly in paper reproduction, whe

safetyarxiv-cs-cl
15 Jul 2026
← Previous
1…8990919293…260
Next →