AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

SCOPE-RL: Optimizing Reasoning Paths Before and After Success

DGX agent

arXiv:2607.11506v2 Announce Type: replace Abstract: Reinforcement learning with verifiable rewards (RLVR) optimizes LLMs using sparse verifiable final-answer rewards. This sparse anchor reliably verif

safetyarxiv-cs-lg
16 Jul 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

ShortOPD: Recovering Pruned LLMs with Short-to-Long On-Policy Distillation

DGX agent

arXiv:2607.13124v1 Announce Type: cross Abstract: Structured pruning is a hardware-friendly way to compress LLMs, but it is mostly validated on multiple-choice recognition tasks, while the same compre

safetyarxiv-cs-ai
16 Jul 2026
Safety

SIVA-RL: Sensitivity-Invariance Visual Alignment for Multimodal Reinforcement Learning

DGX agent

arXiv:2607.13931v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) drives multimodal reasoning, but answer-level correctness does not guarantee that a vision-languag

safetyarxiv-cs-cv
16 Jul 2026
Safety

SteinGate: Tail-Sensitive Safe Reinforcement Learning via Stein Discrepancy

DGX agent

arXiv:2607.13175v1 Announce Type: cross Abstract: Safe reinforcement learning typically enforces safety by bounding expected cumulative costs, a criterion that often fails to detect rare but catastrop

safetyarxiv-cs-ai
16 Jul 2026
Safety

STITCHER: Constrained Trajectory Planning in Complex Environments with Real-Time Motion Primitive Search

DGX agent

arXiv:2510.14893v4 Announce Type: replace Abstract: Autonomous high-speed navigation through large, complex environments requires real-time generation of agile trajectories that are dynamically feasib

safetyarxiv-cs-ro
16 Jul 2026
Safety

Structured Reinforcement Learning for Bayesian Persuasion : Application to Intelligent Interactive Driving

DGX agent

arXiv:2607.13576v1 Announce Type: new Abstract: Interactive driving, wherein an intelligent lead vehicle equipped with real-time traffic data coordinates route choices of connected vehicles, offers a

safetyarxiv-cs-lg
16 Jul 2026
Safety

Temperature Scaling Is Not Enough: Calibration Gaps Under Human Label Distributions

DGX agent

arXiv:2607.13423v1 Announce Type: new Abstract: Temperature scaling is the dominant post-hoc calibration method in modern deep learning. Its theoretical justification rests on an assumption that is ra

safetyarxiv-cs-lg
16 Jul 2026
Safety

The Cafe in Amsterdam: When the Incumbent Becomes the Oracle

DGX agent

arXiv:2607.13393v1 Announce Type: cross Abstract: A field can reformulate its computations freely exactly where its demand is stated independently of any incumbent implementation, and finds itself una

safetyarxiv-cs-ai
16 Jul 2026
Safety

The SIGReg Objective as Variational Free Energy: A Theoretical Active-Inference Account of JEPA World Models

DGX agent

arXiv:2607.13612v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) are the dominant design for latent world models, yet they are usually justified by empirical performa

safetyarxiv-cs-ai
16 Jul 2026
Safety

ThinkBLOX: 3D Indoor Scene Generation with Progressive Reasoning

DGX agent

arXiv:2607.13539v1 Announce Type: new Abstract: While traditional graphics methods often synthesize 3D indoor scenes autoregressively or hierarchically, recent vision-language model (VLM)-based genera

safetyarxiv-cs-cv
16 Jul 2026
Safety

Track and Caption Any Motion: Open-Vocabulary Spatiotemporal Captioning via Trajectory-Conditioned Generation

DGX agent

arXiv:2512.10607v2 Announce Type: replace Abstract: We present TCAM (Track and Caption Any Motion), a generative framework that watches a video and with no text query and no region prompt decides what

safetyarxiv-cs-cv
16 Jul 2026
Safety

Uncertainty-Aware Sequential Decision Rules for Event-Triggered LLM Invocation in Streaming Systems

DGX agent

arXiv:2607.13048v1 Announce Type: cross Abstract: Streaming inference pipelines increasingly pair lightweight fast models with Large Language Models (LLMs) that provide rich semantic understanding at

safetyarxiv-cs-ai
16 Jul 2026
Safety

Vision-Based Obstacle Separation for Strawberry Harvesting in Clusters Using Hierarchical Reinforcement Learning

DGX agent

arXiv:2607.13799v1 Announce Type: new Abstract: Selective harvesting in clustered strawberry environments is challenging because ripe fruits are often occluded by surrounding unripe fruits, making dir

safetyarxiv-cs-ro
16 Jul 2026
Safety

Where Should RL Post-Training Compute Go? Model Size, Search, Learning, and Feedback

DGX agent

arXiv:2607.13389v1 Announce Type: new Abstract: Reinforcement Learning (RL) post-training is increasingly used to adapt foundation models for reasoning, planning, and feedback-driven robot-learning pi

safetyarxiv-cs-lg
16 Jul 2026
Safety

A Neurosymbolic Approach to Natural Language Formalization and Verification

DGX agent

arXiv:2511.09008v2 Announce Type: replace-cross Abstract: Large Language Models perform well at natural language interpretation and reasoning, but their lack of formal correctness guarantees limits th

safetyarxiv-cs-ai
15 Jul 2026
Safety

AAAI-26 Dual Submissions: Novel Challenges

DGX agent

arXiv:2607.11918v1 Announce Type: cross Abstract: Dual submissions, in which identical or substantially similar papers are simultaneously submitted to one or more archival venues, without cross-citati

safetyarxiv-cs-ai
15 Jul 2026
Safety

Accuracy and Normalized Accuracy under Length Bias: Analysis, Guidelines, and a Bayesian Alternative

DGX agent

arXiv:2607.12767v1 Announce Type: new Abstract: Multiple-choice benchmarks that rank candidate completions by conditional log-probability suffer from a length bias: because log-probabilities sum over

safetyarxiv-cs-ai
15 Jul 2026
Safety

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to …

DGX agent

AI agents are already being used to improve the capabilities of our next-generation models. We believe with GPT-Red that we have started to unlock a similar flywheel for safety, where today's models c

safetyopenai--x
15 Jul 2026
Safety

Anomalous Frame Detection Using VLM-Based Description Comparison for Extracting Expert-Specific Actions and Contextual Decision-Making Scenes with Intra-Video Self-Similarity

DGX agent

arXiv:2607.11957v1 Announce Type: new Abstract: Maintenance of critical infrastructures, such as railways and power plants, is essential for ensuring operational safety and reliability. However, the d

safetyarxiv-cs-cv
15 Jul 2026
Safety

Attention Misses Visual Risk: Risk-Adaptive Steering for Multimodal Safety Alignment

DGX agent

arXiv:2510.13698v4 Announce Type: replace Abstract: Even modern AI models often remain vulnerable to multimodal queries in which harmful intent is embedded in images. A widely used approach for safety

safetyarxiv-cs-cv
15 Jul 2026
Safety

Auditable Context-Aware HFMD Forecasting with Structured LLM Agents

DGX agent

arXiv:2511.23276v2 Announce Type: replace Abstract: Effective HFMD surveillance requires forecasts capturing both time-series patterns and contextual drivers such as school calendars, weather, and pol

safetyarxiv-cs-lg
15 Jul 2026
Safety

Beyond Coordinate Gauge: An Audited Protocol for Detecting Donor-Specific Functional Fingerprints after Neural Collapse

DGX agent

arXiv:2607.11967v1 Announce Type: cross Abstract: Independently trained neural networks have no shared neuron-index reference frame, so comparing them requires accounting for coordinate freedom. Neura

safetyarxiv-cs-ai
15 Jul 2026
Safety

Beyond Parallel Tracking: Interactive Multi-Feature Fusion Drives Semantic Reconstruction from Non-invasive Brain Recordings

DGX agent

arXiv:2607.12071v1 Announce Type: new Abstract: Continuous semantic reconstruction from non-invasive neural recordings remains limited by the representational mismatch between semantic feature spaces

safetyarxiv-cs-cl
15 Jul 2026
Safety

Calculating Mutual Information between a Reward Maximizer and its Environment

DGX agent

arXiv:2602.12963v2 Announce Type: replace Abstract: An important question in the field of AI is the extent to which successful behaviour requires an internal representation of the world. In this work,

safetyarxiv-cs-ai
15 Jul 2026
Safety

Calibrated Selective Prediction Using Deep Ensembles for ROI-Based Thyroid Nodule Ultrasound Classification Under Dataset Shift: A Retrospective Evaluation

DGX agent

arXiv:2607.12075v1 Announce Type: cross Abstract: Background: Deep learning models can classify thyroid nodules on ultrasound, but reliable clinical decision support also requires calibrated probabili

safetyarxiv-cs-ai
15 Jul 2026
Safety

Calibration-First Reward-Component Auditing for Reinforcement Learning Control in Smart Greenhouses

DGX agent

arXiv:2607.11959v1 Announce Type: new Abstract: Greenhouse reinforcement learning can test climate-control ideas at a speed and scale that is difficult to achieve with crop experiments alone. For smar

safetyarxiv-cs-ai
15 Jul 2026
Safety

Can Induced Emotion Bias LLM Behaviors in Sequential Decision Making?

DGX agent

arXiv:2607.12631v1 Announce Type: cross Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents in high-stakes domains, understanding contextual factors that may modul

safetyarxiv-cs-ai
15 Jul 2026
Safety

Can LLMs Write Reliable Rubrics? A Meta-Evaluation for Experiment Reproduction

DGX agent

arXiv:2607.12835v1 Announce Type: new Abstract: Rubric-based evaluation is a promising approach for assessing open-ended outputs from LLM-based research agents, particularly in paper reproduction, whe

safetyarxiv-cs-cl
15 Jul 2026
Safety

CASHEW: Stabilizing Multimodal Reasoning via Iterative Trajectory Aggregation

DGX agent

arXiv:2601.08010v3 Announce Type: replace Abstract: Vision-language models achieve strong performance across a wide range of multimodal understanding and reasoning tasks, yet their multi-step reasonin

safetyarxiv-cs-cv
15 Jul 2026
Safety

CGRL: Concept-Guided Pruning and Representation Learning for Whole-Slide Image Classification

DGX agent

arXiv:2607.12556v1 Announce Type: new Abstract: Weakly supervised whole-slide image (WSI) classification is widely used in computational pathology because slide-level labels are easier to obtain than

safetyarxiv-cs-cv
15 Jul 2026
Safety

ChunkFlow: Towards Continuity-Consistent Chunked Policy Learning

DGX agent

arXiv:2607.12992v1 Announce Type: new Abstract: Vision-language action (VLA) models increasingly adopt chunked action heads to satisfy real-time constraints; however, this introduces boundary jitter:

safetyarxiv-cs-ro
15 Jul 2026
Safety

CityBehavEx: A Scalable and Empirically Validated LLM-Assisted Urban Simulation Platform

DGX agent

arXiv:2607.12086v1 Announce Type: new Abstract: Recent LLM-based multi-agent urban simulators can generate semantically rich city routines, but they remain costly to scale and are often weakly validat

safetyarxiv-cs-cl
15 Jul 2026
Safety

Code-MUE: Measuring Code LLMs' Uncertainty through Execution-based Semantic Interaction Graphs

DGX agent

arXiv:2607.12273v1 Announce Type: cross Abstract: As Code Large Language Models (LLMs) become central to modern software engineering, their inherent stochasticity poses significant real-world risks, w

safetyarxiv-cs-ai
15 Jul 2026
Safety

Compos3D: Interactive Part-Based Composition for Creative Control in Generative 3D Models

DGX agent

arXiv:2607.12193v1 Announce Type: cross Abstract: While generative AI has unlocked new opportunities for 3D content creation, current workflows often rely on multiple regenerations, which provides lim

safetyarxiv-cs-cv
15 Jul 2026
Safety

Continuous Policy and Value Iteration for Stochastic Control Problems and Its Convergence

DGX agent

arXiv:2506.08121v2 Announce Type: replace-cross Abstract: We introduce a continuous policy-value iteration algorithm where the approximations of the value function of a stochastic control problem and

safetyarxiv-cs-lg
15 Jul 2026
Safety

Data Safety: Synthetic Data Quality Analysis Using CIFAKE Dataset

DGX agent

arXiv:2607.12165v1 Announce Type: new Abstract: Recently, the societal implementation of high-performance image classification models has expanded rapidly. While these models require vast amounts of t

safetyarxiv-cs-cv
15 Jul 2026
Safety

DenseReward: Dense Reward Learning via Failure Synthesis for Robotic Manipulation

DGX agent

arXiv:2607.13033v1 Announce Type: new Abstract: Reinforcement learning holds great promise for improving robot policies beyond the limits of imitation learning. However, its practical adoption remains

safetyarxiv-cs-ro
15 Jul 2026
Safety

Deployable Human Preference Alignment in Robotics: Learning Representative Rewards from Diverse Human Preferences

DGX agent

arXiv:2607.12466v1 Announce Type: new Abstract: Aligning robot policies with human preferences is essential for deployment to diverse end users. In per-user alignment approach, preference feedback is

safetyarxiv-cs-ro
15 Jul 2026
Safety

Directional Constraints for Efficient Exploration in Safe Reinforcement Learning

DGX agent

arXiv:2607.12784v1 Announce Type: cross Abstract: Reinforcement Learning has revolutionized the landscape of robotic research, allowing robust learning of complex robotic skills in simulation. However

safetyarxiv-cs-lg
15 Jul 2026
Safety

Expert Knowledge-driven Reinforcement Learning for Autonomous Racing via Trajectory Guidance and Dynamics Constraints

DGX agent

arXiv:2603.05842v2 Announce Type: replace Abstract: Reinforcement learning has demonstrated significant potential in the field of autonomous driving. However, it suffers from defects such as training

safetyarxiv-cs-ro
15 Jul 2026
Safety

ExToken: Structured Exploration for Efficient Vision-Language-Action Reinforcement Fine-tuning

DGX agent

arXiv:2607.12931v1 Announce Type: new Abstract: Reinforcement Learning (RL) has demonstrated significant potential for improving Vision-Language-Action (VLA) models on complex manipulation tasks. Howe

safetyarxiv-cs-ro
15 Jul 2026
Safety

ExtraGS: Enhancing Endoscopic View Extrapolation via Diffusion-Guided 3D Gaussian Splatting

DGX agent

arXiv:2607.12785v1 Announce Type: new Abstract: Robot-assisted minimally invasive surgery (MIS) critically depends on reliable endoscopic perception for navigation and safety. However, conventional en

safetyarxiv-cs-cv
15 Jul 2026
Safety

FlowWAM: Optical Flow as a Unified Action Representation for World Action Models

DGX agent

arXiv:2607.13017v1 Announce Type: cross Abstract: World Action Models (WAMs) are able to leverage pretrained video generators for both world modeling and action prediction. However, directly leveragin

safetyarxiv-cs-cv
15 Jul 2026
Safety

From Geometric Recovery to Causal Validation: A Reproducible Audit of Sparse Autoencoder Features, from Superposition Geometry to Causal Inertness

DGX agent

arXiv:2607.12166v1 Announce Type: new Abstract: Sparse autoencoders (SAEs) are the standard for decomposing superposed neural representations into interpretable features, and evaluation relies predomi

safetyarxiv-cs-lg
15 Jul 2026
Safety

From Sentiment to Actionable Insights: Public Sentiment Analysis of Advanced Air Mobility

DGX agent

arXiv:2606.20751v2 Announce Type: replace Abstract: Advanced Air Mobility (AAM) is an emerging low-altitude transportation system whose successful deployment depends on both technological progress and

safetyarxiv-cs-cl
15 Jul 2026
Safety

Function-Aware Fill-in-the-Middle as Mid-Training for Coding Agent Foundation Models

DGX agent

arXiv:2607.12463v1 Announce Type: new Abstract: Coding agents must integrate external tool returns into ongoing reasoning - a capability that standard left-to-right pretraining on code exposes only in

safetyarxiv-cs-ai
15 Jul 2026
Safety

G-SHARE: A Guideline-Based Structured Reasoning Framework for Human-Factor Event Diagnosis

DGX agent

arXiv:2607.11892v1 Announce Type: cross Abstract: Human-factor event diagnosis is essential for learning from operational events in nuclear power plants, yet its quality depends strongly on expert int

safetyarxiv-cs-ai
15 Jul 2026
Safety

GaitSpan: Growing Humanoid Locomotion from Walking to Running

DGX agent

arXiv:2607.12114v1 Announce Type: cross Abstract: A humanoid that can walk should not relearn locomotion from scratch to jog or run. Yet current approaches often obtain gait diversity by prescribing g

safetyarxiv-cs-ai
15 Jul 2026
← Previous
1…4344454647…265
Next →