AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
7 Jul 2026

QSVideo: Query-Conditioned Semantic Temporal Retrieval for Video Understanding

SafetyDGX agent

arXiv:2607.04559v1 Announce Type: new Abstract: The performance of vision-language models (VLMs) in video understanding declines with increasing video duration, as video moments unrelated to the query

R^2PO: Decoupling Rollout and Inference Policies for LLM Reasoning

SafetyDGX agent

arXiv:2601.11960v3 Announce Type: replace-cross Abstract: Existing reinforcement learning methods for LLM reasoning implicitly assume that the policy generating training trajectories should coincide w

RADIANCE: Relative Adaptive Denoising with IP-Adapter for Novel Concept Enhancement

SafetyDGX agent

arXiv:2607.05088v1 Announce Type: new Abstract: Text-to-image (T2I) diffusion models have achieved striking progress but still struggle to synthesize rare concepts involving unusual attribute-object p


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

RADIO1D: Elastic Representations for Condensed Vision Modeling

SafetyDGX agent

arXiv:2607.03624v1 Announce Type: cross Abstract: This paper challenges the assumption that vision-language models (VLMs) require fixed patch-based 2D vision features. Analyzing fine-tuned vision enco

Reason, Reward, Refine: Step-Level Errors Corrections with Structured Feedback for Physics Reasoning in Small Language Models

SafetyDGX agent

arXiv:2607.05199v1 Announce Type: new Abstract: Physics reasoning fails structurally in small language models: an error at any step propagates forward, corrupting every inference that follows. Limited

ReCal3R: Reliability-Calibrated Learning Rates for Streaming 3D Reconstruction

SafetyDGX agent

arXiv:2607.05356v1 Announce Type: new Abstract: Streaming 3D reconstruction relies on a compact recurrent scene state to process long image streams in linear time and bounded memory. However, repeated

REDDIT: Correcting Model-Generated Timestamp Drift in ASR without Forgetting via Replay-Based Distribution Editing

SafetyDGX agent

arXiv:2607.05364v1 Announce Type: cross Abstract: Modern autoregressive ASR systems can emit timestamps as decoded tokens, enabling timestamped transcription without frame-level aligners or inference-

Reference-Induced Consensus for Selective Posed-Reference Visual Localization

SafetyDGX agent

arXiv:2607.04722v1 Announce Type: new Abstract: We present RIC-Loc (Reference-Induced Consensus localization), a scene-training-free posed-reference localizer that is SfM-point-map-free in its main es

Regime-Conditional Stabilisation of LLM-Augmented Cooperative Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2607.04470v1 Announce Type: cross Abstract: Large Language Models (LLMs) offer a natural interface for translating human objectives into reward signals for cooperative multi-agent reinforcement

Reinforcement Learning for Data-Efficient Code-Switched ASR

SafetyDGX agent

arXiv:2607.02757v1 Announce Type: new Abstract: Audio-language models can be prompted for code-switched speech, but their decoding is not optimized for code-switching and often fails at language bound

Relational Multi-Agent Reinforcement Learning for Dynamic Pricing in High-Speed Railway Markets

SafetyDGX agent

arXiv:2607.05179v1 Announce Type: cross Abstract: In liberalised railway systems, operators must set prices dynamically in an environment with partial observability, as they retain private information

Reliability and Identifiability in Persona-Trained Monte Carlo: Variance Decomposition, Stability Bounds, and the Identifiability of Heterogeneous News Reaction

SafetyDGX agent

arXiv:2607.04627v1 Announce Type: new Abstract: Persona-Trained Monte Carlo (PTMC) estimates distributions of market-outcome functionals by repeatedly simulating limit-order-book interaction among K n

Reliability-Aware CT-MRI Registration: A Quality Engineering Framework with Stability Analysis and Risk Classification

SafetyDGX agent

arXiv:2607.02585v1 Announce Type: new Abstract: Multimodal CT-MRI registration is central to image-guided radiotherapy, surgical navigation, and diagnostic workflows, but most pipelines report only ag

Resolving Primitive-Sharing Ambiguity in Long-Tailed TLS-Based Industrial MEP Point Cloud Segmentation via Spatial Context Constraints

SafetyDGX agent

arXiv:2601.19128v2 Announce Type: replace Abstract: In terrestrial laser scanning (TLS)-based mechanical, electrical, and plumbing (MEP) point cloud segmentation, safety-critical components such as re

Rethinking Brain Decoding with CLIP: The Role of Adversarial Robustness

SafetyDGX agent

arXiv:2607.03165v1 Announce Type: new Abstract: Brain decoding aims to uncover neural mechanisms by inferring stimulus-related representations from brain signals. In fMRI studies, this is typically ac

Rethinking On-Policy Self-Distillation for Thinking Models

SafetyDGX agent

arXiv:2607.05184v1 Announce Type: new Abstract: Self-distillation is a promising recipe for self-improvement in language models. In this setting, a model can serve as its own teacher when given privil

Reward-Gated On-Policy Distillation

SafetyDGX agent

arXiv:2607.04037v1 Announce Type: cross Abstract: On-policy distillation is a powerful way to transfer reasoning ability from a strong teacher to a smaller student: the student samples trajectories fr

Reward Granularity in RLVR: Comparing Process and Outcome Reward Structures for Mathematical Reasoning in Small Language Models

SafetyDGX agent

arXiv:2607.02869v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has emerged as a promising paradigm for improving mathematical reasoning in language models. Yet m

Reward Lightning: Fast Video Generation via Homologous Preference Distillation

SafetyDGX agent

arXiv:2607.03960v1 Announce Type: new Abstract: Achieving simultaneous preference alignment and distillation acceleration in video diffusion models remains an open challenge. Existing methods optimize

RL-Ballast: Ship Ballast Water Path Planning and Clog Prediction via Reinforcement Learning

SafetyDGX agent

arXiv:2607.04906v1 Announce Type: new Abstract: Under the Shipping 4.0 paradigm, autonomous and reduced-crew vessels require intelligent internal systems to maintain operational safety and structural

Road-Aware Anomaly Segmentation with Query-Guided Polygons and CLIP in Autonomous Driving

SafetyDGX agent

arXiv:2607.04304v1 Announce Type: new Abstract: Traditional semantic segmentation models operate under a closed-set assumption and struggle to recognize unknown or unexpected objects-an essential capa

RSPO: Reward-Swap Policy Optimization for Multi-Turn LLM Agents

SafetyDGX agent

arXiv:2607.04713v1 Announce Type: cross Abstract: Reinforcement learning holds significant potential for training large language models (LLMs) to handle multi-turn interactive tasks. However, in long-

SABLE: An NDA-Safe Closed-Loop LLM Framework for Analog Circuit Optimization in Industrial EDA Flows

SafetyDGX agent

arXiv:2607.03701v1 Announce Type: cross Abstract: Large language models (LLMs) can propose circuit-optimization decisions, but industrial analog flows cannot expose foundry PDK content, proprietary sc

Safe Inference-Time Alignment via Lagrangian Reward Augmentation

SafetyDGX agent

arXiv:2607.02781v1 Announce Type: cross Abstract: Inference-time alignment steers a frozen language model during decoding using auxiliary reward signals, avoiding the cost of repeated weight updates.

Safe RLHF Beyond Expectation: Stochastic Dominance for Universal Spectral Risk Control

SafetyDGX agent

arXiv:2603.10938v2 Announce Type: replace-cross Abstract: Safe Reinforcement Learning from Human Feedback (RLHF) typically enforces safety through expected cost constraints, but the expectation captur

Sample-Efficient Pareto Front Modeling for Energy-Aware Reinforcement Learning Using Bayesian Optimization

SafetyDGX agent

arXiv:2607.03140v1 Announce Type: new Abstract: Industrial automation increasingly demands control strategies that balance operational performance with strict energy efficiency requirements. A common

Saving GPU Hours in LLM Inference System Development and Online Workloads with Simulation and DBMS-Inspired Cache Replacement Policies

SafetyDGX agent

arXiv:2411.07447v5 Announce Type: replace-cross Abstract: LLMs are increasingly used world-wide from daily tasks to agentic systems and data analytics, requiring significant GPU resources. While LLM i

Scalable Dexterous Robot Learning with AR-based Remote Human-Robot Interactions

SafetyDGX agent

arXiv:2602.07341v2 Announce Type: replace Abstract: This paper focuses on the scalable robot learning for manipulation in the dexterous robot arm-hand systems, where the remote human-robot interaction

Scalable Semantic Steering of Embedding Projections

SafetyDGX agent

arXiv:2607.03978v1 Announce Type: cross Abstract: Low-dimensional projections support interactive visual analysis of high-dimensional data embeddings, but their structure often does not align with ana

Schedulable Job-Level Dependencies for Cause-Effect Chains via Graph Neural Networks

SafetyDGX agent

arXiv:2607.02624v1 Announce Type: cross Abstract: Modern automotive software architectures comprise large sets of mixed-criticality functions executing on shared multi-core platforms with strict real-

SEAM: Smooth Execution of Action-Chunked Motion for Vision-Language-Action Policies

SafetyDGX agent

arXiv:2607.04609v1 Announce Type: new Abstract: Vision-Language-Action (VLA) policies that execute fixed-length action chunks can exhibit multimodal bifurcation: a cross-chunk inconsistency in which a

Securing Multi-Tool AI Agent Chains With Dynamic, Real-Time Compositional Policies

SafetyDGX agent

arXiv:2607.03423v1 Announce Type: cross Abstract: Modern AI agent implementations such as frontier coding agents chain multiple tools at runtime that create a security surface that per-tool guardrails

Self-Improving Diffusion Classifiers with Minority Preference Optimization

SafetyDGX agent

arXiv:2607.03770v1 Announce Type: cross Abstract: Prior studies have demonstrated that diffusion classifiers achieve robust zero-shot classification performance. However, their effectiveness is strong

Self-Reference in Large Language Models: The Introspection Threshold for Recursive Self-Improvement

SafetyDGX agent

arXiv:2607.04277v1 Announce Type: cross Abstract: The pursuit of self-evolving AI raises a critical question: when is autonomous self-improvement sustainable rather than degenerative? Drawing an analo

Semantic Video Communication via Multi-Scale Convolution and Dynamic Routing for Next-Generation Networks

SafetyDGX agent

arXiv:2607.05093v1 Announce Type: new Abstract: The exponential growth of video traffic demands novel semantic communication paradigms that transmit meaning rather than raw bits. We present a generati

Sequential Cohort Selection under Uncertainty

SafetyDGX agent

arXiv:2508.16386v2 Announce Type: replace Abstract: We study the problem of fair cohort selection under uncertainty, motivated by university admissions where applicant outcomes are only partially obse

Shapley-based Data Valuation for LLM Alignment via Sequential Preference Optimization

SafetyDGX agent

arXiv:2512.15765v3 Announce Type: replace Abstract: Data valuation is a natural framework for understanding which preference datasets matter most when aligning a Large Language Model (LLM) using multi

Short-Horizon Position Accuracy of Single-Track Models: Implications for Motion Planning of Autonomous Vehicles

SafetyDGX agent

arXiv:2606.14216v2 Announce Type: replace Abstract: Accurate and computationally efficient vehicle models are essential for motion planning of autonomous vehicles, where positional accuracy directly a

SiamJEPA: On the Role of Siamese Student Encoders in JEPA

SafetyDGX agent

arXiv:2607.04044v1 Announce Type: new Abstract: Recently, Joint Embedding Predictive Architectures (JEPAs) have attracted significant attention in the computer vision and machine learning communities

Silicon Sampling via Cross-Survey Transfer

SafetyDGX agent

arXiv:2607.03091v1 Announce Type: new Abstract: Silicon sampling-using large language models (LLMs) to simulate human survey respondents-has emerged as a promising approach for augmenting traditional

Simple-to-Complex Structured Demonstrations for Vision-Language-Action Learning

SafetyDGX agent

arXiv:2607.04591v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong capabilities in robotic manipulation by integrating visual perception, language understan

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses

SafetyDGX agent

arXiv:2510.15476v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used as interfaces to information, code, and real-world services, making prompt-level security f

SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction

SafetyDGX agent

arXiv:2607.04119v1 Announce Type: cross Abstract: Reconstructing Computer-Aided Design (CAD) modeling sequences from images is crucial for preserving design intent and supporting parametric editing. H

Spatial Attention: Adapting Execution Horizons for Diffusion Policies via Observation Sensitivity

SafetyDGX agent

arXiv:2607.04739v1 Announce Type: new Abstract: Sampling action chunks via generative models has become a widely adopted methodology for robotic learning from demonstration. However, existing methods

SpecGradFilter: A Spectral Gradient Filtering Framework for Taming Federated Heterogeneity

SafetyDGX agent

arXiv:2607.04189v1 Announce Type: new Abstract: Federated Learning (FL) is fundamentally challenged by statistical heterogeneity, where non-identically distributed (non-IID) data induces client drift

Spectral Gradient Descent Mitigates Anisotropy-Driven Misalignment: A Case Study in Phase Retrieval

SafetyDGX agent

arXiv:2601.22652v2 Announce Type: replace-cross Abstract: Spectral gradient methods, such as the Muon optimizer, modify gradient updates by preserving directional information while discarding scale, a

StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models

SafetyDGX agent

arXiv:2603.20659v2 Announce Type: replace Abstract: Large scale pre-training on text and image data along with diverse robot demonstrations has helped Vision Language Action models (VLAs) to generaliz

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training

SafetyDGX agent

arXiv:2607.04963v1 Announce Type: new Abstract: Reinforcement Learning (RL) is the dominant paradigm for training Large Language Model (LLM) agents on long-horizon tasks. However, sparse and delayed r

States are asking for a $1.4 trillion fine of Meta over addicting kids. Basically they are saying Facebook's business model is a result of c…

SafetyDGX agent

States are asking for a $1.4 trillion fine of Meta over addicting kids. Basically they are saying Facebook's business model is a result of crime, and all of Mark Zuckerberg's property should be forfei

Strategic Buying Agents

SafetyDGX agent

arXiv:2607.04708v1 Announce Type: cross Abstract: Agentic AI is shifting online shopping from search toward delegated purchasing, where autonomous buying agents monitor markets and decide when to buy

STRATOS: Bridging the Symbolic-to-Numeric Gap in Spatio-Temporal Text-to-SQL for Meteorological Data

SafetyDGX agent

arXiv:2607.03501v1 Announce Type: cross Abstract: Copernicus, the European Union's Earth observation program, produces petabytes of Earth observation and climate data, offering immense potential for r

Structure-Guided Self-Supervised Matching for One-Shot Medical Landmark Detection

SafetyDGX agent

arXiv:2203.01687v3 Announce Type: replace Abstract: Medical landmark detection usually requires accurate expert annotations, which are laborious and difficult to scale across anatomical regions. In th

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance

SafetyDGX agent

arXiv:2510.08048v4 Announce Type: replace-cross Abstract: Query-product relevance prediction is fundamental to e-commerce search and has become even more critical in the era of AI-powered shopping, wh

Teaming Up with AI: Coordination and Cooperation

SafetyDGX agent

arXiv:2607.03181v1 Announce Type: cross Abstract: Successful diffusion of AI in the workforce hinges on the economic value that AI brings to human endeavors. Bringing AI into the workforce is more tha

Telescope: Improving Zero Shot Detection of LLM Generated Content By Measuring Token Repetition Probability

SafetyDGX agent

arXiv:2607.04061v1 Announce Type: cross Abstract: Distinguishing Large Language Model (LLM) generated text from human writing is a critical and difficult challenge. While LLMs are trained to write lik

Tensor-Train Joint Modeling for Few-Step Discrete Diffusion

SafetyDGX agent

arXiv:2607.03788v1 Announce Type: new Abstract: Discrete diffusion promises orders-of-magnitude faster generation than autoregressive (AR) models for sequential discrete data, yet its full potential o

Text as Partial Constraint: Core-Residual Alignment for Robust Vision-Language Learning

SafetyDGX agent

arXiv:2607.03143v1 Announce Type: cross Abstract: Vision-language alignment powers open-vocabulary recognition, retrieval, and LVLM grounding, yet natural captions are often underspecified, making sim

Text Dictates, Music Decorates: Energy-based Attention for Editable Dance Motion Generation

SafetyDGX agent

arXiv:2606.22726v2 Announce Type: replace Abstract: Choreographic motion generation poses unique challenges for AI, demanding precise semantic control over complex, temporally structured, and expressi

The agent creates, we validate: A Lightweight Framework for Agentic Artifact Generation

SafetyDGX agent

arXiv:2607.02615v1 Announce Type: cross Abstract: Generating structured artifacts with Large Language Models - e.g. database queries, threat framework mappings, entity schemas - is relatively straight

The AI boom is just not sustainable; Apollo’s Torsten Slok joins the chorus.

SafetyDGX agent

The AI boom is just not sustainable; Apollo’s Torsten Slok joins the chorus. Torsten Slok argues that the AI boom can only be seen in the hyperscalers and semiconductor companies and that this is caus

← Previous
1…4445464748…212
Next →