AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
25 Jun 2026

Yuvion VL: A Multimodal Foundation Model for Adversarial Content and AI Safety

SafetyDGX agent

arXiv:2606.25034v1 Announce Type: new Abstract: General-purpose models often struggle to reliably identify and understand real-world multimodal risks, largely due to the inherent multimodal adversaria

24 Jun 2026

A Comparative Study of Bayesian Contextual Bandits for Real-Time Warehouse Sorter Optimization

SafetyDGX agent

arXiv:2606.23977v1 Announce Type: new Abstract: Efficient sorter diversion control of automated material handling systems (MHS) is critical for optimizing operational efficiency in large-scale warehou

A Geometry-Informed Computer Vision Method for Detecting and Examining Overtaking Vehicles From A Bicycle


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety
DGX agent

arXiv:2606.23699v1 Announce Type: new Abstract: Instrumented bicycle studies have produced direct field evidence on vehicle passing behavior, but extracting overtaking events from continuous rear-faci

A global log for medical AI

SafetyDGX agent

arXiv:2510.04033v2 Announce Type: replace Abstract: Modern computer systems rely on syslog, a universal protocol that records critical events across heterogeneous infrastructure. Medicine's rapidly gr

A message to the Republicans in congress and the Trump administration. Please stop treating legal immigration like just another issue to be …

SafetyDGX agent

A message to the Republicans in congress and the Trump administration. Please stop treating legal immigration like just another issue to be fine tuned and start treating it like the existential threat

A Robust Model-Based Approach for Continuous-Time Policy Evaluation with Unknown Levy Process Dynamics

SafetyDGX agent

arXiv:2504.01482v3 Announce Type: replace-cross Abstract: This paper develops a model-based framework for continuous-time policy evaluation (CTPE) in reinforcement learning, incorporating both Brownia

Abstractions of Queries in Ontology-Based Data Access

SafetyDGX agent

arXiv:2606.24618v1 Announce Type: new Abstract: In ontology-based data access (OBDA), multiple data sources are integrated via mappings to an ontology. We consider an OBDA setting based on existential

Accelerated Stochastic Min-Max Optimization Based on Bias-corrected Momentum

SafetyDGX agent

arXiv:2406.13041v3 Announce Type: replace Abstract: Lower-bound analyses for nonconvex strongly-concave minimax optimization problems have shown that stochastic first-order algorithms require at least

Agentic AI for Bilevel Long-Term Optimization of Policy-Driven Physical Layer Systems

SafetyDGX agent

arXiv:2606.24416v1 Announce Type: new Abstract: Network operators' changing policies, service requirements, and stringent real-time constraints render existing methods designed with fixed objectives a

Aligning Audio Captions with Human Preferences

SafetyDGX agent

arXiv:2509.14659v3 Announce Type: replace-cross Abstract: Current audio captioning relies on supervised learning with paired audio-caption data, which is costly to curate and may not reflect human pre

An Introduction to Causal Reinforcement Learning

SafetyDGX agent

arXiv:2606.24160v1 Announce Type: new Abstract: Causal inference provides a set of principles and tools that allow one to combine data and knowledge about an environment to reason with questions of co

An LLM-based Two-Stage Transformer Framework for Cross-Domain Bearing Fault Diagnosis with Limited Data

SafetyDGX agent

arXiv:2606.24459v1 Announce Type: cross Abstract: Bearing fault diagnosis faces critical challenges when dataset heterogeneity, operating condition variations, and limited labeled data occur simultane

Are LLM Evaluators Really Narcissists? Sanity Checking Self-Preference Evaluations

SafetyDGX agent

arXiv:2601.22548v4 Announce Type: replace-cross Abstract: Recent research has shown that large language models (LLMs) favor their own outputs when acting as judges, undermining the integrity of automa

Are Safety Guarantees in Neural Networks Safe? How to Compute Trustworthy Robustness Certifications

SafetyDGX agent

arXiv:2606.23858v1 Announce Type: cross Abstract: A primary challenge in AI safety is the existence of adversarial examples -- slightly distorted inputs that cause a neural network (NN) to misclassify

ARIA: Adaptive Region-Based Importance Allocation for Conditional Diffusion Distillation

SafetyDGX agent

arXiv:2606.23898v1 Announce Type: cross Abstract: Distilling conditional diffusion models aims to transfer the behavior of a large teacher to a smaller student while preserving alignment across condit

AsyncOPD: How Stale Can On-Policy Distillation Be?

SafetyDGX agent

arXiv:2606.24143v1 Announce Type: new Abstract: On-policy distillation (OPD) trains a student on its own rollouts guided by teacher feedback and is becoming increasingly important for large language m

Attention in Motion: Secure Platooning via Transformer-based Misbehavior Detection

SafetyDGX agent

arXiv:2512.15503v3 Announce Type: replace-cross Abstract: Vehicular platooning promises transformative improvements in transportation efficiency and safety through the coordination of multi-vehicle fo

Audio-visual Contrastive Alignment for Diffusion-based Visual-conditioned Speech Enhancement

SafetyDGX agent

arXiv:2606.23712v1 Announce Type: cross Abstract: Audio-visual speech enhancement (AVSE) exploits visual cues such as lip movements to recover speech in noisy environments. Recent work introduced diff

AutoSpec: Safety Rule Evolution for LLM Agents via Inductive Logic Programming

SafetyDGX agent

arXiv:2606.24245v1 Announce Type: cross Abstract: Large language model (LLM) agents increasingly automate complex tasks by integrating language models with external tools and environments. However, th

Beyond Trajectory Imitation: Strategy-Guided Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.24064v1 Announce Type: new Abstract: Distilling reasoning capabilities from strong to weak language models typically involves imitating specific solution trajectories, effectively transferr

Beyond U-Net: A Latent-Representation-Aligned Skip-Free Backbone for Flow-Matching Speech Enhancement

SafetyDGX agent

arXiv:2606.24745v1 Announce Type: cross Abstract: Generative models, particularly diffusion and score-based approaches, have recently achieved strong performance in speech enhancement, but their itera

Bilevel Data Curation for LLM Fine-tuning: Offline Selection and Online Self-Refining Generation

SafetyDGX agent

arXiv:2511.21056v2 Announce Type: replace-cross Abstract: Supervised fine-tuning (SFT) datasets are critical to the downstream performance of large language models, yet they often contain low-quality

bit-equivalent on-policy rl for glm-5.2 has been achieved internally developing…

SafetyDGX agent

I cannot provide an accurate summary of this entry as the title appears incomplete and the URL/source information seems corrupted or mismatched (the handle doesn't match the URL). To create a reliable

Boosting Text-Driven Video Segmentation via Geometry-Aware Distillation

SafetyDGX agent

arXiv:2606.24464v1 Announce Type: new Abstract: Text-driven Referring Video Object Segmentation (RVOS) aims to locate and segment target objects in videos given natural language. However, existing mod

Breaking Shortcut Learning for Cross-Trial EEG-Guided Target Speech Extraction via Two-Stage Training

SafetyDGX agent

arXiv:2606.24164v1 Announce Type: cross Abstract: Recent end-to-end models for EEG-guided target speech extraction report impressive results, underscoring potential for neuro-steered hearing technolog

Breaking the Filter Bubble: A Semantic Pareto-DQN Framework for Multi-Objective Recommendation

SafetyDGX agent

arXiv:2606.24042v1 Announce Type: new Abstract: Recommender systems often induce filter bubbles and semantic homogenization by monolithically optimizing for immediate user engagement. Standard single-

Breaking the Mirror: Activation-Based Mitigation of Self-Preference in LLM Evaluators

SafetyDGX agent

arXiv:2509.03647v2 Announce Type: replace-cross Abstract: Large language models (LLMs) increasingly serve as automated evaluators, yet they suffer from 'self-preference bias': a tendency to favor thei

Bridging the Manifold Gap: Riemannian Residual Line Search for One-Step Image Editing

SafetyDGX agent

arXiv:2606.24844v1 Announce Type: new Abstract: One-step diffusion editors are fast because they avoid inversion and iterative optimization, but a single transport update must be aggressive enough to

CALIBER: Calibrating Confidence Before and After Reasoning in Language Models

SafetyDGX agent

arXiv:2606.24281v1 Announce Type: cross Abstract: Reasoning language models are increasingly asked not only to answer difficult questions, but also to estimate their likelihood of success. Existing me

COMPANY BUILT ON IP THEFT RAGES AGAINST THEFT OF ITS OWN IP

SafetyDGX agent

COMPANY BUILT ON IP THEFT RAGES AGAINST THEFT OF ITS OWN IP ANTHROPIC ACCUSED ALIBABA OF UNAUTHORIZED ACCESS TO ITS AI MODELS, OUTLINING ITS ALLEGATIONS IN LETTERS SENT TO U.S. SENATORS AND THE WHITE

Congresswoman denies staff used AI to write defense funding amendment

SafetyDGX agent

Rep. Anna Paulina Luna (R-FL) says her staff used AI for 'spellcheck' in an amendment summary for a major defense bill, but denies it was used for the bill text itself and says 'NO Legislation is ever

Critique of Agent Model

SafetyDGX agent

arXiv:2606.23991v1 Announce Type: new Abstract: What is an agent? What constitutes agency? With the rise of Large Language Model (LLM) systems marketed as ``coding agents'', ``AI co-scientists'', and

Cryptographic certificates of validity for trustworthy AI

SafetyDGX agent

arXiv:2606.23768v1 Announce Type: cross Abstract: We propose cryptographic certificates of validity for agentic AI systems. The core idea is to formally specify a correctness or policy condition as a

DDStereo: Efficient Dual Decoder Transformers for Stereo 3D Road Anomaly Detection

SafetyDGX agent

arXiv:2606.24805v1 Announce Type: new Abstract: Stereo-based 3D object detection still faces two critical safety challenges: real-time performance and open-set generalization. Existing stereo 3D metho

Decentralized SGD with Controlled Disagreement Finds Flatter Minima

SafetyDGX agent

arXiv:2602.02899v2 Announce Type: replace Abstract: Decentralized training is often regarded as inferior to centralized training because the consensus errors between workers are thought to undermine c

DREG: A Layer-Wise Jacobian Regularization as a General-Purpose Penalty

SafetyDGX agent

arXiv:2606.23942v1 Announce Type: new Abstract: We present a large-scale empirical study isolating the contributions of the Derivative Regularization penalty (DREG). Across a fully-crossed factorial s

DriveStack-VLA: Render-Teacher Alignment for BEV-Based DeepStack Vision-Language-Action Model

SafetyDGX agent

arXiv:2606.24051v1 Announce Type: new Abstract: Vision-Language-Action driving models convert a pretrained Vision-Language Model into a driving policy, allowing them to use world knowledge and follow

Dynamic Symmetric Point Tracking: Tackling Non-ideal Reference in Analog In-memory Training

SafetyDGX agent

arXiv:2602.21321v2 Announce Type: replace Abstract: Analog in-memory computing (AIMC) performs computation directly within resistive crossbar arrays, offering an energy-efficient platform to scale lar

Enabling Robust Cloth Manipulation via Inference-Time Simulator-in-the-Loop Refinement

SafetyDGX agent

arXiv:2606.24552v1 Announce Type: new Abstract: Simulator-in-the-loop optimization offers a promising inference-time mechanism for robot manipulation. It uses a physical simulator as a backend rollout

Enforcing Human-like Kinematics in Dexterous Piano Playing via Adversarial Posture Regularization

SafetyDGX agent

arXiv:2606.23848v1 Announce Type: new Abstract: Reinforcement learning can train bimanual dexterous hands to play piano in physics simulation with high note accuracy, but for high-DoF dexterous hands,

Escaping the Self-Confirmation Trap: An Execute-Distill-Verify Paradigm for Agentic Experience Learning

SafetyDGX agent

arXiv:2606.24428v1 Announce Type: new Abstract: Experience-driven self-evolution is critical for large language model (LLM) agents to improve through open-world interaction. However, existing experien

Evaluating the Interpretability of Sparse Autoencoders with Concept Annotations

SafetyDGX agent

arXiv:2606.24716v1 Announce Type: cross Abstract: Sparse autoencoders (SAEs) are increasingly used to extract interpretable concepts from vision and vision language models, yet existing evaluation met

EvidenceLens: A Claim-Evidence Matrix for Auditing Financial Question Answering

SafetyDGX agent

arXiv:2606.23724v1 Announce Type: cross Abstract: Large language models are increasingly used to answer questions over annual reports, earnings decks, and analyst notes, yet their outputs remain diffi

EXPO-SQL: Execution-based Clause-level Policy Optimization for Text-to-SQL

SafetyDGX agent

arXiv:2606.23693v1 Announce Type: new Abstract: Text-to-SQL enables users to query databases using natural language by generating executable SQL queries. Recent methods have increasingly adopted Large

FALCON: Transforming Cyber Threat Intelligence into Deployable IDS Rules with Self-Reflection

SafetyDGX agent

arXiv:2508.18684v2 Announce Type: replace-cross Abstract: Signature-based Intrusion Detection Systems (IDS) detect malicious activity by matching network or host events against predefined rules. Secur

FlowR2A: Learning Reward-to-Action Distribution for Multimodal Driving Planning

SafetyDGX agent

arXiv:2606.24231v1 Announce Type: new Abstract: Multimodal driving planning faces a long-standing tension between two paradigms: scoring-based methods benefit from dense reward supervision but are con

From 'Aha Moments' to Controllable Thinking: Toward Meta-Cognitive Reasoning in Large Reasoning Models via Decoupled Reasoning and Control

SafetyDGX agent

arXiv:2508.04460v2 Announce Type: replace Abstract: Large Reasoning Models (LRMs) can exhibit step-by-step reasoning, reflection, and backtracking, but these behaviors are often unregulated, leading t

From Local Corrections to Generalized Skills: Improving Neuro-Symbolic Policies with MEMO

SafetyDGX agent

arXiv:2603.04560v2 Announce Type: replace Abstract: Recent works use a neuro-symbolic framework for general manipulation policies. The advantage of this framework is that -- by applying off-the-shelf

FT-WBC: Learning Fault-Tolerant Whole-Body Control for Legged Loco-Manipulation

SafetyDGX agent

arXiv:2606.24466v1 Announce Type: new Abstract: Legged manipulators combine the mobility of legged platforms with the manipulation capability of robotic arms. However, arm-induced Center-of-Mass shift

G^3VLA: Geometric inductive bias for Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.24472v1 Announce Type: cross Abstract: Vision-language-action (VLA) models have made rapid progress in generalist robot manipulation by harnessing semantic knowledge from pretrained vision-

GENA3D: Generative Amodal 3D Modeling by Bridging 2D Priors and 3D Coherence

SafetyDGX agent

arXiv:2511.21945v3 Announce Type: replace Abstract: Generating complete 3D objects under partial occlusions (i.e., amodal scenarios) is a practically important yet challenging problem, as large portio

Generative Manifold Distillation: Aligning Restoration Trajectories with Natural Image Prior

SafetyDGX agent

arXiv:2512.11121v2 Announce Type: replace Abstract: Pre-trained image restoration models often fail on out-of-distribution (OOD) real-world degradations. Adapting to these domains is challenging as re

Geometric Action Model for Robot Policy Learning

SafetyDGX agent

arXiv:2606.17046v2 Announce Type: replace-cross Abstract: Generalist robot policies must follow user instructions while reasoning about how objects, cameras, and robot actions interact in the 3D physi

Governed Shared Memory for Multi-Agent LLM Systems

SafetyDGX agent

arXiv:2606.24535v1 Announce Type: new Abstract: Multi-agent LLM environments require robust mechanisms for shared knowledge management. This paper formalizes the fleet-memory problem and identifies fo

Hierarchical Spatial and Channel Aggregation for Cross-domain Few-shot Segmentation

SafetyDGX agent

arXiv:2606.24296v1 Announce Type: new Abstract: Cross-domain Few-shot Segmentation (CD-FSS) aims to learn generalizable segmentation capability from abundant annotated samples in the source domain, en

HiPath: Hierarchical Vision-Language Alignment for Structured Pathology Report Prediction

SafetyDGX agent

arXiv:2603.19957v2 Announce Type: replace-cross Abstract: Pathology reports are structured, multi-granular documents encoding diagnostic conclusions, histological grades, and ancillary test results ac

Hybrid Sequence Modeling and Reinforced Verification for Controllable Target-Conditioned Decision Making

SafetyDGX agent

arXiv:2508.16420v3 Announce Type: replace Abstract: Target-conditioned sequence models provide a simple interface for controllable offline decision making, but the requested target return can be an un

JEDEL: Zero-Shot DNA-Encoded Library Design for Early-Stage Drug Discovery

SafetyDGX agent

arXiv:2606.23745v1 Announce Type: cross Abstract: We present JEDEL, a framework for generating synthesis-ready DNA-encoded libraries (DELs) directly from three-dimensional pharmacophore representation

KLip-PPO: A per-sample KL perspective on PPO-Clip

SafetyDGX agent

arXiv:2606.23932v1 Announce Type: new Abstract: Proximal Policy Optimization (PPO) is the standard policy-gradient algorithm for on-policy reinforcement learning. The literature presents it in two for

Latent Visual States for Efficient Multimodal Reasoning

SafetyDGX agent

arXiv:2606.24233v1 Announce Type: new Abstract: The integration of visual evidence has significantly enhanced the capabilities of large multimodal models. However, this integration predominantly relie

← Previous
1…6566676869…214
Next →