AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
All
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
Safety

Shapley-based Data Valuation for LLM Alignment via Sequential Preference Optimization

DGX agent

arXiv:2512.15765v3 Announce Type: replace Abstract: Data valuation is a natural framework for understanding which preference datasets matter most when aligning a Large Language Model (LLM) using multi

safetyarxiv-cs-lg
7 Jul 2026
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Short-Horizon Position Accuracy of Single-Track Models: Implications for Motion Planning of Autonomous Vehicles

DGX agent

arXiv:2606.14216v2 Announce Type: replace Abstract: Accurate and computationally efficient vehicle models are essential for motion planning of autonomous vehicles, where positional accuracy directly a

safetyarxiv-cs-ro
7 Jul 2026
Safety

SiamJEPA: On the Role of Siamese Student Encoders in JEPA

DGX agent

arXiv:2607.04044v1 Announce Type: new Abstract: Recently, Joint Embedding Predictive Architectures (JEPAs) have attracted significant attention in the computer vision and machine learning communities

safetyarxiv-cs-cv
7 Jul 2026
Safety

Silicon Sampling via Cross-Survey Transfer

DGX agent

arXiv:2607.03091v1 Announce Type: new Abstract: Silicon sampling-using large language models (LLMs) to simulate human survey respondents-has emerged as a promising approach for augmenting traditional

safetyarxiv-cs-ai
7 Jul 2026
Safety

Simple-to-Complex Structured Demonstrations for Vision-Language-Action Learning

DGX agent

arXiv:2607.04591v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models have demonstrated strong capabilities in robotic manipulation by integrating visual perception, language understan

safetyarxiv-cs-ai
7 Jul 2026
Safety

SoK: Systematizing LLM Prompt Security: Taxonomies, Datasets, and Unified Evaluation of Attacks and Defenses

DGX agent

arXiv:2510.15476v3 Announce Type: replace-cross Abstract: Large Language Models (LLMs) are increasingly used as interfaces to information, code, and real-world services, making prompt-level security f

safetyarxiv-cs-ai
7 Jul 2026
Safety

SOV-CAD: Stepwise Orthographic Views Guided CAD Modeling Sequence Reconstruction

DGX agent

arXiv:2607.04119v1 Announce Type: cross Abstract: Reconstructing Computer-Aided Design (CAD) modeling sequences from images is crucial for preserving design intent and supporting parametric editing. H

safetyarxiv-cs-ai
7 Jul 2026
Safety

Spatial Attention: Adapting Execution Horizons for Diffusion Policies via Observation Sensitivity

DGX agent

arXiv:2607.04739v1 Announce Type: new Abstract: Sampling action chunks via generative models has become a widely adopted methodology for robotic learning from demonstration. However, existing methods

safetyarxiv-cs-ro
7 Jul 2026
Safety

SpecGradFilter: A Spectral Gradient Filtering Framework for Taming Federated Heterogeneity

DGX agent

arXiv:2607.04189v1 Announce Type: new Abstract: Federated Learning (FL) is fundamentally challenged by statistical heterogeneity, where non-identically distributed (non-IID) data induces client drift

safetyarxiv-cs-lg
7 Jul 2026
Safety

Spectral Gradient Descent Mitigates Anisotropy-Driven Misalignment: A Case Study in Phase Retrieval

DGX agent

arXiv:2601.22652v2 Announce Type: replace-cross Abstract: Spectral gradient methods, such as the Muon optimizer, modify gradient updates by preserving directional information while discarding scale, a

safetyarxiv-cs-lg
7 Jul 2026
Safety

StageCraft: Execution Aware Mitigation of Distractor and Obstruction Failures in VLA Models

DGX agent

arXiv:2603.20659v2 Announce Type: replace Abstract: Large scale pre-training on text and image data along with diverse robot demonstrations has helped Vision Language Action models (VLAs) to generaliz

safetyarxiv-cs-ro
7 Jul 2026
Safety

STAPO: Selective Trajectory-Aware Policy Optimization for LLM Agent Training

DGX agent

arXiv:2607.04963v1 Announce Type: new Abstract: Reinforcement Learning (RL) is the dominant paradigm for training Large Language Model (LLM) agents on long-horizon tasks. However, sparse and delayed r

safetyarxiv-cs-ai
7 Jul 2026
Safety

States are asking for a $1.4 trillion fine of Meta over addicting kids. Basically they are saying Facebook's business model is a result of c…

DGX agent

States are asking for a $1.4 trillion fine of Meta over addicting kids. Basically they are saying Facebook's business model is a result of crime, and all of Mark Zuckerberg's property should be forfei

safetygary-marcus--x
7 Jul 2026
Safety

Strategic Buying Agents

DGX agent

arXiv:2607.04708v1 Announce Type: cross Abstract: Agentic AI is shifting online shopping from search toward delegated purchasing, where autonomous buying agents monitor markets and decide when to buy

safetyarxiv-cs-ai
7 Jul 2026
Safety

STRATOS: Bridging the Symbolic-to-Numeric Gap in Spatio-Temporal Text-to-SQL for Meteorological Data

DGX agent

arXiv:2607.03501v1 Announce Type: cross Abstract: Copernicus, the European Union's Earth observation program, produces petabytes of Earth observation and climate data, offering immense potential for r

safetyarxiv-cs-ai
7 Jul 2026
Safety

Structure-Guided Self-Supervised Matching for One-Shot Medical Landmark Detection

DGX agent

arXiv:2203.01687v3 Announce Type: replace Abstract: Medical landmark detection usually requires accurate expert annotations, which are laborious and difficult to scale across anatomical regions. In th

safetyarxiv-cs-cv
7 Jul 2026
Safety

TaoSR-AGRL: Adaptive Guided Reinforcement Learning Framework for E-commerce Search Relevance

DGX agent

arXiv:2510.08048v4 Announce Type: replace-cross Abstract: Query-product relevance prediction is fundamental to e-commerce search and has become even more critical in the era of AI-powered shopping, wh

safetyarxiv-cs-ai
7 Jul 2026
Safety

Teaming Up with AI: Coordination and Cooperation

DGX agent

arXiv:2607.03181v1 Announce Type: cross Abstract: Successful diffusion of AI in the workforce hinges on the economic value that AI brings to human endeavors. Bringing AI into the workforce is more tha

safetyarxiv-cs-ai
7 Jul 2026
Safety

Telescope: Improving Zero Shot Detection of LLM Generated Content By Measuring Token Repetition Probability

DGX agent

arXiv:2607.04061v1 Announce Type: cross Abstract: Distinguishing Large Language Model (LLM) generated text from human writing is a critical and difficult challenge. While LLMs are trained to write lik

safetyarxiv-cs-ai
7 Jul 2026
Safety

Tensor-Train Joint Modeling for Few-Step Discrete Diffusion

DGX agent

arXiv:2607.03788v1 Announce Type: new Abstract: Discrete diffusion promises orders-of-magnitude faster generation than autoregressive (AR) models for sequential discrete data, yet its full potential o

safetyarxiv-cs-lg
7 Jul 2026
Safety

Text as Partial Constraint: Core-Residual Alignment for Robust Vision-Language Learning

DGX agent

arXiv:2607.03143v1 Announce Type: cross Abstract: Vision-language alignment powers open-vocabulary recognition, retrieval, and LVLM grounding, yet natural captions are often underspecified, making sim

safetyarxiv-cs-ai
7 Jul 2026
Safety

Text Dictates, Music Decorates: Energy-based Attention for Editable Dance Motion Generation

DGX agent

arXiv:2606.22726v2 Announce Type: replace Abstract: Choreographic motion generation poses unique challenges for AI, demanding precise semantic control over complex, temporally structured, and expressi

safetyarxiv-cs-ai
7 Jul 2026
Safety

The agent creates, we validate: A Lightweight Framework for Agentic Artifact Generation

DGX agent

arXiv:2607.02615v1 Announce Type: cross Abstract: Generating structured artifacts with Large Language Models - e.g. database queries, threat framework mappings, entity schemas - is relatively straight

safetyarxiv-cs-ai
7 Jul 2026
Safety

The AI boom is just not sustainable; Apollo’s Torsten Slok joins the chorus.

DGX agent

The AI boom is just not sustainable; Apollo’s Torsten Slok joins the chorus. Torsten Slok argues that the AI boom can only be seen in the hyperscalers and semiconductor companies and that this is caus

safetygary-marcus--x
7 Jul 2026
Safety

The Foreign Policy AI Evaluation Gap

DGX agent

arXiv:2607.02955v1 Announce Type: cross Abstract: We argue that AI systems used in conducting foreign policy tasks - broadly enacting 'statecraft' - should be a priority test case for technical AI gov

safetyarxiv-cs-ai
7 Jul 2026
Safety

The ‘Ghost’ in the Database: Recovering Active ADFS Signing Keys via Machine DPAPI

DGX agent

Written by: Shebin Mathew Introduction The 'Golden SAML' technique, first described by CyberArk researchers in 2017, and further detailed by Mandiant researchers in 2021, remains one of the most effec

safetygoogle-cloud-ai
7 Jul 2026
Safety

The Three Regimes of Offline-to-Online Reinforcement Learning

DGX agent

arXiv:2510.01460v4 Announce Type: replace-cross Abstract: Offline-to-online reinforcement learning (RL) has emerged as a practical paradigm that leverages offline datasets for pretraining and online i

safetyarxiv-cs-ai
7 Jul 2026
Safety

TokAN: Accent Normalization Using Self-Supervised Speech Tokens

DGX agent

arXiv:2607.03928v1 Announce Type: cross Abstract: Accent normalization (AN) seeks to convert non-native (L2) accented speech into standard (L1) speech while preserving speaker identity. The current te

safetyarxiv-cs-ai
7 Jul 2026
Safety

Towards Data-Driven Metrics for Social Robot Navigation Benchmarking

DGX agent

arXiv:2509.01251v3 Announce Type: replace Abstract: This paper presents a joint effort towards the development of a data-driven Social Robot Navigation metric to facilitate benchmarking and policy opt

safetyarxiv-cs-ro
7 Jul 2026
Safety

Training Verifiably Robust Agents Using Set-Based Reinforcement Learning

DGX agent

arXiv:2408.09112v2 Announce Type: replace Abstract: Reinforcement learning policies parametrized by deep neural networks have achieved strong performance for continuous control, yet even small input p

safetyarxiv-cs-lg
7 Jul 2026
Safety

Trajectory-Anchor Optimization for Overconfident Thermal Visual Place Recognition: Zero-Leakage OOD Auditing and Kidnapped-Robot Recovery

DGX agent

arXiv:2607.04745v1 Announce Type: cross Abstract: Modern thermal visual place recognition (TIR-VPR) frontends based on foundation models achieve remarkable closed-set retrieval but suffer from an over

safetyarxiv-cs-cv
7 Jul 2026
Safety

Transformer-Based Multi-Agent Reinforcement Learning for Networked Systems with Long-Range Interactions

DGX agent

arXiv:2511.13103v2 Announce Type: replace Abstract: Multi-agent reinforcement learning (MARL) has shown promise for large-scale network control, yet existing methods face two major limitations. First,

safetyarxiv-cs-lg
7 Jul 2026
Safety

Trust Region Policy Distillation

DGX agent

arXiv:2607.04751v1 Announce Type: cross Abstract: Big goals are hard to achieve all at once; breaking them into small steps is wiser. We present Trust Region Policy Distillation (TOP-D), which transfo

safetyarxiv-cs-ai
7 Jul 2026
Safety

Turning Off-Policy Tokens On-Policy: A Plug-in Approach for Improving LLM Alignment

DGX agent

arXiv:2607.04728v1 Announce Type: cross Abstract: Reinforcement learning (RL) post-training for large language models (LLMs) follows a efficient paradigm of 'rollout then update', which inevitably res

safetyarxiv-cs-ai
7 Jul 2026
Safety

Two Black Boxes, One Solver: Encoder Probing and Decoder Attribution for Neural Multi-Attribute VRP under Hard-Mask and Recourse Decoders

DGX agent

arXiv:2607.04487v1 Announce Type: cross Abstract: Neural autoregressive solvers for the Multi-Attribute Vehicle Routing Problem (MAVRP) reach competitive cost but offer no per-step justification, a pr

safetyarxiv-cs-ai
7 Jul 2026
Safety

UI-MOPD: Multi-Platform On-Policy Distillation for Continual GUI Agent Learning

DGX agent

arXiv:2607.04425v1 Announce Type: cross Abstract: Recent advances in multimodal foundation models and agent systems have driven GUI agents from single-platform task execution toward cross-platform int

safetyarxiv-cs-ai
7 Jul 2026
Safety

Uncertainty-Aware Abstention in Large Language Models with Provable Alignment Guarantees

DGX agent

arXiv:2607.04430v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly deployed in question answering (QA) systems, yet they may generate hallucinated or misaligned responses wi

safetyarxiv-cs-cl
7 Jul 2026
Safety

Uncertainty-Aware Last-Layer Adaptation of RETFound for Referable Diabetic Retinopathy Screening Under Dataset Shift

DGX agent

arXiv:2607.02569v1 Announce Type: new Abstract: This paper presents a safety-centered empirical evaluation of uncertainty-aware last-layer adaptation for referable diabetic retinopathy screening using

safetyarxiv-cs-cv
7 Jul 2026
Safety

Uncertainty Quantification for Regression: A Unified Framework based on kernel scores

DGX agent

arXiv:2510.25599v2 Announce Type: replace Abstract: Regression tasks, notably in safety-critical domains, require reliable uncertainty quantification, yet the literature remains largely classification

safetyarxiv-cs-lg
7 Jul 2026
Safety

UNDREAM: Bridging Differentiable Rendering and Photorealistic Simulation for End-to-end Adversarial Attacks

DGX agent

arXiv:2510.16923v3 Announce Type: replace-cross Abstract: Deep learning models deployed in safety critical applications like autonomous driving use simulations to test their robustness against adversa

safetyarxiv-cs-ai
7 Jul 2026
Safety

Virtual Category-Guided Continual Generalized Category Discovery

DGX agent

arXiv:2607.04984v1 Announce Type: new Abstract: Continual Generalized Category Discovery (C-GCD) aims to incrementally identify novel categories from sequential unlabeled data while preserving recogni

safetyarxiv-cs-cv
7 Jul 2026
Safety

Vision Non-Causal Trapezoidal Mamba: Eliminating Directional Scanning in Vision SSMs with Second-Order Dynamics

DGX agent

arXiv:2607.03589v1 Announce Type: new Abstract: State Space Models (SSMs) have emerged as an alternative to Vision Transformers, yet most vision SSMs inherit directional token scanning from causal seq

safetyarxiv-cs-cv
7 Jul 2026
Safety

VISOR++: Universal Visual Inputs based Steering for Large Vision Language Models

DGX agent

arXiv:2509.25533v2 Announce Type: replace-cross Abstract: As Vision Language Models (VLMs) are deployed across safety-critical applications, understanding and controlling their behavioral patterns has

safetyarxiv-cs-ai
7 Jul 2026
Safety

VISOR: Visual Input-based Steering for Output Redirection in Vision-Language Models

DGX agent

arXiv:2508.08521v2 Announce Type: replace-cross Abstract: Vision Language Models (VLMs) are increasingly being used in a broad range of applications, bringing their security and behavioral control to

safetyarxiv-cs-ai
7 Jul 2026
Safety

VLA Grounder: Language-Conditioning Space Optimization for Black-Box VLA Models

DGX agent

arXiv:2607.04517v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models are commonly treated as end-to-end action policies conditioned on natural-language task descriptions. In practice, h

safetyarxiv-cs-ai
7 Jul 2026
Safety

VLM-CASE: Vision-Language Model Enabled Context-Adaptive Safety Envelopes for Anticipatory Safe Autonomous Driving

DGX agent

arXiv:2607.05180v1 Announce Type: cross Abstract: Adverse driving conditions, such as bad weather, remain a principal barrier to autonomous driving because they degrade two things at once: what the ve

safetyarxiv-cs-cv
7 Jul 2026
Safety

Weak-to-Strong Generalization via Direct On-Policy Distillation

DGX agent

arXiv:2607.05394v1 Announce Type: cross Abstract: Reinforcement learning with verifiable rewards (RLVR) is a powerful recipe for improving language-model reasoning, but it is expensive to repeat on ev

safetyarxiv-cs-ai
7 Jul 2026
Safety

When Agents Lie: Premeditation, Persistence, and Exploitation in Repeated Games

DGX agent

arXiv:2607.05132v1 Announce Type: cross Abstract: As large language models are deployed as autonomous agents that communicate intentions before acting, a critical safety question is whether agents tha

safetyarxiv-cs-cl
7 Jul 2026
← Previous
1…5657585960…265
Next →