AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
Safety

RynnValue: Scaling Robotic Value Foundation Models with Temporal Distance

DGX agent

arXiv:2608.09853v1 Announce Type: cross Abstract: General-purpose reward models are increasingly the bottleneck for scaling robot learning, yet the recipe for learning value-related capabilities from

safetyarxiv-cs-cv
11 Aug 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SAFE-CHEM: Uncertainty-Aware Policy Switching for Robust Robotic Chemistry

DGX agent

arXiv:2608.09303v1 Announce Type: cross Abstract: The deployment of autonomous robotic systems in chemistry laboratories is accelerating experimental workflows and providing the foundational data for

safetyarxiv-cs-lg
11 Aug 2026
Safety

Safety Cost of Steering Vectors Is Separable and Reducible

DGX agent

arXiv:2608.08383v1 Announce Type: new Abstract: Steering vectors are a lightweight tool for controlling LLM behavior. However, emerging evidence shows that steering vectors can unintentionally comprom

safetyarxiv-cs-cl
11 Aug 2026
Safety

Satellite Trajectory Optimization via Proximal Policy Optimization for Space Debris Avoidance

DGX agent

arXiv:2608.09628v1 Announce Type: new Abstract: Collision avoidance systems are commonly used to avoid fragmentation events occurring in Low-Earth Orbit (LEO) and Geosynchronous Equatorial Orbit (GEO)

safetyarxiv-cs-lg
11 Aug 2026
Safety

SC-Diff: Semantically Calibrated Diffusion for Visible-to-Infrared Image Translation

DGX agent

arXiv:2608.08555v1 Announce Type: new Abstract: Visible-to-infrared image translation provides a practical way to expand infrared training data using abundant visible images. Diffusion models are prom

safetyarxiv-cs-cv
11 Aug 2026
Safety

Scalable extensions to given-data Sobol' index estimators

DGX agent

arXiv:2509.09078v3 Announce Type: replace-cross Abstract: Given-data methods for variance-based sensitivity analysis have significantly advanced the feasibility of Sobol' index computation for computa

safetyarxiv-cs-lg
11 Aug 2026
Safety

ScaleSense: Cost-Intelligent Scaling Framework via Learned Resource Estimation in Alibaba AnalyticDB

DGX agent

arXiv:2608.07945v1 Announce Type: cross Abstract: Cloud-native serverless data warehouses achieve fine-grained elasticity by decoupling storage from compute, yet determining the optimal resource alloc

safetyarxiv-cs-ai
11 Aug 2026
Safety

SCOUT: Self-Checking and Recovery-Aware Tool-Thought Agents for Ultra-Long Egocentric Video Reasoning

DGX agent

arXiv:2608.07959v1 Announce Type: new Abstract: Ultra-long egocentric video understanding requires reasoning over temporally sparse evidence distributed across hours or days, challenging current multi

safetyarxiv-cs-ai
11 Aug 2026
Safety

Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards

DGX agent

arXiv:2608.07531v1 Announce Type: cross Abstract: Search-augmented language agents should retrieve external information only when necessary and ground their answers in retrieved evidence. Existing ext

safetyarxiv-cs-ai
11 Aug 2026
Safety

Self Supervised Learning from Automatically Generated Demonstrations for Visual Robotic Manipulation

DGX agent

arXiv:2608.07553v1 Announce Type: new Abstract: Robotic manipulation often requires object specific programming, manual data annotation, or calibrated perception pipelines, which limits rapid deployme

safetyarxiv-cs-ro
11 Aug 2026
Safety

SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding

DGX agent

arXiv:2512.09062v2 Announce Type: replace Abstract: Accurate 3D scene interpretation in active construction sites is essential for progress monitoring, safety assessment, and digital twin development.

safetyarxiv-cs-cv
11 Aug 2026
Safety

SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer

DGX agent

arXiv:2602.17632v3 Announce Type: replace-cross Abstract: Modern offline Reinforcement Learning (RL) methods find performant actor-critics, however, fine-tuning these actor-critics online with value-b

safetyarxiv-cs-ai
11 Aug 2026
Safety

Software Engineering for and with GUI Agent

DGX agent

arXiv:2608.09278v1 Announce Type: cross Abstract: GUI agents have advanced rapidly, producing a growing body of frameworks, benchmarks, and applications. However, this growth has outpaced the maturity

safetyarxiv-cs-ai
11 Aug 2026
Safety

Spatiotemporal Context-dependent Personalized Movement Compensation in Delayed Telemanipulation

DGX agent

arXiv:2608.08200v1 Announce Type: new Abstract: Communication delay remains a central challenge in telerobotics, where it disrupts visuomotor coordination and reduces task precision. Motion scaling is

safetyarxiv-cs-ro
11 Aug 2026
Safety

SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning

DGX agent

arXiv:2608.09138v1 Announce Type: cross Abstract: While learned robotic policies hold promise for advancing generalizable manipulation, their practical deployment is often hindered by suboptimal execu

safetyarxiv-cs-ai
11 Aug 2026
Safety

SR-OPSD: Self-Referenced On-Policy Self-Distillation

DGX agent

arXiv:2608.09745v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, provi

safetyarxiv-cs-ai
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Safety

StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoning

DGX agent

arXiv:2608.08326v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective approach for improving multimodal reasoning. However, most existing me

safetyarxiv-cs-ai
11 Aug 2026
Safety

Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of Predicate Representation Granularity

DGX agent

arXiv:2608.00143v2 Announce Type: replace-cross Abstract: Automated attack chain generation is critical for modern cybersecurity, yet manual construction fails to scale as adversary behaviors expand.

safetyarxiv-cs-ai
11 Aug 2026
Safety

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

DGX agent

arXiv:2608.07948v1 Announce Type: new Abstract: VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling comp

safetyarxiv-cs-cv
11 Aug 2026
Safety

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

DGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Sample Complexity of Policy Learning with Mu-Resets

DGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

safetyarxiv-cs-lg
11 Aug 2026
Safety

The Scaling Paradox in Human-AI Collaboration

DGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

DGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Transparency Trap: How AI Disclaimers Create Overconfidence in High-Stakes Decisions

DGX agent

arXiv:2608.07493v1 Announce Type: cross Abstract: Current AI disclaimers often fail to function as intended due to warning habituation and a transparency paradox. As AI-generated information becomes p

safetyarxiv-cs-cl
11 Aug 2026
Safety

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

DGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

safetyarxiv-cs-cl
11 Aug 2026
Safety

There are no lossless transformations of natural-language text

DGX agent

There are no lossless transformations of natural-language text Sophie Alpert shares her 'internal policy on acceptable use of AI writing by engineers'. It's a short read (supporting its own recommenda

safetysimon-willison
11 Aug 2026
Safety

Three Necessary Principles for Self-Supervised Visual Representation Learning

DGX agent

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: sema

safetyarxiv-cs-ai
11 Aug 2026
Safety

ToolUniverse: An open platform for democratizing AI scientists

DGX agent

arXiv:2509.23426v3 Announce Type: replace Abstract: AI scientists are emerging computational systems that serve as collaborative partners in discovery. These systems remain difficult to build because

safetyarxiv-cs-ai
11 Aug 2026
Safety

Topographic Constraints Shape Brain-Like Component Structure in Auditory Models

DGX agent

arXiv:2509.24039v2 Announce Type: replace-cross Abstract: If topography is a fundamental feature of the brain, it should influence both how neurons are arranged in space (i.e. explain brain maps) and

safetyarxiv-cs-ai
11 Aug 2026
Safety

Towards Expressive and Faithful Audio-to-Image Generation: A Unified Multimodal Dataset and Synthesis Framework

DGX agent

arXiv:2608.09529v1 Announce Type: new Abstract: As an important subfield of cross-modal generation, synthesizing static visual content in the form of images from audio, namely audio-to-image (A2I) gen

safetyarxiv-cs-cv
11 Aug 2026
Safety

Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM

DGX agent

arXiv:2511.18721v4 Announce Type: replace-cross Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict 'k-unstable' assumption that

safetyarxiv-cs-ai
11 Aug 2026
Safety

Triple Expert Learning from Noisy Labels for Semi-Supervised Vision Foundation Model Adaptation

DGX agent

arXiv:2608.09052v1 Announce Type: cross Abstract: Semi-supervised adaptation of vision foundation models (VFMs) commonly freezes the pretrained backbone and updates lightweight modules such as LoRA. H

safetyarxiv-cs-ai
11 Aug 2026
Safety

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

DGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

safetyarxiv-cs-ai
11 Aug 2026
Safety

Uncertainty-Aware Variational Reward Factorization via Probabilistic Preference Bases for LLM Personalization

DGX agent

arXiv:2604.00997v2 Announce Type: replace Abstract: Reward factorization personalizes large language models (LLMs) by decomposing rewards into shared basis functions and user-specific weights. Yet, ex

safetyarxiv-cs-cl
11 Aug 2026
Safety

Understanding Reasoning from Pretraining to Post-Training

DGX agent

arXiv:2607.16097v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is l

safetyarxiv-cs-ai
11 Aug 2026
Safety

Unimodality-Promoting Regularized Learning for Ordinal Regression

DGX agent

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and co

safetyarxiv-cs-lg
11 Aug 2026
Safety

UnsDrive: Towards Robust End-to-End Autonomous Driving in Unstructured Scenes

DGX agent

arXiv:2608.09098v1 Announce Type: new Abstract: End-to-end planning has shown strong promise for autonomous driving, but most existing methods are designed for structured urban roads and generalize po

safetyarxiv-cs-ro
11 Aug 2026
Safety

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models

DGX agent

arXiv:2608.08622v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have demonstrated strong performance in open-ended video understanding, yet they remain prone to fluent responses u

safetyarxiv-cs-cv
11 Aug 2026
Safety

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

DGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

safetyarxiv-cs-cv
11 Aug 2026
Safety

Vid2WAM: Distilling Video Diffusion Priors into World Action Models

DGX agent

arXiv:2608.08558v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling future visual dynamics and actions. However, their scalability and generali

safetyarxiv-cs-ro
11 Aug 2026
Safety

VoxZip: Semantic-Anchored Temporal KV Cache Compression for Long-Context Audio Inference

DGX agent

arXiv:2608.08569v1 Announce Type: new Abstract: Recent advancements in Speech Large Language Models have demonstrated remarkable capabilities in understanding complex audio tasks. Despite this progres

safetyarxiv-cs-ai
11 Aug 2026
Safety

WDL-OPD: Weak-Driven On-Policy Distillation via Mixture-Constrained Co-Training

DGX agent

arXiv:2608.09447v1 Announce Type: cross Abstract: On-policy distillation (OPD) aligns a student with a teacher on trajectories sampled from the student itself, reducing the train-test state mismatch o

safetyarxiv-cs-ai
11 Aug 2026
Safety

What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files

DGX agent

arXiv:2608.08453v1 Announce Type: new Abstract: Under the current standard, Agent Skills are SKILL.md files that combine instructions with supporting resources, enabling Large Language Model (LLM) age

safetyarxiv-cs-ai
11 Aug 2026
Safety

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

DGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

safetyarxiv-cs-lg
11 Aug 2026
Safety

Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space

DGX agent

arXiv:2608.07786v1 Announce Type: new Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships t

safetyarxiv-cs-ai
11 Aug 2026
Safety

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

DGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

safetyarxiv-cs-cv
11 Aug 2026
Safety

Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production

DGX agent

arXiv:2608.08471v1 Announce Type: new Abstract: Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new jailbreak techniques and previously un-addressed

safetyarxiv-cs-ai
11 Aug 2026
← Previous
1…56789…263
Next →