AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,596 results
Safety

Safety Cost of Steering Vectors Is Separable and Reducible

DGX agent

arXiv:2608.08383v1 Announce Type: new Abstract: Steering vectors are a lightweight tool for controlling LLM behavior. However, emerging evidence shows that steering vectors can unintentionally comprom

safetyarxiv-cs-cl
11 Aug 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

Satellite Trajectory Optimization via Proximal Policy Optimization for Space Debris Avoidance

DGX agent

arXiv:2608.09628v1 Announce Type: new Abstract: Collision avoidance systems are commonly used to avoid fragmentation events occurring in Low-Earth Orbit (LEO) and Geosynchronous Equatorial Orbit (GEO)

safetyarxiv-cs-lg
11 Aug 2026
Safety

SC-Diff: Semantically Calibrated Diffusion for Visible-to-Infrared Image Translation

DGX agent

arXiv:2608.08555v1 Announce Type: new Abstract: Visible-to-infrared image translation provides a practical way to expand infrared training data using abundant visible images. Diffusion models are prom

safetyarxiv-cs-cv
11 Aug 2026
Safety

Scalable extensions to given-data Sobol' index estimators

DGX agent

arXiv:2509.09078v3 Announce Type: replace-cross Abstract: Given-data methods for variance-based sensitivity analysis have significantly advanced the feasibility of Sobol' index computation for computa

safetyarxiv-cs-lg
11 Aug 2026
Safety

ScaleSense: Cost-Intelligent Scaling Framework via Learned Resource Estimation in Alibaba AnalyticDB

DGX agent

arXiv:2608.07945v1 Announce Type: cross Abstract: Cloud-native serverless data warehouses achieve fine-grained elasticity by decoupling storage from compute, yet determining the optimal resource alloc

safetyarxiv-cs-ai
11 Aug 2026
Safety

SCOUT: Self-Checking and Recovery-Aware Tool-Thought Agents for Ultra-Long Egocentric Video Reasoning

DGX agent

arXiv:2608.07959v1 Announce Type: new Abstract: Ultra-long egocentric video understanding requires reasoning over temporally sparse evidence distributed across hours or days, challenging current multi

safetyarxiv-cs-ai
11 Aug 2026
Safety

Search-G1: Grounded Search Agents via Representation-Based Intrinsic Rewards

DGX agent

arXiv:2608.07531v1 Announce Type: cross Abstract: Search-augmented language agents should retrieve external information only when necessary and ground their answers in retrieved evidence. Existing ext

safetyarxiv-cs-ai
11 Aug 2026
Safety

Self Supervised Learning from Automatically Generated Demonstrations for Visual Robotic Manipulation

DGX agent

arXiv:2608.07553v1 Announce Type: new Abstract: Robotic manipulation often requires object specific programming, manual data annotation, or calibrated perception pipelines, which limits rapid deployme

safetyarxiv-cs-ro
11 Aug 2026
Safety

SIP: Site in Pieces- A Dataset of Disaggregated Construction-Phase 3D Scans for Semantic Segmentation and Scene Understanding

DGX agent

arXiv:2512.09062v2 Announce Type: replace Abstract: Accurate 3D scene interpretation in active construction sites is essential for progress monitoring, safety assessment, and digital twin development.

safetyarxiv-cs-cv
11 Aug 2026
Safety

SMAC: Score-Matched Actor-Critics for Robust Offline-to-Online Transfer

DGX agent

arXiv:2602.17632v3 Announce Type: replace-cross Abstract: Modern offline Reinforcement Learning (RL) methods find performant actor-critics, however, fine-tuning these actor-critics online with value-b

safetyarxiv-cs-ai
11 Aug 2026
Safety

Software Engineering for and with GUI Agent

DGX agent

arXiv:2608.09278v1 Announce Type: cross Abstract: GUI agents have advanced rapidly, producing a growing body of frameworks, benchmarks, and applications. However, this growth has outpaced the maturity

safetyarxiv-cs-ai
11 Aug 2026
Safety

Spatiotemporal Context-dependent Personalized Movement Compensation in Delayed Telemanipulation

DGX agent

arXiv:2608.08200v1 Announce Type: new Abstract: Communication delay remains a central challenge in telerobotics, where it disrupts visuomotor coordination and reduces task precision. Motion scaling is

safetyarxiv-cs-ro
11 Aug 2026
Safety

SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning

DGX agent

arXiv:2608.09138v1 Announce Type: cross Abstract: While learned robotic policies hold promise for advancing generalizable manipulation, their practical deployment is often hindered by suboptimal execu

safetyarxiv-cs-ai
11 Aug 2026
Safety

SR-OPSD: Self-Referenced On-Policy Self-Distillation

DGX agent

arXiv:2608.09745v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, provi

safetyarxiv-cs-ai
11 Aug 2026
Safety

Stealing Reasoning Traces from Proprietary LLM APIs

DGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

safetyarxiv-cs-ai
11 Aug 2026
Safety

StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoning

DGX agent

arXiv:2608.08326v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective approach for improving multimodal reasoning. However, most existing me

safetyarxiv-cs-ai
11 Aug 2026
Safety

Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of Predicate Representation Granularity

DGX agent

arXiv:2608.00143v2 Announce Type: replace-cross Abstract: Automated attack chain generation is critical for modern cybersecurity, yet manual construction fails to scale as adversary behaviors expand.

safetyarxiv-cs-ai
11 Aug 2026
Safety

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

DGX agent

arXiv:2608.07948v1 Announce Type: new Abstract: VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling comp

safetyarxiv-cs-cv
11 Aug 2026
Safety

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

DGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Sample Complexity of Policy Learning with Mu-Resets

DGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

safetyarxiv-cs-lg
11 Aug 2026
Safety

The Scaling Paradox in Human-AI Collaboration

DGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

DGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

safetyarxiv-cs-ai
11 Aug 2026
Safety

The Transparency Trap: How AI Disclaimers Create Overconfidence in High-Stakes Decisions

DGX agent

arXiv:2608.07493v1 Announce Type: cross Abstract: Current AI disclaimers often fail to function as intended due to warning habituation and a transparency paradox. As AI-generated information becomes p

safetyarxiv-cs-cl
11 Aug 2026
Safety

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

DGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

safetyarxiv-cs-cl
11 Aug 2026
Safety

There are no lossless transformations of natural-language text

DGX agent

There are no lossless transformations of natural-language text Sophie Alpert shares her 'internal policy on acceptable use of AI writing by engineers'. It's a short read (supporting its own recommenda

safetysimon-willison
11 Aug 2026
Safety

Three Necessary Principles for Self-Supervised Visual Representation Learning

DGX agent

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: sema

safetyarxiv-cs-ai
11 Aug 2026
Safety

ToolUniverse: An open platform for democratizing AI scientists

DGX agent

arXiv:2509.23426v3 Announce Type: replace Abstract: AI scientists are emerging computational systems that serve as collaborative partners in discovery. These systems remain difficult to build because

safetyarxiv-cs-ai
11 Aug 2026
Safety

Topographic Constraints Shape Brain-Like Component Structure in Auditory Models

DGX agent

arXiv:2509.24039v2 Announce Type: replace-cross Abstract: If topography is a fundamental feature of the brain, it should influence both how neurons are arranged in space (i.e. explain brain maps) and

safetyarxiv-cs-ai
11 Aug 2026
Safety

Towards Expressive and Faithful Audio-to-Image Generation: A Unified Multimodal Dataset and Synthesis Framework

DGX agent

arXiv:2608.09529v1 Announce Type: new Abstract: As an important subfield of cross-modal generation, synthesizing static visual content in the form of images from audio, namely audio-to-image (A2I) gen

safetyarxiv-cs-cv
11 Aug 2026
Safety

Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM

DGX agent

arXiv:2511.18721v4 Announce Type: replace-cross Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict 'k-unstable' assumption that

safetyarxiv-cs-ai
11 Aug 2026
Safety

Triple Expert Learning from Noisy Labels for Semi-Supervised Vision Foundation Model Adaptation

DGX agent

arXiv:2608.09052v1 Announce Type: cross Abstract: Semi-supervised adaptation of vision foundation models (VFMs) commonly freezes the pretrained backbone and updates lightweight modules such as LoRA. H

safetyarxiv-cs-ai
11 Aug 2026
Safety

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

DGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

safetyarxiv-cs-ai
11 Aug 2026
Safety

Uncertainty-Aware Variational Reward Factorization via Probabilistic Preference Bases for LLM Personalization

DGX agent

arXiv:2604.00997v2 Announce Type: replace Abstract: Reward factorization personalizes large language models (LLMs) by decomposing rewards into shared basis functions and user-specific weights. Yet, ex

safetyarxiv-cs-cl
11 Aug 2026
Safety

Understanding Reasoning from Pretraining to Post-Training

DGX agent

arXiv:2607.16097v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is l

safetyarxiv-cs-ai
11 Aug 2026
Safety

Unimodality-Promoting Regularized Learning for Ordinal Regression

DGX agent

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and co

safetyarxiv-cs-lg
11 Aug 2026
Safety

UnsDrive: Towards Robust End-to-End Autonomous Driving in Unstructured Scenes

DGX agent

arXiv:2608.09098v1 Announce Type: new Abstract: End-to-end planning has shown strong promise for autonomous driving, but most existing methods are designed for structured urban roads and generalize po

safetyarxiv-cs-ro
11 Aug 2026
Safety

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models

DGX agent

arXiv:2608.08622v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have demonstrated strong performance in open-ended video understanding, yet they remain prone to fluent responses u

safetyarxiv-cs-cv
11 Aug 2026
Safety

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

DGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

safetyarxiv-cs-cv
11 Aug 2026
Safety

Vid2WAM: Distilling Video Diffusion Priors into World Action Models

DGX agent

arXiv:2608.08558v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling future visual dynamics and actions. However, their scalability and generali

safetyarxiv-cs-ro
11 Aug 2026
Safety

VoxZip: Semantic-Anchored Temporal KV Cache Compression for Long-Context Audio Inference

DGX agent

arXiv:2608.08569v1 Announce Type: new Abstract: Recent advancements in Speech Large Language Models have demonstrated remarkable capabilities in understanding complex audio tasks. Despite this progres

safetyarxiv-cs-ai
11 Aug 2026
Safety

WDL-OPD: Weak-Driven On-Policy Distillation via Mixture-Constrained Co-Training

DGX agent

arXiv:2608.09447v1 Announce Type: cross Abstract: On-policy distillation (OPD) aligns a student with a teacher on trajectories sampled from the student itself, reducing the train-test state mismatch o

safetyarxiv-cs-ai
11 Aug 2026
Safety

What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files

DGX agent

arXiv:2608.08453v1 Announce Type: new Abstract: Under the current standard, Agent Skills are SKILL.md files that combine instructions with supporting resources, enabling Large Language Model (LLM) age

safetyarxiv-cs-ai
11 Aug 2026
Safety

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

DGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

safetyarxiv-cs-lg
11 Aug 2026
Safety

Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space

DGX agent

arXiv:2608.07786v1 Announce Type: new Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships t

safetyarxiv-cs-ai
11 Aug 2026
Safety

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

DGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

safetyarxiv-cs-cv
11 Aug 2026
Safety

Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production

DGX agent

arXiv:2608.08471v1 Announce Type: new Abstract: Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new jailbreak techniques and previously un-addressed

safetyarxiv-cs-ai
11 Aug 2026
Safety

A Practical Evaluation Method for Long-Form Simultaneous Speech-to-Speech Translation

DGX agent

arXiv:2606.15059v2 Announce Type: replace Abstract: Simultaneous speech-to-speech translation (SimulS2ST) enables real-time cross-lingual communication, but existing evaluation has focused largely on

safetyarxiv-cs-cl
10 Aug 2026
Safety

An AI4AI Framework for Visual Token Pruning

DGX agent

arXiv:2608.07193v1 Announce Type: cross Abstract: Visual-token pruning can substantially reduce the inference cost of multimodal large language models (MLLMs), yet existing methods largely rely on fix

safetyarxiv-cs-cv
10 Aug 2026
← Previous
1…56789…263
Next →