AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,598 results
11 Aug 2026

Software Engineering for and with GUI Agent

SafetyDGX agent

arXiv:2608.09278v1 Announce Type: cross Abstract: GUI agents have advanced rapidly, producing a growing body of frameworks, benchmarks, and applications. However, this growth has outpaced the maturity

Spatiotemporal Context-dependent Personalized Movement Compensation in Delayed Telemanipulation

SafetyDGX agent

arXiv:2608.08200v1 Announce Type: new Abstract: Communication delay remains a central challenge in telerobotics, where it disrupts visuomotor coordination and reduces task precision. Motion scaling is

SpeedTuning: Speeding Up Policy Execution with Lightweight Reinforcement Learning

SafetyDGX agent

arXiv:2608.09138v1 Announce Type: cross Abstract: While learned robotic policies hold promise for advancing generalizable manipulation, their practical deployment is often hindered by suboptimal execu


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

SR-OPSD: Self-Referenced On-Policy Self-Distillation

SafetyDGX agent

arXiv:2608.09745v1 Announce Type: cross Abstract: On-policy self-distillation (OPSD) converts feedback into dense token-level supervision on trajectories generated by the policy to be optimized, provi

Stealing Reasoning Traces from Proprietary LLM APIs

SafetyDGX agent

arXiv:2608.09867v1 Announce Type: cross Abstract: Leading large language model providers now conceal their models' step-by-step reasoning, or chain-of-thought, to protect intellectual property and lim

StructReward: Efficient Structured Process Rewards for Self-Correcting Multimodal Reasoning

SafetyDGX agent

arXiv:2608.08326v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective approach for improving multimodal reasoning. However, most existing me

Symbolic Attack Chain Generation from Atomic Red Team Techniques: An Empirical Study of Predicate Representation Granularity

SafetyDGX agent

arXiv:2608.00143v2 Announce Type: replace-cross Abstract: Automated attack chain generation is critical for modern cybersecurity, yet manual construction fails to scale as adversary behaviors expand.

SynVAR: Synergizing Spatial and Semantic Alignment in Visual Autoregressive Model

SafetyDGX agent

arXiv:2608.07948v1 Announce Type: new Abstract: VAR has gained widespread popularity due to its next-scale prediction paradigm. However, it faces substantial performance bottlenecks when handling comp

The Anatomy of a Prompt Injection: A Component Model for Structured Analysis

SafetyDGX agent

arXiv:2608.07808v1 Announce Type: cross Abstract: Four years after prompt injection was first identified in 2022, attacks are still predominantly documented as verbatim strings rather than structured

The Sample Complexity of Policy Learning with Mu-Resets

SafetyDGX agent

arXiv:2608.07772v1 Announce Type: new Abstract: We study policy-based reinforcement learning under the mu-resets interaction protocol of Kakade and Langford [KL02]. This interaction protocol enables t

The Scaling Paradox in Human-AI Collaboration

SafetyDGX agent

arXiv:2608.00818v2 Announce Type: replace Abstract: The discovery of scaling laws has highlighted the extraordinary potential of AI systems with a striking empirical pattern: as AI systems scale, thei

The Theory of Strategic Evolution: Games with Endogenous Players and the Seven Laws of Strategic Replicators

SafetyDGX agent

arXiv:2512.07901v4 Announce Type: replace-cross Abstract: Von Neumann founded both game theory and the theory of self-reproducing automata, but the two programs never merged. Rational players do not c

The Transparency Trap: How AI Disclaimers Create Overconfidence in High-Stakes Decisions

SafetyDGX agent

arXiv:2608.07493v1 Announce Type: cross Abstract: Current AI disclaimers often fail to function as intended due to warning habituation and a transparency paradox. As AI-generated information becomes p

The Voiceprint Fallacy: Why Voices Are Not Unique Biometric Imprints

SafetyDGX agent

arXiv:2608.07980v1 Announce Type: cross Abstract: In recent years, the term voiceprint has regained attention, particularly in technological applications and policy-making contexts, often carrying the

There are no lossless transformations of natural-language text

SafetyDGX agent

There are no lossless transformations of natural-language text Sophie Alpert shares her 'internal policy on acceptable use of AI writing by engineers'. It's a short read (supporting its own recommenda

Three Necessary Principles for Self-Supervised Visual Representation Learning

SafetyDGX agent

arXiv:2608.08309v1 Announce Type: cross Abstract: We argue that learning visual representations without labels requires a training signal jointly complete across three non-overlapping objectives: sema

ToolUniverse: An open platform for democratizing AI scientists

SafetyDGX agent

arXiv:2509.23426v3 Announce Type: replace Abstract: AI scientists are emerging computational systems that serve as collaborative partners in discovery. These systems remain difficult to build because

Topographic Constraints Shape Brain-Like Component Structure in Auditory Models

SafetyDGX agent

arXiv:2509.24039v2 Announce Type: replace-cross Abstract: If topography is a fundamental feature of the brain, it should influence both how neurons are arranged in space (i.e. explain brain maps) and

Towards Expressive and Faithful Audio-to-Image Generation: A Unified Multimodal Dataset and Synthesis Framework

SafetyDGX agent

arXiv:2608.09529v1 Announce Type: new Abstract: As an important subfield of cross-modal generation, synthesizing static visual content in the form of images from audio, namely audio-to-image (A2I) gen

Towards Realistic Guarantees: A Probabilistic Certificate for SmoothLLM

SafetyDGX agent

arXiv:2511.18721v4 Announce Type: replace-cross Abstract: The SmoothLLM defense provides a certification guarantee against jailbreaking attacks, but it relies on a strict 'k-unstable' assumption that

Triple Expert Learning from Noisy Labels for Semi-Supervised Vision Foundation Model Adaptation

SafetyDGX agent

arXiv:2608.09052v1 Announce Type: cross Abstract: Semi-supervised adaptation of vision foundation models (VFMs) commonly freezes the pretrained backbone and updates lightweight modules such as LoRA. H

Unaccountable Delegation, Fading Skills: Mapping the Risks of Workplace AI Agents

SafetyDGX agent

arXiv:2608.08601v1 Announce Type: new Abstract: To anticipate socio-technical risks from AI agents, organizations need taxonomies to classify them. However, existing AI risk taxonomies focus on broad

Uncertainty-Aware Variational Reward Factorization via Probabilistic Preference Bases for LLM Personalization

SafetyDGX agent

arXiv:2604.00997v2 Announce Type: replace Abstract: Reward factorization personalizes large language models (LLMs) by decomposing rewards into shared basis functions and user-specific weights. Yet, ex

Understanding Reasoning from Pretraining to Post-Training

SafetyDGX agent

arXiv:2607.16097v2 Announce Type: replace-cross Abstract: Reinforcement learning (RL) has become central to improving large language models (LLMs) on complex reasoning tasks, yet RL post-training is l

Unimodality-Promoting Regularized Learning for Ordinal Regression

SafetyDGX agent

arXiv:2608.08359v1 Announce Type: new Abstract: Ordinal regression, also called ordinal classification, is classification of ordinal data, in which the underlying target variable is categorical and co

UnsDrive: Towards Robust End-to-End Autonomous Driving in Unstructured Scenes

SafetyDGX agent

arXiv:2608.09098v1 Announce Type: new Abstract: End-to-end planning has shown strong promise for autonomous driving, but most existing methods are designed for structured urban roads and generalize po

VADER: Adaptive Debiasing for Hallucination Mitigation in Video Large Language Models

SafetyDGX agent

arXiv:2608.08622v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have demonstrated strong performance in open-ended video understanding, yet they remain prone to fluent responses u

VANE: Reliable Test-Time Training for Vision-Language-Action Models via Future Visual Representation Prediction

SafetyDGX agent

arXiv:2608.09448v1 Announce Type: cross Abstract: Test-time training (TTT) offers a lightweight way to adapt vision--language--action (VLA) policies from unlabeled deployment streams, but it remains d

Vid2WAM: Distilling Video Diffusion Priors into World Action Models

SafetyDGX agent

arXiv:2608.08558v1 Announce Type: new Abstract: World Action Models (WAMs) improve robot policy learning by jointly modeling future visual dynamics and actions. However, their scalability and generali

VoxZip: Semantic-Anchored Temporal KV Cache Compression for Long-Context Audio Inference

SafetyDGX agent

arXiv:2608.08569v1 Announce Type: new Abstract: Recent advancements in Speech Large Language Models have demonstrated remarkable capabilities in understanding complex audio tasks. Despite this progres

WDL-OPD: Weak-Driven On-Policy Distillation via Mixture-Constrained Co-Training

SafetyDGX agent

arXiv:2608.09447v1 Announce Type: cross Abstract: On-policy distillation (OPD) aligns a student with a teacher on trajectories sampled from the student itself, reducing the train-test state mismatch o

What Keeps Agent Skills from Being Reusable? Evidence from 138K SKILL.md Files

SafetyDGX agent

arXiv:2608.08453v1 Announce Type: new Abstract: Under the current standard, Agent Skills are SKILL.md files that combine instructions with supporting resources, enabling Large Language Model (LLM) age

When Does Trace-Driven Evaluation Mislead MoE Expert Caching? Replay Semantics, Workload Contamination, and Operating Regimes

SafetyDGX agent

arXiv:2608.07911v1 Announce Type: new Abstract: Mixture-of-Experts (MoE) models have outgrown accelerator memory, and offloading expert weights to host memory is now standard. This makes expert cache

Who Built This Model? Tracing LLM Lineage via Spectral Fingerprints in Weight Space

SafetyDGX agent

arXiv:2608.07786v1 Announce Type: new Abstract: Open-weight large language models (LLMs) are increasingly developed through complex, multi-stage pipelines, leading to intricate lineage relationships t

World Tokens: Enhancing Embodied Policies with Training-Time World Modeling

SafetyDGX agent

arXiv:2608.09730v1 Announce Type: new Abstract: Vision-language-action (VLA) models are a widely adopted paradigm for embodied policies. They excel at efficient closed-loop control but do not explicit

Yesterday's Shield, Today's Spear: A Self-Evolving Safety Guardrail in Production

SafetyDGX agent

arXiv:2608.08471v1 Announce Type: new Abstract: Deployed LLM safety guardrails are predominantly static: trained once and frozen at release, while new jailbreak techniques and previously un-addressed

10 Aug 2026

A Practical Evaluation Method for Long-Form Simultaneous Speech-to-Speech Translation

SafetyDGX agent

arXiv:2606.15059v2 Announce Type: replace Abstract: Simultaneous speech-to-speech translation (SimulS2ST) enables real-time cross-lingual communication, but existing evaluation has focused largely on

An AI4AI Framework for Visual Token Pruning

SafetyDGX agent

arXiv:2608.07193v1 Announce Type: cross Abstract: Visual-token pruning can substantially reduce the inference cost of multimodal large language models (MLLMs), yet existing methods largely rely on fix

AutoIntervene: Calibrated Intervention for Action-Chunking Imitation Learning Policies

SafetyDGX agent

arXiv:2608.07065v1 Announce Type: cross Abstract: Action-chunking visuomotor policies learn from demonstrations and improve temporal consistency by predicting short action sequences rather than single

Automated item evaluation: Predicting item acceptance and rejection using LLM-generated critiques

SafetyDGX agent

arXiv:2608.06609v1 Announce Type: new Abstract: Automated item evaluation (AIE) refers to the use of computational methods to assess item quality without requiring manual expert review or field testin

Automated Terminal-to-Housing Assembly System for Flat Ribbon Cable Harness

SafetyDGX agent

arXiv:2608.06996v1 Announce Type: new Abstract: This paper presents a sensor-minimal automated assembly system for bidirectional single-row flat ribbon cable harnesses (FRCHs). Unlike conventional peg

Beyond Post-Hoc Temperature Scaling: Bilevel Optimization for LLM Calibration

SafetyDGX agent

arXiv:2608.07419v1 Announce Type: new Abstract: Preference alignment often makes large language models (LLMs) overconfident and poorly calibrated. Traditional post-hoc temperature scaling is inherentl

bioMoR: Biology-Guided Mixture-of-Recursions for Effective Genomic Learning

SafetyDGX agent

arXiv:2608.06727v1 Announce Type: new Abstract: Transformer models for high-dimensional omics analysis process thousands of genes or pathways, although only a subset requires deep computation. Mixture

Blind to the Pivotal Vote: Aggregate Independence Metrics Miss Where Verification Actually Helps

SafetyDGX agent

arXiv:2608.06940v1 Announce Type: new Abstract: LLM judge panels are a standard evaluation tool, but prior work reports highly correlated panel errors: nine judges provide roughly the effective inform

Bootstrap-Conditioned Action Selection with Tabular Foundation Models

SafetyDGX agent

arXiv:2608.06559v1 Announce Type: new Abstract: Contextual bandits offer a natural framework for sample-efficient personalization, but practical deployment remains difficult under sparse, biased inter

Bypassing Krum: Selection-Aware Backdoor Attacks in Federated Learning

SafetyDGX agent

arXiv:2608.06637v1 Announce Type: cross Abstract: Robust aggregation methods are widely used in federated learning to mitigate the impact of adversarial client behavior. Distance-based aggregation rul

Calibrating WEAT Against Anisotropy: ZCA Whitening as a Geometric Pre-Processing Step for Embedding Association Tests

SafetyDGX agent

arXiv:2608.06908v1 Announce Type: cross Abstract: We propose Zero-phase Component Analysis (ZCA) whitening as a geometric pre-processing step for the Word Embedding Association Test (WEAT). WEAT is a

CASA: Classification Augmented with Safety Attention for Robust Multimodal Alignment

SafetyDGX agent

arXiv:2604.00310v2 Announce Type: replace-cross Abstract: Multimodal large-language models (MLLMs) often experience degraded safety alignment when harmful queries exploit cross-modal interactions. Mod

Cascade: Exploiting SLO-Aware latency budget for fair and high goodput LLM inference serving

SafetyDGX agent

arXiv:2608.06557v1 Announce Type: cross Abstract: The reasoning and agentic capabilities of large language models have expanded the range of applications they support, from short interactive exchanges

CEDAR: Agent-Orchestrated Tree Search for Goal-Directed Optimization of Complex Systems

SafetyDGX agent

arXiv:2608.06871v1 Announce Type: new Abstract: Complex systems, core objects of study in artificial life, model diverse phenomena through nonlinear, feedback-driven interactions that produce emergent

CHIME: A Case for Efficient Long-Context Attention-FC Disaggregated Inference with DIMM-PIM

SafetyDGX agent

arXiv:2504.17584v2 Announce Type: replace-cross Abstract: Attention-FC Disaggregated (AFD) LLM inference systems offload memory-bound Attention operations to memory-rich accelerators (e.g., CPUs, HBM-

Confidence Estimation for Financial Vision-Language Models in Chart and Document Understanding

SafetyDGX agent

arXiv:2608.06532v1 Announce Type: new Abstract: LVLMs are increasingly used to read financial charts, tables, and documents, where a single misread figure can move a decision and the most authoritativ

Corrupting Attention: Evasion-Based Adversarial Attacks on Encoder Attention in Detection Transformers

SafetyDGX agent

arXiv:2608.06674v1 Announce Type: new Abstract: Adversarial vulnerabilities remain a major concern for the safe deployment of neural networks, particularly in object detection, a core task embedded in

Counterfactual Shapley Credit Assignment

SafetyDGX agent

arXiv:2607.16999v2 Announce Type: replace-cross Abstract: The Credit Assignment Problem (CAP) is fundamental to developing efficient and explainable Reinforcement Learning (RL) agents. Existing framew

CreativeInstruct: Scalably Teaching LLMs to Balance Quality, Creativity, and Diversity

SafetyDGX agent

arXiv:2608.07460v1 Announce Type: cross Abstract: While post-training improves the capabilities of large language models (LLMs), it generally lowers their output diversity and creativity, negatively i

CrystalGRPO: Target-Aligned and Coverage-Preserving Reinforcement Learning for Flow-Based Crystal Structure Prediction

SafetyDGX agent

arXiv:2608.06582v1 Announce Type: new Abstract: Flow-based generative models can efficiently produce candidate structures for crystal structure prediction (CSP), but their pretrained objectives do not

DA-Cal: Towards Cross-Domain Calibration in Semantic Segmentation

SafetyDGX agent

arXiv:2602.20860v2 Announce Type: replace Abstract: While existing unsupervised domain adaptation (UDA) methods greatly enhance target domain performance in semantic segmentation, they often neglect n

Degradation-Aware Prompt Learning with Cross-Modal Compensation for Adverse Weather Removal

SafetyDGX agent

arXiv:2608.06939v1 Announce Type: new Abstract: Adverse weather causes diverse and complex image degradations, severely compromising the reliability of computer vision systems. Existing all-in-one res

Density-Functional Excited-State Gradients and Nonadiabatic Couplings on a Consumer GPU from a Contraction-DAG

SafetyDGX agent

arXiv:2608.06536v1 Announce Type: cross Abstract: Nonadiabatic dynamics needs an excited-state gradient and an interstate nonadiabatic coupling matrix element (NACME) at every nuclear geometry, and a

DiDPO: Diff-in-Diff Policy Optimization for Coding Agent Training

SafetyDGX agent

arXiv:2608.07147v1 Announce Type: new Abstract: Reinforcement learning with Verifiable Reward (RLVR) has emerged as a powerful paradigm for training coding agents, where the execution feedback from co

← Previous
1…45678…210
Next →