AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
Safety

Cross-modal Consistency Guidance for Robust Emotion Control in Auto-Regressive TTS Models

DGX agent

arXiv:2510.13293v3 Announce Type: replace Abstract: While Text-to-Speech (TTS) systems enable emotional control via natural-language instructions, expressiveness, naturalness, and speech quality degra

safetyarxiv-cs-cl
20 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

D-CLING: Prior-Preserving Depth-Conditioned Fine-Tuning for Navigation Foundation Models

DGX agent

arXiv:2605.19690v1 Announce Type: new Abstract: Navigation Foundation Models (NFMs) trained on large cross-embodied datasets have demonstrated powerful generalizability in various scenarios. Adopting

safetyarxiv-cs-ro
20 May 2026
Safety

Data-driven Acceleration of MPC with Guarantees

DGX agent

arXiv:2511.13588v2 Announce Type: replace-cross Abstract: Model Predictive Control (MPC) is a powerful framework for optimal control but can be too slow for low-latency applications. We present a data

safetyarxiv-cs-ai
20 May 2026
Safety

DEFLECT: Delay-Robust Execution via Flow-matching Likelihood-Estimated Counterfactual Tuning for VLA Policies

DGX agent

arXiv:2605.19294v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies are typically deployed with asynchronous inference: the robot executes a previously predicted action chunk while

safetyarxiv-cs-ai
20 May 2026
Safety

Dimensional Balance Improves Large Scale Spatiotemporal Prediction Performance

DGX agent

arXiv:2605.18793v1 Announce Type: cross Abstract: Accurate spatiotemporal pattern analysis is critical in fields such as urban traffic, meteorology, and public health monitoring. However, existing met

safetyarxiv-cs-ai
20 May 2026
Safety

Directed Acyclic Graph Convolutional Networks

DGX agent

arXiv:2506.12218v2 Announce Type: replace-cross Abstract: Directed acyclic graphs (DAGs) are central to science and engineering applications including causal inference, scheduling, and neural architec

safetyarxiv-cs-lg
20 May 2026
Safety

Distributional AGI Safety

DGX agent

arXiv:2512.16856v2 Announce Type: replace Abstract: AI safety and alignment research has predominantly been focused on methods for safeguarding individual AI systems, resting on the assumption of an e

safetyarxiv-cs-ai
20 May 2026
Safety

Domain-Adaptive Communication-Rate Optimization for Sim-to-Real Humanoid-Robot Wireless XR Teleoperation

DGX agent

arXiv:2605.19293v1 Announce Type: cross Abstract: Wireless extended reality (XR) teleoperation provides embodied interaction capability for collecting humanoid robot demonstrations, but the large-scal

safetyarxiv-cs-lg
20 May 2026
Safety

Don't Let Bandit Feedback Pull Continual LLM-Recommender Updates Off Target

DGX agent

arXiv:2605.18899v1 Announce Type: cross Abstract: Generative LLM-based recommenders (LLM-Rec) require continual post-deployment updates, yet deployment logs provide only policy-shaped contextual bandi

safetyarxiv-cs-ai
20 May 2026
Safety

Dual-Gated Epistemic Time-Dilation: Autonomous Compute Modulation in Asynchronous MARL

DGX agent

arXiv:2603.23722v2 Announce Type: replace-cross Abstract: While Multi-Agent Reinforcement Learning (MARL) algorithms achieve unprecedented successes across complex continuous domains, their standard d

safetyarxiv-cs-lg
20 May 2026
Safety

DynaSTy: A Framework for SpatioTemporal Node Attribute Prediction in Dynamic Graphs

DGX agent

arXiv:2601.05391v2 Announce Type: replace Abstract: Accurate multistep forecasting of node-level attributes on dynamic graphs is critical for applications ranging from financial trust networks to biol

safetyarxiv-cs-lg
20 May 2026
Safety

DynaTok: Temporally Adaptive and Positional Bias-Aware Token Compression for Video-LLMs

DGX agent

arXiv:2605.19322v1 Announce Type: new Abstract: Recent advances in Video Large Language Models (Video-LLMs) have greatly expanded multimodal reasoning capabilities. However, the massive number of visu

safetyarxiv-cs-cv
20 May 2026
Safety

Efficient Transferable Optimal Transport via Min-Sliced Transport Plans

DGX agent

arXiv:2511.19741v3 Announce Type: replace Abstract: Optimal Transport (OT) offers a powerful framework for finding correspondences between distributions and addressing matching and alignment problems

safetyarxiv-cs-cv
20 May 2026
Safety

ESLD (External Surrogate Latent Defense): A Latent-Space Architecture for Faster, Stronger Prompt-Injection Defense

DGX agent

arXiv:2605.18918v1 Announce Type: cross Abstract: Modern AI assistants are agentic. To answer a single user request, the underlying language model pulls in information from many sources, such as web s

safetyarxiv-cs-ai
20 May 2026
Safety

Exact Linear Attention

DGX agent

arXiv:2605.18848v1 Announce Type: cross Abstract: This paper introduces Exact Linear Attention (ELA), a mechanism that achieves linear computational complexity for Transformer attention by leveraging

safetyarxiv-cs-ai
20 May 2026
Safety

Exploring and Developing a Pre-Model Safeguard with Draft Models

DGX agent

arXiv:2605.19321v1 Announce Type: cross Abstract: Large Language Model (LLM) alignment remains vulnerable to jailbreak attacks that elicit unsafe responses, motivating pre-model and post-model guards.

safetyarxiv-cs-ai
20 May 2026
Safety

Extreme Self-Preference in Language Models

DGX agent

arXiv:2509.26464v2 Announce Type: replace Abstract: Self-preference is a fundamental feature of biological organisms. Since large language models (LLMs) lack sentience, they might be expected to avoid

safetyarxiv-cs-ai
20 May 2026
Safety

Faster-GCG: Efficient Discrete Optimization Jailbreak Attacks against Aligned Large Language Models

DGX agent

arXiv:2410.15362v2 Announce Type: replace-cross Abstract: Aligned Large Language Models (LLMs) have attracted significant attention for their safety, particularly in the context of jailbreak attacks t

safetyarxiv-cs-ai
20 May 2026
Safety

Feature-Space Smoothing: Certified Robustness of Deep Representations

DGX agent

arXiv:2601.16200v3 Announce Type: replace-cross Abstract: Modern deep learning models exhibit strong capabilities across diverse applications, yet remain vulnerable to malicious inputs that induce err

safetyarxiv-cs-cv
20 May 2026
Safety

FlowErase-RL: Rethinking Concept Erasure as Reward Optimization in Flow Matching Models

DGX agent

arXiv:2605.19739v1 Announce Type: new Abstract: Recent advances in flow matching models have significantly improved text-to-image generation quality, but also introduce growing safety risks due to the

safetyarxiv-cs-cv
20 May 2026
Safety

Formal Skill: Programmable Runtime Skills for Efficient and Accurate LLM Agents

DGX agent

arXiv:2605.19604v1 Announce Type: new Abstract: Large Language Model (LLM) agents increasingly act inside real workspaces, where tools and skills determine whether model reasoning becomes reliable act

safetyarxiv-cs-ai
20 May 2026
Safety

From Cumulative Constraints to Adaptive Runtime Safety Control for Nonstationary Reinforcement Learning

DGX agent

arXiv:2605.18841v1 Announce Type: new Abstract: Safety in reinforcement learning is often specified through cumulative cost constraints, but these trajectory-level guarantees do not directly prevent u

safetyarxiv-cs-lg
20 May 2026
Safety

From Refusal to Recovery: A Control-Theoretic Approach to Generative AI Guardrails

DGX agent

arXiv:2510.13727v2 Announce Type: replace Abstract: Generative AI systems are increasingly assisting and acting on behalf of end users in practical settings, from digital shopping assistants to next-g

safetyarxiv-cs-ai
20 May 2026
Safety

GAE Falls Short in Imperfect-Information Self-Play Reinforcement Learning

DGX agent

arXiv:2605.19235v1 Announce Type: new Abstract: Competitive multi-agent reinforcement learning in imperfect-information games requires agents to act under partial observability and against adversarial

safetyarxiv-cs-lg
20 May 2026
Safety

Generative Auto-Bidding with Unified Modeling and Exploration

DGX agent

arXiv:2605.19457v1 Announce Type: new Abstract: Automated bidding is central to modern digital advertising. Early rule-based methods lacked adaptability, while subsequent Reinforcement Learning approa

safetyarxiv-cs-ai
20 May 2026
Safety

Generative-Evaluative Agreement: A Necessary Validity Criterion for LLM-Enabled Adaptive Assessment

DGX agent

arXiv:2605.19529v1 Announce Type: new Abstract: When the same LLM generates assessment items, simulates student responses, and scores them, the validation loop is self-referential. We introduce Genera

safetyarxiv-cs-ai
20 May 2026
Safety

Governing Evolving Memory in LLM Agents: Risks, Mechanisms, and the Stability and Safety Governed Memory (SSGM) Framework

DGX agent

arXiv:2603.11768v2 Announce Type: replace Abstract: Long-term memory has emerged as a foundational component of autonomous Large Language Model (LLM) agents, enabling continuous adaptation, lifelong m

safetyarxiv-cs-ai
20 May 2026
Safety

Graph Neural Planning and Predictive Control for Multi-Robot Communication-Constrained Unlabeled Motion Planning

DGX agent

arXiv:2605.19209v1 Announce Type: new Abstract: The multi-robot unlabeled motion planning problem of concurrently assigning robots to goals and generating safe trajectories is central in many collabor

safetyarxiv-cs-ro
20 May 2026
Safety

Guiding Neuro-Symbolic Scenario Generation with Spatio-Temporal Logic

DGX agent

arXiv:2605.19038v1 Announce Type: cross Abstract: The rapid advancement of autonomous driving (AD) technologies has outpaced the development of robust safety evaluation methods. Conventional testing r

safetyarxiv-cs-lg
20 May 2026
Safety

Hamilton--Jacobi Reachability for Spacecraft Collision Avoidance

DGX agent

arXiv:2605.20138v1 Announce Type: new Abstract: This article presents a Hamilton--Jacobi (HJ) reachability framework for a two--satellite collision avoidance problem operating in the same circular orb

safetyarxiv-cs-ro
20 May 2026
Safety

Hard-Label Black-Box Attacks on 3D Point Clouds

DGX agent

arXiv:2412.00404v2 Announce Type: replace Abstract: With the maturity of depth sensors in various 3D safety-critical applications, 3D point cloud models have been shown to be vulnerable to adversarial

safetyarxiv-cs-cv
20 May 2026
Safety

HeadRank: Decoding-Free Passage Reranking via Preference-Aligned Attention Heads

DGX agent

arXiv:2604.17237v2 Announce Type: replace-cross Abstract: Decoding-free reranking methods that read relevance signals directly from LLM attention weights offer significant latency advantages over auto

safetyarxiv-cs-ai
20 May 2026
Safety

HiDe: Rethinking The Zoom-IN method in High Resolution MLLMs via Hierarchical Decoupling

DGX agent

arXiv:2510.00054v2 Announce Type: replace-cross Abstract: Multimodal Large Language Models (MLLMs) have made significant strides in visual understanding tasks. However, their performance on high-resol

safetyarxiv-cs-ai
20 May 2026
Safety

HOI-PAGE: Zero-Shot Human-Object Interaction Generation with Part Affordance Guidance

DGX agent

arXiv:2506.07209v2 Announce Type: replace-cross Abstract: We present HOI-PAGE, a new approach that prioritizes part-level affordance reasoning to generate high-fidelity 4D human-object interactions (H

safetyarxiv-cs-cv
20 May 2026
Safety

How Do Document Parsers Break? Auditing Structural Vulnerability in Document Intelligence

DGX agent

arXiv:2605.19309v1 Announce Type: new Abstract: Document Layout Analysis (DLA) pipelines provide structured page representations for retrieval-augmented generation, long-document question answering, a

safetyarxiv-cs-cl
20 May 2026
Safety

How does longer temporal context enhance multimodal narrative video processing in the brain?

DGX agent

arXiv:2602.07570v2 Announce Type: replace-cross Abstract: Understanding how humans and artificial intelligence systems process complex narrative videos is a fundamental challenge at the intersection o

safetyarxiv-cs-ai
20 May 2026
Safety

How Does Overparameterization Affect Machine Unlearning of Deep Neural Networks?

DGX agent

arXiv:2503.08633v2 Announce Type: replace Abstract: Machine unlearning is the task of updating a trained model to forget specific training data without retraining from scratch. In this paper, we inves

safetyarxiv-cs-lg
20 May 2026
Safety

Implicit Action Chunking for Smooth Continuous Control

DGX agent

arXiv:2605.19592v1 Announce Type: cross Abstract: Reinforcement learning often produces high-frequency oscillatory control signals that undermine the safety and stability required for physical deploym

safetyarxiv-cs-ai
20 May 2026
Safety

Implicit Bias of Mirror Flow in Homogeneous Neural Networks: Sparse and Dense Feature Learning

DGX agent

arXiv:2605.19458v1 Announce Type: new Abstract: We study the max-margin solutions reached by mirror flow in deep neural networks with homogeneous activation functions. Extending classical results on g

safetyarxiv-cs-lg
20 May 2026
Safety

Improved visual-information-driven model for crowd simulation and its modular application

DGX agent

arXiv:2504.03758v4 Announce Type: replace-cross Abstract: Crowd movement simulation is crucial for pedestrian safety management and facility design. Data-driven models offer the potential to improve r

safetyarxiv-cs-cv
20 May 2026
Safety

Increasing Missingness to Reduce Bias: Richardson-SGD with Missing Data

DGX agent

arXiv:2605.19641v1 Announce Type: cross Abstract: Stochastic gradient methods are central to modern large-scale learning, but their use with incomplete covariates remains delicate since imputation sch

safetyarxiv-cs-lg
20 May 2026
Safety

Interoceptive Divergence in Aesthetic Evaluation and Implications for Human-AI Alignment

DGX agent

arXiv:2605.18759v1 Announce Type: cross Abstract: Artificial intelligence (AI), exemplified by large language models (LLMs), is rapidly approaching and in some cases surpassing human performance acros

safetyarxiv-cs-ai
20 May 2026
Safety

Inverse Design of Metasurface based Absorbers using Physics Guided Conditional Diffusion Models

DGX agent

arXiv:2605.19611v1 Announce Type: new Abstract: Inverse design of metasurfaces for specific electromagnetic responses requires generating geometries that satisfy stringent spectral constraints while m

safetyarxiv-cs-cv
20 May 2026
Safety

Jailbreaking on Text-to-Video Models via Scene Splitting Strategy

DGX agent

arXiv:2509.22292v2 Announce Type: replace-cross Abstract: Along with the rapid advancement of numerous Text-to-Video (T2V) models, growing concerns have emerged regarding their safety risks. While rec

safetyarxiv-cs-ai
20 May 2026
Safety

k-Inductive Neural Barrier Certificates for Unknown Nonlinear Dynamics

DGX agent

arXiv:2605.20108v1 Announce Type: cross Abstract: While conventional (k=1) discrete-time barrier certificate conditions impose strict safety constraints by requiring the function to be non-increasing

safetyarxiv-cs-ai
20 May 2026
Safety

KG-ASG: Collision-Knowledge-Guided Closed-Loop Adversarial Scenario Generation With Primary-Support Attribution

DGX agent

arXiv:2605.18895v1 Announce Type: cross Abstract: Safety validation of autonomous driving systems requires high-risk scenario coverage, clear collision semantics, executable trajectories, and attribut

safetyarxiv-cs-ai
20 May 2026
Safety

Kickstarter retracts stricter rules on mature content after creator backlash, and says it adopted the tougher rules because of its payment processor Stripe (Mariella Moon/Engadget)

DGX agent

Mariella Moon / Engadget: Kickstarter retracts stricter rules on mature content after creator backlash, and says it adopted the tougher rules because of its payment processor Stripe — It explained tha

safetytechmeme
20 May 2026
Safety

Knowing When Not to Predict: Self Supervised Learning and Abstention for Safer DR Screening

DGX agent

arXiv:2605.19133v1 Announce Type: cross Abstract: Self-supervised learning (SSL) is now a standard way to pretrain medical image models, but performance is still mostly judged by downstream accuracy.

safetyarxiv-cs-ai
20 May 2026
← Previous
1…161162163164165…267
Next →