AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,596 results
Safety

MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction

DGX agent

arXiv:2608.10562v1 Announce Type: new Abstract: Not all clicks are equal. Industrial ads ranking decouples conversion probability into click-through rate (CTR) and post-click conversion rate (CVR), ye

safetyarxiv-cs-lg
12 Aug 2026
Safety
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

DGX agent

arXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against

safetyarxiv-cs-ai
12 Aug 2026
Safety

MedUP: Awakening Unified Understanding and Perception in Medical Vision-Language Models

DGX agent

arXiv:2608.10635v1 Announce Type: cross Abstract: Medical Vision-Language Models (Med-VLMs) excel at verbalizing visual content, yet precise visual perception, segmentation, and grounding remain chall

safetyarxiv-cs-ai
12 Aug 2026
Safety

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

DGX agent

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured

safetyarxiv-cs-lg
12 Aug 2026
Safety

MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis

DGX agent

arXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounte

safetyarxiv-cs-ai
12 Aug 2026
Safety

Mitigating Bus Bunching with Reinforcement Learning Enhanced by Semantic Stop Embedding

DGX agent

arXiv:2608.10207v1 Announce Type: new Abstract: Bus bunching degrades service regularity and increases passenger waiting in high-frequency transit. Existing reinforcement-learning-based holding contro

safetyarxiv-cs-ai
12 Aug 2026
Safety

Most biomedical publications show signs of LLM-assisted writing

DGX agent

arXiv:2608.10715v1 Announce Type: cross Abstract: Over the past several years, LLM-powered chatbots and agents have become widely used as a tool for academic writing. LLM-assisted writing can be valua

safetyarxiv-cs-ai
12 Aug 2026
Safety

Multi-View Relational Distillation for Spatial Reasoning with Vision-Language Models

DGX agent

arXiv:2608.10864v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong image and video understanding, yet their visual-spatial representations remain geometrically fragile,

safetyarxiv-cs-cv
12 Aug 2026
Safety

Multiplayer Nash Preference Optimization

DGX agent

arXiv:2509.23102v4 Announce Type: replace Abstract: Reinforcement learning from human feedback (RLHF) has emerged as the standard paradigm for aligning large language models with human preferences. Ho

safetyarxiv-cs-ai
12 Aug 2026
Safety

Navigating the Proximity-Safety Balance: Constraint Decomposition for Human Following in Pedestrian Crowds

DGX agent

arXiv:2608.10056v1 Announce Type: cross Abstract: Following a target human in crowded environments involves an inherent conflict between staying close to the target and navigating safely among surroun

safetyarxiv-cs-ai
12 Aug 2026
Safety

Never Stop Speaking: a Denial-of-Service Attack on End-to-End Speech Language Models

DGX agent

arXiv:2608.10405v1 Announce Type: cross Abstract: Many studies have shown that specially crafted inputs can induce large language models (LLMs) to generate excessively long outputs, resulting in signi

safetyarxiv-cs-ai
12 Aug 2026
Safety

Observational Policy Ranking for SMB Financial Guidance from Multi-Action Accounting Logs

DGX agent

arXiv:2608.10050v1 Announce Type: new Abstract: Small and medium-sized businesses need timely financial guidance, yet historical accounting logs record self-selected and often co-occurring business ch

safetyarxiv-cs-lg
12 Aug 2026
Safety

OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation

DGX agent

arXiv:2603.19201v3 Announce Type: replace Abstract: Contact-rich manipulation tasks, such as wiping and assembly, require accurate perception of contact forces, friction changes, and state transitions

safetyarxiv-cs-ro
12 Aug 2026
Safety

On the Importance of Geometric Nonlinearity and Temperature-Dependent Properties in Multi-Material Thermo-Mechanical Topology Optimization

DGX agent

arXiv:2608.10344v1 Announce Type: cross Abstract: Thermo-mechanical compliant devices are commonly designed with small-strain linear elasticity and temperature-independent material properties, even th

safetyarxiv-cs-lg
12 Aug 2026
Safety

On The Statistical Limits of Self-Improving Agents

DGX agent

arXiv:2510.04399v3 Announce Type: replace Abstract: We develop a learning-theoretic framework for analyzing self-improving agents by decomposing self-modification into five axes. Within this framework

safetyarxiv-cs-ai
12 Aug 2026
Safety

OpenAI VP of Global Policy Ann O'Leary says AI policy in the US is anchored in the states and informed by California's AI transparency law passed in 2025 (Chase DiFeliciantonio/Politico)

DGX agent

Chase DiFeliciantonio / Politico: OpenAI VP of Global Policy Ann O'Leary says AI policy in the US is anchored in the states and informed by California's AI transparency law passed in 2025 — SACRAMENTO

safetytechmeme
12 Aug 2026
Safety

Operationalising Relative Causal Knowledge: Backbone Identifiability from Private Reports on a Shared Outcome

DGX agent

arXiv:2608.10664v1 Announce Type: new Abstract: The Relativity of Causal Knowledge (RCK) explains how a network of agents with different structural causal models can exchange causal knowledge through

safetyarxiv-cs-ai
12 Aug 2026
Safety

Pair-Centric Graph Rewiring for Over-Squashing via Optimal Transport-Guided Communication Alignment

DGX agent

arXiv:2608.10619v1 Announce Type: new Abstract: Message-passing neural networks (MPNNs) often struggle when task-relevant information is distributed across distant regions of a graph, since local prop

safetyarxiv-cs-lg
12 Aug 2026
Safety

Partially Observable Learning for Multi-Platform Dispatch Optimization

DGX agent

arXiv:2608.10897v1 Announce Type: new Abstract: Instant delivery platforms have become a critical component of urban logistics, increasingly relying on crowdsourced couriers to fulfill highly dynamic

safetyarxiv-cs-lg
12 Aug 2026
Safety

Physics-Informed Machine Learning in Prognostics and Health Management: A Systematic Literature Review

DGX agent

arXiv:2608.10047v1 Announce Type: cross Abstract: In modern industry, keeping complex systems reliable, safe, and efficient hinges on Prognostics and Health Management (PHM). Machine Learning (ML) has

safetyarxiv-cs-ai
12 Aug 2026
Safety

Policy Convergence and Divergence Across National and Within Regional AI Strategies: A Policy Design Element Analysis

DGX agent

arXiv:2608.11006v1 Announce Type: cross Abstract: Governments worldwide have responded to the rapid expansion of AI by publishing national and regional AI strategies. Comparing national and regional A

safetyarxiv-cs-ai
12 Aug 2026
Safety

Position Encoding in Transformers: From Absolute and Relative Methods to Rotary Position Embeddings and Long-Context Scaling

DGX agent

arXiv:2608.10021v1 Announce Type: new Abstract: Self-attention models content-dependent interactions between tokens but does not by itself encode token order. Position encoding addresses this limitati

safetyarxiv-cs-cl
12 Aug 2026
Safety

Predicting Space Groups of Double Perovskites by LLM with Dynamic Few-Shot Learning

DGX agent

arXiv:2608.10483v1 Announce Type: new Abstract: Double perovskites (DPs) offer broad compositional tunability, but predicting the space groups (SGs) of stable structures remains difficult because avai

safetyarxiv-cs-ai
12 Aug 2026
Safety

ProbGuard: Calibrated Safety Risk Estimation from LLM Output Distributions

DGX agent

arXiv:2608.10621v1 Announce Type: new Abstract: Recent research on Large Language Model (LLM) safety has widely adopted guardrails to identify unsafe LLM outputs. Existing guardrails typically formula

safetyarxiv-cs-lg
12 Aug 2026
Safety

Procedural Fairness Failures in RLHF from Preference Averaging

DGX agent

arXiv:2608.10126v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) aggregates heterogeneous preferences into a single reward model, assuming preference homogeneity. Wh

safetyarxiv-cs-ai
12 Aug 2026
Safety

ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation

DGX agent

arXiv:2608.10905v1 Announce Type: new Abstract: On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Exi

safetyarxiv-cs-lg
12 Aug 2026
Safety

Rethinking Data Efficiency in Industrial Dense Prediction: Pretraining Coherence, Not Inductive Bias, Determines ViTs Low-Data Advantage

DGX agent

arXiv:2608.10590v1 Announce Type: new Abstract: Vision Transformers (ViTs) are widely believed to require more labeled data than CNNs for industrial dense prediction. Through controlled experiments on

safetyarxiv-cs-cv
12 Aug 2026
Safety

Robust Safety Filtering for Input-Constrained Underactuated Linear Systems

DGX agent

arXiv:2608.10872v1 Announce Type: cross Abstract: We present a robust safety-filtering framework for input-constrained underactuated linear systems subject to unknown disturbances. A baseline H-infty

safetyarxiv-cs-ro
12 Aug 2026
Safety

Robust Sliding Mode and Admittance Control of Underactuated Aerial Manipulators for Contact-Based Inspection

DGX agent

arXiv:2608.10656v1 Announce Type: new Abstract: Contact-based industrial inspection requires aerial platforms to maintain stable interaction while rejecting disturbances. Underactuated aerial manipula

safetyarxiv-cs-ro
12 Aug 2026
Safety

SafeCap: Improving LVLM Safety with Image Captioning Reinforcement Learning

DGX agent

arXiv:2608.10513v1 Announce Type: cross Abstract: Large vision-language models (LVLMs) remain vulnerable to jailbreak attacks that exploit visual inputs to bypass safety alignment inherited from their

safetyarxiv-cs-ai
12 Aug 2026
Safety

SapiensID 2.0: Aligning Human Recognition Foundation Models with Human Perception

DGX agent

arXiv:2608.10497v1 Announce Type: new Abstract: While foundation models have significantly advanced human recognition across diverse modalities, they predominantly rely on static, geometric feature ex

safetyarxiv-cs-cv
12 Aug 2026
Safety

SBCO: Self-Supervised, Verifier-Grounded Harness Optimization For Planning Agents

DGX agent

arXiv:2608.10157v1 Announce Type: new Abstract: Self-improving agents seek to reduce the human engineering effort behind AI systems by enabling them to evolve and self-improve their performance over t

safetyarxiv-cs-ai
12 Aug 2026
Safety

Scheduling Mixed RL Rollouts Beyond Prefix Locality

DGX agent

arXiv:2608.11152v1 Announce Type: cross Abstract: Modern reinforcement learning (RL) post-training pipelines for large language models (LLMs) increasingly combine rollout workloads across multiple dom

safetyarxiv-cs-lg
12 Aug 2026
Safety

SeFaR: Semantic Feature-aware Robustness Testing of Deep Neural Networks

DGX agent

arXiv:2608.10289v1 Announce Type: new Abstract: Deep neural networks are increasingly deployed in safety-critical domains as perception modules, where failures are often caused due to rare and under-r

safetyarxiv-cs-cv
12 Aug 2026
Safety

Selective Prediction Reduces the Negative Effects of Automation Bias Overall but Increases False Negatives

DGX agent

arXiv:2508.07617v2 Announce Type: replace-cross Abstract: AI has the potential to augment human decision making. However, even high-performing models can produce inaccurate predictions when deployed.

safetyarxiv-cs-ai
12 Aug 2026
Safety

Self-Normalized Inference for Constant-Stepsize Temporal-Difference Learning under Markovian Sampling

DGX agent

arXiv:2608.10896v1 Announce Type: cross Abstract: Constant-stepsize temporal-difference (TD) learning is attractive for policy evaluation, but inference from a single Markov trajectory must account fo

safetyarxiv-cs-lg
12 Aug 2026
Safety

Similarity Gates Approve Reversals: A Validity Audit of Embedding-Cosine Thresholds in Agent Systems

DGX agent

arXiv:2608.10216v1 Announce Type: cross Abstract: Agent frameworks ship quality gates that compare text blocks by embedding-cosine similarity and decide at a fixed cutoff. Deduplication filters, seman

safetyarxiv-cs-ai
12 Aug 2026
Safety

Smart Enough to Go Extinct? An Evolutionary Challenge to the Value of General Intelligence and Its Ethical Implications for AGI

DGX agent

arXiv:2608.10730v1 Announce Type: cross Abstract: The pursuit of artificial general intelligence (AGI) rests on a seemingly self-evident premise: that general intelligence, the kind of flexible, domai

safetyarxiv-cs-ai
12 Aug 2026
Safety

SPOTting the Future: Lookahead Explanations for Deep Reinforcement Learning

DGX agent

arXiv:2608.09967v1 Announce Type: new Abstract: Deep reinforcement learning (DRL) agents achieve strong performance in complex environments, yet their decision-making processes remain difficult to int

safetyarxiv-cs-ai
12 Aug 2026
Safety

Status Association Does Not Reliably Predict Decision Leakage

DGX agent

arXiv:2608.10089v1 Announce Type: cross Abstract: Bias evaluations often move too quickly from evidence that a model encodes a social association to claims that the same association will alter consequ

safetyarxiv-cs-ai
12 Aug 2026
Safety

Stay or Stray - A Dynamical Systems Viewpoint of Popularity Bias

DGX agent

arXiv:2608.10474v1 Announce Type: cross Abstract: Popularity bias in recommendation systems arises when a majority user class generates disproportionate interaction data, causing the system to increas

safetyarxiv-cs-lg
12 Aug 2026
Safety

Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning

DGX agent

arXiv:2608.11204v1 Announce Type: cross Abstract: Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g.,

safetyarxiv-cs-ai
12 Aug 2026
Safety

TCAM for Autonomous Deformable Manipulation: The RMC2 Champion System for WBCD 2026 Track 4

DGX agent

arXiv:2608.10718v1 Announce Type: new Abstract: This technical report describes the RMC2 Team's champion solution for the WBCD 2026 Track 4: Deformable Manipulation Challenge. The task requires a robo

safetyarxiv-cs-ro
12 Aug 2026
Safety

Templated or fully Synthetic? Prompt construction as a confound in measuring LLM political stance beyond writing assistance

DGX agent

arXiv:2608.11008v1 Announce Type: new Abstract: Political stance detection in LLMs has long been dominated by closed-ended, multiple-choice political survey questions---originally designed for humans,

safetyarxiv-cs-cl
12 Aug 2026
Safety

The CASE Framework: A Multi-Disciplinary Control Architecture for Governing Enterprise Agentic AI

DGX agent

arXiv:2608.10153v1 Announce Type: new Abstract: Enterprises are deploying autonomous AI agents faster than they can govern them, and prevailing approaches stretch a single discipline, typically DevSec

safetyarxiv-cs-ai
12 Aug 2026
Safety

The GENEA Challenge 2026: A Large-Scale Disentangled Evaluation of Speech-Driven Gesture Generation on the Seamless Interaction Dataset

DGX agent

arXiv:2608.10839v1 Announce Type: new Abstract: This preprint presents the results of the fourth GENEA Challenge, a large-scale human evaluation of five speech-driven gesture-generation systems traine

safetyarxiv-cs-cv
12 Aug 2026
Safety

The Hidden Puppet Master: Predicting Human Belief Change in Manipulative LLM Dialogues

DGX agent

arXiv:2603.20907v5 Announce Type: replace Abstract: As users increasingly turn to LLMs for practical and personal advice, they become vulnerable to subtle steering toward hidden incentives misaligned

safetyarxiv-cs-cl
12 Aug 2026
Safety

The Impact of Operational-Data Fidelity when Assessing Safety-Critical Autonomous-Vehicle Software

DGX agent

arXiv:2608.10025v1 Announce Type: new Abstract: For safety-critical software, data from the software's operational past (e.g. a sequence of success and failure events experienced by the software) can

safetyarxiv-cs-ro
12 Aug 2026
← Previous
1234…263
Next →