AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
Safety

Every Token Counts: Exact Likert-Scale Distributions for Measuring LLM Attitudes and Biases

DGX agent

arXiv:2608.10503v1 Announce Type: new Abstract: As Large Language Models (LLMs) are increasingly deployed as autonomous agents, accurately evaluating their latent values and biases is critical. The NL

safetyarxiv-cs-cl
12 Aug 2026
Safety
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

FedCGR: Federated Cross-Domain Generative Recommendation

DGX agent

arXiv:2608.10929v1 Announce Type: new Abstract: Cross-domain recommendation (CDR) transfers preference knowledge across related domains, but federated deployment makes cross-domain alignment difficult

safetyarxiv-cs-ai
12 Aug 2026
Safety

FoR-SALE: Frame of Reference-guided Spatial Adjustment in LLM-based Diffusion Editing

DGX agent

arXiv:2509.23452v2 Announce Type: replace-cross Abstract: Current text-to-image generation models, even state-of-the-art models, exhibit a significant performance gap when spatial expressions are desc

safetyarxiv-cs-cl
12 Aug 2026
Safety

From Prediction to Incrementality: Causal Optimization for Large-Scale Targeting and Recommendation

DGX agent

arXiv:2608.10182v1 Announce Type: cross Abstract: Large-scale targeting and recommendation systems are typically built around predictive scores fed into heuristic or local allocation. When the busines

safetyarxiv-cs-ai
12 Aug 2026
Safety

Generation-Step-Aware Framework for Cross-Modal Representation and Control in Multilingual Speech-Text Models

DGX agent

arXiv:2601.17387v3 Announce Type: replace Abstract: Multilingual speech-text models rely on cross-modal language alignment to transfer knowledge between speech and text, but it remains unclear whether

safetyarxiv-cs-cl
12 Aug 2026
Safety

Hierarchical Empirical-Bayes Naive Bayes: Minimax Smoothing and Calibration with AODE Extension

DGX agent

arXiv:2608.11162v1 Announce Type: new Abstract: The Naive Bayes (NB) classifier remains a standard choice for categorical data, yet its widely used smoothing rules, such as Laplace, Lidstone, Krichevs

safetyarxiv-cs-lg
12 Aug 2026
Safety

Hip Energized Monopedal Hopping

DGX agent

arXiv:2608.10387v1 Announce Type: new Abstract: We present a novel stepping strategy for pitch unlocked planar monopeds where the reaction torques from stabilizing pitch with a conventional PD + feedf

safetyarxiv-cs-ro
12 Aug 2026
Safety

IADD-TR: Intervention-Aware Dynamics Decoupling with Targeted Regularization for Model-Based Reinforcement Learning

DGX agent

arXiv:2608.10634v1 Announce Type: new Abstract: Model-based reinforcement learning (MBRL), which learns environment dynamics to generate synthetic experience, is a promising approach to sample-efficie

safetyarxiv-cs-lg
12 Aug 2026
Safety

INSIDE the Student's Mind: Jointly Modeling Latent Reasoning and Action in LLM Student Simulators

DGX agent

arXiv:2608.10492v1 Announce Type: new Abstract: Large Language Model (LLM)-based simulators often reproduce observable actions but fail to capture the underlying reasoning behind them. In education, w

safetyarxiv-cs-ai
12 Aug 2026
Safety

IO Factory: Simulating AI-Enabled Influence Campaigns at Scale

DGX agent

arXiv:2608.10920v1 Announce Type: new Abstract: We introduce IO Factory, an AI-driven framework for simulating information and influence campaigns as fully integrated, traceable processes. The threat

safetyarxiv-cs-ai
12 Aug 2026
Safety

Learning in ImaginationLand: Omnidirectional Policies through 3D Generative Models (OP-Gen)

DGX agent

arXiv:2509.06191v2 Announce Type: replace-cross Abstract: Recent 3D generative models, which are capable of generating full object shapes from just a few images, now open up new opportunities in robot

safetyarxiv-cs-cv
12 Aug 2026
Safety

Leveraging Large Language Models for Causal Discovery: a Constraint-based, Argumentation-driven Approach

DGX agent

arXiv:2602.16481v2 Announce Type: replace Abstract: Causal discovery seeks to uncover causal relations from data, typically represented as causal graphs, and is essential for predicting the effects of

safetyarxiv-cs-ai
12 Aug 2026
Safety

LoRCA: LoRA Cycle Adaptation for Histology to HiP-CT Translation with DINOv3

DGX agent

arXiv:2608.10002v1 Announce Type: cross Abstract: Hierarchical Phase-Contrast Tomography (HiP-CT) is a synchrotron based X-ray imaging technique that enables non-destructive, volumetric imaging of int

safetyarxiv-cs-cv
12 Aug 2026
Safety

MARCO: Click-Intent Decomposition for Calibrated Ads Conversion Prediction

DGX agent

arXiv:2608.10562v1 Announce Type: new Abstract: Not all clicks are equal. Industrial ads ranking decouples conversion probability into click-through rate (CTR) and post-click conversion rate (CVR), ye

safetyarxiv-cs-lg
12 Aug 2026
Safety

MarkNull: Model-Agnostic Watermark Removal in AI-Generated Images via On-Manifold Latent Manipulation

DGX agent

arXiv:2608.10166v1 Announce Type: cross Abstract: Digital watermarking has emerged as a critical technique for provenance and copyright attribution in AI-generated imagery, yet its robustness against

safetyarxiv-cs-ai
12 Aug 2026
Safety

MERA: Model Evolution and Routing with Skill Adaptation for Agentic Systems at Scale

DGX agent

arXiv:2608.10333v1 Announce Type: new Abstract: LLM agents execute heterogeneous sequences of model calls within a single task: some invocations require careful reasoning, while others are structured

safetyarxiv-cs-lg
12 Aug 2026
Safety

MIDAS: Mutual Information Disentanglement with Uncertainty-Aware Fusion for Incomplete Multimodal Sentiment Analysis

DGX agent

arXiv:2608.09986v1 Announce Type: new Abstract: Most existing multimodal sentiment analysis approaches assume access to complete multimodal inputs. However, real-world applications frequently encounte

safetyarxiv-cs-ai
12 Aug 2026
Safety

Mitigating Bus Bunching with Reinforcement Learning Enhanced by Semantic Stop Embedding

DGX agent

arXiv:2608.10207v1 Announce Type: new Abstract: Bus bunching degrades service regularity and increases passenger waiting in high-frequency transit. Existing reinforcement-learning-based holding contro

safetyarxiv-cs-ai
12 Aug 2026
Safety

Most biomedical publications show signs of LLM-assisted writing

DGX agent

arXiv:2608.10715v1 Announce Type: cross Abstract: Over the past several years, LLM-powered chatbots and agents have become widely used as a tool for academic writing. LLM-assisted writing can be valua

safetyarxiv-cs-ai
12 Aug 2026
Safety

Multi-View Relational Distillation for Spatial Reasoning with Vision-Language Models

DGX agent

arXiv:2608.10864v1 Announce Type: new Abstract: Vision-language models (VLMs) have achieved strong image and video understanding, yet their visual-spatial representations remain geometrically fragile,

safetyarxiv-cs-cv
12 Aug 2026
Safety

Multiplayer Nash Preference Optimization

DGX agent

arXiv:2509.23102v4 Announce Type: replace Abstract: Reinforcement learning from human feedback (RLHF) has emerged as the standard paradigm for aligning large language models with human preferences. Ho

safetyarxiv-cs-ai
12 Aug 2026
Safety

Observational Policy Ranking for SMB Financial Guidance from Multi-Action Accounting Logs

DGX agent

arXiv:2608.10050v1 Announce Type: new Abstract: Small and medium-sized businesses need timely financial guidance, yet historical accounting logs record self-selected and often co-occurring business ch

safetyarxiv-cs-lg
12 Aug 2026
Safety

OmniVTA: Visuo-Tactile World Modeling for Contact-Rich Robotic Manipulation

DGX agent

arXiv:2603.19201v3 Announce Type: replace Abstract: Contact-rich manipulation tasks, such as wiping and assembly, require accurate perception of contact forces, friction changes, and state transitions

safetyarxiv-cs-ro
12 Aug 2026
Safety

On the Importance of Geometric Nonlinearity and Temperature-Dependent Properties in Multi-Material Thermo-Mechanical Topology Optimization

DGX agent

arXiv:2608.10344v1 Announce Type: cross Abstract: Thermo-mechanical compliant devices are commonly designed with small-strain linear elasticity and temperature-independent material properties, even th

safetyarxiv-cs-lg
12 Aug 2026
Safety

On The Statistical Limits of Self-Improving Agents

DGX agent

arXiv:2510.04399v3 Announce Type: replace Abstract: We develop a learning-theoretic framework for analyzing self-improving agents by decomposing self-modification into five axes. Within this framework

safetyarxiv-cs-ai
12 Aug 2026
Safety

OpenAI VP of Global Policy Ann O'Leary says AI policy in the US is anchored in the states and informed by California's AI transparency law passed in 2025 (Chase DiFeliciantonio/Politico)

DGX agent

Chase DiFeliciantonio / Politico: OpenAI VP of Global Policy Ann O'Leary says AI policy in the US is anchored in the states and informed by California's AI transparency law passed in 2025 — SACRAMENTO

safetytechmeme
12 Aug 2026
Safety

Operationalising Relative Causal Knowledge: Backbone Identifiability from Private Reports on a Shared Outcome

DGX agent

arXiv:2608.10664v1 Announce Type: new Abstract: The Relativity of Causal Knowledge (RCK) explains how a network of agents with different structural causal models can exchange causal knowledge through

safetyarxiv-cs-ai
12 Aug 2026
Safety

Pair-Centric Graph Rewiring for Over-Squashing via Optimal Transport-Guided Communication Alignment

DGX agent

arXiv:2608.10619v1 Announce Type: new Abstract: Message-passing neural networks (MPNNs) often struggle when task-relevant information is distributed across distant regions of a graph, since local prop

safetyarxiv-cs-lg
12 Aug 2026
Safety

Partially Observable Learning for Multi-Platform Dispatch Optimization

DGX agent

arXiv:2608.10897v1 Announce Type: new Abstract: Instant delivery platforms have become a critical component of urban logistics, increasingly relying on crowdsourced couriers to fulfill highly dynamic

safetyarxiv-cs-lg
12 Aug 2026
Safety

Physics-Informed Machine Learning in Prognostics and Health Management: A Systematic Literature Review

DGX agent

arXiv:2608.10047v1 Announce Type: cross Abstract: In modern industry, keeping complex systems reliable, safe, and efficient hinges on Prognostics and Health Management (PHM). Machine Learning (ML) has

safetyarxiv-cs-ai
12 Aug 2026
Safety

Policy Convergence and Divergence Across National and Within Regional AI Strategies: A Policy Design Element Analysis

DGX agent

arXiv:2608.11006v1 Announce Type: cross Abstract: Governments worldwide have responded to the rapid expansion of AI by publishing national and regional AI strategies. Comparing national and regional A

safetyarxiv-cs-ai
12 Aug 2026
Safety

Position Encoding in Transformers: From Absolute and Relative Methods to Rotary Position Embeddings and Long-Context Scaling

DGX agent

arXiv:2608.10021v1 Announce Type: new Abstract: Self-attention models content-dependent interactions between tokens but does not by itself encode token order. Position encoding addresses this limitati

safetyarxiv-cs-cl
12 Aug 2026
Safety

Predicting Space Groups of Double Perovskites by LLM with Dynamic Few-Shot Learning

DGX agent

arXiv:2608.10483v1 Announce Type: new Abstract: Double perovskites (DPs) offer broad compositional tunability, but predicting the space groups (SGs) of stable structures remains difficult because avai

safetyarxiv-cs-ai
12 Aug 2026
Safety

Procedural Fairness Failures in RLHF from Preference Averaging

DGX agent

arXiv:2608.10126v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) aggregates heterogeneous preferences into a single reward model, assuming preference homogeneity. Wh

safetyarxiv-cs-ai
12 Aug 2026
Safety

ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation

DGX agent

arXiv:2608.10905v1 Announce Type: new Abstract: On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Exi

safetyarxiv-cs-lg
12 Aug 2026
Safety

Robust Sliding Mode and Admittance Control of Underactuated Aerial Manipulators for Contact-Based Inspection

DGX agent

arXiv:2608.10656v1 Announce Type: new Abstract: Contact-based industrial inspection requires aerial platforms to maintain stable interaction while rejecting disturbances. Underactuated aerial manipula

safetyarxiv-cs-ro
12 Aug 2026
Safety

SapiensID 2.0: Aligning Human Recognition Foundation Models with Human Perception

DGX agent

arXiv:2608.10497v1 Announce Type: new Abstract: While foundation models have significantly advanced human recognition across diverse modalities, they predominantly rely on static, geometric feature ex

safetyarxiv-cs-cv
12 Aug 2026
Safety

SBCO: Self-Supervised, Verifier-Grounded Harness Optimization For Planning Agents

DGX agent

arXiv:2608.10157v1 Announce Type: new Abstract: Self-improving agents seek to reduce the human engineering effort behind AI systems by enabling them to evolve and self-improve their performance over t

safetyarxiv-cs-ai
12 Aug 2026
Safety

Scheduling Mixed RL Rollouts Beyond Prefix Locality

DGX agent

arXiv:2608.11152v1 Announce Type: cross Abstract: Modern reinforcement learning (RL) post-training pipelines for large language models (LLMs) increasingly combine rollout workloads across multiple dom

safetyarxiv-cs-lg
12 Aug 2026
Safety

Selective Prediction Reduces the Negative Effects of Automation Bias Overall but Increases False Negatives

DGX agent

arXiv:2508.07617v2 Announce Type: replace-cross Abstract: AI has the potential to augment human decision making. However, even high-performing models can produce inaccurate predictions when deployed.

safetyarxiv-cs-ai
12 Aug 2026
Safety

Self-Normalized Inference for Constant-Stepsize Temporal-Difference Learning under Markovian Sampling

DGX agent

arXiv:2608.10896v1 Announce Type: cross Abstract: Constant-stepsize temporal-difference (TD) learning is attractive for policy evaluation, but inference from a single Markov trajectory must account fo

safetyarxiv-cs-lg
12 Aug 2026
Safety

Status Association Does Not Reliably Predict Decision Leakage

DGX agent

arXiv:2608.10089v1 Announce Type: cross Abstract: Bias evaluations often move too quickly from evidence that a model encodes a social association to claims that the same association will alter consequ

safetyarxiv-cs-ai
12 Aug 2026
Safety

Stay or Stray - A Dynamical Systems Viewpoint of Popularity Bias

DGX agent

arXiv:2608.10474v1 Announce Type: cross Abstract: Popularity bias in recommendation systems arises when a majority user class generates disproportionate interaction data, causing the system to increas

safetyarxiv-cs-lg
12 Aug 2026
Safety

Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning

DGX agent

arXiv:2608.11204v1 Announce Type: cross Abstract: Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g.,

safetyarxiv-cs-ai
12 Aug 2026
Safety

TCAM for Autonomous Deformable Manipulation: The RMC2 Champion System for WBCD 2026 Track 4

DGX agent

arXiv:2608.10718v1 Announce Type: new Abstract: This technical report describes the RMC2 Team's champion solution for the WBCD 2026 Track 4: Deformable Manipulation Challenge. The task requires a robo

safetyarxiv-cs-ro
12 Aug 2026
Safety

Templated or fully Synthetic? Prompt construction as a confound in measuring LLM political stance beyond writing assistance

DGX agent

arXiv:2608.11008v1 Announce Type: new Abstract: Political stance detection in LLMs has long been dominated by closed-ended, multiple-choice political survey questions---originally designed for humans,

safetyarxiv-cs-cl
12 Aug 2026
Safety

The GENEA Challenge 2026: A Large-Scale Disentangled Evaluation of Speech-Driven Gesture Generation on the Seamless Interaction Dataset

DGX agent

arXiv:2608.10839v1 Announce Type: new Abstract: This preprint presents the results of the fourth GENEA Challenge, a large-scale human evaluation of five speech-driven gesture-generation systems traine

safetyarxiv-cs-cv
12 Aug 2026
Safety

The Parser Already Knows: Lightweight Bias Correction in Constrained Decoding

DGX agent

arXiv:2608.10137v1 Announce Type: new Abstract: Grammar Constrained Decoding (GCD) forces Language Models (LMs) to produce syntactically valid outputs by masking out non-conforming tokens at each step

safetyarxiv-cs-cl
12 Aug 2026
← Previous
1…6869707172…302
Next →