AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,314 results
Safety

Policy Convergence and Divergence Across National and Within Regional AI Strategies: A Policy Design Element Analysis

DGX agent

arXiv:2608.11006v1 Announce Type: cross Abstract: Governments worldwide have responded to the rapid expansion of AI by publishing national and regional AI strategies. Comparing national and regional A

safetyarxiv-cs-ai
12 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Position Encoding in Transformers: From Absolute and Relative Methods to Rotary Position Embeddings and Long-Context Scaling

DGX agent

arXiv:2608.10021v1 Announce Type: new Abstract: Self-attention models content-dependent interactions between tokens but does not by itself encode token order. Position encoding addresses this limitati

safetyarxiv-cs-cl
12 Aug 2026
Safety

Predicting Space Groups of Double Perovskites by LLM with Dynamic Few-Shot Learning

DGX agent

arXiv:2608.10483v1 Announce Type: new Abstract: Double perovskites (DPs) offer broad compositional tunability, but predicting the space groups (SGs) of stable structures remains difficult because avai

safetyarxiv-cs-ai
12 Aug 2026
Safety

Procedural Fairness Failures in RLHF from Preference Averaging

DGX agent

arXiv:2608.10126v1 Announce Type: cross Abstract: Reinforcement Learning from Human Feedback (RLHF) aggregates heterogeneous preferences into a single reward model, assuming preference homogeneity. Wh

safetyarxiv-cs-ai
12 Aug 2026
Safety

ReOrder-OPD:Reliability-Aware Prompt Ordering for On-Policy Distillation

DGX agent

arXiv:2608.10905v1 Announce Type: new Abstract: On-policy distillation (OPD) applies token-level teacher supervision to student-generated trajectories, but this supervision is not always reliable. Exi

safetyarxiv-cs-lg
12 Aug 2026
Safety

Robust Sliding Mode and Admittance Control of Underactuated Aerial Manipulators for Contact-Based Inspection

DGX agent

arXiv:2608.10656v1 Announce Type: new Abstract: Contact-based industrial inspection requires aerial platforms to maintain stable interaction while rejecting disturbances. Underactuated aerial manipula

safetyarxiv-cs-ro
12 Aug 2026
Safety

SapiensID 2.0: Aligning Human Recognition Foundation Models with Human Perception

DGX agent

arXiv:2608.10497v1 Announce Type: new Abstract: While foundation models have significantly advanced human recognition across diverse modalities, they predominantly rely on static, geometric feature ex

safetyarxiv-cs-cv
12 Aug 2026
Safety

SBCO: Self-Supervised, Verifier-Grounded Harness Optimization For Planning Agents

DGX agent

arXiv:2608.10157v1 Announce Type: new Abstract: Self-improving agents seek to reduce the human engineering effort behind AI systems by enabling them to evolve and self-improve their performance over t

safetyarxiv-cs-ai
12 Aug 2026
Safety

Scheduling Mixed RL Rollouts Beyond Prefix Locality

DGX agent

arXiv:2608.11152v1 Announce Type: cross Abstract: Modern reinforcement learning (RL) post-training pipelines for large language models (LLMs) increasingly combine rollout workloads across multiple dom

safetyarxiv-cs-lg
12 Aug 2026
Safety

Selective Prediction Reduces the Negative Effects of Automation Bias Overall but Increases False Negatives

DGX agent

arXiv:2508.07617v2 Announce Type: replace-cross Abstract: AI has the potential to augment human decision making. However, even high-performing models can produce inaccurate predictions when deployed.

safetyarxiv-cs-ai
12 Aug 2026
Safety

Self-Normalized Inference for Constant-Stepsize Temporal-Difference Learning under Markovian Sampling

DGX agent

arXiv:2608.10896v1 Announce Type: cross Abstract: Constant-stepsize temporal-difference (TD) learning is attractive for policy evaluation, but inference from a single Markov trajectory must account fo

safetyarxiv-cs-lg
12 Aug 2026
Safety

Status Association Does Not Reliably Predict Decision Leakage

DGX agent

arXiv:2608.10089v1 Announce Type: cross Abstract: Bias evaluations often move too quickly from evidence that a model encodes a social association to claims that the same association will alter consequ

safetyarxiv-cs-ai
12 Aug 2026
Safety

Stay or Stray - A Dynamical Systems Viewpoint of Popularity Bias

DGX agent

arXiv:2608.10474v1 Announce Type: cross Abstract: Popularity bias in recommendation systems arises when a majority user class generates disproportionate interaction data, causing the system to increas

safetyarxiv-cs-lg
12 Aug 2026
Safety

Surgical WAM: A World-Action Model for Data-Efficient Surgical Robot Learning

DGX agent

arXiv:2608.11204v1 Announce Type: cross Abstract: Learning reliable surgical manipulation policies is bottlenecked by the scarcity of action-labeled demonstrations: teleoperated surgical robot (e.g.,

safetyarxiv-cs-ai
12 Aug 2026
Safety

TCAM for Autonomous Deformable Manipulation: The RMC2 Champion System for WBCD 2026 Track 4

DGX agent

arXiv:2608.10718v1 Announce Type: new Abstract: This technical report describes the RMC2 Team's champion solution for the WBCD 2026 Track 4: Deformable Manipulation Challenge. The task requires a robo

safetyarxiv-cs-ro
12 Aug 2026
Safety

Templated or fully Synthetic? Prompt construction as a confound in measuring LLM political stance beyond writing assistance

DGX agent

arXiv:2608.11008v1 Announce Type: new Abstract: Political stance detection in LLMs has long been dominated by closed-ended, multiple-choice political survey questions---originally designed for humans,

safetyarxiv-cs-cl
12 Aug 2026
Safety

The GENEA Challenge 2026: A Large-Scale Disentangled Evaluation of Speech-Driven Gesture Generation on the Seamless Interaction Dataset

DGX agent

arXiv:2608.10839v1 Announce Type: new Abstract: This preprint presents the results of the fourth GENEA Challenge, a large-scale human evaluation of five speech-driven gesture-generation systems traine

safetyarxiv-cs-cv
12 Aug 2026
Safety

The Parser Already Knows: Lightweight Bias Correction in Constrained Decoding

DGX agent

arXiv:2608.10137v1 Announce Type: new Abstract: Grammar Constrained Decoding (GCD) forces Language Models (LMs) to produce syntactically valid outputs by masking out non-conforming tokens at each step

safetyarxiv-cs-cl
12 Aug 2026
Safety

Towards Color-Faithful Low-Light Image Enhancement via Adaptive Color Debiasing and Saturation Rectification

DGX agent

arXiv:2608.10512v1 Announce Type: new Abstract: Low-light imaging often introduces color bias caused by the low signal-to-noise ratio and the image formation process. Although recent low-light image e

safetyarxiv-cs-cv
12 Aug 2026
Safety

TRACE-GS: On-Policy Trajectory Distillation with Privileged Geometric Conditioning for Sparse-View 3DGS Restoration

DGX agent

arXiv:2608.10286v1 Announce Type: new Abstract: We present TRACE-GS, an on-policy trajectory distillation framework that leverages privileged geometric conditioning at training time, thereby adapting

safetyarxiv-cs-cv
12 Aug 2026
Safety

UniMod: Enhancing Multi-Modal Medical Diagnosis through Cross-Modality and Within-Modality Alignment

DGX agent

arXiv:2608.10316v1 Announce Type: new Abstract: Multi-modal learning combining medical images and clinical text is promising for disease diagnosis. However, standard multi-modal training leads to shor

safetyarxiv-cs-cv
12 Aug 2026
Safety

Watching Synthetic Videos: Aligning Cross-modal Representations with Visual Synthesis for Zero-shot Video Captioning

DGX agent

arXiv:2608.11013v1 Announce Type: new Abstract: Text-only training is a popular paradigm in zero-shot video captioning, where the video distribution is not available to the model during training, lead

safetyarxiv-cs-cv
12 Aug 2026
Safety

What DINO saw: ALiBi positional encoding reduces positional bias in Vision Transformers

DGX agent

arXiv:2603.16840v2 Announce Type: replace Abstract: Vision transformers (ViTs) - especially feature foundation models like DINOv2 - learn rich representations useful for many downstream tasks. However

safetyarxiv-cs-cv
12 Aug 2026
Safety

What We Know about Responsible AI Practices in Industry: A Half Decade of Empirical Research

DGX agent

arXiv:2608.10431v1 Announce Type: cross Abstract: Responsible AI (RAI) has become a central concern for technology companies, regulators, and the public. How industry practitioners interpret, implemen

safetyarxiv-cs-ai
12 Aug 2026
Safety

A Distribution Mapping Approach to Counterfactually Fair Reinforcement Learning

DGX agent

arXiv:2608.08743v1 Announce Type: cross Abstract: Reinforcement learning (RL) seeks to optimize sequential decisions to maximize population-level benefits over time. However, when deployed in high-sta

safetyarxiv-cs-lg
11 Aug 2026
Safety

A Dynamic-Semantics Framework for Grounding Human Referring Expressions in Visual Perceptual Data

DGX agent

arXiv:2608.08663v1 Announce Type: cross Abstract: Humans converge on shared names for novel, hard-to-describe objects through repeated interaction, a process psycholinguists call lexical entrainment.

safetyarxiv-cs-cv
11 Aug 2026
Safety

A Structural Dynamics Graph World Model: Unified Modeling, Constrained Rollout, and Interpretable Calibration

DGX agent

arXiv:2608.08689v1 Announce Type: new Abstract: The state evolution of a complex system arises jointly from object laws, relational propagation, domain conservation, and unmodeled error. Forcing all s

safetyarxiv-cs-ai
11 Aug 2026
Safety

Action- and Language-Conditioned Video Assessment for Embodied Control

DGX agent

arXiv:2608.08273v1 Announce Type: cross Abstract: Vision-based embodied agents executing multi-step natural language instructions require feedback mechanisms that assess task progress over complete tr

safetyarxiv-cs-cv
11 Aug 2026
Safety

Adaptive Supervised Anchoring for On-Policy Self-Distillation

DGX agent

arXiv:2608.07935v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) adapts a language model by distilling guidance from a frozen teacher on trajectories sampled from the student. Its ef

safetyarxiv-cs-lg
11 Aug 2026
Safety

Agentic Visual Reasoning in Whole-Slide Pathology Images via Active Perception

DGX agent

arXiv:2608.08648v1 Announce Type: new Abstract: Whole-slide visual reasoning requires identifying sparse diagnostic evidence in gigapixel pathology slides and integrating observations across spatial s

safetyarxiv-cs-cv
11 Aug 2026
Model Releases

Anchor-Based AI Approach for Pre-Crash Object Detection Utilizing Micro-Doppler Signatures in Automotive Radar

DGX agent

arXiv:2608.08701v1 Announce Type: new Abstract: Advanced automated driving presents significant potential to improve modern automotive safety systems, but it depends highly on the reliable activation

model-releasesarxiv-cs-ro
11 Aug 2026
Safety

ArchAgent v2: A Case Study with the Data Prefetching Championship

DGX agent

arXiv:2608.09874v1 Announce Type: new Abstract: Agentic artificial intelligence has shown great promise in automating algorithm design, but scaling similar techniques to computer microarchitecture dis

safetyarxiv-cs-ai
11 Aug 2026
Safety

Artificial Leviathan: Exploring Social Evolution of LLM Agents Through the Lens of Hobbesian Social Contract Theory

DGX agent

arXiv:2406.14373v3 Announce Type: replace Abstract: The emergence of Large Language Models (LLMs) and advancements in Artificial Intelligence (AI) offer an opportunity for computational social science

safetyarxiv-cs-ai
11 Aug 2026
Safety

Auditing Instruction-Trajectory Mismatches in Multimodal Robot Demonstrations

DGX agent

arXiv:2608.07895v1 Announce Type: cross Abstract: Robot demonstration datasets used to train vision-language-action policies can contain a subtle but harmful failure mode: trajectories that are behavi

safetyarxiv-cs-lg
11 Aug 2026
Safety

Autonomy Reshapes How Personalization Affects Privacy Concerns and Trust in LLM Agents

DGX agent

arXiv:2510.04465v3 Announce Type: replace-cross Abstract: LLM agents require personal information for personalization in order to effectively act on users' behalf, but this raises privacy concerns tha

safetyarxiv-cs-ai
11 Aug 2026
Safety

Beyond Aggregate Calibration: Decomposing Income-Conditional Recall Disparities in Automated Credit Default Prediction

DGX agent

arXiv:2608.08202v1 Announce Type: new Abstract: Data-centric curation pipelines frequently rely on model confidence scores to flag and filter noisy or mislabeled training instances. Evaluating this fi

safetyarxiv-cs-lg
11 Aug 2026
Safety

Beyond Binary: Continuous State Optimization with Graph-Structured Objectives

DGX agent

arXiv:2608.09366v1 Announce Type: new Abstract: Large-scale learning systems often face the challenge of balancing multiple, potentially competing objectives, such as fairness, accuracy, and latency.

safetyarxiv-cs-lg
11 Aug 2026
Safety

Beyond cognacy

DGX agent

arXiv:2507.03005v3 Announce Type: replace Abstract: Computational phylogenetics has become an established tool in historical linguistics, with many language families now analyzed using likelihood-base

safetyarxiv-cs-cl
11 Aug 2026
Safety

Beyond Solvability: Task Learnability as a Static Prior for LLM RL Post-Training

DGX agent

arXiv:2608.09217v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a central post-training paradigm for eliciting reasoning capabilities in large language models, yet uniform tas

safetyarxiv-cs-ai
11 Aug 2026
Safety

CDGC-Net: 3D Medical Image Segmentation with Cooperative Dual-Scale Self-Attention and Grouped Channel Modeling

DGX agent

arXiv:2608.08575v1 Announce Type: cross Abstract: Accurate 3D medical image segmentation requires the integration of long-range anatomical context with fine boundary detail. Existing methods often mod

safetyarxiv-cs-ai
11 Aug 2026
Safety

CIFA: Contextual-Intersectional Fairness Auditing for Hidden Subgroup Discovery in Face Analysis

DGX agent

arXiv:2608.09669v1 Announce Type: new Abstract: Fairness evaluation in computer vision commonly relies on aggregate accuracy and demographic subgroup analysis. However, visual models are also sensitiv

safetyarxiv-cs-cv
11 Aug 2026
Safety

CLAM: Causal Spatial Disaggregation to Infer Local Effects From Coarse Data

DGX agent

arXiv:2608.08064v1 Announce Type: new Abstract: Learning fine-grained spatial patterns from coarse-resolution data is challenging, especially in causal settings where high-resolution effects must be i

safetyarxiv-cs-lg
11 Aug 2026
Safety

Coarse-to-Fine Registration of Jawbone CT and Intraoral Scan Data Using GeDi and ICP with Pseudo-IOS Ground Truth

DGX agent

arXiv:2608.07564v1 Announce Type: cross Abstract: In digital dentistry and oral surgery, the registration of jawbone CT and intraoral scanner (IOS) data is essential for integrating internal bone stru

safetyarxiv-cs-ai
11 Aug 2026
Safety

Concept-Guided Spatial Regularization for World Models in Atari Pong

DGX agent

arXiv:2607.15142v2 Announce Type: replace Abstract: World models are usually evaluated as components of model-based reinforcement learning (MBRL) systems, leaving their standalone reliability understu

safetyarxiv-cs-ai
11 Aug 2026
Safety

Confusion-Geometry Rebalancing for Long-Tailed Adversarial Training

DGX agent

arXiv:2608.09688v1 Announce Type: cross Abstract: Adversarial training under long tailed distributions suffers from a dual imbalance: the class imbalance skews the training objective toward head class

safetyarxiv-cs-ai
11 Aug 2026
Safety

Contextual Value Alignment via Multilayer Combinatorial Fusion

DGX agent

arXiv:2608.07642v1 Announce Type: new Abstract: Aligning large language models (LLMs) with human values remains a major challenge, especially for trustworthy AI. While existing approaches such as RLHF

safetyarxiv-cs-ai
11 Aug 2026
Safety

Control-Oriented Scenario Tree Construction through Reinforcement Learning

DGX agent

arXiv:2608.09335v1 Announce Type: new Abstract: Multistage stochastic model predictive control (MPC) handles uncertainty by optimizing over a scenario tree, a finite branching approximation of future

safetyarxiv-cs-ai
11 Aug 2026
Safety

Coordinated incentives in AI-generated misinformation governance

DGX agent

arXiv:2608.07070v1 Announce Type: cross Abstract: With the rapid diffusion of AI-generated content, AI-driven misinformation is becoming increasingly pervasive and difficult to govern, undermining inf

safetyarxiv-cs-ai
11 Aug 2026
← Previous
1…5960616263…257
Next →