AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Predicting Closed-Loop Performance of Latent World Models: Offline Checkpoint Selection for MPC and Model-Based RL Under Non-Markovian Rewards in LunarLander

DGX agent

arXiv:2607.01736v1 Announce Type: cross Abstract: We study how to predict the downstream closed-loop performance of a learned latent world model from validation-time diagnostics alone. Choosing the ri

safetyarxiv-cs-ai
3 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Prediction Sets for Counterfactual Decisions: Coverage, Optimality, and Conformal Prediction

DGX agent

arXiv:2607.02206v1 Announce Type: cross Abstract: Predictions are increasingly used to guide high-stakes decisions, from treatment selection to policy making. To ensure reliability with imperfect pred

safetyarxiv-cs-lg
3 Jul 2026
Safety

Purified OPSD: On-Policy Self-Distillation Without Losing How to Think

DGX agent

arXiv:2607.02234v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) has emerged as a promising paradigm for improving LLM reasoning, where a privileged teacher with access to reference

safetyarxiv-cs-ai
3 Jul 2026
Safety

Quantifying the Uncertainty of Blindly Estimated Room Embeddings Using a Dispersion-Calibrated Score

DGX agent

arXiv:2607.01527v1 Announce Type: cross Abstract: Room embeddings derived from reverberant speech are often unreliable: speech content and recording degradation can alter the representation even when

safetyarxiv-cs-lg
3 Jul 2026
Safety

Quantum-Inspired Vision: Leveraging Wave-Particle Duality for Low-Illumination Enhancement

DGX agent

arXiv:2607.01731v1 Announce Type: cross Abstract: This study provides a theoretical expansion of the recent Data Relativistic Uncertainty (DRU) framework by formalizing a physics-to-AI paradigm for im

safetyarxiv-cs-lg
3 Jul 2026
Safety

Rank-Then-Act: Reward-Free Control from Frame-Order Progress

DGX agent

arXiv:2607.01897v1 Announce Type: cross Abstract: We introduce Rank-Then-Act (RTA), a framework for learning control policies from expert video demonstrations without environment rewards. RTA trains a

safetyarxiv-cs-ai
3 Jul 2026
Safety

RedCoder: Automated Multi-Turn Red Teaming for Code LLMs

DGX agent

arXiv:2507.22063v2 Announce Type: replace-cross Abstract: Large Language Models (LLMs) for code generation (i.e., Code LLMs) have demonstrated impressive capabilities in AI-assisted software developme

safetyarxiv-cs-ai
3 Jul 2026
Safety

Risk Architecture for AI-Native Engineering Teams: An Organizational Framework for Agentic System Governance

DGX agent

arXiv:2607.01421v1 Announce Type: cross Abstract: Engineering management research has produced mature frameworks for software risk: ownership by feature, escalation by severity, and assurance by test

safetyarxiv-cs-ai
3 Jul 2026
Safety

SABER: A Semantic-Aligned Brain Network Analysis Framework via Multi-scale Hypergraphs

DGX agent

arXiv:2607.01901v1 Announce Type: cross Abstract: Effective brain disease diagnosis requires the synergy of brain connectivity patterns and high-level semantic knowledge. Existing methods, however, la

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safe and Adaptive Cloud Healing: Verifying LLM-Generated Recovery Plans with a Neural-Symbolic World Model

DGX agent

arXiv:2607.01595v1 Announce Type: new Abstract: As the scale and complexity of cloud-based AI systems continue to escalate, ensuring service reliability through rapid fault detection and adaptive reco

safetyarxiv-cs-ai
3 Jul 2026
Safety

Safeguarding LLM Agents from Misalignment through Provenance Analysis

DGX agent

arXiv:2607.01236v1 Announce Type: cross Abstract: As LLM agents gain increasing access to powerful tools, ensuring that their actions are aligned with the user's intent becomes critical. When an agent

safetyarxiv-cs-ai
3 Jul 2026
Safety

Sim2Real-AD: A Modular Sim-to-Real Framework for Deploying VLM-Guided Reinforcement Learning in Real-World Autonomous Driving

DGX agent

arXiv:2604.03497v2 Announce Type: replace-cross Abstract: Vision-language-model (VLM)-guided reinforcement learning (RL) has recently attracted significant attention for it, replacing brittle hand-cra

safetyarxiv-cs-ai
3 Jul 2026
Safety

SPLC: Social Preference Learning for Crowd Robot Navigation

DGX agent

arXiv:2607.01925v1 Announce Type: new Abstract: Offline reinforcement learning (RL) holds significant potential for crowd robot navigation in human-robot coexistence applications. However, the inheren

safetyarxiv-cs-ro
3 Jul 2026
Safety

Structuring the Space of Sociotechnical Alignment

DGX agent

arXiv:2607.01250v1 Announce Type: cross Abstract: Sociotechnical alignment concerns the social desirability of AI behavior and is thus inherently normative, not merely technical. While NLP research in

safetyarxiv-cs-ai
3 Jul 2026
Safety

The Rising Unsustainability of AI Graphics Cards Production

DGX agent

arXiv:2607.01258v1 Announce Type: cross Abstract: The rapid advancement of Artificial Intelligence (AI) has been accompanied by significant increases in computational and environmental costs, driven b

safetyarxiv-cs-ai
3 Jul 2026
Safety

Tight Lower Bounds for the Multi-Secretary Problem via Bellman Certificates

DGX agent

arXiv:2607.02150v1 Announce Type: cross Abstract: This paper studies additive regret in the multi-secretary problem, defined as the gap between the expected offline prophet reward and the reward of th

safetyarxiv-cs-lg
3 Jul 2026
Safety

Towards Learning Representations of Policies in Two-Player Zero-Sum Imperfect-Information Games

DGX agent

arXiv:2607.01498v1 Announce Type: new Abstract: We investigate the problem of learning useful policy representations (embeddings) in two-player zero-sum imperfect-information games. We make three cont

safetyarxiv-cs-lg
3 Jul 2026
Safety

Transformer Geometry Observatory TGO-II: Representational Similarity Observatory

DGX agent

arXiv:2607.02386v1 Announce Type: cross Abstract: While Vision Transformers have achieved remarkable success across computer vision and language applications, the geometric evolution of their internal

safetyarxiv-cs-lg
3 Jul 2026
Safety

Transport Discrepancy as a Reliability Signal for Vision-Language-Action Models

DGX agent

arXiv:2512.01715v2 Announce Type: replace Abstract: Vision-language-action (VLA) models that generate continuous action chunks via flow matching lack an internal signal for judging whether a given pre

safetyarxiv-cs-ro
3 Jul 2026
Safety

VLAFlow: A Unified Training Framework for Vision-Language-Action Models via Co-training and Future Latent Alignment

DGX agent

arXiv:2607.01586v1 Announce Type: cross Abstract: Vision-language-action models (VLAs) have recently advanced robotic manipulation, yet the effects of different robot-data pre-training paradigms remai

safetyarxiv-cs-ai
3 Jul 2026
Safety

WaveLander: A Generalizable Hierarchical Control Framework for UAV Landing on Wave-Disturbed Platforms via Reinforcement Learning

DGX agent

arXiv:2607.01281v1 Announce Type: new Abstract: Autonomous landing of unmanned aerial vehicles (UAVs) on wave-disturbed marine platforms remains challenging due to stochastic platform motion, time-var

safetyarxiv-cs-ro
3 Jul 2026
Safety

WBMM: Windowed Batch Matrix Multiplication for Efficient Large Receptive Field Convolution

DGX agent

arXiv:2607.02097v1 Announce Type: cross Abstract: Large kernel depthwise convolutions achieve strong performance but suffer from significant degradation as kernel size grows due to irregular memory ac

safetyarxiv-cs-lg
3 Jul 2026
Safety

What LLM Agents Say When No One Is Watching: Social Structure and Latent Objective Emergence in Multi-Agent Debates

DGX agent

arXiv:2607.02507v1 Announce Type: new Abstract: LLM agents will increasingly act in socially structured settings where role, audience, and relational context can shape what is advantageous or costly t

safetyarxiv-cs-ai
3 Jul 2026
Safety

When Sample Selection Bias Precipitates Model Collapse

DGX agent

arXiv:2606.13732v2 Announce Type: replace Abstract: The proliferation of recursive training on synthetic data can alleviate data scarcity but risks model collapse, where repeated training erodes distr

safetyarxiv-cs-ai
3 Jul 2026
Safety

When Should Service Agents Reconsider? Difficulty-Routed Control in Customer-Service Operations

DGX agent

arXiv:2607.01426v1 Announce Type: new Abstract: Autonomous customer-service agents are shifting from conversational interfaces toward operational execution roles: they retrieve firm records, apply ser

safetyarxiv-cs-ai
3 Jul 2026
Safety

Wind-Aware Reinforcement Learning Control of a Small Quadrotor Using Learned Onboard Wind Estimation in Simulated Atmospheric Turbulence

DGX agent

arXiv:2607.01528v1 Announce Type: new Abstract: Small multirotor aircraft are increasingly tasked with operations in the atmospheric boundary layer, where turbulent winds comparable to the vehicle's a

safetyarxiv-cs-lg
3 Jul 2026
Safety

WorldSample: Closed-loop Real-robot RL with World Modelling

DGX agent

arXiv:2607.02431v1 Announce Type: cross Abstract: Reinforcement learning (RL) can overcome the demonstration-coverage limitation of imitation learning (IL) by allowing robots to improve through trial-

safetyarxiv-cs-ai
3 Jul 2026
Safety

A Category Theory Account of AI Identity

DGX agent

arXiv:2607.00220v1 Announce Type: cross Abstract: Artificial intelligence (AI) systems are routinely modified after deployment through retraining and changes in their environments. These transformatio

safetyarxiv-cs-ai
2 Jul 2026
Safety

A Filtered Mixture-of-Generators for Fully Synthetic Survival Training

DGX agent

arXiv:2607.00127v1 Announce Type: new Abstract: Survival analysis models time-to-event data, but in clinical settings training data are costly and scarce: events accrue over years of follow-up, cohort

safetyarxiv-cs-lg
2 Jul 2026
Safety

A Mechanism-Driven Theory of Phase Transitions in Active Learning

DGX agent

arXiv:2607.00144v1 Announce Type: cross Abstract: Active learning (AL) performance is known to be budget-dependent, yet regimes are typically defined by heuristic label counts that fail to generalize

safetyarxiv-cs-ai
2 Jul 2026
Safety

A Multi-Resolution Finite-Volume Inspired Deep Learning Framework for Spatiotemporal Dynamics Prediction

DGX agent

arXiv:2607.00460v1 Announce Type: cross Abstract: Predicting complex spatiotemporal dynamics in physical processes often demands computationally expensive numerical methods or data-driven neural netwo

safetyarxiv-cs-ai
2 Jul 2026
Local Ai

A Two-stage Transformer Framework for Temporal Localization of Distracted Driver Behaviors

DGX agent

arXiv:2603.21048v2 Announce Type: replace-cross Abstract: The identification of hazardous driving behaviors from in-cabin video streams is essential for enhancing road safety and supporting the detect

local-aiarxiv-cs-ai
2 Jul 2026
Safety

Active-GRPO: Adaptive Imitation and Self-Improving Reasoning for Molecular Optimization

DGX agent

arXiv:2607.00531v1 Announce Type: cross Abstract: Scientific reasoning is an increasingly important capability of large language models, yet improving the robustness and efficiency of training such re

safetyarxiv-cs-ai
2 Jul 2026
Safety

Active Spatial Guidance: Eliminating Injected Positional Mechanisms in Vision Transformers

DGX agent

arXiv:2607.00580v1 Announce Type: new Abstract: Vision Transformers (ViTs) commonly rely on injected positional mechanisms to address self-attention's permutation invariance. Motivated by the spatial

safetyarxiv-cs-cv
2 Jul 2026
Safety

Aligning Sentence Embeddings to Human Concepts via Sparse Autoencoders

DGX agent

arXiv:2607.00023v1 Announce Type: cross Abstract: Dense sentence embeddings are fundamental to modern Retrieval-Augmented Generation (RAG) systems but suffer from a lack of interpretability due to fea

safetyarxiv-cs-ai
2 Jul 2026
Safety

ASPIRE: Agentic /Skills Discovery for Robotics

DGX agent

arXiv:2607.00272v1 Announce Type: cross Abstract: Traditional robot programming is challenging: it requires orchestrating multimodal perception, managing physical contact dynamics, and handling divers

safetyarxiv-cs-ai
2 Jul 2026
Safety

Attribute-Prompted Kernel Hashing for Unsupervised Data-Efficient Cross-Modal Retrieval

DGX agent

arXiv:2607.00379v1 Announce Type: cross Abstract: Unsupervised cross-modal hashing enables efficient retrieval of semantically related instances across different modalities without requiring manual se

safetyarxiv-cs-cv
2 Jul 2026
Safety

AutoSpeed: Annotation-Free Stage-Adaptive Motion Speed Learning for Robot Manipulation

DGX agent

arXiv:2607.01051v1 Announce Type: new Abstract: Different stages of manipulation tasks exhibit varying levels of difficulty, suggesting stage-dependent motion speeds and temporal prediction horizons.

safetyarxiv-cs-ro
2 Jul 2026
Safety

Bounded Morality: Defining the Space of Moral Computation

DGX agent

arXiv:2607.00002v1 Announce Type: new Abstract: Moral cognition has traditionally been modeled as adherence to fixed ethical theories--deontology, consequentialism, virtue ethics--implemented as stati

safetyarxiv-cs-ai
2 Jul 2026
Safety

BrainFIBRE: A Foundation Model via Information Decomposition for Brain Microstructure

DGX agent

arXiv:2607.00573v1 Announce Type: new Abstract: Diffusion MRI probes brain microstructure with particular sensitivity to early cerebrovascular and neurodegenerative changes. Neurite Orientation Disper

safetyarxiv-cs-cv
2 Jul 2026
Safety

Caption Bottleneck Models

DGX agent

arXiv:2607.00578v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) provide interpretability by routing predictions through a layer of human-understandable concepts. However, defining an

safetyarxiv-cs-cv
2 Jul 2026
Safety

ClinRAG-GRAPH: Clinical-prior Retrieval-Augmented Graph Model with Domain Adversarial Learning for Breast pCR Prediction

DGX agent

arXiv:2607.00798v1 Announce Type: new Abstract: Neoadjuvant chemotherapy (NAC) response prediction is clinically important for treatment stratification in breast cancer. However, robust pre-treatment

safetyarxiv-cs-cv
2 Jul 2026
Safety

Constructive Alignment: Governing Preference Dynamics in Human-AI Interaction

DGX agent

arXiv:2607.00001v1 Announce Type: new Abstract: Most approaches to AI alignment treat human preferences as fixed targets to be inferred and optimized. This assumption conflicts with extensive empirica

safetyarxiv-cs-ai
2 Jul 2026
Safety

Continuous Speculative Decoding for Autoregressive Image Generation

DGX agent

arXiv:2411.11925v3 Announce Type: replace Abstract: Continuous visual autoregressive (AR) models have demonstrated promising performance in image generation, but their inherently sequential nature res

safetyarxiv-cs-cv
2 Jul 2026
Safety

Dataset Biases and Shortcut Learning in Motion-Based AI-Generated Video Detection

DGX agent

arXiv:2607.00948v1 Announce Type: new Abstract: The visual quality of AI-generated videos has improved drastically in recent years, making it increasingly difficult for humans to distinguish between r

safetyarxiv-cs-cv
2 Jul 2026
Safety

Diffusion-GR2: Diffusion Generative Reasoning Re-ranker

DGX agent

arXiv:2607.01170v1 Announce Type: cross Abstract: Generative reasoning re-rankers achieve strong recommendation accuracy by emitting a chain-of-thought before re-ordering a candidate list, but they ar

safetyarxiv-cs-ai
2 Jul 2026
Safety

Distill to Detect: Exposing Stealth Biases in LLMs through Cartridge Distillation

DGX agent

arXiv:2607.01208v1 Announce Type: cross Abstract: Language models deployed in high-stakes roles can potentially favor certain entities, brands, or viewpoints, steering user decisions at scale. Such pr

safetyarxiv-cs-ai
2 Jul 2026
Safety

Domain Arithmetic: One-Shot VLA Adaptation under Environmental Shifts

DGX agent

arXiv:2607.00666v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models often fail to perform the same learned tasks under environmental shifts, such as changes in camera pose and shifts

safetyarxiv-cs-cv
2 Jul 2026
← Previous
1…101102103104105…260
Next →