AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,600 results
5 Aug 2026

🚨The endgame of commoditization and falling prices and lack of a moat that I warned was inevitable in August 2023 took three years. But tha…

SafetyDGX agent

🚨The endgame of commoditization and falling prices and lack of a moat that I warned was inevitable in August 2023 took three years. But that moment has come. TBD whether OpenAI and Anthropic can survi

The Evolutionary Origin of Values: implications for AI alignment, sentience and existential risk

SafetyDGX agent

arXiv:2608.03361v1 Announce Type: cross Abstract: AI systems based on Large Language Models (LLMs) have prompted fears that they may harbor hidden goals, seek to dominate or eliminate humanity, or eve

The Signal Horizon: Local Blindness and the Contraction of Pauli-Weight Spectra in Noisy Quantum Encodings

SafetyDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2602.14735v2 Announce Type: replace-cross Abstract: The performance of quantum classifiers is typically analyzed through global state distinguishability or the trainability of variational models

Tired Actor: Fatigue-Informed Character Control

SafetyDGX agent

arXiv:2608.03528v1 Announce Type: new Abstract: Replicating human behavior with physics simulation has been a long-expected goal in character animation. Existing efforts have achieved impressive perfo

Toward Certified Functional Safety for Industrial Humanoid Robots: The Fail-Passive Gap and a Feasibility Study

SafetyDGX agent

arXiv:2608.02809v1 Announce Type: new Abstract: Industrial humanoid robots are constrained less by locomotion or manipulation capability than by the immaturity of functional safety certification for l

Track4Action: Distilling World-Centric 3D Tracker into Vision-Language-Action Policies

SafetyDGX agent

arXiv:2608.03727v1 Announce Type: new Abstract: Action labels tell a vision-language-action (VLA) policy which robot commands to imitate, but not how those commands change the 3D world. The aligned de

TurnSight: Turn-Level Hindsight Self-Distillation for Tool-Integrated Reasoning

SafetyDGX agent

arXiv:2608.04007v1 Announce Type: cross Abstract: Tool-Integrated Reasoning (TIR) enables LLMs to solve complex tasks through iterative tool interactions. However, existing reinforcement learning meth

Unified Visuomotor Targets: Supervising VLAs Beyond Physical Actions

SafetyDGX agent

arXiv:2608.03563v1 Announce Type: new Abstract: VLA models are trained to predict robot actions from visual and language observations. This is a natural choice, but it creates a mismatch: VLMs encode

ValueFormer: A Causal Transformer Value Function with Stage-Aware Labels for Semi-Autonomous Vision-Language-Action Policies

SafetyDGX agent

arXiv:2608.02958v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) policies trained by behavior cloning fail silently: from the action stream alone, a collapsing rollout looks much like on

Welcome to the thunderdome of commoditization

SafetyDGX agent

Welcome to the thunderdome of commoditization 🚨The endgame of commoditization and falling prices and lack of a moat that I warned was inevitable in August 2023 took three years. But that moment has co

What is the Right Embedding Space for Contrastive Learning in REC?

SafetyDGX agent

arXiv:2505.22850v2 Announce Type: replace Abstract: Referring Expression Counting (REC) requires distinguishing visually similar objects described by fine-grained text cues. Existing methods tackle th

When Oracle Conditioning Misleads Deployment: Conditioning-Availability Bias in Echocardiographic Segmentation

SafetyDGX agent

arXiv:2608.03342v1 Announce Type: cross Abstract: Conditional segmentation models may be trained and evaluated with auxiliary signals cleaner than those available at deployment. We study this protocol

When Policies Change Probabilities: Modular Decision-Making for LLM Code Review

SafetyDGX agent

arXiv:2608.02677v1 Announce Type: cross Abstract: LLM code reviewers often estimate patch risk and make approval decisions in one prompt. A probability should depend on evidence; costs should determin

When Search Teaches Style: Causal Internalization of Tactical Priors in AlphaZero

SafetyDGX agent

arXiv:2504.14636v3 Announce Type: replace-cross Abstract: AlphaZero is normally evaluated as one agent: a policy-value network fused with Monte Carlo tree search. That fusion hides a causal question.

When Teachers Mislead: Spurious-Signal-Aware On-Policy Distillation

SafetyDGX agent

arXiv:2608.03632v1 Announce Type: new Abstract: On-Policy distillation (OPD) transfers teacher capabilities by supervising student-sampled trajectories with dense token-level teacher signals. Recent s

White House, AI firms keep safety framework talks private

SafetyDGX agent

The White House met with representatives from leading artificial intelligence companies today to discuss a safety framework for the government to review frontier models prior to launch, although there

Yes, the AIs were given a cybersecurity challenge, with internet access enabled and safety filters disabled. But the extent to which Mythos …

SafetyDGX agent

Yes, the AIs were given a cybersecurity challenge, with internet access enabled and safety filters disabled. But the extent to which Mythos 5 pursued its mission (fake identities, social engineering,

ZK-SR117: A Chunked Zero-Knowledge Attestation Design for Aggregated Fair-Lending Metrics, with a Control Mapping toward Full SR 11-7 Coverage

SafetyDGX agent

arXiv:2608.02664v1 Announce Type: cross Abstract: Deploying ML models in regulated decision-making (credit underwriting, fraud detection, loan approval) requires demonstrating fairness and robustness

4 Aug 2026

3D-CovDiffusion: 3D-Aware Diffusion Policy for Coverage Path Planning

SafetyDGX agent

arXiv:2510.03011v2 Announce Type: replace Abstract: Diffusion models have shown strong potential for robot skill learning, yet their role in coverage path planning remains underexplored. In industrial

A Constitution-Grid Instrument for Data-Efficient RL Alignment (C-Guard)

SafetyDGX agent

arXiv:2608.00180v1 Announce Type: new Abstract: Conflicting objectives are general in RL alignment, and training on them data-efficiently is hard. Training a safety guard with RL means optimizing two

A Forward-Inverse Dynamic Game Framework for Enhanced Multi-Agent Trajectory Planning

SafetyDGX agent

arXiv:2608.01636v1 Announce Type: new Abstract: This paper studies feedback Nash equilibrium (FBNE) seeking for multi-agent trajectory planning in nonlinear dynamical systems with unknown agents' obje

A Heuristic Perspective on Debiasing Language Models

SafetyDGX agent

arXiv:2608.00622v1 Announce Type: new Abstract: Language models (LMs) often acquire various biases during pre-training and may express them in interactions, potentially causing social harm. Existing m

A pioneer of AI-in-biology just published a sharp rebuttal to the 'AI will cure cancer' narrative. The core argument: It's a measurement pro…

SafetyDGX agent

A pioneer of AI-in-biology just published a sharp rebuttal to the 'AI will cure cancer' narrative. The core argument: It's a measurement problem. Over 90% of drugs that enter clinical trials still fai

A Spectral Filtering Approach to Regret Analysis of Distributed Online Control for Linear Dynamical Systems

SafetyDGX agent

arXiv:2608.02375v1 Announce Type: cross Abstract: This paper studies the distributed online control problem over a network of linear time-invariant (LTI) systems in the presence of adversarial disturb

ACE-GraphRAG: Agentic Context Engineering for Hierarchical GraphRAG

SafetyDGX agent

arXiv:2608.01269v1 Announce Type: new Abstract: Hierarchical Graph Retrieval-Augmented Generation (GraphRAG) organizes corpus knowledge at multiple levels of granularity, yet fixed context constructio

AdaThinkV: Adaptive Thinking for Token-Efficient Video Reasoning

SafetyDGX agent

arXiv:2608.01980v1 Announce Type: new Abstract: Chain-of-thought (CoT) reasoning can improve performance on difficult video questions but often wastes decoding tokens on simple ones. We study whether

Advancing Relevance Measurement with Vision-Language Models for Web-Scale Search

SafetyDGX agent

arXiv:2608.02446v1 Announce Type: cross Abstract: Relevance evaluation plays a crucial role in personalized search systems, serving as a guardrail alongside user engagement metrics to ensure that sear

AffordTrajDP: Dynamic Affordance-Guided Visuomotor Policy Learning for Robotic Manipulation

SafetyDGX agent

arXiv:2608.01603v1 Announce Type: new Abstract: Affordance-guided imitation learning has shown impressive performance in robotic manipulation tasks by compressing visual perception into task-specific

AI-Based Thesis Assessment: An Empirical Study of Human Evaluation Priorities and Their Impact on Automated Assessment

SafetyDGX agent

arXiv:2608.00717v1 Announce Type: cross Abstract: Rubric-based AI systems for thesis assessment use criterion weights to assign different levels of importance to evaluation criteria. These weights are

AlphaG-OPD: Reliability-Gated Sibling Counterfactuals for On-Policy Distillation in Symbolic Alpha Factor Discovery

SafetyDGX agent

arXiv:2608.01303v1 Announce Type: new Abstract: Symbolic alpha factor discovery can score a completed expression, but it provides no direct label for the structural decisions that produced it. Generat

An AI-Based Decision-Support Pipeline for Day-Ahead Photovoltaic Forecasting

SafetyDGX agent

arXiv:2608.02088v1 Announce Type: new Abstract: Reliable photovoltaic (PV) forecasts are needed for low-carbon energy systems, but newly deployed sites often have short, imperfect records. This makes

Analytic Planning under Uncertainty with Moment Closure

SafetyDGX agent

arXiv:2608.02519v1 Announce Type: new Abstract: Effective model-based reinforcement learning in stochastic environments requires planning that accounts for predictive uncertainty. Propagating full sta

Announcing Cloudflare Wallets: the programmable wallet for the agentic Internet

SafetyDGX agent

Cloudflare Wallets will provide AI agents with native payments and verifiable identity on the web. Using the x402 protocol, agents can autonomously purchase APIs and content within clear safety guardr

Any House Any Task: Scalable Long-Horizon Planning for Abstract Human Tasks

SafetyDGX agent

arXiv:2602.12244v2 Announce Type: replace Abstract: Open world language conditioned task planning is crucial for robots operating in large-scale household environments. While many recent works attempt

ARMOR: A Robust Self-Supervised Framework for Root Cause Analysis in Microservices under Missing Modality

SafetyDGX agent

arXiv:2603.25538v3 Announce Type: replace Abstract: Automated incident management is critical for microservice reliability. While recent unified frameworks leverage multimodal data for joint optimizat

ARMOR: Robust Reinforcement Learning-based Control for UAVs under Physical Attacks

SafetyDGX agent

arXiv:2506.22423v2 Announce Type: replace Abstract: Unmanned Aerial Vehicles (UAVs) depend on onboard sensors for perception, navigation, and control. However, these sensors are susceptible to physica

ASTRA: Asynchronous Spatio-Temporal Reconstruction via Trajectory Alignment

SafetyDGX agent

arXiv:2608.02006v1 Announce Type: new Abstract: Dynamic 3D scene reconstruction has achieved remarkable success under the assumption of strictly synchronized multi-camera inputs. However, in real-worl

Automated ECG Interval Measurement and Wave Delineation Using Fast Fourier Convolution ResNet

SafetyDGX agent

arXiv:2608.00058v1 Announce Type: cross Abstract: Accurate measurement of ECG intervals, including PR, QRS duration, and QT/QTc, is central to cardiac diagnosis, yet the published ECG delineation lite

Averaging Bias: Human Faithfulness Annotations are not Locally Faithful

SafetyDGX agent

arXiv:2608.00205v1 Announce Type: new Abstract: Evaluation of faithfulness of text summarization treats a model generated summary as faithful only if every of its sentences is supported by the source

Benign Overfitting in Linear Classifiers with a Bias Term

SafetyDGX agent

arXiv:2511.12840v2 Announce Type: replace-cross Abstract: Overparameterized models often generalize well even when they interpolate noisy training data. This is known as benign overfitting. For linear

Beyond On-Policy Exploration: Integrating External Policy Rollouts for Reinforcement Learning in Diffusion Language Models

SafetyDGX agent

arXiv:2608.01717v1 Announce Type: new Abstract: Recent reinforcement learning methods for diffusion large language models (dLLMs) commonly rely on on-policy rollouts generated by the target dLLM itsel

Beyond Random Partitioning: Unsupervised Spatio-Temporal Stratification for Cohort Balancing in Longitudinal Medical Imaging

SafetyDGX agent

arXiv:2608.00073v1 Announce Type: new Abstract: Rigorous dataset partitioning is a foundational, yet frequently overlooked, prerequisite for reliable deep learning in longitudinal medical imaging. Nai

Beyond VLAs: How World Action Models Reshape Robot Manipulation

SafetyDGX agent

World Action Models (WAMs) use video-based world modeling instead of vision‑language backbones, giving robots a learned physics engine that supports zero‑shot transfer to new tasks, environments, and

BiCAA: Bidirectional Credit Assignment for Search-Augmented Agent

SafetyDGX agent

arXiv:2608.01321v1 Announce Type: new Abstract: Multi-step search is a fundamental capability for search agents, enabling them to iteratively acquire, refine, and integrate external evidence for compl

Breaking the Statistical Similarity Trap in Extreme Convection Detection

SafetyDGX agent

arXiv:2509.09195v2 Announce Type: replace-cross Abstract: Current evaluation metrics for deep learning weather models create a 'Statistical Similarity Trap', rewarding blurry predictions while missing

Bridging the Sim-to-Real Gap in Parallel-Link Leg Mechanisms via Simulator-Side Dynamics Normalization

SafetyDGX agent

arXiv:2608.01697v1 Announce Type: new Abstract: This paper addresses the sim-to-real gap in dynamics arising when a parallel-link mechanism is represented by a serial-tree surrogate in simulation. Con

CAAT: Contact-Aware Attention Scaling and Tactile Masking for Data-Efficient Contact-Rich Manipulation

SafetyDGX agent

arXiv:2608.01102v1 Announce Type: new Abstract: In contact-rich manipulation, visual observations primarily guide motion in free space, whereas tactile observations become particularly informative dur

CalibBEV: LiDAR-Camera Calibration via BEV Alignment

SafetyDGX agent

arXiv:2608.02309v1 Announce Type: new Abstract: We present CalibBEV, a novel Bird's Eye View (BEV) alignment approach for LiDAR-camera calibration. Our method unifies LiDAR and camera data into a shar

can anyone with sufficient training in mathematics refuse this provocation? no ad hominem please; I have seen enough of that and am looking …

SafetyDGX agent

can anyone with sufficient training in mathematics refuse this provocation? no ad hominem please; I have seen enough of that and am looking for real counterarguments, only. 🙏 OpenAI and Anthropic each

can anyone with sufficient training in mathematics refute this provocation? no ad hominem please; I have seen enough of that and am looking …

SafetyDGX agent

can anyone with sufficient training in mathematics refute this provocation? no ad hominem please; I have seen enough of that and am looking for real counterarguments, only. 🙏 OpenAI and Anthropic each

CAVE: Competence-Aware Visual Boundary Evidence Alignment for Video Temporal Grounding

SafetyDGX agent

arXiv:2608.02078v1 Announce Type: new Abstract: Large vision-language models (LVLMs) have achieved substantial performance gains in Video Temporal Grounding (VTG) through reinforcement learning (RL).

Characterizing Bias in Post-Bandit Inference under Index Algorithms

SafetyDGX agent

arXiv:2608.01069v1 Announce Type: new Abstract: Bandit algorithms generate data for downstream inference, but adaptive sampling biases post-bandit sample means. We analyze this bias for stable index a

ChordVideo: One-Step, Training-Free, Temporally Consistent Video Editing via Low-Energy Transport

SafetyDGX agent

arXiv:2608.00769v1 Announce Type: new Abstract: One-step text-to-image models enable training-free, inversion-free editing with only 1--2 network function evaluations (NFE), while ChordEdit stabilizes

CoLI: A Reproducible Platform for Continuum Robot Learning via Monolithic 3D Printing and Isomorphic Teleoperation

SafetyDGX agent

arXiv:2606.20389v2 Announce Type: replace Abstract: Continuum robots offer strong potential for manipulation tasks due to their high degrees of freedom, compliant structures, and operational safety. H

Coming this weekend: What chess teaches us about Generative AI and its limitations Image with impossible position and bishops of two sizes p…

SafetyDGX agent

Gary Marcus tweeted on August 4, 2026 at 3:53 PM that a forthcoming weekend presentation will explore how chess illustrates the limitations of generative AI systems—highlighting an image produced by C

Comparing and Modeling Argumentation in German Political Communication across Arenas

SafetyDGX agent

arXiv:2608.00288v1 Announce Type: new Abstract: Deliberation, involving the formulation and exchange of arguments, forms an integral part of political decision making in democracies. Argumentation pat

Conformal bandits: bringing statistical validity and reward efficiency under weak arm separability

SafetyDGX agent

arXiv:2512.09850v2 Announce Type: replace Abstract: We introduce Conformal Bandits, a novel framework integrating Conformal Prediction (CP) into bandit problems, a classic paradigm for sequential deci

Constrained Co-Design for Photonic Bayesian Neural Networks

SafetyDGX agent

arXiv:2608.02229v1 Announce Type: new Abstract: Classical neural networks frequently produce overconfident predictions on ambiguous or out-of-distribution (OOD) data, a liability that grows with each

Convex Neural Energy Elements: Monolithic Finite-Element Assembly of Geometry-Parameterized Neural Operators with Stability and Error Guarantees

SafetyDGX agent

arXiv:2608.02036v1 Announce Type: new Abstract: Extending the neural-operator element method from individually trained, fixed-geometry neural elements to a library of reusable, geometry-parameterized

Counting the Cost of War Under Satellite Embargo: Zero-Shot Estimation of Impacted Infrastructure

SafetyDGX agent

arXiv:2608.00119v1 Announce Type: new Abstract: Rapid estimation of impacted structures - critical for conflict-zone humanitarian response - is frequently hindered by post-strike satellite data embarg

← Previous
1…1213141516…210
Next →