AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
23 Jun 2026

AI Capex pushback from GS: 'If frontier intelligence can increasingly be developed in the East at a fraction of the cost incurred in the Wes…

SafetyDGX agent

AI Capex pushback from GS: 'If frontier intelligence can increasingly be developed in the East at a fraction of the cost incurred in the West … then the largest capital allocators are also the ones mo

AI data centres are hungry for land, water & power. I’m calling on every major AI company to publicly disclose the full environmental impact…

SafetyDGX agent

AI data centres are hungry for land, water & power. I’m calling on every major AI company to publicly disclose the full environmental impact of its systems – as a matter of transparency No more hidden

ALOE: Action-Level Off-Policy Evaluation for Vision-Language-Action Model Post-Training

Safety
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
DGX agent

arXiv:2602.12691v3 Announce Type: replace Abstract: We study how to improve large foundation vision-language-action (VLA) systems through human-in-the-loop reinforcement learning (RL) in real-world en

Although I supported Musk in his suit against OpenAI, and admire what he did for electric cars, I am afraid my considered overall view is no…

SafetyDGX agent

Although I supported Musk in his suit against OpenAI, and admire what he did for electric cars, I am afraid my considered overall view is not very different from this: Candidly I have no idea why anyo

APEX: Action Priors Enable Efficient Exploration for Robust Motion Tracking on Legged Robots

SafetyDGX agent

arXiv:2505.10022v4 Announce Type: replace Abstract: Learning natural, animal-like locomotion from demonstrations has become a core paradigm in legged robotics. While motion tracking can reproduce refe

ARGUSTRACK: A Multi-View Annotation System for Multi-Object Tracking

SafetyDGX agent

arXiv:2606.20687v1 Announce Type: new Abstract: Multi-Camera Multi-Target (MCMT) tracking has emerged as a critical capability for applications ranging from autonomous driving to animal behavior monit

ARP: Enhancing Quantized Skill Abstractions via Visual Alignment and Iterative Refinement for Robotic Manipulation

SafetyDGX agent

arXiv:2606.22480v1 Announce Type: new Abstract: Learning visuomotor policies for long-horizon manipulation remains a fundamental challenge. Recent skill-based imitation learning methods based on discr

Backpropagating Through Simulation: Analytic Policy Gradients for Sample and Learning Efficient Differentiable Continuous Control

SafetyDGX agent

arXiv:2606.21525v1 Announce Type: new Abstract: Model-free reinforcement learning algorithms such as Proximal Policy Optimization (PPO) treat the environment as a black box, estimating policy gradient

Balancing Performance and Diversity in GRPO Autoregressive Text-to-Image Post-Training

SafetyDGX agent

arXiv:2606.21498v1 Announce Type: cross Abstract: Autoregressive text-to-image (T2I) generation has recently advanced rapidly, yet aligning generated images with human preferences remains challenging.

BARD-MARL: Byzantine-Agent Detection for Learned Communication in Multi-Agent Reinforcement Learning

SafetyDGX agent

arXiv:2606.20701v1 Announce Type: cross Abstract: Learned communication improves coordination in cooperative multi-agent reinforcement learning, but it also creates a trust problem: a trained policy m

Behavioral and Representational Evidence of Binomial Ordering Preferences in Large Language Models

SafetyDGX agent

arXiv:2606.21645v1 Announce Type: cross Abstract: Large language models (LLMs) can readily reproduce conventional expressions, yet their ability to model gradient frequency distributions remains under

B[FM]^2: Brain Foundation Model via Flow Matching with SplitUNet

SafetyDGX agent

arXiv:2606.20812v1 Announce Type: new Abstract: EEG foundation models can learn generalizable representations from large-scale EEG corpora to enable single-backbone transfer across diverse clinical an

BLENDS: Bayesian Learning-Enhanced Deep Smoothing for GNSS-Denied Environments

SafetyDGX agent

arXiv:2606.22456v1 Announce Type: new Abstract: Maintaining accurate navigation during GNSS outages remains a significant challenge for autonomous systems relying on low-cost inertial sensors. While c

Boosting CVaR Policy Optimization with Quantile Gradients

SafetyDGX agent

arXiv:2601.22100v3 Announce Type: replace Abstract: Optimizing Conditional Value-at-risk (CVaR) using policy gradient (a.k.a CVaR-PG) faces significant challenges of sample inefficiency. This ineffici

Bridge the Gaps: Heterogeneous Attributed Graph Clustering via Quaternion Representation Learning

SafetyDGX agent

arXiv:2606.23199v1 Announce Type: new Abstract: Attributed graph clustering partitions nodes by jointly exploiting node attributes and graph topology. It remains challenging due to attribute heterogen

Bypassing Minimization Bias: A Shift-Invariant Variance Estimator for Off-Equilibrium Local Learning Coefficients

SafetyDGX agent

arXiv:2606.22389v1 Announce Type: new Abstract: Singular Learning Theory leverages the Local Learning Coefficient (LLC) to quantify the geometry of neural network loss landscapes. However, mean-energy

Can LLMs Control Readability? A Multi-Dimensional Evaluation Framework for CEFR-Controlled Arabic Generation

SafetyDGX agent

arXiv:2606.21981v1 Announce Type: cross Abstract: While Large Language Models (LLMs) can generate fluent Arabic text, their ability to reliably control readability levels remains unclear. We propose a

Causal Gaussian Processes for Robust Treatment Effect Evaluation with Unobserved Confounding

SafetyDGX agent

arXiv:2606.21809v1 Announce Type: new Abstract: The presence of confounding bias poses a key challenge in policy evaluation, as the target causal effects of actions are not identifiable (i.e., underde

Causally Fair Node Classification on Non-IID Graph Data

SafetyDGX agent

arXiv:2505.01652v2 Announce Type: replace Abstract: Fair machine learning seeks to identify and mitigate biases in predictions against unfavorable populations characterized by demographic attributes,

CFPO: Counterfactual Policy Optimization for Multimodal Reasoning

SafetyDGX agent

arXiv:2606.23206v1 Announce Type: new Abstract: Large Vision-Language Models (LVLMs) have demonstrated remarkable capabilities in multimodal reasoning. However, prevailing reinforcement learning (RL)

Chain-of-Goals Hierarchical Policy for Long-Horizon Offline Goal-Conditioned RL

SafetyDGX agent

arXiv:2602.03389v2 Announce Type: replace Abstract: Offline goal-conditioned reinforcement learning remains challenging for long-horizon tasks. While hierarchical approaches mitigate this issue by dec

CLAR: Learning 3D Representations for Robotic Manipulation by Fusing Masked Reconstruction with Multi-Level Contrastive Alignment

SafetyDGX agent

arXiv:2507.08262v2 Announce Type: replace-cross Abstract: The spatial information inherent in 3D point clouds is crucial for robotic manipulation. However, existing 3D pre-training methods face a fund

ClayBuddy: A Framework, Evaluation, & Mitigation of Coding Agent Failures

SafetyDGX agent

arXiv:2606.19380v2 Announce Type: replace-cross Abstract: Software engineering and deployment are increasingly delegated to AI coding agents. The scale of their adoption is surfacing rare, but highly

Closed-Loop Verbal Reinforcement Learning for Task-Level Robotic Planning

SafetyDGX agent

arXiv:2603.22169v2 Announce Type: replace Abstract: We propose a new Verbal Reinforcement Learning (VRL) framework for interpretable task-level planning in mobile robotic systems operating under execu

Cohort-Anchored Foundation Models for Electronic Health Records: From Risk Scores to Auditable Peer Cohorts

SafetyDGX agent

arXiv:2606.21885v1 Announce Type: new Abstract: Foundation models have achieved remarkable performance across medical question answering, imaging, and electronic health record (EHR) tasks, yet reliabl

Concept Alignment Contrast and Long-Short Prompt Memory for Test-Time Adaptation of SAM3 in Medical Image Segmentation

SafetyDGX agent

arXiv:2606.22963v1 Announce Type: new Abstract: Concept segmentation models like Segment Anything Model 3 (SAM3) show strong generalization on natural images, yet their performance degrades in medical

Conditional Flow Matching for Visually-Guided Acoustic Highlighting

SafetyDGX agent

arXiv:2602.03762v3 Announce Type: replace-cross Abstract: Visually-guided acoustic highlighting seeks to rebalance audio in alignment with the accompanying video, creating a coherent audio-visual expe

Confidence-Uncertainty Boundary Calibration for Bayesian Deep Learning in Medical Image Analysis

SafetyDGX agent

arXiv:2602.11973v2 Announce Type: replace Abstract: In critical decision support systems based on medical imaging, the reliability of AI-assisted decision-making is as relevant as predictive accuracy.

Continuous Behavioral Authentication via Multi-Expert BERT Log Analysis for Secure Data Sharing

SafetyDGX agent

arXiv:2606.21900v1 Announce Type: cross Abstract: Continuous authentication for mobile and zero-trust systems requires nonintrusive evidence confirming the enrolled user-device context remains valid a

CoorDex: Coordinating Body and Hand Priors for Continuous Dexterous Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.23680v1 Announce Type: cross Abstract: Humanoid loco-manipulation is often simplified into a stop-and-go process: walking to an object, stopping to manipulate it, and then resuming locomoti

Counsel: A Meta-Evaluation Dataset for Agentic Tasks

SafetyDGX agent

arXiv:2606.21627v1 Announce Type: cross Abstract: As agentic systems tackle increasingly complex multi-step tasks, evaluating their trajectories presents a major bottleneck - human annotation of a sin

Counterfactual learning of new adaptive instructional policies using logged data

SafetyDGX agent

arXiv:2606.23015v1 Announce Type: new Abstract: Optimizing instructional policies in Intelligent Tutoring Systems (ITS) typically requires costly online experimentation or student simulators that may

Cross-Modal Corroboration for Annotation-Free Wildlife Monitoring

SafetyDGX agent

arXiv:2606.21613v1 Announce Type: new Abstract: Scaling wildlife monitoring for real-world conservation deployments requires automated analysis of smart sensors that operate under severe annotation sc

Customizing Video Portraits via Identity-ActionDecoupling

SafetyDGX agent

arXiv:2606.22347v1 Announce Type: new Abstract: Identity-Preserving Text-to-Video Generation (IPT2V) seeks to synthesize a temporally coherent video from a reference image and a textual description, w

Darwin Mobile Agent: A Roadmap for Self-Evolution

SafetyDGX agent

arXiv:2606.20622v1 Announce Type: cross Abstract: The goal of artificial intelligence is to create agents capable of general, adaptive behaviour in open-ended environments. Guided by the 'Bitter Lesso

DASIP: Dynamic Test-Time Compute Scaling for Robot Control with Stochastic Interpolant Policies

SafetyDGX agent

arXiv:2511.20906v2 Announce Type: replace Abstract: Diffusion- and flow-based policies deliver state-of-the-art performance on long-horizon robotic manipulation and imitation learning tasks. However,

Data-Driven Image Registration and Deformation Modeling for Image-Guided Neurosurgery: A Systematic Review

SafetyDGX agent

arXiv:2602.10155v2 Announce Type: replace-cross Abstract: Accurate compensation of brain deformation is critical for reliable image-guided neurosurgery. Surgical manipulation and tumor resection induc

DataMIL: Selecting Data for Robot Imitation Learning with Datamodels

SafetyDGX agent

arXiv:2505.09603v2 Announce Type: replace-cross Abstract: Recently, the robotics community has amassed ever larger and more diverse datasets to train generalist policies. However, while these policies

Deep Learning for Individual Heterogeneity

SafetyDGX agent

arXiv:2010.14694v4 Announce Type: replace-cross Abstract: This paper integrates deep neural networks (DNNs) into structural models to increase flexibility and capture rich heterogeneity while preservi

Deep RL for Fast Long-Horizon Operations Scheduling on NASA's Carruthers Geocorona Observatory Mission

SafetyDGX agent

arXiv:2606.22159v1 Announce Type: cross Abstract: Spacecraft operations scheduling is a highly constrained, long-horizon combinatorial optimization problem that traditionally relies on heuristics, con

DeformX: A Versatile Co-Simulation Framework for Deformable Linear Objects

SafetyDGX agent

arXiv:2606.22116v1 Announce Type: new Abstract: Deformable linear objects (DLOs) such as wires, cables, and ropes are common in robotic manipulation tasks, yet simulating them with both visual realism

Delta-Diffusion: Modeling Longitudinal Brain Amyloid-PET Trajectories via Conditional Poisson Diffusion Bridge

SafetyDGX agent

arXiv:2606.22216v1 Announce Type: cross Abstract: While longitudinal brain PET imaging is the gold standard for quantifying the spatiotemporal accumulation of Beta-amyloid, its widespread clinical uti

Dense Reward for Multi-View 3D Reasoning with Global Maps and Local Views

SafetyDGX agent

arXiv:2606.23557v1 Announce Type: new Abstract: Multi-view 3D Visual Question Answering (MV3D-VQA) requires integrating partial observations into a coherent 3D scene representation and selecting infor

Distribution-Aware Diffusion-LLM for Robust Ultra-Long-Term Time Series Forecasting

SafetyDGX agent

arXiv:2606.23391v1 Announce Type: new Abstract: Time series forecasting is a fundamental machine learning task. Recent work has explored Large Language Models (LLMs) for this purpose due to their stro

DiT-Reward: Generative Representations for Text-to-Image Reward Modeling

SafetyDGX agent

arXiv:2606.23626v1 Announce Type: new Abstract: Can representations learned for image generation also support the evaluation of generated images? We study text-to-image reward prediction as a downstre

Don't Tell the Answer, Truly Guide the Reasoning During RL Rollouts

SafetyDGX agent

arXiv:2510.09388v2 Announce Type: replace Abstract: Reinforcement Learning (RL) has become a key driver for enhancing the long chain-of-thought (CoT) reasoning capabilities of Large Language Models (L

DT-GOL: Dual-Track Geometric Online Learning in Nonstationary Environment with Label Delay

SafetyDGX agent

arXiv:2606.22950v1 Announce Type: new Abstract: Online learning is crucial for handling complex data streams in big data applications. Recent research has begun to focus on dynamic scenarios, i.e., no

Dual-Attention Convolution Experts for Sparse Tensor Completion

SafetyDGX agent

arXiv:2606.21427v1 Announce Type: new Abstract: Tensor factorization (TF) has been widely adopted for high-dimensional sparse data completion tasks. Despite significant progress, neural TF methods oft

dVLA-RL: Reinforcement Learning over Denoising Trajectories for Discrete Diffusion Vision-Language-Action Models

SafetyDGX agent

arXiv:2606.23623v1 Announce Type: new Abstract: Vision-Language-Action (VLA) models have established a powerful paradigm for generalist robotic manipulation by grounding control into the semantic reas

EasyUUV: An LLM-Enhanced Universal and Lightweight Sim-to-Real Reinforcement Learning Framework for UUV Attitude Control

SafetyDGX agent

arXiv:2510.22126v3 Announce Type: replace Abstract: Despite recent advances in Unmanned Underwater Vehicle (UUV) attitude control, existing methods still struggle with generalizability, robustness to

Efficient Reinforcement Finetuning via Adaptive Curriculum Learning

SafetyDGX agent

arXiv:2504.05520v4 Announce Type: replace Abstract: Reinforcement finetuning (RFT) has shown great potential for enhancing the mathematical reasoning capabilities of large language models (LLMs), but

Efficient Training of Boltzmann Generators Using Off-Policy Log-Dispersion Regularization

SafetyDGX agent

arXiv:2602.03729v2 Announce Type: replace Abstract: Sampling from unnormalized probability densities is a central challenge in computational science. Boltzmann generators are generative models that en

Encoder-Decoder Manifold Alignment for Idempotent Generation

SafetyDGX agent

arXiv:2606.22304v1 Announce Type: new Abstract: Recently, several learning paradigms have been introduced to enforce idempotency in generative models. The goal is to ensure that repeated application o

Enhancing IMU-Based Online Handwriting Recognition via Contrastive Learning with Zero Inference Overhead

SafetyDGX agent

arXiv:2602.07049v2 Announce Type: replace Abstract: Online handwriting recognition using inertial measurement units opens up handwriting on paper as input for digital devices. Doing it on edge hardwar

Enhancing LLMs for Graph Tasks via Graph-aware LoRA Generation

SafetyDGX agent

arXiv:2606.22429v1 Announce Type: new Abstract: Graph neural networks (GNNs) tightly couple their input-output parameters to dataset-specific feature spaces and target sets, exhibiting limited transfe

EPSVec: Efficient and Private Synthetic Data Generation via Dataset Vectors

SafetyDGX agent

arXiv:2602.21218v2 Announce Type: replace-cross Abstract: High-quality data is essential for modern machine learning, yet many valuable corpora are sensitive and cannot be freely shared. Synthetic dat

EvoRubrics: Dynamic Rubrics as Rewards via Adversarial Co-Evolution for LLM Reinforcement Learning

SafetyDGX agent

arXiv:2606.23038v1 Announce Type: new Abstract: Rubric-based rewards offer interpretable and fine-grained optimization signals for reinforcement learning in open-ended tasks where verifiable answers a

Expected Free Energy-based Planning as Variational Inference

SafetyDGX agent

arXiv:2606.20658v1 Announce Type: cross Abstract: Planning under uncertainty requires agents to balance goal achievement with information gathering. Active inference addresses this through the Expecte

Extraction and Analysis of Multimodal Concepts in Vision Language Models through Sparse Autoencoders

SafetyDGX agent

arXiv:2606.21197v1 Announce Type: new Abstract: Vision Language Models (VLMs) have demonstrated impressive performance in tasks requiring joint understanding of images and text, such as image captioni

Fair Transit Stop Placement: A Clustering Perspective and Beyond

SafetyDGX agent

arXiv:2602.06776v2 Announce Type: replace-cross Abstract: We study the transit stop placement (TrSP) problem in general metric spaces, where agents travel between source-destination pairs and may eith

← Previous
1…105106107108109…242
Next →