AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlog
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
Safety

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

DGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

safetyarxiv-cs-ai
26 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

CyBOKClaw: Human-in-the-Loop CyBOK Mapping for Cybersecurity Curriculum

DGX agent

arXiv:2605.24663v1 Announce Type: cross Abstract: This paper presents CyBOKClaw, an interpretable human-in-the-loop retrieval framework for mapping cybersecurity keywords or phrases (KWoPs) to the Cyb

safetyarxiv-cs-ai
26 May 2026
Safety

D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation

DGX agent

arXiv:2605.25022v1 Announce Type: cross Abstract: Dataset distillation (DD) aims to compress large-scale datasets into compact synthetic sets while preserving training efficacy. However, existing stud

safetyarxiv-cs-ai
26 May 2026
Safety

Dale meets Langevin: A Multiplicative Denoising Diffusion Model

DGX agent

arXiv:2510.02730v2 Announce Type: replace Abstract: Exponentiated gradient descent (EGD), a biologically motivated optimisation algorithm that respects Dale's law, produces log-normally distributed sy

safetyarxiv-cs-lg
26 May 2026
Safety

Data-Driven Optimization of Tactile Sensor Configurations for Efficient Dexterous Manipulation

DGX agent

arXiv:2409.20473v3 Announce Type: replace Abstract: Tactile sensing is critical for learning-based dexterous manipulation, yet principled guidelines for sensor placement remain largely absent. While d

safetyarxiv-cs-ro
26 May 2026
Safety

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care

DGX agent

arXiv:2510.08350v3 Announce Type: replace-cross Abstract: Objective: Enteral nutrition (EN) delivery in the ICU remains suboptimal due to limited personalization and uncertainty regarding appropriate

safetyarxiv-cs-ai
26 May 2026
Safety

DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading

DGX agent

arXiv:2605.25527v1 Announce Type: new Abstract: This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradien

safetyarxiv-cs-lg
26 May 2026
Safety

DeGRe: Dense-supervised Generative Reranking for Recommendation

DGX agent

arXiv:2605.25749v1 Announce Type: cross Abstract: In multi-stage recommender systems, reranking optimizes overall utility by capturing intra-list contextual dependencies, yet its central challenge lie

safetyarxiv-cs-ai
26 May 2026
Safety

Delta Energy: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization

DGX agent

arXiv:2510.11296v3 Announce Type: replace-cross Abstract: Recent approaches for vision-language models (VLMs) have shown remarkable success in achieving fast downstream adaptation. When applied to rea

safetyarxiv-cs-lg
26 May 2026
Safety

Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL

DGX agent

arXiv:2605.24001v1 Announce Type: cross Abstract: Recent advances in one-step text-to-image generation have enabled real-time synthesis with remarkable efficiency and quality. Previous reinforcement l

safetyarxiv-cs-ai
26 May 2026
Safety

Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs

DGX agent

arXiv:2605.23975v1 Announce Type: new Abstract: Audio large language models (Audio LLMs) exhibit systematic failures in transcribing code-switching speech despite strong multilingual capabilities. Foc

safetyarxiv-cs-cl
26 May 2026
Safety

Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2603.18444v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective post-training paradigm for improving the reasoning capabilit

safetyarxiv-cs-ai
26 May 2026
Safety

Discovering Lexical Gaps Using Embeddings from Multilingual LLMs

DGX agent

arXiv:2605.24310v1 Announce Type: new Abstract: Lexical gaps are words that do not exist in certain languages. They pose challenges for building multilingual lexical resources, for machine translation

safetyarxiv-cs-cl
26 May 2026
Safety

Discrete diffusion samplers and bridges: Off-policy algorithms and applications in latent spaces

DGX agent

arXiv:2602.05961v2 Announce Type: replace Abstract: Sampling from a distribution p(x) propto e^{-E(x)} known up to a normalising constant is an important and challenging problem in statistics. Recent

safetyarxiv-cs-lg
26 May 2026
Safety

DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection

DGX agent

arXiv:2605.24639v1 Announce Type: cross Abstract: With the widespread application of drones in recent years, object detection of aerial images has attracted increasing attention, especially open-vocab

safetyarxiv-cs-ai
26 May 2026
Safety

Disentangled Double Machine Learning for Accurate Causal Effect Estimation

DGX agent

arXiv:2605.24808v1 Announce Type: cross Abstract: Confounding bias is a key challenge in causal effect estimation from observational data. Double Machine Learning (DML) addresses this issue by estimat

safetyarxiv-cs-ai
26 May 2026
Safety

Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion

DGX agent

arXiv:2601.21670v3 Announce Type: replace-cross Abstract: Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from domi

safetyarxiv-cs-lg
26 May 2026
Safety

Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal Models

DGX agent

arXiv:2603.17044v2 Announce Type: replace-cross Abstract: Unified multimodal models share a language model backbone for both understanding and generating images. Can DPO align both capabilities simult

safetyarxiv-cs-ai
26 May 2026
Safety

Document Classification Pattern Recognition via Information Fusion: A Systematic Review of Multimodal and Multiview Representation Approaches

DGX agent

arXiv:2605.23910v1 Announce Type: cross Abstract: Information fusion is used widely to improve document classification by the integration of multiple data sources (multimodal) or representations (mult

safetyarxiv-cs-ai
26 May 2026
Safety

DVAO: Dynamic Variance-adaptive Advantage Optimization for Multi-reward Reinforcement Learning

DGX agent

arXiv:2605.25604v1 Announce Type: new Abstract: Reinforcement Learning has become a standard paradigm for aligning Large Language Models with human intent and task requirements. While Group Relative P

safetyarxiv-cs-cl
26 May 2026
Safety

Dynamic Dual-Granularity Skill Bank for Agentic RL

DGX agent

arXiv:2603.28716v2 Announce Type: replace Abstract: Agentic RL can benefit substantially from reusable experience, yet existing skill-based methods mainly extract trajectory-level guidance and often l

safetyarxiv-cs-ai
26 May 2026
Safety

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

DGX agent

arXiv:2605.24924v1 Announce Type: new Abstract: Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency

safetyarxiv-cs-ro
26 May 2026
Safety

Dynamic Relational Priming Improves Transformer in Multivariate Time Series

DGX agent

arXiv:2509.12196v2 Announce Type: replace-cross Abstract: Standard attention mechanisms in transformers employ static token representations that remain unchanged across all pair-wise computations in e

safetyarxiv-cs-ai
26 May 2026
Safety

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

DGX agent

arXiv:2512.04733v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (AD) systems increasingly adopt vision-language-action (VLA) models, yet they typically ignore the passenger's e

safetyarxiv-cs-ai
26 May 2026
Safety

ECHO: Terminal Agents Learn World Models for Free

DGX agent

arXiv:2605.24517v1 Announce Type: cross Abstract: CLI agents are the closest thing language models have to an embodied setting: the model emits commands, the terminal executes them, and the returned s

safetyarxiv-cs-cl
26 May 2026
Safety

ECo-MoE: Embodiment-Conditioned Mixture of Experts Increases the Evolvability of Robots

DGX agent

arXiv:2605.24225v1 Announce Type: new Abstract: In this paper, we introduce a model of evolution and learning in robots that co-optimizes a distribution of latent design vectors (genotypes) and a mixt

safetyarxiv-cs-ro
26 May 2026
Safety

EMA: Effort Metric Attention for Anatomical Effort-Guided Human Motion Diffusion

DGX agent

arXiv:2605.24566v1 Announce Type: cross Abstract: Human motion diffusion models can synthesize action sequences from text, but controlling motion intensity remains challenging. Existing approaches rel

safetyarxiv-cs-lg
26 May 2026
Safety

Emergent Analogical Reasoning in Transformers

DGX agent

arXiv:2602.01992v4 Announce Type: replace Abstract: Analogy is a central faculty of human intelligence, enabling abstract patterns discovered in one domain to be applied to another. Despite its centra

safetyarxiv-cs-ai
26 May 2026
Safety

EPPC-OASIS: Ontology-Aware Adaptation and Structured Inference Refinement for Electronic Patient-Provider Communication Mining in Secure Messages

DGX agent

arXiv:2605.24172v1 Announce Type: new Abstract: Secure patient-provider messages contain clinically important communication behaviors that are difficult to characterize manually at scale. The Electron

safetyarxiv-cs-ai
26 May 2026
Safety

Eureka: Intelligent Feature Engineering for Enterprise AI Cloud Resource Demand Prediction

DGX agent

arXiv:2605.25297v1 Announce Type: cross Abstract: Effective features are crucial for predictive model performance, but creating them often requires domain expertise, limiting scalability across applic

safetyarxiv-cs-ai
26 May 2026
Safety

Europe’s sovereign AI moment arrives in Heilbronn next week. At TECH by Handelsblatt 2026, Cohere CEO and Co-founder, @aidangomez, will join…

DGX agent

Europe’s sovereign AI moment arrives in Heilbronn next week. At TECH by Handelsblatt 2026, Cohere CEO and Co-founder, @aidangomez, will join leaders across business, policy, and industry to discuss ho

safetycohere--x
26 May 2026
Safety

even by the standards of the last few years, we are in some truly insane territory here.

DGX agent

even by the standards of the last few years, we are in some truly insane territory here. “Unserious, empty, hallucinatory, and borderline dishonest” - the prospectus for a company that the S&P 500 is

safetygary-marcus--x
26 May 2026
Safety

Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat

DGX agent

arXiv:2605.25091v1 Announce Type: new Abstract: As modern air combat evolves toward beyond-visual-range (BVR) multi-aircraft cooperative engagements, autonomous decision-making for unmanned combat aer

safetyarxiv-cs-ai
26 May 2026
Safety

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs

DGX agent

arXiv:2605.24345v1 Announce Type: new Abstract: In online reinforcement learning, data scarcity creates epistemic uncertainty that makes robustness important early in learning, whereas sufficient expl

safetyarxiv-cs-lg
26 May 2026
Safety

Extracting Training Data from Diffusion Language Models via Infilling

DGX agent

arXiv:2605.24173v1 Announce Type: cross Abstract: Memorization in large language models has been studied almost exclusively through prefix-conditioned extraction, a natural choice for autoregressive m

safetyarxiv-cs-ai
26 May 2026
Safety

Extreme Region Policy Distillation

DGX agent

arXiv:2605.25582v1 Announce Type: cross Abstract: Reinforcement learning for large language models faces a fundamental trade-off between sample efficiency and asymptotic performance: strictly on-polic

safetyarxiv-cs-ai
26 May 2026
Safety

Factored Latent Action World Models

DGX agent

arXiv:2602.16229v2 Announce Type: replace Abstract: Learning latent actions from action-free video has emerged as a powerful paradigm for scaling up controllable world model learning. Latent actions p

safetyarxiv-cs-lg
26 May 2026
Safety

FairJudge: Abstention-Aware Multimodal Judges for Fairness and Alignment Evaluation in Text-to-Image Models

DGX agent

arXiv:2510.22827v3 Announce Type: replace-cross Abstract: Evaluating text-to-image (T2I) systems requires judging not only whether an image matches a prompt, but also whether socially salient attribut

safetyarxiv-cs-lg
26 May 2026
Safety

Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges

DGX agent

arXiv:2605.23970v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as automatic judges for summarization and dialogue evaluation. Prior work has documented biases such

safetyarxiv-cs-cl
26 May 2026
Safety

Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning

DGX agent

arXiv:2605.24286v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning is useful for monitoring language models only when the reasoning trace faithfully reflects the computation that produ

safetyarxiv-cs-cl
26 May 2026
Safety

Few-Shot Neural Differentiable Simulator: Real-to-Sim Rigid-Contact Modeling

DGX agent

arXiv:2603.06218v2 Announce Type: replace Abstract: Accurate physics simulation is essential for robotic learning and control, yet analytical simulators often fail to capture complex contact dynamics,

safetyarxiv-cs-ro
26 May 2026
Safety

Flat Minima and Generalization: Insights from Stochastic Convex Optimization

DGX agent

arXiv:2511.03548v2 Announce Type: replace Abstract: Understanding the generalization behavior of learning algorithms is a central goal of learning theory. A recently emerging explanation is that learn

safetyarxiv-cs-lg
26 May 2026
Safety

From Reasoning to Code: GRPO Optimization for Underrepresented Languages

DGX agent

arXiv:2506.11027v3 Announce Type: replace-cross Abstract: Generating accurate and executable code using Large Language Models (LLMs) remains a significant challenge for underrepresented programming la

safetyarxiv-cs-ai
26 May 2026
Safety

From Simulation to Enaction: Post-trained language models recognize and react to their own generations

DGX agent

arXiv:2605.25459v1 Announce Type: cross Abstract: Language models are pretrained as passive predictors with no incentive to model the consequences of their own outputs. Post-training changes this: a m

safetyarxiv-cs-ai
26 May 2026
Safety

FusionCore: A 23-State Unscented Kalman Filter for IMU, Wheel Encoder, GPS, and Visual SLAM Fusion in ROS 2

DGX agent

arXiv:2605.25239v1 Announce Type: new Abstract: We present FusionCore, an open-source ROS 2 sensor fusion package that fuses IMU, wheel encoder odometry, GPS, and Visual SLAM pose into a single 100 Hz

safetyarxiv-cs-ro
26 May 2026
Safety

GeMPO: Generalized Measure Matching for Online Diffusion Reinforcement Learning

DGX agent

arXiv:2603.10250v2 Announce Type: replace Abstract: A commonly used family of RL algorithms for diffusion policies conducts softmax reweighting over samples from the behavior policy, which often induc

safetyarxiv-cs-lg
26 May 2026
Safety

Generative OOD-regularized Model-based Policy Optimization

DGX agent

arXiv:2605.24405v1 Announce Type: cross Abstract: We study sequential decision-making with offline reinforcement learning (RL). Traditional offline RL policies may result in out-of-distribution (OOD)

safetyarxiv-cs-ai
26 May 2026
Safety

Generative Visual Code Mobile World Models

DGX agent

arXiv:2602.01576v2 Announce Type: replace-cross Abstract: Mobile Graphical User Interface (GUI) World Models (WMs) offer a promising path for improving mobile GUI agent performance at train- and infer

safetyarxiv-cs-ai
26 May 2026
← Previous
1…182183184185186…302
Next →