AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
12,435 results
Safety

Characterizing Linear Alignment Across Language Models

DGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

safetyarxiv-cs-ai
26 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems

DGX agent

arXiv:2605.25290v1 Announce Type: cross Abstract: Online experiments in ads, recommendation, and member-experience systems are often planned before the dominant interference mechanism is known. A trea

safetyarxiv-cs-lg
26 May 2026
Safety

Clustering as Reasoning: A k-Means Interpretation of Chain-of-Thought Graph Learning

DGX agent

arXiv:2605.24867v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) on text-attributed graphs (TA

safetyarxiv-cs-ai
26 May 2026
Safety

CODESKILL: Learning Self-Evolving Skills for Coding Agents

DGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

safetyarxiv-cs-ai
26 May 2026
Safety

ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation

DGX agent

arXiv:2605.25553v1 Announce Type: cross Abstract: Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with th

safetyarxiv-cs-ro
26 May 2026
Safety

Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection

DGX agent

arXiv:2605.24294v1 Announce Type: cross Abstract: Android malware detectors often degrade after deployment because of concept drift, while full retraining at each maintenance step is costly. We propos

safetyarxiv-cs-ai
26 May 2026
Safety

Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2602.08499v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an effective paradigm for improving the reasoning capabilities of large language mode

safetyarxiv-cs-ai
26 May 2026
Safety

Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning

DGX agent

arXiv:2605.25977v1 Announce Type: cross Abstract: This paper provides an empirical implementation of the creative quality metric proposed in Calibrated Surprise (Zou & Xu, 2026a). The question this pa

safetyarxiv-cs-ai
26 May 2026
Safety

CROCS: A Two-Stage Clustering Framework for Behaviour-Centric Consumer Segmentation with Smart Meter Data

DGX agent

arXiv:2601.10494v2 Announce Type: replace-cross Abstract: With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Managemen

safetyarxiv-cs-lg
26 May 2026
Safety

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents

DGX agent

arXiv:2605.25511v1 Announce Type: new Abstract: Recent advancements in Reinforcement Learning (RL), particularly Group Relative Policy Optimization (GRPO), have significantly enhanced the reasoning ca

safetyarxiv-cs-cl
26 May 2026
Safety

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

DGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

safetyarxiv-cs-ai
26 May 2026
Safety

CyBOKClaw: Human-in-the-Loop CyBOK Mapping for Cybersecurity Curriculum

DGX agent

arXiv:2605.24663v1 Announce Type: cross Abstract: This paper presents CyBOKClaw, an interpretable human-in-the-loop retrieval framework for mapping cybersecurity keywords or phrases (KWoPs) to the Cyb

safetyarxiv-cs-ai
26 May 2026
Safety

D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation

DGX agent

arXiv:2605.25022v1 Announce Type: cross Abstract: Dataset distillation (DD) aims to compress large-scale datasets into compact synthetic sets while preserving training efficacy. However, existing stud

safetyarxiv-cs-ai
26 May 2026
Safety

Dale meets Langevin: A Multiplicative Denoising Diffusion Model

DGX agent

arXiv:2510.02730v2 Announce Type: replace Abstract: Exponentiated gradient descent (EGD), a biologically motivated optimisation algorithm that respects Dale's law, produces log-normally distributed sy

safetyarxiv-cs-lg
26 May 2026
Safety

Data-Driven Optimization of Tactile Sensor Configurations for Efficient Dexterous Manipulation

DGX agent

arXiv:2409.20473v3 Announce Type: replace Abstract: Tactile sensing is critical for learning-based dexterous manipulation, yet principled guidelines for sensor placement remain largely absent. While d

safetyarxiv-cs-ro
26 May 2026
Safety

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care

DGX agent

arXiv:2510.08350v3 Announce Type: replace-cross Abstract: Objective: Enteral nutrition (EN) delivery in the ICU remains suboptimal due to limited personalization and uncertainty regarding appropriate

safetyarxiv-cs-ai
26 May 2026
Safety

DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading

DGX agent

arXiv:2605.25527v1 Announce Type: new Abstract: This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradien

safetyarxiv-cs-lg
26 May 2026
Safety

DeGRe: Dense-supervised Generative Reranking for Recommendation

DGX agent

arXiv:2605.25749v1 Announce Type: cross Abstract: In multi-stage recommender systems, reranking optimizes overall utility by capturing intra-list contextual dependencies, yet its central challenge lie

safetyarxiv-cs-ai
26 May 2026
Safety

Delta Energy: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization

DGX agent

arXiv:2510.11296v3 Announce Type: replace-cross Abstract: Recent approaches for vision-language models (VLMs) have shown remarkable success in achieving fast downstream adaptation. When applied to rea

safetyarxiv-cs-lg
26 May 2026
Safety

Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL

DGX agent

arXiv:2605.24001v1 Announce Type: cross Abstract: Recent advances in one-step text-to-image generation have enabled real-time synthesis with remarkable efficiency and quality. Previous reinforcement l

safetyarxiv-cs-ai
26 May 2026
Safety

Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs

DGX agent

arXiv:2605.23975v1 Announce Type: new Abstract: Audio large language models (Audio LLMs) exhibit systematic failures in transcribing code-switching speech despite strong multilingual capabilities. Foc

safetyarxiv-cs-cl
26 May 2026
Safety

Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards

DGX agent

arXiv:2603.18444v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective post-training paradigm for improving the reasoning capabilit

safetyarxiv-cs-ai
26 May 2026
Safety

Discovering Lexical Gaps Using Embeddings from Multilingual LLMs

DGX agent

arXiv:2605.24310v1 Announce Type: new Abstract: Lexical gaps are words that do not exist in certain languages. They pose challenges for building multilingual lexical resources, for machine translation

safetyarxiv-cs-cl
26 May 2026
Safety

Discrete diffusion samplers and bridges: Off-policy algorithms and applications in latent spaces

DGX agent

arXiv:2602.05961v2 Announce Type: replace Abstract: Sampling from a distribution p(x) propto e^{-E(x)} known up to a normalising constant is an important and challenging problem in statistics. Recent

safetyarxiv-cs-lg
26 May 2026
Safety

DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection

DGX agent

arXiv:2605.24639v1 Announce Type: cross Abstract: With the widespread application of drones in recent years, object detection of aerial images has attracted increasing attention, especially open-vocab

safetyarxiv-cs-ai
26 May 2026
Safety

Disentangled Double Machine Learning for Accurate Causal Effect Estimation

DGX agent

arXiv:2605.24808v1 Announce Type: cross Abstract: Confounding bias is a key challenge in causal effect estimation from observational data. Double Machine Learning (DML) addresses this issue by estimat

safetyarxiv-cs-ai
26 May 2026
Safety

Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion

DGX agent

arXiv:2601.21670v3 Announce Type: replace-cross Abstract: Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from domi

safetyarxiv-cs-lg
26 May 2026
Safety

Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal Models

DGX agent

arXiv:2603.17044v2 Announce Type: replace-cross Abstract: Unified multimodal models share a language model backbone for both understanding and generating images. Can DPO align both capabilities simult

safetyarxiv-cs-ai
26 May 2026
Safety

Document Classification Pattern Recognition via Information Fusion: A Systematic Review of Multimodal and Multiview Representation Approaches

DGX agent

arXiv:2605.23910v1 Announce Type: cross Abstract: Information fusion is used widely to improve document classification by the integration of multiple data sources (multimodal) or representations (mult

safetyarxiv-cs-ai
26 May 2026
Safety

DVAO: Dynamic Variance-adaptive Advantage Optimization for Multi-reward Reinforcement Learning

DGX agent

arXiv:2605.25604v1 Announce Type: new Abstract: Reinforcement Learning has become a standard paradigm for aligning Large Language Models with human intent and task requirements. While Group Relative P

safetyarxiv-cs-cl
26 May 2026
Safety

Dynamic Dual-Granularity Skill Bank for Agentic RL

DGX agent

arXiv:2603.28716v2 Announce Type: replace Abstract: Agentic RL can benefit substantially from reusable experience, yet existing skill-based methods mainly extract trajectory-level guidance and often l

safetyarxiv-cs-ai
26 May 2026
Safety

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

DGX agent

arXiv:2605.24924v1 Announce Type: new Abstract: Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency

safetyarxiv-cs-ro
26 May 2026
Safety

Dynamic Relational Priming Improves Transformer in Multivariate Time Series

DGX agent

arXiv:2509.12196v2 Announce Type: replace-cross Abstract: Standard attention mechanisms in transformers employ static token representations that remain unchanged across all pair-wise computations in e

safetyarxiv-cs-ai
26 May 2026
Safety

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

DGX agent

arXiv:2512.04733v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (AD) systems increasingly adopt vision-language-action (VLA) models, yet they typically ignore the passenger's e

safetyarxiv-cs-ai
26 May 2026
Safety

ECHO: Terminal Agents Learn World Models for Free

DGX agent

arXiv:2605.24517v1 Announce Type: cross Abstract: CLI agents are the closest thing language models have to an embodied setting: the model emits commands, the terminal executes them, and the returned s

safetyarxiv-cs-cl
26 May 2026
Safety

ECo-MoE: Embodiment-Conditioned Mixture of Experts Increases the Evolvability of Robots

DGX agent

arXiv:2605.24225v1 Announce Type: new Abstract: In this paper, we introduce a model of evolution and learning in robots that co-optimizes a distribution of latent design vectors (genotypes) and a mixt

safetyarxiv-cs-ro
26 May 2026
Safety

EMA: Effort Metric Attention for Anatomical Effort-Guided Human Motion Diffusion

DGX agent

arXiv:2605.24566v1 Announce Type: cross Abstract: Human motion diffusion models can synthesize action sequences from text, but controlling motion intensity remains challenging. Existing approaches rel

safetyarxiv-cs-lg
26 May 2026
Safety

Emergent Analogical Reasoning in Transformers

DGX agent

arXiv:2602.01992v4 Announce Type: replace Abstract: Analogy is a central faculty of human intelligence, enabling abstract patterns discovered in one domain to be applied to another. Despite its centra

safetyarxiv-cs-ai
26 May 2026
Safety

EPPC-OASIS: Ontology-Aware Adaptation and Structured Inference Refinement for Electronic Patient-Provider Communication Mining in Secure Messages

DGX agent

arXiv:2605.24172v1 Announce Type: new Abstract: Secure patient-provider messages contain clinically important communication behaviors that are difficult to characterize manually at scale. The Electron

safetyarxiv-cs-ai
26 May 2026
Safety

Eureka: Intelligent Feature Engineering for Enterprise AI Cloud Resource Demand Prediction

DGX agent

arXiv:2605.25297v1 Announce Type: cross Abstract: Effective features are crucial for predictive model performance, but creating them often requires domain expertise, limiting scalability across applic

safetyarxiv-cs-ai
26 May 2026
Safety

Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat

DGX agent

arXiv:2605.25091v1 Announce Type: new Abstract: As modern air combat evolves toward beyond-visual-range (BVR) multi-aircraft cooperative engagements, autonomous decision-making for unmanned combat aer

safetyarxiv-cs-ai
26 May 2026
Safety

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs

DGX agent

arXiv:2605.24345v1 Announce Type: new Abstract: In online reinforcement learning, data scarcity creates epistemic uncertainty that makes robustness important early in learning, whereas sufficient expl

safetyarxiv-cs-lg
26 May 2026
Safety

Extracting Training Data from Diffusion Language Models via Infilling

DGX agent

arXiv:2605.24173v1 Announce Type: cross Abstract: Memorization in large language models has been studied almost exclusively through prefix-conditioned extraction, a natural choice for autoregressive m

safetyarxiv-cs-ai
26 May 2026
Safety

Extreme Region Policy Distillation

DGX agent

arXiv:2605.25582v1 Announce Type: cross Abstract: Reinforcement learning for large language models faces a fundamental trade-off between sample efficiency and asymptotic performance: strictly on-polic

safetyarxiv-cs-ai
26 May 2026
Safety

Factored Latent Action World Models

DGX agent

arXiv:2602.16229v2 Announce Type: replace Abstract: Learning latent actions from action-free video has emerged as a powerful paradigm for scaling up controllable world model learning. Latent actions p

safetyarxiv-cs-lg
26 May 2026
Safety

FairJudge: Abstention-Aware Multimodal Judges for Fairness and Alignment Evaluation in Text-to-Image Models

DGX agent

arXiv:2510.22827v3 Announce Type: replace-cross Abstract: Evaluating text-to-image (T2I) systems requires judging not only whether an image matches a prompt, but also whether socially salient attribut

safetyarxiv-cs-lg
26 May 2026
Safety

Faithful or Fabricated? A Causal Framework for Rationalization Bias in LLM Judges

DGX agent

arXiv:2605.23970v1 Announce Type: new Abstract: Large language models (LLMs) are increasingly used as automatic judges for summarization and dialogue evaluation. Prior work has documented biases such

safetyarxiv-cs-cl
26 May 2026
Safety

Faithfulness as Information Flow: Evaluating and Training Faithful Chain-of-Thought Reasoning

DGX agent

arXiv:2605.24286v1 Announce Type: cross Abstract: Chain-of-thought (CoT) reasoning is useful for monitoring language models only when the reasoning trace faithfully reflects the computation that produ

safetyarxiv-cs-cl
26 May 2026
← Previous
1…159160161162163…260
Next →