AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,487 results
26 May 2026

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to t…

SafetyDGX agent

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to take a big chunk of your retirement, and unless you call your

Authority Inversion in LLM-Mediated Ubiquitous Systems: When Models Trust Users Over Sensors

SafetyDGX agent

arXiv:2605.23938v1 Announce Type: new Abstract: Large language models (LLMs) increasingly fuse heterogeneous inputs in ubiquitous systems. Yet, how LLMs implicitly allocate authority when sensor measu

Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction

SafetyDGX agent
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

arXiv:2512.15605v4 Announce Type: replace Abstract: Autoregressive models (ARMs) currently constitute the dominant paradigm for large language models (LLMs). Energy-based models (EBMs) represent anoth

AvAtar: Learning to Align via Active Optimal Transport

SafetyDGX agent

arXiv:2605.24395v1 Announce Type: new Abstract: Alignment plays a fundamental role in many machine learning problems, such as multi-network analysis, multimodal learning, and point cloud registration.

Balancing Fairness, Privacy, and Accuracy: A Multitask Adversarial Framework for Centralized Data-Driven Systems

SafetyDGX agent

arXiv:2605.24458v1 Announce Type: cross Abstract: The integration of fairness and privacy in centralized data-driven applications is critical, especially as these systems increasingly influence sector

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

SafetyDGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

SafetyDGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

SafetyDGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization

SafetyDGX agent

arXiv:2605.25129v1 Announce Type: new Abstract: Diffusion models have shown promise in learning to solve constraint optimization problems. However, they are mostly restricted to problems with binary v

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

SafetyDGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

SafetyDGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

Causal methods for LLM development and evaluation

SafetyDGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

Characterizing Linear Alignment Across Language Models

SafetyDGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems

SafetyDGX agent

arXiv:2605.25290v1 Announce Type: cross Abstract: Online experiments in ads, recommendation, and member-experience systems are often planned before the dominant interference mechanism is known. A trea

Clustering as Reasoning: A k-Means Interpretation of Chain-of-Thought Graph Learning

SafetyDGX agent

arXiv:2605.24867v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) on text-attributed graphs (TA

CODESKILL: Learning Self-Evolving Skills for Coding Agents

SafetyDGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation

SafetyDGX agent

arXiv:2605.25553v1 Announce Type: cross Abstract: Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with th

Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection

SafetyDGX agent

arXiv:2605.24294v1 Announce Type: cross Abstract: Android malware detectors often degrade after deployment because of concept drift, while full retraining at each maintenance step is costly. We propos

Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2602.08499v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an effective paradigm for improving the reasoning capabilities of large language mode

corruption

SafetyDGX agent

corruption Elon Musk has used Trump’s war in Iran to quintuple the amount he’s charging the Pentagon for SpaceX services. https://www.thedailybeast.com/elon-musks-spacex-uses-trumps-war-to-squeeze-mor

Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning

SafetyDGX agent

arXiv:2605.25977v1 Announce Type: cross Abstract: This paper provides an empirical implementation of the creative quality metric proposed in Calibrated Surprise (Zou & Xu, 2026a). The question this pa

CROCS: A Two-Stage Clustering Framework for Behaviour-Centric Consumer Segmentation with Smart Meter Data

SafetyDGX agent

arXiv:2601.10494v2 Announce Type: replace-cross Abstract: With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Managemen

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents

SafetyDGX agent

arXiv:2605.25511v1 Announce Type: new Abstract: Recent advancements in Reinforcement Learning (RL), particularly Group Relative Policy Optimization (GRPO), have significantly enhanced the reasoning ca

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

SafetyDGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

CyBOKClaw: Human-in-the-Loop CyBOK Mapping for Cybersecurity Curriculum

SafetyDGX agent

arXiv:2605.24663v1 Announce Type: cross Abstract: This paper presents CyBOKClaw, an interpretable human-in-the-loop retrieval framework for mapping cybersecurity keywords or phrases (KWoPs) to the Cyb

D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation

SafetyDGX agent

arXiv:2605.25022v1 Announce Type: cross Abstract: Dataset distillation (DD) aims to compress large-scale datasets into compact synthetic sets while preserving training efficacy. However, existing stud

Dale meets Langevin: A Multiplicative Denoising Diffusion Model

SafetyDGX agent

arXiv:2510.02730v2 Announce Type: replace Abstract: Exponentiated gradient descent (EGD), a biologically motivated optimisation algorithm that respects Dale's law, produces log-normally distributed sy

Data-Driven Optimization of Tactile Sensor Configurations for Efficient Dexterous Manipulation

SafetyDGX agent

arXiv:2409.20473v3 Announce Type: replace Abstract: Tactile sensing is critical for learning-based dexterous manipulation, yet principled guidelines for sensor placement remain largely absent. While d

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care

SafetyDGX agent

arXiv:2510.08350v3 Announce Type: replace-cross Abstract: Objective: Enteral nutrition (EN) delivery in the ICU remains suboptimal due to limited personalization and uncertainty regarding appropriate

DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading

SafetyDGX agent

arXiv:2605.25527v1 Announce Type: new Abstract: This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradien

DeGRe: Dense-supervised Generative Reranking for Recommendation

SafetyDGX agent

arXiv:2605.25749v1 Announce Type: cross Abstract: In multi-stage recommender systems, reranking optimizes overall utility by capturing intra-list contextual dependencies, yet its central challenge lie

Delta Energy: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization

SafetyDGX agent

arXiv:2510.11296v3 Announce Type: replace-cross Abstract: Recent approaches for vision-language models (VLMs) have shown remarkable success in achieving fast downstream adaptation. When applied to rea

Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL

SafetyDGX agent

arXiv:2605.24001v1 Announce Type: cross Abstract: Recent advances in one-step text-to-image generation have enabled real-time synthesis with remarkable efficiency and quality. Previous reinforcement l

Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs

SafetyDGX agent

arXiv:2605.23975v1 Announce Type: new Abstract: Audio large language models (Audio LLMs) exhibit systematic failures in transcribing code-switching speech despite strong multilingual capabilities. Foc

Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2603.18444v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective post-training paradigm for improving the reasoning capabilit

Discovering Lexical Gaps Using Embeddings from Multilingual LLMs

SafetyDGX agent

arXiv:2605.24310v1 Announce Type: new Abstract: Lexical gaps are words that do not exist in certain languages. They pose challenges for building multilingual lexical resources, for machine translation

Discrete diffusion samplers and bridges: Off-policy algorithms and applications in latent spaces

SafetyDGX agent

arXiv:2602.05961v2 Announce Type: replace Abstract: Sampling from a distribution p(x) propto e^{-E(x)} known up to a normalising constant is an important and challenging problem in statistics. Recent

DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection

SafetyDGX agent

arXiv:2605.24639v1 Announce Type: cross Abstract: With the widespread application of drones in recent years, object detection of aerial images has attracted increasing attention, especially open-vocab

Disentangled Double Machine Learning for Accurate Causal Effect Estimation

SafetyDGX agent

arXiv:2605.24808v1 Announce Type: cross Abstract: Confounding bias is a key challenge in causal effect estimation from observational data. Double Machine Learning (DML) addresses this issue by estimat

Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion

SafetyDGX agent

arXiv:2601.21670v3 Announce Type: replace-cross Abstract: Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from domi

Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal Models

SafetyDGX agent

arXiv:2603.17044v2 Announce Type: replace-cross Abstract: Unified multimodal models share a language model backbone for both understanding and generating images. Can DPO align both capabilities simult

Document Classification Pattern Recognition via Information Fusion: A Systematic Review of Multimodal and Multiview Representation Approaches

SafetyDGX agent

arXiv:2605.23910v1 Announce Type: cross Abstract: Information fusion is used widely to improve document classification by the integration of multiple data sources (multimodal) or representations (mult

DVAO: Dynamic Variance-adaptive Advantage Optimization for Multi-reward Reinforcement Learning

SafetyDGX agent

arXiv:2605.25604v1 Announce Type: new Abstract: Reinforcement Learning has become a standard paradigm for aligning Large Language Models with human intent and task requirements. While Group Relative P

Dynamic Dual-Granularity Skill Bank for Agentic RL

SafetyDGX agent

arXiv:2603.28716v2 Announce Type: replace Abstract: Agentic RL can benefit substantially from reusable experience, yet existing skill-based methods mainly extract trajectory-level guidance and often l

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

SafetyDGX agent

arXiv:2605.24924v1 Announce Type: new Abstract: Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency

Dynamic Relational Priming Improves Transformer in Multivariate Time Series

SafetyDGX agent

arXiv:2509.12196v2 Announce Type: replace-cross Abstract: Standard attention mechanisms in transformers employ static token representations that remain unchanged across all pair-wise computations in e

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2512.04733v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (AD) systems increasingly adopt vision-language-action (VLA) models, yet they typically ignore the passenger's e

ECHO: Terminal Agents Learn World Models for Free

SafetyDGX agent

arXiv:2605.24517v1 Announce Type: cross Abstract: CLI agents are the closest thing language models have to an embodied setting: the model emits commands, the terminal executes them, and the returned s

ECo-MoE: Embodiment-Conditioned Mixture of Experts Increases the Evolvability of Robots

SafetyDGX agent

arXiv:2605.24225v1 Announce Type: new Abstract: In this paper, we introduce a model of evolution and learning in robots that co-optimizes a distribution of latent design vectors (genotypes) and a mixt

EMA: Effort Metric Attention for Anatomical Effort-Guided Human Motion Diffusion

SafetyDGX agent

arXiv:2605.24566v1 Announce Type: cross Abstract: Human motion diffusion models can synthesize action sequences from text, but controlling motion intensity remains challenging. Existing approaches rel

Emergent Analogical Reasoning in Transformers

SafetyDGX agent

arXiv:2602.01992v4 Announce Type: replace Abstract: Analogy is a central faculty of human intelligence, enabling abstract patterns discovered in one domain to be applied to another. Despite its centra

EPPC-OASIS: Ontology-Aware Adaptation and Structured Inference Refinement for Electronic Patient-Provider Communication Mining in Secure Messages

SafetyDGX agent

arXiv:2605.24172v1 Announce Type: new Abstract: Secure patient-provider messages contain clinically important communication behaviors that are difficult to characterize manually at scale. The Electron

Eureka: Intelligent Feature Engineering for Enterprise AI Cloud Resource Demand Prediction

SafetyDGX agent

arXiv:2605.25297v1 Announce Type: cross Abstract: Effective features are crucial for predictive model performance, but creating them often requires domain expertise, limiting scalability across applic

Europe’s sovereign AI moment arrives in Heilbronn next week. At TECH by Handelsblatt 2026, Cohere CEO and Co-founder, @aidangomez, will join…

SafetyDGX agent

Europe’s sovereign AI moment arrives in Heilbronn next week. At TECH by Handelsblatt 2026, Cohere CEO and Co-founder, @aidangomez, will join leaders across business, policy, and industry to discuss ho

even by the standards of the last few years, we are in some truly insane territory here.

SafetyDGX agent

even by the standards of the last few years, we are in some truly insane territory here. “Unserious, empty, hallucinatory, and borderline dishonest” - the prospectus for a company that the S&P 500 is

Evolutionary Enhanced Multi-Agent Reinforcement Learning for Cooperative Air Combat

SafetyDGX agent

arXiv:2605.25091v1 Announce Type: new Abstract: As modern air combat evolves toward beyond-visual-range (BVR) multi-aircraft cooperative engagements, autonomous decision-making for unmanned combat aer

Evolving Robustness--Exploration Trade-off in Online Reinforcement Learning via Quantile Bayesian Risk MDPs

SafetyDGX agent

arXiv:2605.24345v1 Announce Type: new Abstract: In online reinforcement learning, data scarcity creates epistemic uncertainty that makes robustness important early in learning, whereas sufficient expl

Extracting Training Data from Diffusion Language Models via Infilling

SafetyDGX agent

arXiv:2605.24173v1 Announce Type: cross Abstract: Memorization in large language models has been studied almost exclusively through prefix-conditioned extraction, a natural choice for autoregressive m

Extreme Region Policy Distillation

SafetyDGX agent

arXiv:2605.25582v1 Announce Type: cross Abstract: Reinforcement learning for large language models faces a fundamental trade-off between sample efficiency and asymptotic performance: strictly on-polic

Factored Latent Action World Models

SafetyDGX agent

arXiv:2602.16229v2 Announce Type: replace Abstract: Learning latent actions from action-free video has emerged as a powerful paradigm for scaling up controllable world model learning. Latent actions p

← Previous
1…145146147148149…242
Next →