AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,548
  • Agents7,263
  • Applications5,198
  • Concepts5
  • Hardware1,751
  • Industry6,096
  • Local Ai4,728
  • Model Releases22,555
  • Research19,193
  • Safety12,813
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent
84,548Total entries
1Added by human
84,547Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,813 results
26 May 2026

AI-Driven Controlled Environment Agriculture as Resilient Infrastructure for U.S. Fresh-Produce Supply Chains

SafetyDGX agent

arXiv:2605.23946v1 Announce Type: cross Abstract: Climate volatility, regional production concentration, labor constraints, cyber risk, and dependence on long-distance fresh-produce supply chains expo

Analogies between Transformer Layers and Power Method

SafetyDGX agent

arXiv:2605.25619v1 Announce Type: new Abstract: In the paper we show that there is an analogy between the operations occurring in a layer of a transformer (projections and layer normalizations, disreg

Anatomy-Anchored Self-Supervision: Distilling Vision Foundation Models for Invariant Ultrasound Representation

SafetyDGX agent

arXiv:2605.25402v1 Announce Type: cross Abstract: Self-supervised pre-training paradigm has gained increasing prominence for learning transferable representations in medical imaging, yet existing meth


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to t…

SafetyDGX agent

and you, through stock index funds, and because of some recent rule changes, will basically be forced to buy this garbage. Elon’s about to take a big chunk of your retirement, and unless you call your

Anthropic: fantasy versus reality.

SafetyDGX agent

Gary Marcus critiques Anthropic's claims about their AI capabilities, likely contrasting inflated marketing promises with the actual technical limitations and performance of their systems. The post pr

AnyScene: Towards Highly Controllable Driving Scene Generation at Anywhere and Beyond

SafetyDGX agent

arXiv:2605.26113v1 Announce Type: new Abstract: Generating high-fidelity and controllable synthetic data is critical for advancing end-to-end autonomous driving, particularly for addressing the long t

Approximating Safety Feedback Without a Safety Oracle via Model Predictive Control

SafetyDGX agent

arXiv:2510.20955v2 Announce Type: replace Abstract: Safe decision-making algorithms for control of mobile robots often require the existence of feedback to verify the safety of proposed actions. This

Auditing medical multi-agent AI reveals risks of false consensus

SafetyDGX agent

arXiv:2510.10185v2 Announce Type: replace-cross Abstract: Large language models are increasingly being assembled into medical multi-agent systems that emulate multidisciplinary consultation through sp

Authority Inversion in LLM-Mediated Ubiquitous Systems: When Models Trust Users Over Sensors

SafetyDGX agent

arXiv:2605.23938v1 Announce Type: new Abstract: Large language models (LLMs) increasingly fuse heterogeneous inputs in ubiquitous systems. Yet, how LLMs implicitly allocate authority when sensor measu

Autoregressive Language Models are Secretly Energy-Based Models: Insights into the Lookahead Capabilities of Next-Token Prediction

SafetyDGX agent

arXiv:2512.15605v4 Announce Type: replace Abstract: Autoregressive models (ARMs) currently constitute the dominant paradigm for large language models (LLMs). Energy-based models (EBMs) represent anoth

AvAtar: Learning to Align via Active Optimal Transport

SafetyDGX agent

arXiv:2605.24395v1 Announce Type: new Abstract: Alignment plays a fundamental role in many machine learning problems, such as multi-network analysis, multimodal learning, and point cloud registration.

Balancing Fairness, Privacy, and Accuracy: A Multitask Adversarial Framework for Centralized Data-Driven Systems

SafetyDGX agent

arXiv:2605.24458v1 Announce Type: cross Abstract: The integration of fairness and privacy in centralized data-driven applications is critical, especially as these systems increasingly influence sector

Beyond Killer Robots: General AI Attitudes and Public Support for Military AI in Nine Countries

SafetyDGX agent

arXiv:2605.25196v1 Announce Type: cross Abstract: AI-enabled military systems are a fixture of modern military conflict. Applications vary from autonomous drones for surveillance and attack to AI-supp

Beyond the Proxy: Trajectory-Distilled Guidance for Offline GFlowNet Training

SafetyDGX agent

arXiv:2505.20110v3 Announce Type: replace-cross Abstract: Generative Flow Networks (GFlowNets) excel at sampling diverse, high-reward objects. In many practical applications where active reward querie

Beyond the Target: From Imitation to Collaboration in Speculative Decoding

SafetyDGX agent

arXiv:2605.24793v1 Announce Type: new Abstract: Speculative decoding (SPD) accelerates large language model (LLM) inference by letting a smaller draft model propose multiple future tokens that are ver

Blocked Gibbs meets Diffusion Transformers: Unsupervised Learning for Constraint Optimization

SafetyDGX agent

arXiv:2605.25129v1 Announce Type: new Abstract: Diffusion models have shown promise in learning to solve constraint optimization problems. However, they are mostly restricted to problems with binary v

Bridging the Gap: Enabling Soft Actor Critic for High Performance Legged Locomotion

SafetyDGX agent

arXiv:2605.24975v1 Announce Type: cross Abstract: Proximal Policy Optimization (PPO) has become the de facto standard for training legged robots, thanks to its robustness and scalability in massively

Can Large Language Models Resolve Semantic Discrepancy in Self-Destructive Subcultures? Evidence from Jirai Kei

SafetyDGX agent

arXiv:2601.05004v2 Announce Type: replace Abstract: Self-destructive behaviors are linked to complex psychological states and can be challenging to diagnose. These behaviors may be even harder to iden

Capability and Robustness Cannot Both Be Free: An Information-Theoretic Bound for Vision-Language-Action Models

SafetyDGX agent

arXiv:2605.25889v1 Announce Type: cross Abstract: Vision-Language-Action (VLA) models are increasingly deployed on real robots, where each predicted action is executed and each failure carries a safet

Causal methods for LLM development and evaluation

SafetyDGX agent

arXiv:2605.25998v1 Announce Type: new Abstract: Large language model (LLM) development is currently driven by large-scale empirical iteration over data mixtures, reward models, routing strategies, and

Certified Robustness from Approximate Gaussian Mixture Structures in Pretrained Latent Spaces

SafetyDGX agent

arXiv:2605.25352v1 Announce Type: cross Abstract: Deep learning models are vulnerable to adversarial perturbations, raising important concerns for safety-critical deployment. Empirical defenses can ac

Characterizing Linear Alignment Across Language Models

SafetyDGX agent

arXiv:2603.18908v4 Announce Type: replace Abstract: Language models increasingly appear to learn similar representations, despite differences in training objectives, architectures, and data modalities

Choosing Online Experiment Designs under Interference in Ads, Recommendations, and Member-Experience Systems

SafetyDGX agent

arXiv:2605.25290v1 Announce Type: cross Abstract: Online experiments in ads, recommendation, and member-experience systems are often planned before the dominant interference mechanism is known. A trea

Clustering as Reasoning: A k-Means Interpretation of Chain-of-Thought Graph Learning

SafetyDGX agent

arXiv:2605.24867v1 Announce Type: new Abstract: Chain-of-Thought (CoT) prompting has shown promise in enhancing the reasoning capabilities of large language models (LLMs) on text-attributed graphs (TA

CODESKILL: Learning Self-Evolving Skills for Coding Agents

SafetyDGX agent

arXiv:2605.25430v1 Announce Type: new Abstract: Coding agents produce rich trajectories while solving software-engineering tasks. To enable agent self-evolution, these trajectories can be distilled in

ComPose: A Unified Completion-Pose Framework for Robust Category-Level Object Pose Estimation

SafetyDGX agent

arXiv:2605.25553v1 Announce Type: cross Abstract: Category-level object pose estimation aims to predict the pose and size of arbitrary objects in specific categories. Existing methods struggle with th

Concept Drift Adaptation Using Self-Supervised and Reinforcement Learning In Android Malware Detection

SafetyDGX agent

arXiv:2605.24294v1 Announce Type: cross Abstract: Android malware detectors often degrade after deployment because of concept drift, while full retraining at each maintenance step is costly. We propos

Contextual Rollout Bandits for Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2602.08499v2 Announce Type: replace-cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) is an effective paradigm for improving the reasoning capabilities of large language mode

Contractual Skills: A GovernSpec Design Framework for Enterprise AI Agents

SafetyDGX agent

arXiv:2605.22634v2 Announce Type: replace-cross Abstract: Skills have become a practical packaging mechanism for agent instructions, workflows, scripts, and reference materials. In enterprise settings

corruption

SafetyDGX agent

corruption Elon Musk has used Trump’s war in Iran to quintuple the amount he’s charging the Pentagon for SpaceX services. https://www.thedailybeast.com/elon-musks-spacex-uses-trumps-war-to-squeeze-mor

Counterfactually Safe Reinforcement Learning

SafetyDGX agent

arXiv:2605.25114v1 Announce Type: cross Abstract: Reinforcement learning algorithms are generally designed to maximize the expected return across a population. However, a policy that is optimal on ave

Creative Quality Alignment: Expert Tacit Knowledge Transfer via Chain-of-Thought Fine-Tuning

SafetyDGX agent

arXiv:2605.25977v1 Announce Type: cross Abstract: This paper provides an empirical implementation of the creative quality metric proposed in Calibrated Surprise (Zou & Xu, 2026a). The question this pa

CROCS: A Two-Stage Clustering Framework for Behaviour-Centric Consumer Segmentation with Smart Meter Data

SafetyDGX agent

arXiv:2601.10494v2 Announce Type: replace-cross Abstract: With grid operators confronting rising uncertainty from renewable integration and a broader push toward electrification, Demand-Side Managemen

CRPO: Character-centric Group Relative Policy Optimization for Role-aware Reasoning in Role-playing Agents

SafetyDGX agent

arXiv:2605.25511v1 Announce Type: new Abstract: Recent advancements in Reinforcement Learning (RL), particularly Group Relative Policy Optimization (GRPO), have significantly enhanced the reasoning ca

Cultivating Machine Intelligence: The OMEGA Shift from Top-Down Optimization to Autopoietic Cognitive Ecologies

SafetyDGX agent

arXiv:2605.25062v1 Announce Type: cross Abstract: The dominant artificial intelligence paradigm trains neural architectures via gradient descent against proxy objectives and reinforcement learning fro

CyBOKClaw: Human-in-the-Loop CyBOK Mapping for Cybersecurity Curriculum

SafetyDGX agent

arXiv:2605.24663v1 Announce Type: cross Abstract: This paper presents CyBOKClaw, an interpretable human-in-the-loop retrieval framework for mapping cybersecurity keywords or phrases (KWoPs) to the Cyb

D3S2: Diffusion-Guided Dataset Distillation for Semantic Segmentation

SafetyDGX agent

arXiv:2605.25022v1 Announce Type: cross Abstract: Dataset distillation (DD) aims to compress large-scale datasets into compact synthetic sets while preserving training efficacy. However, existing stud

Dale meets Langevin: A Multiplicative Denoising Diffusion Model

SafetyDGX agent

arXiv:2510.02730v2 Announce Type: replace Abstract: Exponentiated gradient descent (EGD), a biologically motivated optimisation algorithm that respects Dale's law, produces log-normally distributed sy

Data-Driven Optimization of Tactile Sensor Configurations for Efficient Dexterous Manipulation

SafetyDGX agent

arXiv:2409.20473v3 Announce Type: replace Abstract: Tactile sensing is critical for learning-based dexterous manipulation, yet principled guidelines for sensor placement remain largely absent. While d

DBPnet: Damper Characteristics-Based Bayesian Physics-Informed Neural Network for Wheel Load Estimation

SafetyDGX agent

arXiv:2605.24860v1 Announce Type: cross Abstract: Advanced driver assistance systems (ADAS) play an important role in modern automotive intelligence, significantly enhancing vehicle safety and stabili

DeepEN: A Deep Reinforcement Learning Framework for Personalized Enteral Nutrition in Critical Care

SafetyDGX agent

arXiv:2510.08350v3 Announce Type: replace-cross Abstract: Objective: Enteral nutrition (EN) delivery in the ICU remains suboptimal due to limited personalization and uncertainty regarding appropriate

DeepSeekMath Meets Order Book: Group-Aware Policy Optimization for High-Frequency Directional Trading

SafetyDGX agent

arXiv:2605.25527v1 Announce Type: new Abstract: This paper studies reinforcement learning for high-frequency trading on limit order books by pairing an Order-Flow-based state model with policy-gradien

DeGRe: Dense-supervised Generative Reranking for Recommendation

SafetyDGX agent

arXiv:2605.25749v1 Announce Type: cross Abstract: In multi-stage recommender systems, reranking optimizes overall utility by capturing intra-list contextual dependencies, yet its central challenge lie

Delta Energy: Optimizing Energy Change During Vision-Language Alignment Improves both OOD Detection and OOD Generalization

SafetyDGX agent

arXiv:2510.11296v3 Announce Type: replace-cross Abstract: Recent approaches for vision-language models (VLMs) have shown remarkable success in achieving fast downstream adaptation. When applied to rea

Diff-Instruct with Diffused Reward: Towards Principled One-step Generator RL

SafetyDGX agent

arXiv:2605.24001v1 Announce Type: cross Abstract: Recent advances in one-step text-to-image generation have enabled real-time synthesis with remarkable efficiency and quality. Previous reinforcement l

Direct Preference Optimization for English-Mandarin Code-Switching Speech Recognition in Audio LLMs

SafetyDGX agent

arXiv:2605.23975v1 Announce Type: new Abstract: Audio large language models (Audio LLMs) exhibit systematic failures in transcribing code-switching speech despite strong multilingual capabilities. Foc

Discounted Beta-Bernoulli Reward Estimation for Sample-Efficient Reinforcement Learning with Verifiable Rewards

SafetyDGX agent

arXiv:2603.18444v2 Announce Type: replace-cross Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as an effective post-training paradigm for improving the reasoning capabilit

Discovering Lexical Gaps Using Embeddings from Multilingual LLMs

SafetyDGX agent

arXiv:2605.24310v1 Announce Type: new Abstract: Lexical gaps are words that do not exist in certain languages. They pose challenges for building multilingual lexical resources, for machine translation

Discrete diffusion samplers and bridges: Off-policy algorithms and applications in latent spaces

SafetyDGX agent

arXiv:2602.05961v2 Announce Type: replace Abstract: Sampling from a distribution p(x) propto e^{-E(x)} known up to a normalising constant is an important and challenging problem in statistics. Recent

DisDop: Distillation with Domain Priors for Open-Vocabulary Aerial Object Detection

SafetyDGX agent

arXiv:2605.24639v1 Announce Type: cross Abstract: With the widespread application of drones in recent years, object detection of aerial images has attracted increasing attention, especially open-vocab

Disentangled Double Machine Learning for Accurate Causal Effect Estimation

SafetyDGX agent

arXiv:2605.24808v1 Announce Type: cross Abstract: Confounding bias is a key challenge in causal effect estimation from observational data. Double Machine Learning (DML) addresses this issue by estimat

Diverse via bounded Agreement: Geometric Regularization for Multimodal Fusion

SafetyDGX agent

arXiv:2601.21670v3 Announce Type: replace-cross Abstract: Multimodal fusion is often treated as an optimization-balancing problem, where training signals are adjusted to prevent one modality from domi

Do Understanding and Generation Fight? A Diagnostic Study of DPO for Unified Multimodal Models

SafetyDGX agent

arXiv:2603.17044v2 Announce Type: replace-cross Abstract: Unified multimodal models share a language model backbone for both understanding and generating images. Can DPO align both capabilities simult

Document Classification Pattern Recognition via Information Fusion: A Systematic Review of Multimodal and Multiview Representation Approaches

SafetyDGX agent

arXiv:2605.23910v1 Announce Type: cross Abstract: Information fusion is used widely to improve document classification by the integration of multiple data sources (multimodal) or representations (mult

DVAO: Dynamic Variance-adaptive Advantage Optimization for Multi-reward Reinforcement Learning

SafetyDGX agent

arXiv:2605.25604v1 Announce Type: new Abstract: Reinforcement Learning has become a standard paradigm for aligning Large Language Models with human intent and task requirements. While Group Relative P

Dynamic Dual-Granularity Skill Bank for Agentic RL

SafetyDGX agent

arXiv:2603.28716v2 Announce Type: replace Abstract: Agentic RL can benefit substantially from reusable experience, yet existing skill-based methods mainly extract trajectory-level guidance and often l

Dynamic Neural Koopman Distillation for Real-Time Robot Control Using Diffusion Models

SafetyDGX agent

arXiv:2605.24924v1 Announce Type: new Abstract: Diffusion models excel at generating diverse and multimodal trajectories for robotic planning, yet their iterative denoising process introduces latency

Dynamic Optimization and Safety Indicator Injection for Jailbreaking Text-to-Image Models with Multimodal Safety Filters

SafetyDGX agent

arXiv:2505.18979v2 Announce Type: replace Abstract: Text-to-image (T2I) models can generate not-safe-for-work (NSFW) content, motivating multi-stage safety pipelines with both text and image filters.

Dynamic Relational Priming Improves Transformer in Multivariate Time Series

SafetyDGX agent

arXiv:2509.12196v2 Announce Type: replace-cross Abstract: Standard attention mechanisms in transformers employ static token representations that remain unchanged across all pair-wise computations in e

E3AD: An Emotion-Aware Vision-Language-Action Model for Human-Centric End-to-End Autonomous Driving

SafetyDGX agent

arXiv:2512.04733v2 Announce Type: replace-cross Abstract: End-to-end autonomous driving (AD) systems increasingly adopt vision-language-action (VLA) models, yet they typically ignore the passenger's e

← Previous
1…115116117118119…214
Next →