AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,490 results
9 Jun 2026

Know More, Know Clearer: A Meta-Cognitive Framework for Knowledge Augmentation in Large Language Models

SafetyDGX agent

arXiv:2602.12996v2 Announce Type: replace-cross Abstract: Knowledge augmentation has significantly enhanced the performance of Large Language Models (LLMs) in knowledge-intensive tasks. However, exist

LAEI: Layered Autonomous Edge Intelligence Framework for Robust UAV Swarm Operations

SafetyDGX agent

arXiv:2606.09099v1 Announce Type: new Abstract: Autonomous UAV swarms require scalable coordination mechanisms that maintain mission performance under limited communication, environmental uncertainty,

Latent Diffusion Policy: Shaping Latent Spaces for Diffusion-Based Robotic Manipulation

SafetyDGX agent

arXiv:2606.08657v1 Announce Type: cross Abstract: Diffusion-based visuomotor policies operating directly in raw action spaces conflate scene comprehension with trajectory generation within a single de

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Latent Spherical Flow Policy for Reinforcement Learning with Combinatorial Actions

SafetyDGX agent

arXiv:2601.22211v2 Announce Type: replace Abstract: Reinforcement learning (RL) with combinatorial action spaces remains challenging because feasible action sets are exponentially large and governed b

Layer-wise Derivative Controlled Networks Achieve Competitive Accuracy and Gradient Stability Across Data Regimes

SafetyDGX agent

arXiv:2606.07908v1 Announce Type: new Abstract: Derivative-controlled networks based on ChainzRule (CR) combine cubic polynomial layers with a lightweight forward-mode per-layer Jacobian penalty (DREG

LCAM: A Framework for Diagnosing Interactional Alignment Failures in Con-versational AI

SafetyDGX agent

arXiv:2606.08131v1 Announce Type: cross Abstract: Conversational AI is increasingly used for advice, interpretation, reassurance, and decision support in contexts where users may be vulnerable, uncert

Learning Quantized Continuous Controllers for Integer Hardware

SafetyDGX agent

arXiv:2511.07046v4 Announce Type: replace-cross Abstract: Deploying continuous-control reinforcement learning policies on embedded hardware requires meeting tight latency and power budgets. Small FPGA

Light-WAM: Efficient World Action Models with State-Fusion Action Decoding

SafetyDGX agent

arXiv:2606.08242v1 Announce Type: new Abstract: World Action Models (WAMs) extend robot policy learning by incorporating future prediction as an additional training objective, encouraging the policy t

LiteVSR: Lightweight Adaptation of Frozen Diffusion Transformers for Video Super-Resolution

SafetyDGX agent

arXiv:2606.09250v1 Announce Type: new Abstract: Adapting large-scale pre-trained video generators for Video Super-Resolution (VSR) in novel domains remains computationally prohibitive. Methods that re

LogNEO: A GPT-Neo Reinforcement Learning Framework for Accurate Real-Time Log Anomaly Detection

SafetyDGX agent

arXiv:2606.08153v1 Announce Type: cross Abstract: Detecting anomalies in large-scale system logs is critical for the reliability and security of modern computing infrastructure. We present LogNEO, a l

MaskAlign: Token-Subset Representation Alignment for Efficient Diffusion Training

SafetyDGX agent

arXiv:2606.08788v1 Announce Type: new Abstract: Representation alignment with pretrained vision models has recently shown strong potential for accelerating diffusion transformer training. By aligning

MC-PDD: Masked Corpus-Level Pretraining Data Detection for Black-Box Large Language Models

SafetyDGX agent

arXiv:2606.07996v1 Announce Type: cross Abstract: Pretraining is fundamental to the development of Large Language Models (LLMs), yet the opacity of pretraining data complicates model analysis and rais

me January 2025 @politico, timing slightly off: we may soon see “the largest cyberattack in history, taking down, at least for a little whil…

SafetyDGX agent

me January 2025 @politico, timing slightly off: we may soon see “the largest cyberattack in history, taking down, at least for a little while, some sizeable piece of the world’s infrastructure… genera

MEC-Cox: Machine-Learning-Assisted Generalized Entropy Calibration for ATT Marginal Hazard-Ratio Estimation

SafetyDGX agent

arXiv:2606.08305v1 Announce Type: cross Abstract: Externally controlled survival trials are increasingly used when concurrent randomized controls are infeasible, particularly in oncology and rare-dise

Mind Your Steps: A General Learning Framework for Accurate Humanoid Foothold Tracking

SafetyDGX agent

arXiv:2606.08253v1 Announce Type: cross Abstract: Enabling humanoid robots to operate in complex, dynamic environments remains a critical challenge, fundamentally limited by the ability to navigate ro

MIRAGE: Metadata-Integrated Repository Analysis and Guided Enhancement for MSR Datasets

SafetyDGX agent

arXiv:2606.07611v1 Announce Type: cross Abstract: This paper proposes an improved approach to the analysis of Mining Software Repositories (MSR) datasets via metadata enrichment, FAIRness assessment,

mllm-shap: A Shapley Value Explainability Platform for Text-Audio Multimodal Large Language Models

SafetyDGX agent

arXiv:2606.07531v1 Announce Type: cross Abstract: We introduce mllm-shap, an open-source Python framework designed to extend Shapley Value (SV) explainability from text-only Large Language Models to M

Momentum for Reasoning: Dense Intrinsic Signals in Policy Optimization

SafetyDGX agent

arXiv:2606.08815v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) has emerged as a powerful paradigm for eliciting long-chain reasoning in large language models. Ho

Monday’s gains are Tuesday’s losses?

SafetyDGX agent

This post likely discusses the volatility and unpredictability of financial markets, examining whether gains achieved on one trading day are frequently reversed or lost the following day. Gary Marcus

MotionWAM: Towards Foundation World Action Models for Real-Time Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.09215v1 Announce Type: new Abstract: World Action Models (WAMs) couple a video dynamics prior to the policy and have shown encouraging results on tabletop manipulation, but iterative denois

Multimodal Generative Engine Optimization: Rank Manipulation for Vision-Language Model Rankers

SafetyDGX agent

arXiv:2601.12263v2 Announce Type: replace-cross Abstract: Vision-Language Models (VLMs) integrate visual and textual knowledge into unified representations that increasingly underpin modern retrieval

Muses: Designing, Composing, Generating Nonexistent Fantasy 3D Creatures without Training

SafetyDGX agent

arXiv:2601.03256v2 Announce Type: replace Abstract: We present Muses, the first training-free method for fantastic 3D creature generation in a feed-forward paradigm. Previous methods, which rely on pa

NeuroAlign: Hierarchical Multimodal Fusion of Dynamic and Structural Neuroimaging for MCI Analysis

SafetyDGX agent

arXiv:2606.07635v1 Announce Type: cross Abstract: Multimodal neuroimaging fusion of functional MRI (fMRI) and diffusion tensor imaging (DTI) provides complementary information for cognitive impairment

Neutrality Bites: Gender Representation in AI-Generated Animal Stories

SafetyDGX agent

arXiv:2606.07969v1 Announce Type: cross Abstract: Gender bias in AI-generated stories is a well-documented problem. While much attention has been paid to reducing or mitigating this bias, it is not al

no

SafetyDGX agent

no Hmm, @GaryMarcus , do you think they’d be able to achieve *any* of these goals? 😁 “Currently at OpenAI we have three main goals… 1. Build an automated AI researcher. 2. Accelerate the economy. 3. G

No Modality Left Behind: Adapting to Missing Modalities via Knowledge Distillation for Brain Tumor Segmentation

SafetyDGX agent

arXiv:2509.15017v2 Announce Type: replace Abstract: Accurate brain tumor segmentation is essential for preoperative evaluation and personalized treatment. Multi-modal MRI is widely used due to its abi

Nope. I think the positive press for @pontifex is because the Pope cares about humanity, and the consequence of technology for humanity. And…

SafetyDGX agent

Nope. I think the positive press for @pontifex is because the Pope cares about humanity, and the consequence of technology for humanity. And that the tech bros mostly don’t. He’s also sharper on philo

NoRD: A Data-Efficient Vision-Language-Action Model that Drives without Reasoning

SafetyDGX agent

arXiv:2602.21172v3 Announce Type: replace Abstract: Vision-Language-Action (VLA) models are advancing autonomous driving by replacing modular pipelines with unified end-to-end architectures. However,

Normality Calibration in Semi-supervised Graph Anomaly Detection

SafetyDGX agent

arXiv:2510.02014v3 Announce Type: replace Abstract: Graph anomaly detection (GAD) has attracted growing interest for its crucial ability to uncover irregular patterns in broad applications. Semi-super

OASIS: From Simulation Data Collection to Real-World Humanoid Loco-Manipulation

SafetyDGX agent

arXiv:2606.08548v1 Announce Type: new Abstract: Recent progress in robot manipulation has been largely driven by learning from large-scale demonstrations. For humanoid robot loco-manipulation tasks, h

omega-EVA: Envision, Verify, and Act with Latent Interactive World Models

SafetyDGX agent

arXiv:2606.09457v1 Announce Type: new Abstract: Embodied policies typically map current observations directly to actions, leaving candidate-action consequences implicit. World models provide predictiv

Operationalising the Superficial Alignment Hypothesis via Task Complexity

SafetyDGX agent

arXiv:2602.15829v2 Announce Type: replace Abstract: The superficial alignment hypothesis (SAH) posits that large language models learn most of their knowledge during pre-training, and that post-traini

Optimal Fair Aggregation of Crowdsourced Noisy Labels using Demographic Parity Constraints

SafetyDGX agent

arXiv:2601.23221v2 Announce Type: replace Abstract: As acquiring reliable ground-truth labels is usually costly, or infeasible, crowdsourcing and aggregation of noisy human annotations is the typical

OrderDP: A Theoretically Guaranteed Lossless Dynamic Data Pruning Framework

SafetyDGX agent

arXiv:2606.08574v1 Announce Type: cross Abstract: Data pruning (DP), as an oft-stated strategy to alleviate heavy training burdens, reduces the volume of training samples according to a well-defined p

Outage Detection in Self-Healing Smart Grids Using Reinforcement Learning with Spectral Graph Neural Networks

SafetyDGX agent

arXiv:2606.07583v1 Announce Type: cross Abstract: Self-healing smart grids can quickly adjust their network configuration during outages to minimize power disruptions. During an outage, several action

PAEC: Position-Aware Entropy Calibration for LLM Reasoning in RLVR

SafetyDGX agent

arXiv:2606.08543v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards (RLVR) improves large language model reasoning but often suffers from rapid policy-entropy collapse, wher

PAFO: Pareto Fairness Optimization for Personalized Reward Modeling

SafetyDGX agent

arXiv:2606.07988v1 Announce Type: new Abstract: Large language models (LLMs) increasingly rely on reward models to align their outputs with diverse user preferences. While personalized reward models a

PairAlign: A Framework for Sequence Tokenization via Self-Alignment with Applications to Audio Tokenization

SafetyDGX agent

arXiv:2605.06582v2 Announce Type: replace Abstract: Many operations on sensory data -- comparison, memory, retrieval, and reasoning -- are naturally expressed over discrete symbolic structures. In lan

PairWise Image Finder: An Open-source Tool for Finding Visually Aligned Street-Level Image Pairs for Urban Perception Studies

SafetyDGX agent

arXiv:2606.08795v1 Announce Type: new Abstract: Change detection and scene recognition techniques have been widely applied to Street View Imagery (SVI) to understand changes in scenes across the years

Path Planning Using Deep Deterministic Policy Gradient: A Reinforcement Learning Approach

SafetyDGX agent

arXiv:2606.07855v1 Announce Type: new Abstract: Path-planning for autonomous vehicles in threat-laden environments is a fundamental challenge because the problem is nonlinear and nonconvex even in sim

PathoSage: Towards Multi-Source Evidence Adjudication in Pathology via Experience-Aware Agentic Workflow

SafetyDGX agent

arXiv:2606.07549v1 Announce Type: new Abstract: Recent advances in Multimodal Large Language Models (MLLMs) and agent workflows have shown strong promise for computational pathology, yet reliable patc

Payoff scaling shapes cooperation in LLM agents across languages

SafetyDGX agent

arXiv:2601.19082v2 Announce Type: replace Abstract: Large language models (LLMs) are increasingly deployed as autonomous agents that negotiate, coordinate, and act on behalf of users. Whether they coo

PBSD: Privileged Bayesian Self-Distillation for Long-Horizon Credit Assignment

SafetyDGX agent

arXiv:2606.09348v1 Announce Type: new Abstract: Long-horizon agentic tasks pose a fundamental credit assignment challenge for outcome-base reinforcement learning: trajectory-level rewards verify final

Petri Net Modeling and Deadlock-Free Scheduling of Attachable Heterogeneous AGV Systems

SafetyDGX agent

arXiv:2508.00724v2 Announce Type: replace-cross Abstract: The increasing demand for flexible automation has accelerated the adoption of heterogeneous automated guided vehicles (AGVs). This work invest

Physically Consistent Null Space Alignment for Detection of Low-Magnitude False Data Injection Attacks

SafetyDGX agent

arXiv:2606.08473v1 Announce Type: new Abstract: False data injection attacks (FDIAs) introducing small measurement perturbations can still cause large deviations in power system state estimation when

Polaffini: A feature-based approach for robust affine and polyaffine image registration

SafetyDGX agent

arXiv:2602.17337v2 Announce Type: replace Abstract: In this work we present Polaffini, a robust and versatile framework for anatomically grounded registration. Medical image registration is dominated

PriFT: Prior-Support Guided Supervised Fine-Tuning

SafetyDGX agent

arXiv:2606.09396v1 Announce Type: cross Abstract: Supervised fine-tuning (SFT) is an efficient approach for downstream task adaptation and often serves as the initialization stage for reinforcement le

Prisma-World: Camera-Controllable Multi-Agent Video World Model

SafetyDGX agent

arXiv:2606.09507v1 Announce Type: new Abstract: Video world models have made rapid progress in generating controllable visual experiences, but most of them still simulate the world from a single obser

PROBE-Web: An Interactive System for Probing Evaluation Landscapes of Knowledge Graph Completion Models

SafetyDGX agent

arXiv:2606.08926v1 Announce Type: new Abstract: Knowledge graph completion (KGC) models are commonly evaluated using rank-based metrics such as MRR and Hits@K, despite different users often requiring

Property-Informed Diffusion-Based Text-to-Microstructure Generation

SafetyDGX agent

arXiv:2606.08150v1 Announce Type: new Abstract: Designing 3D metamaterial microstructures that meet the intended functions remains a major challenge, as it typically requires domain expertise, iterati

Proxy Reward Internalization and Mechanistic Exploitation: A Learned Precursor to Reward Hacking and Its Generalization

SafetyDGX agent

arXiv:2606.09711v1 Announce Type: new Abstract: Reward hacking is usually studied after it becomes visible, once a model earns high proxy reward while failing the intended task. We instead study what

PRPO: Perception-Reinforced Policy Optimization via Token-Level Dynamic Advantage Reshaping

SafetyDGX agent

arXiv:2606.08708v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become an effective paradigm for improving the reasoning capability of Large Vision-Language M

PTDL:Multi-Terrain Fall Recovery via Phase-Terrain Decoupled Learning

SafetyDGX agent

arXiv:2606.08922v1 Announce Type: new Abstract: Humanoid robots can fall on slopes, gravel, and uneven ground in unstructured environments. We target integrated fall recovery and locomotion: rebuildin

Q-VGM: Q-Guided Value-Gradient Matching for Flow-Matching VLA Policies

SafetyDGX agent

arXiv:2606.08015v1 Announce Type: new Abstract: We propose Q-Guided Value-Gradient Matching (Q-VGM), an off-policy reinforcement learning (RL) method that tackles a long-standing challenge in fine-tun

QnRL: Quantum-Native Reinforcement Learning

SafetyDGX agent

arXiv:2606.08276v1 Announce Type: cross Abstract: Quantum reinforcement learning (QRL) is a promising approach to learn effective decision strategies across several applications with stochastic enviro

Quantifying Uncertainty in Space Debris Capture with Active Tether-Net Systems Caused by Noisy Observations

SafetyDGX agent

arXiv:2606.07580v1 Announce Type: cross Abstract: As Low Earth Orbit has grown more crowded with space debris, the need for reliable and efficient debris removal solutions becomes more urgent. An acti

Quantitative Promise Theory: Intentionality and Inference in Autonomous Agents

SafetyDGX agent

arXiv:2606.08552v1 Announce Type: new Abstract: I discuss some quantitative representations of Promise Theory for processes involving autonomous agents. Agent models are common in software systems, ma

ReCoVLA: VLM-Guided Reward Compilation for Failure Recovery in Vision-Language-Action Policies

SafetyDGX agent

arXiv:2606.09630v1 Announce Type: cross Abstract: Vision-language-action (VLA) policies provide strong priors for language-conditioned manipulation, but remain brittle in off-nominal states requiring

ReGIL: Retrieval-Guided Imitation Learning from a Single Demonstration

SafetyDGX agent

arXiv:2606.09381v1 Announce Type: new Abstract: Learning robot manipulation policies with deep neural networks from a single demonstration remains highly challenging, as even small deviations from the

Region-Wise Correspondence Prediction between Manga Line Art Images

SafetyDGX agent

arXiv:2509.09501v4 Announce Type: replace Abstract: Understanding region-wise correspondences between manga line art images is fundamental for high-level manga processing, supporting downstream tasks

← Previous
1…116117118119120…242
Next →