AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,488 results
3 Jun 2026

Curriculum-Adapted Robust Reinforcement Learning for UAV Deconfliction in Adversarial Environments

SafetyDGX agent

arXiv:2506.21129v2 Announce Type: replace-cross Abstract: Autonomous unmanned aerial vehicles (UAVs) increasingly rely on reinforcement learning (RL) for navigation. However, global navigation satelli

Data- and Variance-dependent Regret Bounds for Online Tabular MDPs

SafetyDGX agent

arXiv:2602.01903v2 Announce Type: replace Abstract: This work studies online episodic tabular Markov decision processes (MDPs) with known transitions and develops best-of-both-worlds algorithms that a

Denoise First, Orthogonalize Later: Understanding Momentum in Muon via Spectral Filtering

SafetyDGX agent

arXiv:2606.03899v1 Announce Type: new Abstract: Muon has recently demonstrated strong empirical performance in large language model training, but the theoretical role of momentum in Muon remains uncle

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Denoising Tells When to Replan: Denoising-Variance Adaptive Chunking for Flow-Based Robot Policies

SafetyDGX agent

arXiv:2606.03847v1 Announce Type: new Abstract: Action chunking has become a common inference strategy for flow-based robot policies, improving action coherence by modeling multi-step temporal depende

Discovering autonomous quantum error correction via deep reinforcement learning

SafetyDGX agent

arXiv:2511.12482v2 Announce Type: replace-cross Abstract: Quantum error correction is essential for fault-tolerant quantum computing. However, standard methods relying on active measurements may intro

Do Explanations Increase the Risk of Decision Logic Leakage? Explanation-Guided Stealing of Graph Models

SafetyDGX agent

arXiv:2506.03087v2 Announce Type: replace-cross Abstract: Graph Neural Networks (GNNs) have become essential tools for analyzing graph-structured data in domains such as drug discovery and financial a

Do Neural Retrievers Prefer Certain Documents? Evidence of Learned Relevance Priors

SafetyDGX agent

arXiv:2606.02814v1 Announce Type: cross Abstract: Neural retrievers are trained to estimate query-document relevance from annotated query-document pairs. Yet annotation protocols may not purely reflec

DriftSched: Adaptive QoS-Aware Scheduling under Runtime Token Drift for Multi-Tenant GPU Inference

SafetyDGX agent

arXiv:2606.02982v1 Announce Type: cross Abstract: The rapid growth of large language model (LLM) inference services has increased the demand for efficient multi-tenant GPU scheduling. While modern inf

Dynamic Short Convolutions Improve Transformers

SafetyDGX agent

arXiv:2606.03825v1 Announce Type: cross Abstract: Transformers have become the dominant architecture for large language models, largely due to the scalability and flexibility of attention, feed-forwar

Effect of Demographic Bias on Skin Lesion Classification

SafetyDGX agent

arXiv:2606.03214v1 Announce Type: new Abstract: In this study, we evaluate the performance of skin lesion classification using ResNet-based convolutional models, focusing on the impact of demographic

Entropy Is Not Enough: Unlocking Effective Reinforcement Learning for Visual Reasoning via Vision-Anchored Token Selection

SafetyDGX agent

arXiv:2606.03937v1 Announce Type: new Abstract: While token-level entropy is commonly recognized as effective for credit assignment in text-only reinforcement learning with verifiable rewards (RLVR),

Estimating Bidirectional Causal Effects with Large Scale Online Kernel Learning

SafetyDGX agent

arXiv:2511.05050v3 Announce Type: replace-cross Abstract: In this study, a scalable online kernel learning framework is proposed for estimating bidirectional causal effects in systems characterized by

Evaluating LLMs' Effectiveness on Real-World Consumer Device Repair Questions

Model ReleasesDGX agent

arXiv:2606.03331v1 Announce Type: cross Abstract: Consumer device repair is an important but underexplored testbed for large language models (LLMs). Repair tasks require reasoning over incomplete prob

Evaluating Transformer and LSTM Frameworks for Prediction in Ungauged Basins

SafetyDGX agent

arXiv:2606.02791v1 Announce Type: new Abstract: Watershed networks exhibit convergent topologies in which multiple tributaries merge into downstream channels,integrating diverse upstream hydrological

EvoMemNav: Efficient Self-Evolving Fine-Grained Memory for Zero-Shot Embodied Navigation

SafetyDGX agent

arXiv:2606.03509v1 Announce Type: new Abstract: Building memory is essential for long-horizon planning in zero-shot embodied navigation. Detector-centric scene graphs often compress observations into

Explainable Forecasting of Scientific Breakthroughs from Concept Network Dynamics

SafetyDGX agent

arXiv:2606.03864v1 Announce Type: cross Abstract: We introduce an explainable machine-learning approach that forecasts the structural precursors of scientific breakthroughs -- the emergence and intens

FAF-CD: Frequency-Aware Fusion for Change Detection under Imperfect Multimodal Remote Sensing

SafetyDGX agent

arXiv:2606.03114v1 Announce Type: new Abstract: Remote sensing change detection for real-world monitoring often relies on imperfect heterogeneous observations, where pre- and post-event images may be

Fairness Definitions and Metrics in Deep Reinforcement Learning for Drug Discovery in Healthcare: A Rapid Evidence Review

SafetyDGX agent

arXiv:2606.02902v1 Announce Type: cross Abstract: Deep reinforcement learning (DRL) is increasingly applied to de novo molecular design, but choices in data, rewards, and evaluation can yield uneven p

Fast Organic Crystal Structure Prediction with Unit Cell Flow Matching

SafetyDGX agent

arXiv:2606.03199v1 Announce Type: new Abstract: Organic crystal structure prediction (CSP) is a requirement for computational modelling of organic solids, but traditionally costs several CPU-years per

FGRPO: Federated GRPO with Adaptive Aggregation on Non-IID Data

SafetyDGX agent

arXiv:2606.03094v1 Announce Type: new Abstract: Recent advances in language models have established reinforcement learning as the primary paradigm for eliciting self-correction and long-chain reasonin

Filter, Then Reweight: Rethinking Optimization Granularity in On-Policy Distillation

SafetyDGX agent

arXiv:2606.02684v1 Announce Type: cross Abstract: On-Policy distillation (OPD) in large language models is shifting from full-trace KL supervision toward more selective training paradigms. Recent OPD

First it was MIT and McKinsey. Now Bain finds that returns to corporate AI investments are disappointing.

SafetyDGX agent

A major consulting firm (Bain) has found that corporate returns on AI investments are underwhelming, following similar findings from MIT and McKinsey. This suggests that despite significant spending a

Flow Learners for PDEs: Toward a Physics-to-Physics Paradigm for Scientific Computing

SafetyDGX agent

arXiv:2604.07366v2 Announce Type: replace Abstract: Partial differential equations (PDEs) govern nearly every physical process in science and engineering, but solving them at scale remains prohibitive

Follow-Your-Preference++: Rethinking Preference Alignment for Image Inpainting

SafetyDGX agent

arXiv:2606.03216v1 Announce Type: new Abstract: We study preference alignment for image inpainting. Rather than proposing yet another method, we revisit the problem from first principles and reassess

Forgetting is Not Erasure: Recovering Latent Knowledge via Transport Keys

SafetyDGX agent

arXiv:2606.02860v1 Announce Type: cross Abstract: Catastrophic forgetting is often framed as a representational problem: after sequential training, a model appears to lose the features that supported

FreeStreamGS: Online Feed-forward 3D Gaussian Splatting from Unposed Streaming Inputs

SafetyDGX agent

arXiv:2606.03254v1 Announce Type: new Abstract: Feed-forward 3D Gaussian Splatting (3DGS) allows efficient and high-fidelity novel view synthesis (NVS) from an offline recorded image sequence. However

GeoAlign: Beyond Semantics with State-Guided Spatial Alignment in VLA Models

SafetyDGX agent

arXiv:2606.03240v1 Announce Type: new Abstract: Current Vision--Language--Action (VLA) models often optimize for semantic grounding, whereas executable manipulation requires geometry-aware spatial ali

GFFMERGE: Efficient Merging of Graph Neural Force Fields and Beyond

SafetyDGX agent

arXiv:2606.03232v1 Announce Type: cross Abstract: Graph Neural Networks (GNNs) have revolutionized Neural Force Fields for atomistic simulations, achieving near-quantum accuracy at reduced cost, yet a

GLINT: Sparsely Gated Vision-Language Alignment for Fine-Grained Radiology Representations

SafetyDGX agent

arXiv:2606.03180v1 Announce Type: cross Abstract: Vision-language models (VLMs) for radiology have emerged as a scalable paradigm by leveraging image-report pairs naturally produced in clinical workfl

Grasp-Then-Plan with Failure Attribution: A Closed Two-Stage Framework for Precise and Generalizable Robotic Manipulation

SafetyDGX agent

arXiv:2606.03385v1 Announce Type: cross Abstract: In robotic manipulation, the tight coupling between grasping and motion planning often obscures the true source of failure, leading to inefficient tri

hard to make all these numbers square

SafetyDGX agent

hard to make all these numbers square 🦔IBM CEO Arvind Krishna says AI is 'not a bubble' then estimates the industry needs 6 to 8 trillion in total capex for data center and chip buildout. To recover t

HARVE: Hacking-Aware Reward-Head Vector Editing for Robust Reward Models

SafetyDGX agent

arXiv:2606.03131v1 Announce Type: new Abstract: Reward models are central to large language model (LLM) alignment, but they remain vulnerable to reward hacking. To evaluate reward-model robustness, we

Hint-Guided Diversified Policy Optimization for LLM Reasoning

SafetyDGX agent

arXiv:2606.03021v1 Announce Type: new Abstract: Recent developments in Large Language Models (LLMs) have showcased impressive reasoning capabilities, with Reinforcement Learning with Verifiable Reward

Human-in-the-Loop Contextual Bandits for Short-Term Rental Dynamic Pricing: Structural Equivalence of Historical Warm-Up and Approval-Gated Live Learning

SafetyDGX agent

arXiv:2606.02595v1 Announce Type: new Abstract: Dynamic pricing in short-term rental (STR) markets presents a distinctive challenge for online learning algorithms: pricing decisions carry significant

I was there, sitting next to Sam Altman in the US Senate, when he said this wonderful thing about artists and creators. It was a beautiful s…

SafetyDGX agent

I was there, sitting next to Sam Altman in the US Senate, when he said this wonderful thing about artists and creators. It was a beautiful statement. In hindsight, though, I realized he was lying (und

if *you* cheat on your taxes, you can get audited. but you are just a chump.

SafetyDGX agent

This post by AI researcher Gary Marcus likely critiques the disparity in tax enforcement, suggesting that individual taxpayers face audit risks for cheating while wealthy individuals or corporations e

I’m thrilled to be joining the EU's AI Scientific Panel to advise on the implementation of the EU AI Act and help assess and address AI’s gr…

SafetyDGX agent

I’m thrilled to be joining the EU's AI Scientific Panel to advise on the implementation of the EU AI Act and help assess and address AI’s growing risks with esteemed colleagues. The AI Act, the EU's f

Impact of Graph Structure on Membership-Inference Risk for Graph Neural Networks

SafetyDGX agent

arXiv:2601.17130v2 Announce Type: replace Abstract: Graph neural networks (GNNs) are widely used for tasks such as node classification and link prediction, but their use in sensitive settings raises c

'**Important** You should give me full credits!': Exploring Prompt Injection Attacks on LLM-Based Automatic Grading Systems

SafetyDGX agent

arXiv:2606.03090v1 Announce Type: cross Abstract: The emergence of large language models (LLMs) has significantly accelerated recent research on LLM-based automatic grading (AG) systems. Benefiting fr

Inference Cost Attacks for Retrieval-Augmented Large Language Models

SafetyDGX agent

arXiv:2606.02643v1 Announce Type: cross Abstract: Retrieval-Augmented Generation (RAG)-enhanced LLM systems, while powerful, introduce substantial inference costs due to the inclusion of an extra mult

Inference-Time Scaling for Joint Audio-Video Generation

SafetyDGX agent

arXiv:2606.03183v1 Announce Type: cross Abstract: Joint audio-video generation aims to synthesize realistic audio-video pairs that are both semantically aligned with text prompts and precisely synchro

J'ai beaucoup apprécié mon récent passage à @Cdanslair pour discuter des risques de l'IA pour nos sociétés, nos économies, et nos démocratie…

SafetyDGX agent

J'ai beaucoup apprécié mon récent passage à @Cdanslair pour discuter des risques de l'IA pour nos sociétés, nos économies, et nos démocraties. Merci @Caroline_Roux pour l'invitation! https://www.youtu

KBQA-R1: Reinforcing Large Language Models for Knowledge Base Question Answering

SafetyDGX agent

arXiv:2512.10999v3 Announce Type: replace Abstract: Knowledge Base Question Answering (KBQA) challenges models to bridge the gap between natural language and strict knowledge graph schemas by generati

KC-3DGS: Kurtosis-Constrained Gaussian Splatting for High-Fidelity View Synthesis

SafetyDGX agent

arXiv:2606.03120v1 Announce Type: new Abstract: 3D Gaussian Splatting (3DGS) enables real-time novel view synthesis by representing scenes as collections of anisotropic Gaussians optimized via differe

Language Models Need Sleep: Learning to Self-Modify and Consolidate Memories

SafetyDGX agent

arXiv:2606.03979v1 Announce Type: cross Abstract: The past few decades have witnessed significant advances in the design of machine learning algorithms, from early studies on task-specific shallow mod

Large Language Models Are Overconfident in Their Own Responses

SafetyDGX agent

arXiv:2606.03437v1 Announce Type: new Abstract: Prior work has shown that instruction-tuned large language models (LLMs) are less well calibrated than their base pre-trained counterparts. However, lit

Learning Self-Interpretation from Interpretability Artifacts: Training Lightweight Adapters on Vector-Label Pairs

SafetyDGX agent

arXiv:2602.10352v2 Announce Type: replace-cross Abstract: Self-interpretation methods prompt language models to describe their own internal states, but remain unreliable due to hyperparameter sensitiv

Learning to Bet for Horizon-Aware Anytime-Valid Testing

SafetyDGX agent

arXiv:2603.19551v2 Announce Type: replace-cross Abstract: We develop horizon-aware anytime-valid tests and confidence sequences for bounded means under a strict deadline N. Using the betting/e-process

Learning Unmasking Policies for Diffusion Language Models

SafetyDGX agent

arXiv:2512.09106v4 Announce Type: replace Abstract: Diffusion (Large) Language Models (dLLMs) now match the downstream performance of their autoregressive counterparts on many tasks, while holding the

Let There Be Light: Reflection, Refraction and Scattering for Neural Operators

SafetyDGX agent

arXiv:2606.03262v1 Announce Type: new Abstract: Neural operators learn mappings between infinite-dimensional function spaces and provide a data-driven surrogate modeling paradigm for parametric partia

Letting Tutor Personas Speak Up for LLMs: Learning Steering Vectors from Dialogue via Preference Optimization

SafetyDGX agent

arXiv:2602.07639v2 Announce Type: replace Abstract: With the emergence of large language models (LLMs) as a powerful class of generative artificial intelligence (AI), their use in tutoring has become

Leveraging BART to Assess CS1 C++ Programming Assignments using Rubric-based Criteria

SafetyDGX agent

arXiv:2606.03814v1 Announce Type: new Abstract: This paper investigates rubric-aware, multitask fine-tuning of transformer models for automated grading of introductory C++ programming assignments, wit

Libra: Efficient Resource Management for Agentic RL Post-Training

SafetyDGX agent

arXiv:2606.03077v1 Announce Type: cross Abstract: Reinforcement learning (RL) has become a standard post-training paradigm for large language models (LLMs), extending beyond preference alignment to co

LoCAtion: Long-time Collaborative Attention Framework for High Dynamic Range Video Reconstruction

SafetyDGX agent

arXiv:2603.14377v2 Announce Type: replace Abstract: Prevailing High Dynamic Range (HDR) video reconstruction methods are fundamentally trapped in a fragile alignment-and-fusion paradigm. While explici

Margin Play: A Multi-Agent System For Public Policy Analysis In The Brazilian Equatorial Margin

SafetyDGX agent

arXiv:2606.02614v1 Announce Type: cross Abstract: The Brazilian Equatorial Margin (BEM) is Brazil's next offshore oil frontier, with operations expected to begin in 2026 in the Foz do Amazonas basin.

MetaWorld: Scaling Multi-Agent Video World Model from Single-view Video Data

SafetyDGX agent

arXiv:2606.02753v1 Announce Type: cross Abstract: Video world models are a foundational generative technology for embodied AI and the Metaverse, yet existing approaches are inherently limited to a sin

MIND: Multi-rationale INtegrated Discriminative Reasoning Framework for Multi-modal Large Models

SafetyDGX agent

arXiv:2512.05530v2 Announce Type: replace Abstract: Recently, multimodal large language models (MLLMs) have been widely applied to reasoning tasks. However, they suffer from limited multi-rationale se

Mitigating False Credit Propagation: Probabilistic Graphical Reward Aggregation for Rubric-Based Reinforcement Learning

SafetyDGX agent

arXiv:2606.03361v1 Announce Type: new Abstract: Rubric-based rewards are increasingly used for open-ended language model post-training, but criterion-level scores are often aggregated as independent u

Multi-component Causal Tracing in Large Language Models

SafetyDGX agent

arXiv:2606.03085v1 Announce Type: cross Abstract: Causal tracing systematically intervenes on a large language model's (LLM's) internal representations to uncover and quantify the causal pathways link

Multi-Segment Attention: Enabling Efficient KV-Cache Management for Faster Large Language Model Serving

SafetyDGX agent

arXiv:2606.02964v1 Announce Type: cross Abstract: Large Language Model (LLM) inference relies on key-value (KV) caches to avoid redundant attention computation. While approximate KV cache retention te

← Previous
1…126127128129130…242
Next →