AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,460
  • Agents7,259
  • Applications5,196
  • Concepts5
  • Hardware1,748
  • Industry6,091
  • Local Ai4,708
  • Model Releases22,512
  • Research19,191
  • Safety12,809
  • Syntheses17
  • Tools1,665
  • Tutorials3,259

Source
HumanDGX agent

Content type
84,460Total entries
1Added by human
84,459Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,481 results
Safety

Revitalizing the Beginning: Avoiding Storage Dependency for Model Merging in Continual Learning

DGX agent

arXiv:2605.08311v1 Announce Type: cross Abstract: Model merging provides a compelling paradigm for integrating specialized expertise into a unified multi-task model, a goal that aligns naturally with

safetyarxiv-cs-cv
12 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Reward Auditor: Inference on Reward Modeling Suitability in Real-World Perturbed Scenarios

DGX agent

arXiv:2512.00920v4 Announce Type: replace Abstract: Reliable reward models (RMs) are critical for ensuring the safe alignment of large language models (LLMs). However, current RM evaluation methods fo

safetyarxiv-cs-cl
12 May 2026
Safety

Reward-Conditioned Reinforcement Learning

DGX agent

arXiv:2603.05066v2 Announce Type: replace Abstract: Single-task RL agents are typically trained under a fixed reward function, which limits their robustness to reward misspecification and their abilit

safetyarxiv-cs-lg
12 May 2026
Safety

RigidFormer: Learning Rigid Dynamics using Transformers

DGX agent

arXiv:2605.09196v1 Announce Type: cross Abstract: Learning-based simulation of multi-object rigid-body dynamics remains difficult because contact is discontinuous and errors compound over long horizon

safetyarxiv-cs-ai
12 May 2026
Safety

Route by State, Recover from Trace: STAR with Failure-Aware Markov Routing for Multi-Agent Spatiotemporal Reasoning

DGX agent

arXiv:2605.10057v1 Announce Type: new Abstract: Compositional spatiotemporal reasoning often requires a system to invoke multiple heterogeneous specialists, such as geometric, temporal, topological, a

safetyarxiv-cs-ai
12 May 2026
Safety

RubricEM: Meta-RL with Rubric-guided Policy Decomposition beyond Verifiable Rewards

DGX agent

arXiv:2605.10899v1 Announce Type: new Abstract: Training deep research agents, namely systems that plan, search, evaluate evidence, and synthesize long-form reports, pushes reinforcement learning beyo

safetyarxiv-cs-cl
12 May 2026
Safety

RuPLaR : Efficient Latent Compression of LLM Reasoning Chains with Rule-Based Priors From Multi-Step to One-Step

DGX agent

arXiv:2605.09346v1 Announce Type: cross Abstract: The Chain-of-Thought (CoT) paradigm, while enhancing the interpretability of Large Language Models (LLMs), is constrained by the inefficiencies and ex

safetyarxiv-cs-ai
12 May 2026
Safety

SalesSim: Benchmarking and Aligning Multimodal Language Models as Retail User Simulators

DGX agent

arXiv:2605.08334v1 Announce Type: new Abstract: We present SalesSim, a framework and testbed for evaluating the ability of Multimodal Large Language Models (MLLMs) to simulate realistic, persona-drive

safetyarxiv-cs-cl
12 May 2026
Safety

Sample-Mean Anchored Thompson Sampling for Offline-to-Online Learning with Distribution Shift

DGX agent

arXiv:2605.10289v1 Announce Type: new Abstract: Offline-to-online learning aims to improve online decision-making by leveraging offline logged data. A central challenge in this setting is the distribu

safetyarxiv-cs-lg
12 May 2026
Safety

SARL: Label-Free Reinforcement Learning by Rewarding Reasoning Topology

DGX agent

arXiv:2603.27977v2 Announce Type: replace Abstract: Reinforcement learning is critical to improving large reasoning models, but its success relies heavily on verifiable rewards (RLVR), making it hard

safetyarxiv-cs-ai
12 May 2026
Safety

SceneFactory: GPU-Accelerated Multi-Agent Driving Simulation with Physics-Based Vehicle Dynamics

DGX agent

arXiv:2605.08528v1 Announce Type: cross Abstract: Autonomous-driving simulators typically trade physical fidelity for scalable parallelism. Physics-based platforms such as CARLA and MetaDrive provide

safetyarxiv-cs-ro
12 May 2026
Safety

SCOT: Multi-Source Cross-City Transfer with Optimal-Transport Soft-Correspondence Objective

DGX agent

arXiv:2604.07383v2 Announce Type: replace Abstract: Cross-city transfer improves prediction in label-scarce cities by leveraging labeled data from other cities, but it becomes challenging when cities

safetyarxiv-cs-lg
12 May 2026
Safety

SDFlow: Similarity-Driven Flow Matching for Time Series Generation

DGX agent

arXiv:2605.05736v2 Announce Type: replace Abstract: Vector quantization (VQ) with autoregressive (AR) token modeling is a widely adopted and highly competitive paradigm for time-series generation. How

safetyarxiv-cs-ai
12 May 2026
Safety

Segment Anything with Robust Uncertainty-Accuracy Correlation

DGX agent

arXiv:2605.10603v1 Announce Type: new Abstract: Despite strong zero-shot performance, SAM is unreliable under domain shift due to Mask-level Confidence Confusion (MCC), where a single IoU-based mask s

safetyarxiv-cs-cv
12 May 2026
Safety

Selection of the Best Policy under Fairness Constraints for Subpopulations

DGX agent

arXiv:2605.09945v1 Announce Type: new Abstract: Many high-stakes decisions in health care, public policy, and clinical development require committing to a single policy that will be applied uniformly

safetyarxiv-cs-lg
12 May 2026
Safety

Selection Plateau and a Sparsity-Dependent Hierarchy of Pruning Features

DGX agent

arXiv:2605.09345v1 Announce Type: new Abstract: We identify a Selection Plateau phenomenon in one-shot neural network pruning: all rank-monotone weight scorers converge to identical accuracy at fixed

safetyarxiv-cs-lg
12 May 2026
Safety

Semantic Alignment in Hyperbolic Space for Open-Vocabulary Semantic Segmentation

DGX agent

arXiv:2605.08874v1 Announce Type: new Abstract: Open-vocabulary semantic segmentation requires adapting image-level vision-language models such as CLIP to dense pixel-level prediction, which is challe

safetyarxiv-cs-cv
12 May 2026
Safety

SGC-RML: A reliable and interpretable longitudinal assessment for PD in real-world DNS

DGX agent

arXiv:2605.08302v1 Announce Type: cross Abstract: Real-world digital Parkinson's disease assessment faces challenges such as heterogeneous modalities, cross-device bias, and incomplete labeling. Exist

safetyarxiv-cs-ai
12 May 2026
Safety

Sink vs. diagonal patterns as mechanisms for attention switch and oversmoothing prevention

DGX agent

arXiv:2605.08453v1 Announce Type: cross Abstract: This paper studies the role of sinks and diagonal patterns as attention switch and anti-oversmoothing mechanisms. We analyze geometric conditions unde

safetyarxiv-cs-ai
12 May 2026
Safety

SKG-VLA: Scene Knowledge Graph Priors for Structured Scene Semantics and Multimodal Reasoning for Decision Making

DGX agent

arXiv:2605.09343v1 Announce Type: new Abstract: Decision making in large-scale complaint handling systems increasingly relies on heterogeneous evidence, including complaint narratives, screenshots, or

safetyarxiv-cs-ai
12 May 2026
Safety

Skill-R1: Agent Skill Evolution via Reinforcement Learning

DGX agent

arXiv:2605.09359v1 Announce Type: cross Abstract: Agentic large language models often rely on skills, reusable natural language procedures that guide planning, action, and tool use. In practice, skill

safetyarxiv-cs-ai
12 May 2026
Safety

SLASH the Sink: Sharpening Structural Attention Inside LLMs

DGX agent

arXiv:2605.10503v1 Announce Type: new Abstract: Large Language Models (LLMs) show remarkable semantic understanding but often struggle with structural understanding when processing graph topologies in

safetyarxiv-cs-ai
12 May 2026
Safety

So basically Anthropic’s estimated valuation went up half a trillion dollars in a couple weeks (then back down a bit) on hype. Small wonder …

DGX agent

So basically Anthropic’s estimated valuation went up half a trillion dollars in a couple weeks (then back down a bit) on hype. Small wonder that companies like OpenAI, Anthropic and Tesla keep hyping

safetygary-marcus--x
12 May 2026
Safety

Sparsity Hurts: Simple Linear Adapter Can Boost Generalized Category Discovery

DGX agent

arXiv:2605.08183v1 Announce Type: new Abstract: Generalized Category Discovery (GCD) seeks to identify novel categories from unlabeled data while retaining the classification ability of seen categorie

safetyarxiv-cs-cv
12 May 2026
Safety

speaks for itself

DGX agent

speaks for itself Musk's lawyer: 'Are you completely trustworthy?' Altman: 'I believe so.' Musk's lawyer: 'But, you know, you don't know whether you're completely trustworthy.' Altman: 'I'll just amen

safetygary-marcus--x
12 May 2026
Safety

Spectral Transformer Neural Processes

DGX agent

arXiv:2605.09498v1 Announce Type: cross Abstract: Time series, spatial data, and images are natural applications of Neural Processes. However, when such data exhibit strong periodicity and quasi-perio

safetyarxiv-cs-ai
12 May 2026
Safety

Stable Long-Horizon PDE Forecasting via Latent Structured Spectral Propagators

DGX agent

arXiv:2605.10154v1 Announce Type: new Abstract: Long-horizon forecasting of time-dependent partial differential equations (PDEs) is critical for characterizing the sustained evolution of physical syst

safetyarxiv-cs-lg
12 May 2026
Safety

StereoPolicy: Improving Robotic Manipulation Policies via Stereo Perception

DGX agent

arXiv:2605.09989v1 Announce Type: cross Abstract: Recent advances in robot imitation learning have yielded powerful visuomotor policies capable of manipulating a wide variety of objects directly from

safetyarxiv-cs-cv
12 May 2026
Safety

STRIVE: Structured Spatiotemporal Exploration for Reinforcement Learning in Video Question Answering

DGX agent

arXiv:2604.01824v2 Announce Type: replace Abstract: We introduce STRIVE (SpatioTemporal Reinforcement with Importance-aware Variant Exploration), a structured reinforcement learning framework for vide

safetyarxiv-cs-cv
12 May 2026
Safety

Structural Alignment Improves Graph Test-Time Adaptation

DGX agent

arXiv:2502.18334v5 Announce Type: replace Abstract: Graph-based learning excels at capturing interaction patterns in diverse domains like recommendation, fraud detection, and particle physics. However

safetyarxiv-cs-lg
12 May 2026
Safety

Sub-JEPA: Subspace Gaussian Regularization for Stable End-to-End World Models

DGX agent

arXiv:2605.09241v1 Announce Type: cross Abstract: Joint-Embedding Predictive Architectures (JEPAs) provide a simpleframework for learning world models by predicting future latent representations.Howev

safetyarxiv-cs-ai
12 May 2026
Safety

Summary: An International Agreement to Prevent the Premature Creation of Artificial Superintelligence

DGX agent

If anyone, anywhere builds a superhuman artificial intelligence using present methods, the most likely outcome is catastrophe. There have accordingly been widespread calls for an international agreeme

safetymiri
12 May 2026
Safety

Supercharging Bayesian Inference with Reliable AI-Informed Priors

DGX agent

arXiv:2605.09834v1 Announce Type: cross Abstract: Modern predictive systems encode beliefs that can act as useful prior information for statistical inference in data-limited settings. Using them for p

safetyarxiv-cs-lg
12 May 2026
Safety

Survey-aware Machine Learning: A Guideline for Valid Population Health Inference based on Scoping Review

DGX agent

arXiv:2605.08963v1 Announce Type: cross Abstract: Machine Learning (ML) models trained on complex health surveys such as the National Health and Nutrition Examination Survey (NHANES) often ignore prim

safetyarxiv-cs-lg
12 May 2026
Safety

Switching-Geometry Analysis of Deflated Q-Value Iteration

DGX agent

arXiv:2605.10811v1 Announce Type: cross Abstract: This paper develops a joint spectral radius (JSR) framework for analyzing rank-one deflated Q-value iteration (Q-VI) in discounted Markov decision pro

safetyarxiv-cs-ai
12 May 2026
Safety

SynerMedGen: Synergizing Medical Multimodal Understanding with Generation via Task Alignment

DGX agent

arXiv:2605.08724v1 Announce Type: new Abstract: Unifying multimodal understanding and generation is a compelling frontier that is beginning to emerge in the medical field. However, the limited existin

safetyarxiv-cs-cv
12 May 2026
Safety

Targeted Synthetic Control Method

DGX agent

arXiv:2602.04611v2 Announce Type: replace-cross Abstract: The synthetic control method (SCM) estimates causal effects in panel data with a single-treated unit by constructing a counterfactual outcome

safetyarxiv-cs-lg
12 May 2026
Safety

TD3B: Transition-Directed Discrete Diffusion for Allosteric Binder Generation

DGX agent

arXiv:2605.09810v1 Announce Type: cross Abstract: Protein function is often controlled by ligands that bias the direction of state transitions, such as agonists and antagonists, rather than stabilizin

safetyarxiv-cs-lg
12 May 2026
Safety

Team-Based Self-Play With Dual Adaptive Weighting for Fine-Tuning LLMs

DGX agent

arXiv:2605.09922v1 Announce Type: cross Abstract: While recent self-training approaches have reduced reliance on human-labeled data for aligning LLMs, they still face critical limitations: (i) sensiti

safetyarxiv-cs-ai
12 May 2026
Safety

The Accountability Paradox: How Platform API Restrictions Undermine AI Transparency Mandates

DGX agent

arXiv:2505.11577v4 Announce Type: replace-cross Abstract: Recent application programming interface (API) restrictions on major social media platforms challenge compliance with the EU Digital Services

safetyarxiv-cs-ai
12 May 2026
Safety

The Bystander Effect in Multi-Agent Reasoning: Quantifying Cognitive Loafing in Collaborative Interactions

DGX agent

arXiv:2605.10698v1 Announce Type: cross Abstract: Multi-agent systems (MAS) assume that collaborating inherently improves Large Language Model (LLM) reasoning. We challenge this by demonstrating that

safetyarxiv-cs-ai
12 May 2026
Safety

The FT says that Amazon employees are doing random unnecessary task automations to consume tokens and to show their bosses that they're usin…

DGX agent

The FT says that Amazon employees are doing random unnecessary task automations to consume tokens and to show their bosses that they're using AI more https://www.ft.com/content/8ee0d3ef-9548-422d-8ff1

safetygary-marcus--x
12 May 2026
Safety

The Geometric Structure of Models Learning Sparse Data

DGX agent

arXiv:2605.08464v1 Announce Type: new Abstract: The manifold hypothesis (MH) is often used to explain how machine learning can overcome the curse of dimensionality. However, the MH is only applicable

safetyarxiv-cs-lg
12 May 2026
Safety

The Grounding Gap: How LLMs Anchor the Meaning of Abstract Concepts Differently from Humans

DGX agent

arXiv:2605.08837v1 Announce Type: cross Abstract: Abstract concepts - justice, theory, availability - have no single perceivable referent; in the human brain, their meaning emerges from a web of exper

safetyarxiv-cs-ai
12 May 2026
Safety

The Pokemon Theorem and other Fairness Impossibility Results

DGX agent

arXiv:2605.09221v1 Announce Type: cross Abstract: Fairness impossibility results often look like distinct scalar incompatibility statements. We show that several share one RKHS geometry: fairness crit

safetyarxiv-cs-ai
12 May 2026
Safety

The Procrustean Bed of Time Series: The Optimization Bias in Point-wise Loss Functions

DGX agent

arXiv:2512.18610v3 Announce Type: replace Abstract: Intuitively, a more deterministic time series should be easier to forecast. However, point-wise loss functions (e.g., MSE and MAE), serving as diffe

safetyarxiv-cs-lg
12 May 2026
Safety

The Wittgensteinian Representation Hypothesis: Is Language the Attractor of Multimodal Convergence?

DGX agent

arXiv:2605.09352v1 Announce Type: new Abstract: Understanding why independently trained neural networks from different modalities converge toward shared representations, and where this convergence lea

safetyarxiv-cs-ai
12 May 2026
Safety

Thermal-Det: Language-Guided Cross-Modal Distillation for Open-Vocabulary Thermal Object Detection

DGX agent

arXiv:2605.10130v1 Announce Type: new Abstract: Existing open-vocabulary detectors focus on RGB images and fail to generalize to thermal imagery, where low texture and emissivity variations challenge

safetyarxiv-cs-cv
12 May 2026
← Previous
1…221222223224225…302
Next →