AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
Safety

Overcoming Environmental Meta-Stationarity in MARL via Adaptive Curriculum and Counterfactual Group Advantage

DGX agent

arXiv:2506.07548v2 Announce Type: replace-cross Abstract: Multi-agent reinforcement learning (MARL) has reached competitive performance on cooperative tasks against scripted adversaries, yet most meth

safetyarxiv-cs-ro
7 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

Overcoming reward signal challenges: Verifiable rewards-based reinforcement learning with GRPO on SageMaker AI

DGX agent

In this post, you will learn how to implement reinforcement learning with verifiable rewards (RLVR) to introduce verification and transparency into reward signals to improve training performance. This

safetyaws-ml-blog
7 May 2026
Safety

PhySe-RPO: Physics and Semantics Guided Relative Policy Optimization for Diffusion-Based Surgical Smoke Removal

DGX agent

arXiv:2603.22844v4 Announce Type: replace Abstract: Surgical smoke severely degrades intraoperative video quality, obscuring anatomical structures and limiting surgical perception. Existing learning-b

safetyarxiv-cs-ai
7 May 2026
Safety

POMA-3D: The Point Map Way to 3D Scene Understanding

DGX agent

arXiv:2511.16567v3 Announce Type: replace Abstract: In this paper, we introduce POMA-3D, the first self-supervised 3D representation model learned from point maps. Point maps encode explicit 3D coordi

safetyarxiv-cs-cv
7 May 2026
Safety

Preference-Based Self-Distillation: Beyond KL Matching via Reward Regularization

DGX agent

arXiv:2605.05040v1 Announce Type: new Abstract: On-policy distillation is an efficient alternative to reinforcement learning, offering dense token-level training signals. However, its reliance on a st

safetyarxiv-cs-lg
7 May 2026
Safety

ProFit: Leveraging High-Value Signals in SFT via Probability-Guided Token Selection

DGX agent

arXiv:2601.09195v3 Announce Type: replace Abstract: Supervised fine-tuning (SFT) is a fundamental post-training strategy to align Large Language Models (LLMs) with human intent. However, traditional S

safetyarxiv-cs-cl
7 May 2026
Safety

Provable imitation learning for control of instability in partially-observed Vlasov--Poisson equations

DGX agent

arXiv:2605.05081v1 Announce Type: new Abstract: We consider the stabilization of Vlasov--Poisson plasma dynamics, a central control problem in nuclear fusion. Our focus is the gap between what an idea

safetyarxiv-cs-lg
7 May 2026
Safety

Purdah and Patriarchy: Evaluating and Mitigating South Asian Biases in Open-Ended Multilingual LLM Generations

DGX agent

arXiv:2505.18466v2 Announce Type: replace Abstract: Evaluations of Large Language Models (LLMs) often overlook intersectional and culturally specific biases, particularly in underrepresented multiling

safetyarxiv-cs-cl
7 May 2026
Safety

Quantifying Trust: Financial Risk Management for Trustworthy AI Agents

DGX agent

arXiv:2604.03976v2 Announce Type: replace Abstract: Prior work on trustworthy AI emphasizes model-internal properties such as bias mitigation, adversarial robustness, and interpretability. As AI syste

safetyarxiv-cs-ai
7 May 2026
Safety

Reinforcement Learning for Compositional Generalization with Outcome-Level Optimization

DGX agent

arXiv:2605.04920v1 Announce Type: cross Abstract: Compositional generalization refers to correctly interpret novel combinations of known primitives, which remains a major challenge. Existing approache

safetyarxiv-cs-cl
7 May 2026
Safety

Rollout Pass-Rate Control: Steering Binary-Reward RL Toward Its Most Informative Regime

DGX agent

arXiv:2605.05112v1 Announce Type: new Abstract: SWE-bench-style agentic reinforcement learning relies on expensive stateful trajectories, yet substantial compute is wasted on sampled rollout groups wi

safetyarxiv-cs-lg
7 May 2026
Safety

S1-MMAlign: A Large-Scale, Multi-Disciplinary Dataset for Scientific Figure-Text Understanding

DGX agent

arXiv:2601.00264v2 Announce Type: replace Abstract: Multimodal learning has revolutionized general domain tasks, yet its application in scientific discovery is hindered by the profound semantic gap be

safetyarxiv-cs-cv
7 May 2026
Safety

Scalable inference of spatial regions and temporal signatures from time series

DGX agent

arXiv:2605.05008v1 Announce Type: cross Abstract: Regionalization aims to partition a spatial domain into contiguous regions that share similar characteristics, enabling more effective spatial analysi

safetyarxiv-cs-lg
7 May 2026
Safety

Scalable Multi Agent Diffusion Policies for Coverage Control

DGX agent

arXiv:2509.17244v2 Announce Type: replace Abstract: We propose MADP, a novel diffusion-model-based approach for collaboration in decentralized robot swarms. MADP leverages diffusion models to generate

safetyarxiv-cs-ro
7 May 2026
Safety

Scalable Policy Maximization Under Network Interference

DGX agent

arXiv:2505.18118v2 Announce Type: replace-cross Abstract: Many interventions, such as vaccines in clinical trials or coupons in online marketplaces, must be assigned sequentially without full knowledg

safetyarxiv-cs-lg
7 May 2026
Safety

Sequential Strategic Classification with Multi-Stage Selective Classifiers

DGX agent

arXiv:2605.04202v1 Announce Type: new Abstract: Strategic classification studies the problem where self-interested individuals or agents manipulate their response to obtain favorable decision outcomes

safetyarxiv-cs-lg
7 May 2026
Safety

Structural Equivalence and Learning Dynamics in Delayed MARL

DGX agent

arXiv:2605.04345v1 Announce Type: new Abstract: We formally establish the equivalence between Observation Delay (OD) and Action Delay (AD) in cooperative partially observable multi-agent systems using

safetyarxiv-cs-lg
7 May 2026
Safety

Temporal Structure Matters for Efficient Test-Time Adaptation in Wearable Human Activity Recognition

DGX agent

arXiv:2605.04617v1 Announce Type: new Abstract: Wearable human activity recognition (WHAR) models often suffer from performance degradation under real-world cross-user distribution shifts. Test-time a

safetyarxiv-cs-cv
7 May 2026
Safety

The balcony solar boom is coming to the US

DGX agent

Dozens of US states are considering legislation to allow people to install plug-in solar systems, often called balcony solar. These small arrays require little to no setup and could help cut emissions

safetymit-tech-review
7 May 2026
Safety

The illustration on this story is quite funny, but the study itself has quite big implications I think. The point is that there might be bet…

DGX agent

The illustration on this story is quite funny, but the study itself has quite big implications I think. The point is that there might be better ways to design AI systems so that they don’t simply do e

safetygary-marcus--x
7 May 2026
Safety

Theories are proven by accurate predictions. I first read @GaryMarcus in the late ’90s, and his ideas helped shape my deterministic ML frame…

DGX agent

Theories are proven by accurate predictions. I first read @GaryMarcus in the late ’90s, and his ideas helped shape my deterministic ML frameworks, GAIuS & KATO. He’s been proven right repeatedly, yet

safetygary-marcus--x
7 May 2026
Safety

“they hadn’t figured out how OpenAI would pay for it” may turn out to be the epitaph for an entire era. 🪦 scoop from @anissagardizy8 @thein…

DGX agent

“they hadn’t figured out how OpenAI would pay for it” may turn out to be the epitaph for an entire era. 🪦 scoop from @anissagardizy8 @theinformation, and credit her with the great line “OpenAI has mad

safetygary-marcus--x
7 May 2026
Safety

Threshold-Guided Optimization for Visual Generative Models

DGX agent

arXiv:2605.04653v1 Announce Type: new Abstract: Aligning large visual generative models with human feedback is often performed through pairwise preference optimization. While such approaches are conce

safetyarxiv-cs-lg
7 May 2026
Safety

Time series causal discovery with variable lags

DGX agent

arXiv:2605.04081v1 Announce Type: new Abstract: Causal Bayesian Networks (CBNs) are a powerful tool for reasoning under uncertainty about complex real-world problems. Such problems evolve over time, r

safetyarxiv-cs-lg
7 May 2026
Safety

To Fuse or to Drop? Dual-Path Learning for Resolving Modality Conflicts in Multimodal Emotion Recognition

DGX agent

arXiv:2605.04877v1 Announce Type: cross Abstract: Multimodal emotion recognition (MER) benefits from combining text, audio, and vision, yet standard fusion often fails when modalities conflict. Crucia

safetyarxiv-cs-lg
7 May 2026
Safety

Towards General Preference Alignment: Diffusion Models at Nash Equilibrium

DGX agent

arXiv:2605.04494v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has been popular for aligning text-to-image (T2I) diffusion models with human preferences. As a main

safetyarxiv-cs-cv
7 May 2026
Safety

Training-Time Batch Normalization Reshapes Local Partition Geometry in Piecewise-Affine Networks

DGX agent

arXiv:2605.04946v1 Announce Type: new Abstract: Batch normalization (BN) is central to modern deep networks, but its effect on the realized function during training remains less understood than its op

safetyarxiv-cs-lg
7 May 2026
Safety

Two years ago. I stand by both predictions.

DGX agent

Gary Marcus reflects on predictions he made two years prior and reaffirms his confidence in their accuracy. The post suggests Marcus is reviewing his track record on forecasts, likely related to artif

safetygary-marcus--x
7 May 2026
Safety

UAV-VL-R1: Generalizing Vision-Language Models via Supervised Fine-Tuning and Multi-Stage GRPO for UAV Visual Reasoning

DGX agent

arXiv:2508.11196v2 Announce Type: replace Abstract: Recent advances in vision-language models (VLMs) have demonstrated strong generalization in natural image tasks. However, their performance often de

safetyarxiv-cs-cv
7 May 2026
Safety

UI2Code^N: UI-to-Code Generation as Interactive Visual Optimization

DGX agent

arXiv:2511.08195v3 Announce Type: replace Abstract: UI-to-code aims to translate UI screenshots into executable front-end code. Despite progress with vision-language models (VLMs), most existing metho

safetyarxiv-cs-cv
7 May 2026
Safety

ULF-Loc: Unbiased Landmark Feature for Robust Visual Localization with 3D Gaussian Splatting

DGX agent

arXiv:2605.04730v1 Announce Type: new Abstract: Visual localization is a core technology for augmented reality and autonomous navigation. Recent methods combine the efficient rendering of 3D Gaussian

safetyarxiv-cs-cv
7 May 2026
Safety

Uncertainty-Aware Exploratory Direct Preference Optimization for Multimodal Large Language Models

DGX agent

arXiv:2605.04874v1 Announce Type: cross Abstract: Direct Preference Optimization (DPO) has proven to be an effective solution for mitigating hallucination in Multimodal Large Language Models (MLLMs) b

safetyarxiv-cs-cl
7 May 2026
Safety

Unifying Dynamical Systems and Graph Theory to Mechanistically Understand Computation in Neural Networks

DGX agent

arXiv:2605.03598v2 Announce Type: cross Abstract: Understanding how biological and artificial neural networks implement computation from connectivity is a central problem in neuroscience and machine l

safetyarxiv-cs-ai
7 May 2026
Safety

UniMoCo: Unified Modality Completion for Robust Multi-Modal Embeddings

DGX agent

arXiv:2505.11815v2 Announce Type: replace Abstract: Current vision-language models have been explored for multi-modal embedding tasks like information retrieval. However, they face significant challen

safetyarxiv-cs-cv
7 May 2026
Safety

Using Common Random Numbers for Simulation-based Planning with Rollouts

DGX agent

arXiv:2605.04732v1 Announce Type: new Abstract: Simulation-based planning with rollouts is a widely-deployed technique for decision making in stochastic environments. The primary instrument of simulat

safetyarxiv-cs-lg
7 May 2026
Safety

Variance Matters: Improving Domain Adaptation via Stratified Sampling

DGX agent

arXiv:2512.05226v2 Announce Type: replace Abstract: Domain shift remains a key challenge in deploying machine learning models to the real world. Unsupervised domain adaptation (UDA) aims to address th

safetyarxiv-cs-lg
7 May 2026
Safety

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying t…

DGX agent

What happens when AIs become smarter than us? Why would they keep humans around if given the choice? Our new paper argues that only trying to control AIs is a limited strategy, and that a stable, mutu

safetydan-hendrycks--x
7 May 2026
Safety

When Life Gives You BC, Make Q-functions: Extracting Q-values from Behavior Cloning for On-Robot Reinforcement Learning

DGX agent

arXiv:2605.05172v1 Announce Type: new Abstract: Behavior Cloning (BC) has emerged as a highly effective paradigm for robot learning. However, BC lacks a self-guided mechanism for online improvement af

safetyarxiv-cs-ro
7 May 2026
Safety

Why Expert Alignment Is Hard: Evidence from Subjective Evaluation

DGX agent

arXiv:2605.04972v1 Announce Type: new Abstract: Aligning large language models with expert judgment is especially difficult in subjective evaluation tasks, where experts may disagree, rely on tacit cr

safetyarxiv-cs-cl
7 May 2026
Safety

Wow this paper has been “just published” many times for nearly a year. I have called out at least two other “influencers” for the same thing…

DGX agent

Wow this paper has been “just published” many times for nearly a year. I have called out at least two other “influencers” for the same thing on the same paper, including another one earlier this week.

safetygary-marcus--x
7 May 2026
Safety

yep, really embarassing

DGX agent

yep, really embarassing One of the richest VCs, Marc Andreessen, has shown again that he thinks generative AI works perfectly well if you use the right prompts, thus ignoring the statistical nature of

safetygary-marcus--x
7 May 2026
Safety

100% stood the test of time: “What Ilya saw” was Sam’s bad behavior, not AGI.

DGX agent

100% stood the test of time: “What Ilya saw” was Sam’s bad behavior, not AGI. When Altman got fired “What did Ilya see?” became a wildly popular conspiracy meme. I think we can safely say now that wha

safetygary-marcus--x
6 May 2026
Safety

2nd episode of The Roman Forum is an interview with AI Safety/Governance expert Connor Leahy @NPCollapse. Connor is a great speaker and is l…

DGX agent

2nd episode of The Roman Forum is an interview with AI Safety/Governance expert Connor Leahy @NPCollapse. Connor is a great speaker and is lobbying to get government to ban Superintelligence. My first

safetyconnor-leahy--x
6 May 2026
Safety

A Robust Unsupervised Domain Adaptation Framework for Medical Image Classification Using RKHS-MMD

DGX agent

arXiv:2605.03787v1 Announce Type: new Abstract: Labeling medical images is a major bottleneck in the field of medical imaging, as it requires domain-specific expertise, and it gets further complicated

safetyarxiv-cs-cv
6 May 2026
Safety

A Universal Reproducing Kernel Hilbert Space from Polynomial Alignment and IMQ Distance

DGX agent

arXiv:2605.03262v1 Announce Type: new Abstract: We introduce the Yat kernel $k_{b,arepsilon}(mathbf{w},mathbf{x})=frac{(mathbf{w}^opmathbf{x}+b)^2}{|mathbf{x}-mathbf{w}|^2+arepsilon},qquad bge 0, arep

safetyarxiv-cs-lg
6 May 2026
Safety

A US appeals court strikes down a 2023 FCC rule banning broadband access discrimination based on income, race, and more; Chair Brendan Carr welcomes the ruling (Jon Brodkin/Ars Technica)

DGX agent

Jon Brodkin / Ars Technica: A US appeals court strikes down a 2023 FCC rule banning broadband access discrimination based on income, race, and more; Chair Brendan Carr welcomes the ruling — An appeals

safetytechmeme
6 May 2026
Safety

Adaptive 3D-RoPE: Physics-Aligned Rotary Positional Encoding for Wireless Foundation Models

DGX agent

arXiv:2605.00968v1 Announce Type: cross Abstract: Positional encoding plays a pivotal role in determin?ing the extrapolation and generalization performance of wireless foundation models for channel st

safetyarxiv-cs-ai
6 May 2026
Safety

ADAPTS: Agentic Decomposition for Automated Protocol-agnostic Tracking of Symptoms

DGX agent

arXiv:2605.03212v1 Announce Type: cross Abstract: Modeling latent clinical constructs from unconstrained clinical interactions is a unique challenge in affective computing. We present ADAPTS (Agentic

safetyarxiv-cs-cl
6 May 2026
← Previous
1…230231232233234…302
Next →