AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlog
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

Search: “safety”

GridTimelineEvolution
14,478 results
Safety

MPD^2-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis

DGX agent

arXiv:2605.08024v1 Announce Type: new Abstract: Learning-to-defer (L2D) can make glaucoma screening safer by routing difficult/uncertain cases to humans, yet standard formulations overlook expert avai

safetyarxiv-cs-ai
11 May 2026
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
Safety

Multi-environment Invariance Learning with Missing Data

DGX agent

arXiv:2601.07247v2 Announce Type: replace-cross Abstract: Learning models that can handle distribution shifts is a key challenge in domain generalization. Invariance learning, an approach that focuses

safetyarxiv-cs-lg
11 May 2026
Safety

Multi-Environment POMDPs with Finite-Horizon Objectives

DGX agent

arXiv:2605.07537v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are systems in which one agent interacts with a stochastic environment, and receives only partia

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation

DGX agent

arXiv:2603.16876v2 Announce Type: replace-cross Abstract: We propose MARL-Rad, a multi-modal multi-agent reinforcement learning framework for radiology report generation that trains the entire agentic

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Objective Multi-Agent Bandits: From Learning Efficiency to Fairness Optimization

DGX agent

arXiv:2605.06864v1 Announce Type: new Abstract: We study multi-objective multi-agent multi-armed bandits (MO-MA-MAB) under stochastic rewards, where agents observe heterogeneous reward vectors and com

safetyarxiv-cs-lg
11 May 2026
Safety

Mythos found a single vulnerability in cURL (along with three false positives, and one issue they classified as a bug). The founder/lead dev…

DGX agent

Mythos identified one genuine vulnerability in cURL while also reporting three false positives and one issue classified as a bug during their security analysis. The post references cURL's founder or l

safetygary-marcus--x
11 May 2026
Safety

@NameInteger @GaryMarcus @geoffreyhinton Hinton's argument was about encoding: LLMs don't encode text as text, but as weighted matrices. It …

DGX agent

@NameInteger @GaryMarcus @geoffreyhinton Hinton's argument was about encoding: LLMs don't encode text as text, but as weighted matrices. It makes no difference. Memorised data is memorised data no mat

safetygary-marcus--x
11 May 2026
Safety

No Forgetting Learning: Buffer-free Continual Learning Classification

DGX agent

arXiv:2503.04638v3 Announce Type: replace Abstract: Most Continual Learning (CL) methods maintain performance on earlier tasks by storing exemplars in a replay buffer, introducing memory overhead that

safetyarxiv-cs-lg
11 May 2026
Safety

NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models

DGX agent

arXiv:2605.07794v1 Announce Type: new Abstract: World Action Models (WAMs) are an emerging family of policies that tie robot action generation to future-observation modeling. In this work, we focus on

safetyarxiv-cs-ro
11 May 2026
Safety

Not even surprised by horrific stories like these anymore. The mission of getting LLMs aligned with human values has largely been a failure.

DGX agent

Not even surprised by horrific stories like these anymore. The mission of getting LLMs aligned with human values has largely been a failure. NEW: ChatGPT advised the FSU shooter that a mass shooting w

safetygary-marcus--x
11 May 2026
Safety

OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search

DGX agent

arXiv:2604.03675v2 Announce Type: replace Abstract: Agentic search enables language models to solve knowledge-intensive tasks by adaptively acquiring external evidence over multiple steps. Reinforceme

safetyarxiv-cs-ai
11 May 2026
Safety

Object Hallucination-Free Reinforcement Unlearning for Vision-Language Models

DGX agent

arXiv:2605.08031v1 Announce Type: new Abstract: Vision-language models (VLMs) raise growing concerns about privacy, copyright, and bias, motivating machine unlearning to remove sensitive knowledge. Ho

safetyarxiv-cs-cv
11 May 2026
Safety

Offline Policy Optimization with Posterior Sampling

DGX agent

arXiv:2605.07393v1 Announce Type: new Abstract: A fundamental challenge in model-based offline reinforcement learning (RL) lies in the trade-off between generalization and robustness against exploitat

safetyarxiv-cs-ai
11 May 2026
Safety

oh. my. god. 😱

DGX agent

oh. my. god. 😱 FT Exclusive: NHS England has granted external staff from companies including Palantir “unlimited access” to identifiable patient data while working on a part of its flagship data platf

safetygary-marcus--x
11 May 2026
Safety

On the Meta-Design of Allocation Problems

DGX agent

arXiv:2602.08786v4 Announce Type: replace-cross Abstract: There is an extensive literature that studies how to find optimal policies in resource allocation problems, taking the underlying design param

safetyarxiv-cs-lg
11 May 2026
Safety

On Training in Imagination

DGX agent

arXiv:2605.06732v1 Announce Type: new Abstract: State-of-the-art model-based reinforcement learning methods train policies on imagined rollouts. These rollouts are trajectories generated by a learned

safetyarxiv-cs-lg
11 May 2026
Safety

One estimate of how much annual revenue AI needs to “make sense”: 1.6 trillion. That’s four times what Google made in its best year. (total …

DGX agent

One estimate of how much annual revenue AI needs to “make sense”: 1.6 trillion. That’s four times what Google made in its best year. (total revenue so far is perhaps on order of 100 billion.) In 2024

safetygary-marcus--x
11 May 2026
Safety

One Token Per Frame: Reconsidering Visual Bandwidth in World Models for VLA Policy

DGX agent

arXiv:2605.07931v1 Announce Type: cross Abstract: Vision-language-action (VLA) models increasingly rely on auxiliary world modules to plan over long horizons, yet how such modules should be parameteri

safetyarxiv-cs-ai
11 May 2026
Safety

Online Allocation with Unknown Shared Supply

DGX agent

arXiv:2605.07080v1 Announce Type: new Abstract: Many real-world resource allocation systems, such as humanitarian logistics and vaccine distribution, must preposition limited supply across multiple lo

safetyarxiv-cs-ai
11 May 2026
Safety

Openclaw token consumption fell by half in a month, per openrouter data What happened?

DGX agent

Openclaw token consumption dropped 50% within a month according to OpenRouter usage data, as reported by Gary Marcus. The post likely discusses potential causes for this significant decline, such as c

safetygary-marcus--x
11 May 2026
Safety

Optimal Recourse Summaries via Bi-Objective Decision Tree Learning

DGX agent

arXiv:2605.07598v1 Announce Type: new Abstract: Actionable Recourse provides individuals with actions they can take to change an unfavorable classifier outcome. While useful at the instance level, it

safetyarxiv-cs-lg
11 May 2026
Safety

PACEvolve++: Improving Test-time Learning for Evolutionary Search Agents

DGX agent

arXiv:2605.07039v1 Announce Type: new Abstract: Large language models have become drivers of evolutionary search, but most systems rely on a fixed, prompt-elicited policy to sample next candidates. Th

safetyarxiv-cs-lg
11 May 2026
Safety

Pan-FM: A Pan-Organ Foundation Model with Saliency-Guided Masking for Missing Robustness

DGX agent

arXiv:2605.07055v1 Announce Type: cross Abstract: Foundation models (FMs) have shown great promise in medical imaging, but most FMs are trained on unimodal data within isolated domains, such as brain

safetyarxiv-cs-ai
11 May 2026
Safety

PaT: Planning-after-Trial for Efficient Test-Time Code Generation

DGX agent

arXiv:2605.07248v1 Announce Type: new Abstract: Beyond training-time optimization, scaling test-time computation has emerged as a key paradigm to extend the reasoning capabilities of Large Language Mo

safetyarxiv-cs-cl
11 May 2026
Safety

Persistent-Transient Policy Evaluation for Markov Chains via Minimal Peripheral Quotients

DGX agent

arXiv:2602.00474v2 Announce Type: replace-cross Abstract: We study fixed-policy evaluation for finite Markov chains that may be reducible and periodic. Classical evaluation methods with gain and bias

safetyarxiv-cs-lg
11 May 2026
Safety

Physical Simulators as Do-Operators: Causal Discovery under Latent Confounders for AI-for-Science

DGX agent

arXiv:2605.07467v1 Announce Type: cross Abstract: Existing interventional causal discovery methods -- IGSP, DCDI, ENCO -- assume causal sufficiency (no latent confounders) and rely on virtual interven

safetyarxiv-cs-ai
11 May 2026
Safety

Physics-Based Benchmarking Metrics for Multimodal Synthetic Images

DGX agent

arXiv:2511.15204v3 Announce Type: replace-cross Abstract: Current state of the art measures like BLEU, CIDEr, VQA score, SigLIP-2 and CLIPScore are often unable to capture semantic or structural accur

safetyarxiv-cs-ai
11 May 2026
Safety

PLOT: Progressive Localization via Optimal Transport in Neural Causal Abstraction

DGX agent

arXiv:2605.06979v1 Announce Type: cross Abstract: Causal abstraction offers a principled framework for mechanistic interpretability, aligning a high-level causal model with the low-level computation r

safetyarxiv-cs-ai
11 May 2026
Safety

POETS: Uncertainty-Aware LLM Optimization via Compute-Efficient Policy Ensembles

DGX agent

arXiv:2605.07775v1 Announce Type: cross Abstract: Balancing exploration and exploitation is a core challenge in sequential decision-making and black-box optimization. We introduce POETS (extbf{Po}licy

safetyarxiv-cs-ai
11 May 2026
Safety

Position: Mechanistic Interpretability Must Disclose Identification Assumptions for Causal Claims

DGX agent

arXiv:2605.08012v1 Announce Type: cross Abstract: Mechanistic interpretability papers increasingly use causal vocabulary: circuits, mediators, causal abstraction, monosemanticity. Such claims require

safetyarxiv-cs-ai
11 May 2026
Safety

Post-training makes large language models less human-like

DGX agent

arXiv:2605.07632v1 Announce Type: cross Abstract: Large language models (LLMs) are increasingly used as surrogates for human participants, but it remains unclear which models best capture human behavi

safetyarxiv-cs-ai
11 May 2026
Safety

ProtoSSL: Interpretable Prototype Learning from Unlabeled Time-Series Data

DGX agent

arXiv:2605.06943v1 Announce Type: new Abstract: In time-series domains where both predictive performance and interpretability are essential, deep neural networks achieve strong results but provide lim

safetyarxiv-cs-lg
11 May 2026
Safety

Proxy3D: Efficient 3D Representations for Vision-Language Models via Semantic Clustering and Alignment

DGX agent

arXiv:2605.08064v1 Announce Type: new Abstract: Spatial intelligence in vision-language models (VLMs) attracts research interest with the practical demand to reason in the 3D world.Despite promising r

safetyarxiv-cs-cv
11 May 2026
Safety

Prune-OPD: Efficient and Reliable On-Policy Distillation for Long-Horizon Reasoning

DGX agent

arXiv:2605.07804v1 Announce Type: cross Abstract: On-policy distillation (OPD) leverages dense teacher rewards to enhance reasoning models. However, scaling OPD to long-horizon tasks exposes a critica

safetyarxiv-cs-ai
11 May 2026
Safety

Q-MMR: Off-Policy Evaluation via Recursive Reweighting and Moment Matching

DGX agent

arXiv:2605.06474v2 Announce Type: replace-cross Abstract: We present a novel theoretical framework, Q-MMR, for off-policy evaluation in finite-horizon MDPs. Q-MMR learns a set of scalar weights, one f

safetyarxiv-cs-ai
11 May 2026
Safety

R-GTD: A Geometric Analysis of Gradient Temporal-Difference Learning in Singular Regimes

DGX agent

arXiv:2601.20599v2 Announce Type: replace-cross Abstract: Gradient temporal-difference (GTD) learning algorithms are widely used for off-policy policy evaluation with function approximation. However,

safetyarxiv-cs-ai
11 May 2026
Safety

Radiologist-Guided Causal Concept Bottleneck Models for Chest X-Ray Interpretation

DGX agent

arXiv:2605.07785v1 Announce Type: new Abstract: Concept Bottleneck Models (CBMs) in medical imaging aim to improve model interpretability by predicting intermediate clinical concepts before final diag

safetyarxiv-cs-cv
11 May 2026
Safety

Reason to Play: Behavioral and Brain Alignment Between Frontier LRMs and Human Game Learners

DGX agent

arXiv:2605.08019v1 Announce Type: new Abstract: Humans rapidly learn abstract knowledge when encountering novel environments and flexibly deploy this knowledge to guide efficient and intelligent actio

safetyarxiv-cs-ai
11 May 2026
Safety

ReasonEdit: Towards Interpretable Image Editing Evaluation via Reinforcement Learning

DGX agent

arXiv:2605.07477v1 Announce Type: new Abstract: Recent text-guided image editing (TIE) models have achieved remarkable progress, however, many edited results still suffer from artifacts, unintended mo

safetyarxiv-cs-cv
11 May 2026
Safety

Receipts: https://open.substack.com/pub/garymarcus/p/deconstructing-geoffrey-hintons-weakest?r=8tdk6&utm_medium=ios

DGX agent

Gary Marcus analyzes and critiques Geoffrey Hinton's arguments regarding weaknesses in deep learning and artificial neural networks. The article likely examines specific technical or conceptual claims

safetygary-marcus--x
11 May 2026
Safety

ReCLIP++: Learn to Rectify the Bias of CLIP for Unsupervised Semantic Segmentation

DGX agent

arXiv:2408.06747v4 Announce Type: replace Abstract: Recent works utilize CLIP to perform the challenging unsupervised semantic segmentation task where only images without annotations are available. Ho

safetyarxiv-cs-cv
11 May 2026
Safety

Reflections and New Directions for Human-Centered Large Language Models

DGX agent

arXiv:2605.06901v1 Announce Type: new Abstract: Large Language Models (LLMs) are increasingly shaping the private and professional lives of users, with numerous applications in business, education, fi

safetyarxiv-cs-cl
11 May 2026
Safety

Reinforcement Learning for Exponential Utility: Algorithms and Convergence in Discounted MDPs

DGX agent

arXiv:2605.08053v1 Announce Type: new Abstract: Reinforcement learning (RL) for exponential-utility optimization in discounted Markov decision processes (MDPs) lacks principled value-based algorithms.

safetyarxiv-cs-lg
11 May 2026
Safety

RELO: Reinforcement Learning to Localize for Visual Object Tracking

DGX agent

arXiv:2605.07379v1 Announce Type: cross Abstract: Conventional visual object trackers localize targets using handcrafted spatial priors, often in the form of heatmaps. Such priors provide only surroga

safetyarxiv-cs-ai
11 May 2026
Safety

Repeated Deceptive Path Planning against Learnable Observer

DGX agent

arXiv:2605.07174v1 Announce Type: new Abstract: We study the problem of deceptive path planning (DPP), where an agent aims to conceal its true destination from external observers. While existing work

safetyarxiv-cs-ai
11 May 2026
Safety

reply to Hinton’s reply to me, for additional context:

DGX agent

reply to Hinton’s reply to me, for additional context: Dear @geoffreyhinton, I literally never said that AI systems “JUST regurgitate”; that’s plainly false. I don’t believe it, and I didn’t say it. (

safetygary-marcus--x
11 May 2026
Safety

Resource-Element Energy Difference for Noncoherent Over-the-Air Federated Learning

DGX agent

arXiv:2605.07263v1 Announce Type: cross Abstract: Over-the-air federated learning (OTA-FL) reduces uplink latency by exploiting waveform superposition, but conventional analog aggregation schemes typi

safetyarxiv-cs-ai
11 May 2026
Safety

Response-G1: Explicit Scene Graph Modeling for Proactive Streaming Video Understanding

DGX agent

arXiv:2605.07575v1 Announce Type: cross Abstract: Proactive streaming video understanding requires Video-LLMs to decide when to respond as a video unfolds, a task where existing methods often fall sho

safetyarxiv-cs-ai
11 May 2026
← Previous
1…225226227228229…302
Next →