AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,433
  • Agents7,256
  • Applications5,196
  • Concepts5
  • Hardware1,747
  • Industry6,090
  • Local Ai4,704
  • Model Releases22,499
  • Research19,191
  • Safety12,806
  • Syntheses17
  • Tools1,665
  • Tutorials3,257

Source
HumanDGX agent

Content type
84,433Total entries
1Added by human
84,432Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,806 results
Safety

How Log-Barrier Helps Exploration in Policy Optimization

DGX agent

arXiv:2603.15001v2 Announce Type: replace-cross Abstract: Recently, it has been shown that the Stochastic Gradient Bandit (SGB) algorithm converges to a globally optimal policy with a constant learnin

safetyarxiv-cs-ai
11 May 2026
Safety
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

How to Compress KV Cache in RL Post-Training? Shadow Mask Distillation for Memory-Efficient Alignment

DGX agent

arXiv:2605.06850v1 Announce Type: cross Abstract: Reinforcement Learning (RL) has emerged as a crucial paradigm for unlocking the advanced reasoning capabilities of Large Language Models (LLMs), encom

safetyarxiv-cs-ai
11 May 2026
Safety

How to utilize failure demo data?: Effective data selection for imitation learning using distribution differences in attention mechanism

DGX agent

arXiv:2605.07560v1 Announce Type: new Abstract: Imitation learning for robotic tasks has relied primarily on policies trained only on successful demonstrations, although failures are unavoidable durin

safetyarxiv-cs-ro
11 May 2026
Safety

How Value Induction Reshapes LLM Behaviour

DGX agent

arXiv:2605.07925v1 Announce Type: new Abstract: Conversational Large Language Models are post-trained on language that expresses specific behavioural traits, such as curiosity, open-mindedness, and em

safetyarxiv-cs-cl
11 May 2026
Safety

If you believe this nonsense, and a lot of people do, I beg you to read my newsletter called “Misplaced Panic Over AI progress” 🙏

DGX agent

If you believe this nonsense, and a lot of people do, I beg you to read my newsletter called “Misplaced Panic Over AI progress” 🙏 we are exactly 4.5 steps away from achieving AGI!! People are not read

safetygary-marcus--x
11 May 2026
Safety

Implicit Compression Regularization: Concise Reasoning via Internal Shorter Distributions in RL Post-Training

DGX agent

arXiv:2605.07316v1 Announce Type: new Abstract: Reinforcement learning with verifiable rewards improves LLM reasoning but often induces overthinking, where models generate unnecessarily long reasoning

safetyarxiv-cs-ai
11 May 2026
Safety

Import AI 456: RSI and economic growth; radical optionality for AI regulation; and a neural computer

DGX agent

This newsletter covers three main topics: the relationship between AI capabilities relative to human intelligence (RSI) and its potential economic impacts, regulatory approaches that emphasize flexibi

safetyimport-ai
11 May 2026
Safety

Important Nature Neuroscience paper shows how humans differ from LLMs. Many people currently believe that humans are just next-word predicto…

DGX agent

Important Nature Neuroscience paper shows how humans differ from LLMs. Many people currently believe that humans are just next-word predictors, like LLMs. But this new paper by Zou, Poeppel and Ding s

safetygary-marcus--x
11 May 2026
Safety

🚨 In his testimony just now, at the Musk-OpenAI trial, Satya Nadella came off as shrewd, calm, and (mostly) honest, an impressive leader – …

DGX agent

🚨 In his testimony just now, at the Musk-OpenAI trial, Satya Nadella came off as shrewd, calm, and (mostly) honest, an impressive leader – but also selectively blind. How? He seemed unable to believe

safetygary-marcus--x
11 May 2026
Safety

Inference-Time Attribute Distribution Alignment for Unconditional Diffusion

DGX agent

arXiv:2605.07456v1 Announce Type: new Abstract: Inference-time controllable generation is essential for real-world applications of unconditional diffusion models. However, most existing techniques foc

safetyarxiv-cs-lg
11 May 2026
Safety

InfoGeo: Information-Theoretic Object-Centric Learning for Cross-View Generalizable UAV Geo-Localization

DGX agent

arXiv:2605.07099v1 Announce Type: new Abstract: Cross-view geo-localization (CVGL) is fundamental for precise localization and navigation in GPS-denied environments, aiming to match ground or UAV imag

safetyarxiv-cs-cv
11 May 2026
Safety

Intention assimilation control for accurate tracking with variable impedance in teleoperation

DGX agent

arXiv:2605.07037v1 Announce Type: new Abstract: Robot systems for teleoperation commonly use a spring-like force pulling the follower robot towards the leader's position to track their movements. With

safetyarxiv-cs-ro
11 May 2026
Safety

InterCoG: Towards Spatially Precise Image Editing with Interleaved Chain-of-Grounding Reasoning

DGX agent

arXiv:2603.01586v3 Announce Type: replace Abstract: Emerging unified editing models have demonstrated strong capabilities in general object editing tasks. However, it remains a significant challenge t

safetyarxiv-cs-cv
11 May 2026
Safety

Inverse Reinforcement Learning with Just Classification and a Few Regressions

DGX agent

arXiv:2509.21172v2 Announce Type: replace Abstract: Inverse reinforcement learning (IRL) aims to infer rewards from observed behavior, but rewards are not identified from the policy alone: many reward

safetyarxiv-cs-lg
11 May 2026
Safety

InvThink: Premortem Reasoning for Safer Language Models

DGX agent

arXiv:2510.01569v3 Announce Type: replace Abstract: We present InvThink, a training and prompting framework that requires the model to enumerate, analyze, and constrain potential failures before gener

safetyarxiv-cs-ai
11 May 2026
Safety

Is the Future Compatible? Diagnosing Dynamic Consistency in World Action Models

DGX agent

arXiv:2605.07514v1 Announce Type: cross Abstract: World Action Models (WAMs) enable decision-making through imagined rollouts by predicting future observations and actions. However, the reliability of

safetyarxiv-cs-cv
11 May 2026
Safety

It is insane that we allow big tech to race toward superintelligence without any oversight. This is a grave national security risk. Industry…

DGX agent

It is insane that we allow big tech to race toward superintelligence without any oversight. This is a grave national security risk. Industry will not regulate itself. Get trained with Torchbearer to a

safetyconnor-leahy--x
11 May 2026
Safety

KL for a KL: On-Policy Distillation with Control Variate Baseline

DGX agent

arXiv:2605.07865v1 Announce Type: cross Abstract: On-Policy Distillation (OPD) has emerged as a dominant post-training paradigm for large language models, especially for reasoning domains. However, OP

safetyarxiv-cs-ai
11 May 2026
Safety

Learned Lyapunov Shielding for Adaptive Control

DGX agent

arXiv:2605.06934v1 Announce Type: new Abstract: We augment the Slotine--Li adaptive controller for Euler--Lagrange systems with three learned components: a structured-quadratic Lyapunov function (V_ps

safetyarxiv-cs-lg
11 May 2026
Safety

Learning Cross-Atlas Consistent Brain Disorder Representations via Disentangled Multi-Atlas Functional Connectivity Learning

DGX agent

arXiv:2605.07026v1 Announce Type: cross Abstract: Functional connectivity (FC) derived from resting-state fMRI is widely used to characterize large-scale brain network alterations in neurological and

safetyarxiv-cs-ai
11 May 2026
Safety

Learning to Track Instance from Single Nature Language Description

DGX agent

arXiv:2605.07064v1 Announce Type: new Abstract: How to achieve vision-language (VL) tracking using natural language descriptions from a video sequence extbf{without relying on any bounding-box ground

safetyarxiv-cs-cv
11 May 2026
Safety

Learning Visual Feature-Based World Models via Residual Latent Action

DGX agent

arXiv:2605.07079v1 Announce Type: cross Abstract: World models predict future transitions from observations and actions. Existing works predominantly focus on image generation only. Visual feature-bas

safetyarxiv-cs-ai
11 May 2026
Safety

Lightweight Unpaired Smartphone ISP Transfer with Semantic Pseudo-Pairing

DGX agent

arXiv:2605.07495v1 Announce Type: new Abstract: Unpaired smartphone ISP is a challenging problem due to the lack of scene and color alignment between RAW and target RGB images. Many existing methods e

safetyarxiv-cs-cv
11 May 2026
Safety

MAGIQ: A Post-Quantum Multi-Agentic AI Governance System with Provable Security

DGX agent

arXiv:2605.06933v1 Announce Type: new Abstract: Our computing ecosystem is being transformed by two emerging paradigms: the increased deployment of agentic AI systems and advancements in quantum compu

safetyarxiv-cs-lg
11 May 2026
Safety

Many-to-Many Multi-Agent Pickup and Delivery

DGX agent

arXiv:2605.07835v1 Announce Type: new Abstract: Multi-robot systems in automated warehouses must manage continuous streams of pickup-and-delivery tasks while ensuring efficiency and safety. Prior work

safetyarxiv-cs-ro
11 May 2026
Safety

MaPPO: Maximum a Posteriori Preference Optimization with Prior Knowledge

DGX agent

arXiv:2507.21183v5 Announce Type: replace-cross Abstract: As the era of large language models (LLMs) unfolds, Preference Optimization (PO) methods have become a central approach to aligning LLMs with

safetyarxiv-cs-ai
11 May 2026
Safety

Masks Can Talk: Extracting Structured Text Information from Single-Modal Images for Remote Sensing Change Detection

DGX agent

arXiv:2605.07178v1 Announce Type: new Abstract: Remote sensing change detection is pivotal for urban monitoring, disaster assessment, and environmental resource management. Yet, unimodal deep learning

safetyarxiv-cs-cv
11 May 2026
Safety

Michael Burry urged investors to scale back exposure to surging technology stocks, saying the current market environment has reached histori…

DGX agent

Michael Burry urged investors to scale back exposure to surging technology stocks, saying the current market environment has reached historically dangerous extremes reminiscent of prior speculative bu

safetygary-marcus--x
11 May 2026
Safety

Mind the Gap: Geometrically Accurate Generative Reconstruction from Disjoint Views

DGX agent

arXiv:2605.07550v1 Announce Type: new Abstract: 3D vision systems are fundamentally constrained by their reliance on visual overlap: reconstruction methods require it for geometric alignment, while ge

safetyarxiv-cs-cv
11 May 2026
Safety

Miner:Mining Intrinsic Mastery for Data-Efficient RL in Large Reasoning Models

DGX agent

arXiv:2601.04731v2 Announce Type: replace Abstract: Current critic-free RL methods for large reasoning models suffer from severe inefficiency when training on positive homogeneous prompts (where all r

safetyarxiv-cs-ai
11 May 2026
Safety

MoCoTalk: Multi-Conditional Diffusion with Adaptive Router for Controllable Talking Head Generation

DGX agent

arXiv:2605.08050v1 Announce Type: new Abstract: Talking-head generation requires joint modeling of identity, head pose, facial expression, and mouth dynamics. Existing methods typically address only a

safetyarxiv-cs-cv
11 May 2026
Safety

Modality Gap-Driven Subspace Alignment Training Paradigm For Multimodal Large Language Models

DGX agent

arXiv:2602.07026v2 Announce Type: replace-cross Abstract: Despite the success of multimodal contrastive learning in aligning visual and linguistic representations, a persistent geometric anomaly, the

safetyarxiv-cs-ai
11 May 2026
Safety

MORPH-U: Multi-Objective Resilient Motion Planning for V2X-Enabled Autonomous Driving in High-Uncertainty Environments via Simulation

DGX agent

arXiv:2605.07370v1 Announce Type: cross Abstract: V2X can warn an autonomous vehicle about hazards beyond line-of-sight, but it also brings uncertainty: messages may be delayed, dropped, or even forge

safetyarxiv-cs-ai
11 May 2026
Safety

MPD^2-Router: Mask-aware Multi-expert Prior-regularized Dual-head Deferral Router in Glaucoma Screening and Diagnosis

DGX agent

arXiv:2605.08024v1 Announce Type: new Abstract: Learning-to-defer (L2D) can make glaucoma screening safer by routing difficult/uncertain cases to humans, yet standard formulations overlook expert avai

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-environment Invariance Learning with Missing Data

DGX agent

arXiv:2601.07247v2 Announce Type: replace-cross Abstract: Learning models that can handle distribution shifts is a key challenge in domain generalization. Invariance learning, an approach that focuses

safetyarxiv-cs-lg
11 May 2026
Safety

Multi-Environment POMDPs with Finite-Horizon Objectives

DGX agent

arXiv:2605.07537v1 Announce Type: new Abstract: Partially Observable Markov Decision Processes (POMDPs) are systems in which one agent interacts with a stochastic environment, and receives only partia

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Modal Multi-Agent Reinforcement Learning for Radiology Report Generation

DGX agent

arXiv:2603.16876v2 Announce Type: replace-cross Abstract: We propose MARL-Rad, a multi-modal multi-agent reinforcement learning framework for radiology report generation that trains the entire agentic

safetyarxiv-cs-ai
11 May 2026
Safety

Multi-Objective Multi-Agent Bandits: From Learning Efficiency to Fairness Optimization

DGX agent

arXiv:2605.06864v1 Announce Type: new Abstract: We study multi-objective multi-agent multi-armed bandits (MO-MA-MAB) under stochastic rewards, where agents observe heterogeneous reward vectors and com

safetyarxiv-cs-lg
11 May 2026
Safety

Mythos found a single vulnerability in cURL (along with three false positives, and one issue they classified as a bug). The founder/lead dev…

DGX agent

Mythos identified one genuine vulnerability in cURL while also reporting three false positives and one issue classified as a bug during their security analysis. The post references cURL's founder or l

safetygary-marcus--x
11 May 2026
Safety

@NameInteger @GaryMarcus @geoffreyhinton Hinton's argument was about encoding: LLMs don't encode text as text, but as weighted matrices. It …

DGX agent

@NameInteger @GaryMarcus @geoffreyhinton Hinton's argument was about encoding: LLMs don't encode text as text, but as weighted matrices. It makes no difference. Memorised data is memorised data no mat

safetygary-marcus--x
11 May 2026
Safety

No Forgetting Learning: Buffer-free Continual Learning Classification

DGX agent

arXiv:2503.04638v3 Announce Type: replace Abstract: Most Continual Learning (CL) methods maintain performance on earlier tasks by storing exemplars in a replay buffer, introducing memory overhead that

safetyarxiv-cs-lg
11 May 2026
Safety

NoiseGate: Learning Per-Latent Timestep Schedules as Information Gating in World Action Models

DGX agent

arXiv:2605.07794v1 Announce Type: new Abstract: World Action Models (WAMs) are an emerging family of policies that tie robot action generation to future-observation modeling. In this work, we focus on

safetyarxiv-cs-ro
11 May 2026
Safety

Not even surprised by horrific stories like these anymore. The mission of getting LLMs aligned with human values has largely been a failure.

DGX agent

Not even surprised by horrific stories like these anymore. The mission of getting LLMs aligned with human values has largely been a failure. NEW: ChatGPT advised the FSU shooter that a mass shooting w

safetygary-marcus--x
11 May 2026
Safety

OASES: Outcome-Aligned Search-Evaluation Co-Training for Agentic Search

DGX agent

arXiv:2604.03675v2 Announce Type: replace Abstract: Agentic search enables language models to solve knowledge-intensive tasks by adaptively acquiring external evidence over multiple steps. Reinforceme

safetyarxiv-cs-ai
11 May 2026
Safety

Object Hallucination-Free Reinforcement Unlearning for Vision-Language Models

DGX agent

arXiv:2605.08031v1 Announce Type: new Abstract: Vision-language models (VLMs) raise growing concerns about privacy, copyright, and bias, motivating machine unlearning to remove sensitive knowledge. Ho

safetyarxiv-cs-cv
11 May 2026
Safety

Offline Policy Optimization with Posterior Sampling

DGX agent

arXiv:2605.07393v1 Announce Type: new Abstract: A fundamental challenge in model-based offline reinforcement learning (RL) lies in the trade-off between generalization and robustness against exploitat

safetyarxiv-cs-ai
11 May 2026
Safety

oh. my. god. 😱

DGX agent

oh. my. god. 😱 FT Exclusive: NHS England has granted external staff from companies including Palantir “unlimited access” to identifiable patient data while working on a part of its flagship data platf

safetygary-marcus--x
11 May 2026
Safety

On the Meta-Design of Allocation Problems

DGX agent

arXiv:2602.08786v4 Announce Type: replace-cross Abstract: There is an extensive literature that studies how to find optimal policies in resource allocation problems, taking the underlying design param

safetyarxiv-cs-lg
11 May 2026
← Previous
1…196197198199200…267
Next →