AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,707 results
4 May 2026

Prompt-Induced Score Variance in Zero-Shot Binary Vision-Language Safety Classification

SafetyDGX agent

arXiv:2605.00326v1 Announce Type: new Abstract: Single-prompt first-token probabilities from zero-shot vision-language model (VLM) safety classifiers are treated as decision scores, but we show they a

Provable and scalable quantum Gaussian processes for quantum learning

SafetyDGX agent

arXiv:2605.00099v1 Announce Type: cross Abstract: Despite rapid recent advances in quantum machine learning, the field is in many ways stuck. Existing approaches can exhibit serious limitations, and w

Recovering Hidden Reward in Diffusion-Based Policies

SafetyDGX agent

arXiv:2605.00623v1 Announce Type: new Abstract: This paper introduces EnergyFlow, a framework that unifies generative action modeling with inverse reinforcement learning by parameterizing a scalar ene


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Reinforcement Learning for LLM Post-Training: A Survey

SafetyDGX agent

arXiv:2407.16216v3 Announce Type: replace Abstract: Large language models (LLMs) trained via pretraining and supervised fine-tuning (SFT) can still produce harmful and misaligned outputs, or struggle

Reinforcement Learning with LLM-Guided Action Spaces for Synthesizable Lead Optimization

SafetyDGX agent

arXiv:2604.07669v2 Announce Type: replace Abstract: Lead optimization in drug discovery requires improving therapeutic properties while ensuring that molecular modifications correspond to feasible syn

Reinforcement Learning with Markov Risk Measures and Multipattern Risk Approximation

SafetyDGX agent

arXiv:2605.00654v1 Announce Type: new Abstract: For a risk-averse finite-horizon Markov Decision Problem, we introduce a special class of Markov coherent risk measures, called mini-batch measures. We

ReLay: Personalized LLM-Generated Plain-Language Summaries for Better Understanding, but at What Cost?

SafetyDGX agent

arXiv:2605.00468v1 Announce Type: new Abstract: Plain Language Summaries (PLS) aim to make research accessible to lay readers, but they are typically written in a one-size-fits-all style that ignores

ResRL: Boosting LLM Reasoning via Negative Sample Projection Residual Reinforcement Learning

SafetyDGX agent

arXiv:2605.00380v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) enhances reasoning of Large Language Models (LLMs) but usually exhibits limited generation diver

Resting Neurons, Active Insights: Robustify Activation Sparsity for Large Language Models

SafetyDGX agent

arXiv:2512.12744v3 Announce Type: replace Abstract: Activation sparsity offers a compelling route to accelerate large language model (LLM) inference by selectively suppressing hidden activations, yet

SAGA: Workflow-Atomic Scheduling for AI Agent Inference on GPU Clusters

SafetyDGX agent

arXiv:2605.00528v1 Announce Type: cross Abstract: AI agents execute tens to hundreds of chained LLM calls per task, yet GPU schedulers treat each call as independent, discarding gigabytes of intermedi

SAVGO: Learning State-Action Value Geometry with Cosine Similarity for Continuous Control

SafetyDGX agent

arXiv:2605.00787v1 Announce Type: new Abstract: While representation and similarity learning have improved the sample efficiency of Reinforcement Learning (RL), they are rarely used to shape policy up

Scaling Video Understanding via Compact Latent Multi-Agent Collaboration

SafetyDGX agent

arXiv:2605.00444v1 Announce Type: new Abstract: Multi-modal large language models (MLLMs) advance vision language understanding but face inherent limitations in long-video tasks due to bounded percept

SIMON: Saliency-aware Integrative Multi-view Object-centric Neural Decoding

SafetyDGX agent

arXiv:2605.00401v1 Announce Type: new Abstract: Recent EEG-to-image retrieval methods leverage pretrained vision encoders and foveation-inspired priors, but typically assume a fixed, center-focused vi

Soft Graph Diffusion Transformer for MIMO Detection

SafetyDGX agent

arXiv:2605.00449v1 Announce Type: cross Abstract: Learning-based MIMO detection has shown strong empirical performance, yet existing methods typically rely on fixed-depth architectures without explici

Soft-MSM: Differentiable Context-Aware Elastic Alignment for Time Series

SafetyDGX agent

arXiv:2605.00069v1 Announce Type: new Abstract: Elastic distances like dynamic time warping (DTW) are central to time series machine learning because they compare sequences under local temporal misali

Stable-GFlowNet: Toward Diverse and Robust LLM Red-Teaming via Contrastive Trajectory Balance

SafetyDGX agent

arXiv:2605.00553v1 Announce Type: new Abstract: Large Language Model (LLM) Red-Teaming, which proactively identifies vulnerabilities of LLMs, is an essential process for ensuring safety. Finding effec

Statistical Impossibility and Possibility of Aligning LLMs with Human Preferences: From Condorcet Paradox to Nash Equilibrium

SafetyDGX agent

arXiv:2503.10990v2 Announce Type: replace-cross Abstract: Aligning large language models (LLMs) with diverse human preferences is critical for ensuring fairness and informed outcomes when deploying th

The Determinism of Randomness: Latent Space Degeneracy in Diffusion Model

SafetyDGX agent

arXiv:2511.07756v4 Announce Type: replace Abstract: Diffusion models initialize generation from an isotropic Gaussian latent, yet changing only the random seed can substantially alter prompt faithfuln

Towards A Generative Protein Evolution Machine with DPLM-Evo

SafetyDGX agent

arXiv:2605.00182v1 Announce Type: new Abstract: Proteins are shaped by gradual evolution under biophysical and functional constraints. Protein language models learn rich evolutionary constraints from

Uniform-Correct Policy Optimization: Breaking RLVR's Indifference to Diversity

SafetyDGX agent

arXiv:2605.00365v1 Announce Type: cross Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has achieved substantial gains in single-attempt accuracy (Pass@1) on reasoning tasks, yet often

UniVidX: A Unified Multimodal Framework for Versatile Video Generation via Diffusion Priors

SafetyDGX agent

arXiv:2605.00658v1 Announce Type: new Abstract: Recent progress has shown that video diffusion models (VDMs) can be repurposed for diverse multimodal graphics tasks. However, existing methods often tr

Unlearning What Matters: Token-Level Attribution for Precise Language Model Unlearning

SafetyDGX agent

arXiv:2605.00364v1 Announce Type: new Abstract: Machine unlearning has emerged as a critical capability for addressing privacy, safety, and regulatory concerns in large language models (LLMs). Existin

Unlocking Zero-Shot Geospatial Reasoning via Indirect Rewards

SafetyDGX agent

arXiv:2510.00072v2 Announce Type: replace Abstract: Training robust reasoning vision-language models (VLMs) in rare domains (such as geospatial) is fundamentally constrained by supervision scarcity. W

Unpaired Image Deraining Using Reward-Guided Self-Reinforcement Strategy

SafetyDGX agent

arXiv:2605.00719v1 Announce Type: new Abstract: Unsupervised deraining has attracted attention for its ability to learn the real-world distribution of rain without paired supervision. However, the lac

VGR: Visual Grounded Reasoning

SafetyDGX agent

arXiv:2506.11991v3 Announce Type: replace-cross Abstract: In the field of multimodal chain-of-thought (CoT) reasoning, existing approaches predominantly rely on reasoning on pure language space, which

VLBiMan: Vision-Language Anchored One-Shot Demonstration Enables Generalizable Bimanual Robotic Manipulation

SafetyDGX agent

arXiv:2509.21723v4 Announce Type: replace Abstract: Achieving generalizable bimanual manipulation requires systems that can learn efficiently from minimal human input while adapting to real-world unce

VR founder Jaron Lanier says a future where people are paid for the data they give AI is better than one where everyone depends on billionai…

SafetyDGX agent

VR founder Jaron Lanier says a future where people are paid for the data they give AI is better than one where everyone depends on billionaires to survive 'artificial intelligence is a marketing term,

Wasserstein Distributionally Robust Regret Optimization for Reinforcement Learning from Human Feedback

SafetyDGX agent

arXiv:2605.00155v1 Announce Type: cross Abstract: Reinforcement learning from human feedback (RLHF) has become a core post-training step for aligning large language models, yet the reward signal used

Way stronger. Today was huge for Musk, bad for OpenAI.

SafetyDGX agent

Way stronger. Today was huge for Musk, bad for OpenAI. The case against OpenAI is getting markedly stronger now that Musk is off the stand. Why? Musk’s lawyer is interrogating OpenAI founder Greg Broc

Well-known, well-respected AI expert Stuart Russell is now testifying at the Musk-OpenAI trial – despite OpenAI’s BS attempts to exclude him…

SafetyDGX agent

Well-known, well-respected AI expert Stuart Russell is now testifying at the Musk-OpenAI trial – despite OpenAI’s BS attempts to exclude him. Read the backstory here: A new filing just dropped in the

What Physics do Data-Driven MoCap-to-Radar Models Learn?

SafetyDGX agent

arXiv:2605.00018v1 Announce Type: new Abstract: Data-driven MoCap-to-radar models generate plausible micro-Doppler spectrograms, but do they actually learn the underlying physics? We introduce a physi

Why do these influencers always say “just published” for papers published last year? This paper is great, and I have written extensively abo…

SafetyDGX agent

Why do these influencers always say “just published” for papers published last year? This paper is great, and I have written extensively about it in my newsletter. But c’mon. Apple didn’t “just publis

World Model for Robot Learning: A Comprehensive Survey

SafetyDGX agent

arXiv:2605.00080v1 Announce Type: cross Abstract: World models, which are predictive representations of how environments evolve under actions, have become a central component of robot learning. They s

Wow. Greg Brockman had a 10M side deal with Altman even in the early nonprofit days, which not disclosed to Elon or (I believe) in nonprofit…

SafetyDGX agent

Gary Marcus reported that Greg Brockman, OpenAI's President, had a $10 million side deal with Sam Altman during OpenAI's early nonprofit period that was not disclosed to Elon Musk or documented in non

3 May 2026

A profile of BlackBerry's QNX division, whose operating system controls safety features in 275M cars and accounts for half of BlackBerry's revenue (Ben Cohen/Wall Street Journal)

SafetyDGX agent

Ben Cohen / Wall Street Journal: A profile of BlackBerry's QNX division, whose operating system controls safety features in 275M cars and accounts for half of BlackBerry's revenue — John Wall has spen

agreed

SafetyDGX agent

agreed The core argument of my latest article 'There is no AI race' was that the 'AI race' between the US and China is a manufactured narrative concocted by US tech companies to serve their corporate

AI will create more jobs than any other technology in history. The doomers' fundamental error isn't just the lump of labor fallacy. It's dee…

SafetyDGX agent

AI will create more jobs than any other technology in history. The doomers' fundamental error isn't just the lump of labor fallacy. It's deeper than that. They assume a finite problem space. This is t

Altman’s bait-and-switch:

SafetyDGX agent

Altman’s bait-and-switch: So, my problem with this is, Altman basically admits he's been running a confidence game for years about this singularity stuff, then pivots when it becomes inconvenient, and

An amazing time capsule of the unevenness of current AI: https://clocks.brianmoore.com

SafetyDGX agent

Brian Moore's interactive clock comparison tool demonstrates the significant performance gaps and inconsistencies across different AI systems when performing the seemingly simple task of displaying ti

“In summary, there is very little evidence for LLMs benefiting patients or doctors for health outcomes” - Dr. @EricTopol Read his full revie…

SafetyDGX agent

“In summary, there is very little evidence for LLMs benefiting patients or doctors for health outcomes” - Dr. @EricTopol Read his full review here: https://open.substack.com/pub/erictopol/p/the-parado

It’s been almost six weeks since Jensen Huang claimed we had achieved AGI. Why don’t I feel any different? https://www.forbes.com/sites/anto…

SafetyDGX agent

It’s been almost six weeks since Jensen Huang claimed we had achieved AGI. Why don’t I feel any different? https://www.forbes.com/sites/antoniopequenoiv/2026/03/23/nvidias-jensen-huang-says-he-thinks-

Joseph Gordon-Levitt says “almost all” AI systems are “built on mass theft,” arguing that companies using large language models “shouldn’t b…

SafetyDGX agent

Joseph Gordon-Levitt argues that most AI systems are fundamentally built on 'mass theft,' contending that companies deploying large language models without proper authorization or compensation for tra

Old days: “Bigshot chief scientist of major corporation can’t handle criticism of the work he hypes” New days: “Same bigshot no longer hypes…

SafetyDGX agent

Old days: “Bigshot chief scientist of major corporation can’t handle criticism of the work he hypes” New days: “Same bigshot no longer hypes LLMs; can’t handle the fact that someone else noticed flaws

Thank god, indeed, that @ylecun publicly stood up and admitted he was wrong, to fight for the greater good. Oh wait.

SafetyDGX agent

Thank god, indeed, that @ylecun publicly stood up and admitted he was wrong, to fight for the greater good. Oh wait. @GaryMarcus thank god he publicly stood up, admitted he was wrong and you were righ

that’s right.

SafetyDGX agent

that’s right. 😂 Finally, the great LeCun conversion arc is complete! “LLM agents = disaster” – said the guy who once defended Galactica like it was the Second Coming. Gary, you didn’t change… the time

Why is the AI backlash growing? Outside of coding (where there is clear value), and a handful of other domains (e.g. brainstorming), Generat…

SafetyDGX agent

Why is the AI backlash growing? Outside of coding (where there is clear value), and a handful of other domains (e.g. brainstorming), Generative AI has been a net negative for society. GenAI has been u

2 May 2026

5 months later, I would say that AI is definitely on track to produce 90% of all bullshit within a few years. Congratulations LLM Industry, …

SafetyDGX agent

5 months later, I would say that AI is definitely on track to produce 90% of all bullshit within a few years. Congratulations LLM Industry, you are changing the world! Claim AI will generate 90% of kn

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compa…

SafetyDGX agent

Are AIs about to subjugate humanity? In the debate about catastrophic AI risk, evolutionary scenarios receive far too little attention compared to largely speculative arguments about “instrumental con

Dawkins earned this; Dennett did not.

SafetyDGX agent

This post likely refers to an academic honor or recognition that Richard Dawkins received but Daniel Dennett did not, posted as a commentary by cognitive scientist Gary Marcus on X (formerly Twitter).

Fact: I could be rich if I posted a lot of bullshit about AI.

SafetyDGX agent

Gary Marcus comments on financial incentives in AI discourse, suggesting that posting sensationalized or false claims about artificial intelligence could generate significant revenue. This reflects Ma

If only Richard Dawkins had seen this.

SafetyDGX agent

If only Richard Dawkins had seen this. 1/2 Why AI is unlikely to become conscious – my 2026 @TEDTalks is now online. What do you think about the prospects for 'conscious AI'? https://www.ted.com/talks

If you read Dawkins on AI and consciouness I hope you will take the time to also watch @anilkseth’s lucid TED talk on the same topic (link b…

SafetyDGX agent

If you read Dawkins on AI and consciouness I hope you will take the time to also watch @anilkseth’s lucid TED talk on the same topic (link below). @RichardDawkins Hi @RichardDawkins - the illusion of

Indeed, it does seem strange. Far more skepticism seems to have been applied to the first idea than the second.

SafetyDGX agent

Indeed, it does seem strange. Far more skepticism seems to have been applied to the first idea than the second. @garydanteleigh @GaryMarcus I just can't wrap my head around the fact that this guy is s

omg, conscious AI, something BIG is happening, blah blah blah

SafetyDGX agent

Gary Marcus, a prominent AI researcher and skeptic, posted commentary on claims about conscious AI systems. The post likely discusses concerns about overstated assertions regarding machine consciousne

People are really enjoying our full workshops showing end to end walkthroughs of real production workflows! This is a rare double header wit…

SafetyDGX agent

People are really enjoying our full workshops showing end to end walkthroughs of real production workflows! This is a rare double header with @braintrust's Giran Moodley and @OussamaHaff walking thoug

suno can spit bars 😵‍💫 https://open.spotify.com/track/2ZqWAGhGt13WAt4krXv8bV?si=c953d55900a8455f ♫ Bias Alignment I’m deriving the varianc…

SafetyDGX agent

suno can spit bars 😵‍💫 https://open.spotify.com/track/2ZqWAGhGt13WAt4krXv8bV?si=c953d55900a8455f ♫ Bias Alignment I’m deriving the variance, bias alignment is delicate Scaling parameters, balancing gr

The Simpsons, on the trillion dollar buildout for GenAI.

SafetyDGX agent

Gary Marcus discusses concerns about the massive financial investment being directed toward generative AI development and deployment, likely examining whether the trillion-dollar spending is justified

writing a substack dissecting Richard Dawkins’ latest was not how I had planned to spend my Saturday morning 😢

SafetyDGX agent

Gary Marcus expressed frustration on X about unexpectedly spending his Saturday morning writing a Substack article critiquing Richard Dawkins' recent work or statements. The post suggests Marcus felt

1 May 2026

A High-Throughput Compute-Efficient POMDP Hide-And-Seek-Engine (HASE) for Multi-Agent Operations

SafetyDGX agent

arXiv:2604.27162v1 Announce Type: cross Abstract: Reinforcement Learning (RL) algorithms exhibit high sample complexity, particularly when applied to Decentralized Partially Observable Markov Decision

A Woman with a Knife or A Knife with a Woman? Measuring Directional Bias Amplification in Image Captions

SafetyDGX agent

arXiv:2503.07878v5 Announce Type: replace-cross Abstract: When we train models on biased datasets, they not only reproduce data biases, but can worsen them at test time - a phenomenon called bias ampl

← Previous
1…168169170171172…212
Next →