AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,570
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,566
  • Research19,194
  • Safety12,816
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
All
84,570Total entries
1Added by human
84,569Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,816 results
Safety

TUX: Measuring Human--AI Tacit Understanding

DGX agent

arXiv:2605.30930v1 Announce Type: cross Abstract: As large language models (LLMs) increasingly act as collaborative partners, human--AI alignment is often evaluated through explicit task success, accu

safetyarxiv-cs-ai
1 Jun 2026
Safety

Ubiquity of Emergent Hebbian Dynamics in Regularized Learning

DGX agent
Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters

arXiv:2505.18069v3 Announce Type: replace Abstract: Hebbian and anti-Hebbian plasticity are widely observed in the brain and are classically modeled as mechanistic, local homosynaptic rules stabilized

safetyarxiv-cs-lg
1 Jun 2026
Safety

Uncertainty-Aware and Temporally Regulated Expert Advice in Reinforcement Learning for Autonomous Driving

DGX agent

arXiv:2605.30576v1 Announce Type: new Abstract: Exploration in reinforcement learning for autonomous driving is inherently unsafe: agents must experience novel behaviors to learn, yet exploration can

safetyarxiv-cs-ai
1 Jun 2026
Safety

Unfolding Generative Flows with Koopman Operators: Trajectory-Preserving Linearization

DGX agent

arXiv:2506.22304v3 Announce Type: replace-cross Abstract: Continuous Normalizing Flows (CNFs) enable elegant generative modeling but remain bottlenecked by their iterative nature requiring costly samp

safetyarxiv-cs-cv
1 Jun 2026
Safety

UniAudio-Token: Empowering Semantic Speech Tokenizers with General Audio Perception

DGX agent

arXiv:2605.31521v1 Announce Type: new Abstract: Semantic speech tokenizers have become a widely used interface for Audio-LLMs, owing to their compact single-codebook design and strong linguistic align

safetyarxiv-cs-cl
1 Jun 2026
Safety

UniRTL: Unifying Code and Graph for Robust RTL Representation Learning

DGX agent

arXiv:2605.31040v1 Announce Type: new Abstract: Developing effective representations for register transfer level (RTL) designs is crucial for accelerating the hardware design workflow. Existing approa

safetyarxiv-cs-lg
1 Jun 2026
Safety

Unsupervised Defect Detection for Surgical Instruments

DGX agent

arXiv:2509.21561v2 Announce Type: replace Abstract: Ensuring the safety of surgical instruments requires reliable detection of visual defects. However, manual inspection is prone to error, and existin

safetyarxiv-cs-cv
1 Jun 2026
Safety

Used Car Salesbots? Honesty and Credulity of LLMs as Bargaining Agents under Partial Information

DGX agent

arXiv:2605.31445v1 Announce Type: cross Abstract: In this work we study agents in simulated bargaining scenarios, where a buyer and a seller communicate through a text channel and attempt to negotiate

safetyarxiv-cs-ai
1 Jun 2026
Safety

UXR PoV for Neuroinclusive Emotion Regulation

DGX agent

arXiv:2605.31131v1 Announce Type: cross Abstract: Attention-deficit/hyperactivity disorder (ADHD) is a psychiatric disorder which presents itself in individuals through patterns of developmentally ina

safetyarxiv-cs-ai
1 Jun 2026
Safety

Value Functions as Supermartingale Certificates

DGX agent

arXiv:2605.31524v1 Announce Type: new Abstract: Certification methods for stochastic systems provide sufficient proof rules, based on real-valued supermartingale certificates, to determine the almost-

safetyarxiv-cs-lg
1 Jun 2026
Safety

VeriGate: Verifier-Gated Step-Level Supervision for GRPO

DGX agent

arXiv:2605.30451v1 Announce Type: new Abstract: Group Relative Policy Optimization (GRPO) is an effective recipe for training reasoning models with verifier-based outcome rewards, but its supervision

safetyarxiv-cs-lg
1 Jun 2026
Safety

Vision-Language Models Suppress Female Representations Under Ambiguous Input

DGX agent

arXiv:2605.31556v1 Announce Type: cross Abstract: Alignment teaches vision-language models (VLMs) to avoid expressing demographic biases, and when gender is clearly visible they largely succeed. Far l

safetyarxiv-cs-ai
1 Jun 2026
Safety

Wall-OSS-0.5 Technical Report

DGX agent

arXiv:2605.30877v1 Announce Type: new Abstract: Large-scale Vision-Language-Action (VLA) pretraining is increasingly adopted as the foundation for robot policies, yet the evidence for pretrained VLAs

safetyarxiv-cs-ro
1 Jun 2026
Safety

@webisticsdawg @GaryMarcus Personally I think everyone w a 401(k) should be absolutely livid. I'm disgusted by this, it's so brazen. How dar…

DGX agent

@webisticsdawg @GaryMarcus Personally I think everyone w a 401(k) should be absolutely livid. I'm disgusted by this, it's so brazen. How dare the richest man in the world pick working Americans' pocke

safetygary-marcus--x
1 Jun 2026
Safety

What Am I Missing? Question-Answering as Hidden State Probing

DGX agent

arXiv:2605.31561v1 Announce Type: new Abstract: Test-time reasoning has become a significant field of study since the introduction of chain-of-thought reasoning in large language models (LLMs). Howeve

safetyarxiv-cs-cl
1 Jun 2026
Safety

What if the skeptics are right and superintelligence is impossible? Great! Then a proactive ban costs us nothing, prevents massive compute w…

DGX agent

What if the skeptics are right and superintelligence is impossible? Great! Then a proactive ban costs us nothing, prevents massive compute waste, and hurts no one. But if they are wrong? We face an un

safetyconnor-leahy--x
1 Jun 2026
Safety

When are LLMs Sufficient Policy Optimizers for Sequential RL Tasks?

DGX agent

arXiv:2605.30719v1 Announce Type: cross Abstract: We study when large language models (LLMs) can serve as effective black-box policy optimizers for reinforcement learning (RL) tasks, i.e., when can we

safetyarxiv-cs-ai
1 Jun 2026
Safety

which is NOT new; see this quote from 5 years ago. the fact that all this is still true says a lot.

DGX agent

which is NOT new; see this quote from 5 years ago. the fact that all this is still true says a lot. This was right five years ago, and still is: “Large scale pretrained models are certainly likely to

safetygary-marcus--x
1 Jun 2026
Safety

Who Gets Credit or Blame? Attributing Accountability in Modern AI Systems

DGX agent

arXiv:2506.00175v5 Announce Type: replace-cross Abstract: Modern AI systems are typically developed through multiple stages-pretraining, fine-tuning rounds, and subsequent adaptation or alignment, whe

safetyarxiv-cs-ai
1 Jun 2026
Safety

Why Linear Recurrent Memory Works in Partially Observable Reinforcement Learning

DGX agent

arXiv:2605.31261v1 Announce Type: cross Abstract: The family of linear recurrent neural networks has shown strong performance as recurrent memory units in partially observable reinforcement learning.

safetyarxiv-cs-ai
1 Jun 2026
Safety

World Action Verifier: Self-Improving World Models via Forward-Inverse Asymmetry

DGX agent

arXiv:2604.01985v2 Announce Type: replace-cross Abstract: General-purpose world models promise scalable policy evaluation, optimization, and planning, yet achieving the required level of robustness re

safetyarxiv-cs-ai
1 Jun 2026
Safety

World2Act: Latent Action Post-Training from World Model Dynamics

DGX agent

arXiv:2603.10422v2 Announce Type: replace Abstract: World Models (WMs) offer a promising mechanism for post-training Vision-Language-Action (VLA) policies by providing dynamics priors that improve gen

safetyarxiv-cs-cv
1 Jun 2026
Safety

Your Teacher Can't Help You Here: Combating Supervision Fidelity Decay in On-Policy Distillation

DGX agent

arXiv:2605.30833v1 Announce Type: cross Abstract: On-policy distillation transfers reasoning capabilities by training a student model on its own generated trajectories using token-level feedback from

safetyarxiv-cs-ai
1 Jun 2026
Safety

ZAPS-DA: Zero-Phase Action Policy Smoothing with Decoupled Actor for Continuous Control in Reinforcement Learning

DGX agent

arXiv:2605.30612v1 Announce Type: cross Abstract: Continuous control policies trained with off-policy reinforcement learning frequently exhibit high-frequency action jitter, rendering direct deploymen

safetyarxiv-cs-lg
1 Jun 2026
Safety

Zero Collapse: A Failure Mode of Policy Gradient Methods in Discontinuous Reward Environments

DGX agent

arXiv:2605.30896v1 Announce Type: new Abstract: Bidding in repeated auctions is a central challenge for reinforcement learning (RL), combining continuous control with the strategic complexities of dig

safetyarxiv-cs-lg
1 Jun 2026
Safety

Anyone remember how I said in January 2025 that AI was going to be stumble and be dubbed “too big to fail”, along with cries for bailouts? I…

DGX agent

Anyone remember how I said in January 2025 that AI was going to be stumble and be dubbed “too big to fail”, along with cries for bailouts? If that call was correct – which increasingly seems likely, a

safetygary-marcus--x
31 May 2026
Safety

further discussion here: https://open.substack.com/pub/garymarcus/p/the-pope-appears-to-understand-ai?r=8tdk6&utm_campaign=post-expanded-sha…

DGX agent

Gary Marcus discusses the Pope's understanding and perspective on artificial intelligence, examining statements or positions the religious leader has taken regarding AI technology and its implications

safetygary-marcus--x
31 May 2026
Safety

I am absolutely with @GaryMarcus and the Pope on this. (Not something you expect to say everyday). We are not creating beings. Systems don’t…

DGX agent

I am absolutely with @GaryMarcus and the Pope on this. (Not something you expect to say everyday). We are not creating beings. Systems don’t exp. grief nor hope, hold a value construct, or the ability

safetygary-marcus--x
31 May 2026
Safety

Listen, I’m a big “the index is the index” guy, but they are openly looting the coffers. This is 100% fraud.

DGX agent

Listen, I’m a big “the index is the index” guy, but they are openly looting the coffers. This is 100% fraud. Rule changes for the SpaceX SPCX IPO: Index providers waived the profitability requirement

safetygary-marcus--x
31 May 2026
Safety

@ParValue26 @Hedgeye Let’s be clear what this actually is: the administration and the world’s richest man working together to screw over ord…

DGX agent

@ParValue26 @Hedgeye Let’s be clear what this actually is: the administration and the world’s richest man working together to screw over ordinary investors to goose returns on the IPO for the wealthie

safetygary-marcus--x
31 May 2026
Safety

serious accusation. does this fit with people’s experience?

DGX agent

serious accusation. does this fit with people’s experience? Is Anthropic altering model performance to force costly upgrades? Chapter Co-Founder and CEO @CobiBGantz outlines a shift his team recently

safetygary-marcus--x
31 May 2026
Safety

SpaceX being rammed into indices with no profit requirements, seasoning, and generally looser constraints is economic terrorism. Index track…

DGX agent

SpaceX being rammed into indices with no profit requirements, seasoning, and generally looser constraints is economic terrorism. Index trackers will eat the loss when reality catches up and retail inv

safetygary-marcus--x
31 May 2026
Safety

The backlash against AI - generated content is so visceral - expect the 'human-authored' certification gain momentum, esp for fiction. https…

DGX agent

Gary Marcus argues that consumer backlash against AI-generated content will be increasingly visceral, particularly in creative fields like fiction, leading to growing demand for 'human-authored' certi

safetygary-marcus--x
31 May 2026
Safety

Three companies are about to IPO at a higher combined value than all 2,600 dot-com IPOs from 1995 to 2000 combined. – SpaceX, OpenAI, Anthro…

DGX agent

Three companies are about to IPO at a higher combined value than all 2,600 dot-com IPOs from 1995 to 2000 combined. – SpaceX, OpenAI, Anthropic (2026): ~3.75 trillion – Every dot-com IPO from 1995–200

safetygary-marcus--x
31 May 2026
Safety

Weird how the Pope seems to understand AI better than @geoffreyhinton, but I am 100% with the Pope on this. We are NOT creating beings. The …

DGX agent

Weird how the Pope seems to understand AI better than @geoffreyhinton, but I am 100% with the Pope on this. We are NOT creating beings. The Pope is right. We are creating interactive fiction that is t

safetygary-marcus--x
31 May 2026
Safety

weird the way this tweet was getting a ton of traffic and then just stopped. 🤷‍♂️

DGX agent

weird the way this tweet was getting a ton of traffic and then just stopped. 🤷‍♂️ I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthro

safetygary-marcus--x
31 May 2026
Safety

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the o…

DGX agent

AI safety can't happen behind closed doors! Super cool to see that the @AISecurityInst is releasing its evals, datasets, and models in the open on @huggingface, so researchers everywhere can scrutiniz

safetyclem-delangue--x
30 May 2026
Safety

calling someone a retard when you don’t how apostrophes work 🙄

DGX agent

This post likely critiques the irony of someone insulting another person's intelligence while making a basic grammatical error (omitting an apostrophe in 'don't'). Gary Marcus, a cognitive scientist a

safetygary-marcus--x
30 May 2026
Safety

genius reply, full of reasoned intellectual argument, from the kind of retail investor who is probably going to get burned on the SpaceX IPO…

DGX agent

genius reply, full of reasoned intellectual argument, from the kind of retail investor who is probably going to get burned on the SpaceX IPO. Going to MIT and NYU may have meant something in the past,

safetygary-marcus--x
30 May 2026
Safety

i am only blocking tesla supporters that come at me with insult rather than argument. but that’s a lot of them. and a not great sign for the…

DGX agent

Gary Marcus discusses his moderation approach on social media, noting that he blocks Tesla supporters primarily when they resort to insults rather than substantive arguments, and observes that this oc

safetygary-marcus--x
30 May 2026
Safety

I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthropic, Openai, and Googl…

DGX agent

I honestly think Elon’s best days are behind him: BYD is crushing Tesla in EVs. Waymo is crushing Tesla in AVs. Anthropic, Openai, and Google are crushing Xai on AI. The SpaceX S-1 is so ridiculous th

safetygary-marcus--x
30 May 2026
Safety

“I'm more concerned about the lack of intellectual diversity within the frontier AI commentariat/research world. This improved a lot over th…

DGX agent

“I'm more concerned about the lack of intellectual diversity within the frontier AI commentariat/research world. This improved a lot over the last two years, but we're still far from a healthy ecosyst

safetygary-marcus--x
30 May 2026
Safety

is having a four month lead a sustainable multitrillion dollar business model?

DGX agent

is having a four month lead a sustainable multitrillion dollar business model? We took another look at the capability gap between open-weight and proprietary models. Since the start of the year, open-

safetygary-marcus--x
30 May 2026
Safety

it is the single most clever move Elon ever pulled off

DGX agent

it is the single most clever move Elon ever pulled off Rule changes for the SpaceX SPCX IPO: Index providers waived the profitability requirement and cut the seasoning window from 90 days to 5. This f

safetygary-marcus--x
30 May 2026
Safety

More of my thoughts on the topic in @Corriere: https://www.corriere.it/esteri/26_maggio_29/yoshua-bengio-ai-pope-6aa95de2-33f7-4f73-8905-34b…

DGX agent

More of my thoughts on the topic in @Corriere: https://www.corriere.it/esteri/26_maggio_29/yoshua-bengio-ai-pope-6aa95de2-33f7-4f73-8905-34b1abedbxlk.shtml “Like nuclear energy, AI must be at the serv

safetyyoshua-bengio--x
30 May 2026
Safety

My feed is suddenly filled w Elon supporters who can’t understand the difference between an insult and a genuine counterargument. As always,…

DGX agent

Gary Marcus expresses frustration about an influx of Elon Musk supporters on his social media feed who conflate personal insults with substantive counterarguments. The post appears to critique a lack

safetygary-marcus--x
30 May 2026
Safety

so much for multitrillion dollar candy companies

DGX agent

so much for multitrillion dollar candy companies Note to all staff: Turns out that super tasty AI candy we've been putting out in large bowls isn't free and it isn't necessarily leading to better prod

safetygary-marcus--x
30 May 2026
Safety

this won’t end well. it will end with a bailout.

DGX agent

this won’t end well. it will end with a bailout. Cash flow no longer covers the AI capex bill, so hyperscalers are funding it with record debt: Hyperscaler bond issuance has soared to $150 billion YTD

safetygary-marcus--x
30 May 2026
← Previous
1…130131132133134…267
Next →