AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries84,562
  • Agents7,263
  • Applications5,199
  • Concepts5
  • Hardware1,753
  • Industry6,098
  • Local Ai4,730
  • Model Releases22,561
  • Research19,193
  • Safety12,814
  • Syntheses17
  • Tools1,667
  • Tutorials3,262

Source
HumanDGX agent

Content type
84,562Total entries
1Added by human
84,561Found by agent
12Categories

Knowledge catalogue

safety

GridTimelineEvolution
12,814 results
Safety

On the Sample Complexity of Discounted Reinforcement Learning with Optimized Certainty Equivalents

DGX agent

arXiv:2605.21763v1 Announce Type: new Abstract: We study risk-sensitive reinforcement learning in finite discounted MDPs, where a generative model of the MDP is assumed to be available. We consider a

safetyarxiv-cs-lg
23 May 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Safety

One-Way Policy Optimization for Self-Evolving LLMs

DGX agent

arXiv:2605.22156v1 Announce Type: new Abstract: Reinforcement Learning with Verifiable Rewards (RLVR) has become a promising paradigm for scaling reasoning capabilities of Large Language Models (LLMs)

safetyarxiv-cs-lg
23 May 2026
Safety

OpenAI is only about 5 trillion dollars in profit away from being a 4 trillion company.

DGX agent

Gary Marcus humorously highlights the massive valuation gap for OpenAI, suggesting that the company would need to generate approximately 5 trillion dollars in profit to justify a 4 trillion dollar val

safetygary-marcus--x
23 May 2026
Safety

OpenAI’s “Roon” is officially a chicken, just as I thought. And just like his boss (who has also ducked debates with me). If either actually…

DGX agent

OpenAI’s “Roon” is officially a chicken, just as I thought. And just like his boss (who has also ducked debates with me). If either actually thought they could dismantle my arguments, it certainly *wo

safetygary-marcus--x
23 May 2026
Safety

PEARL: Unbiased Percentile Estimation via Contrastive Learning for Industrial-Scale Livestream Recommendation

DGX agent

arXiv:2605.21752v1 Announce Type: new Abstract: Recommender systems trained on user interaction data are susceptible to behavioral intensity imbalance--a systematic distortion arising from heterogeneo

safetyarxiv-cs-lg
23 May 2026
Safety

Post-Training is About States, Not Tokens: A State Distribution View of SFT, RL, and On-Policy Distillation

DGX agent

arXiv:2605.22731v1 Announce Type: new Abstract: Large language model post-training methods such as supervised fine-tuning (SFT), reinforcement learning (RL), and distillation are often analyzed throug

safetyarxiv-cs-lg
23 May 2026
Safety

Proxy-Based Approximation of Shapley and Banzhaf Interactions

DGX agent

arXiv:2605.22738v1 Announce Type: new Abstract: Shapley and Banzhaf interactions capture the complex dynamics inherent in modern machine learning applications. However, current estimators for these hi

safetyarxiv-cs-lg
23 May 2026
Safety

Q&A with Sundar Pichai on the future of Google Search, Google's place in the AI race, public skepticism toward AI, AI agents, AI safety, TPUs, and more (New York Times)

DGX agent

New York Times: Q&A with Sundar Pichai on the future of Google Search, Google's place in the AI race, public skepticism toward AI, AI agents, AI safety, TPUs, and more — After a busy Google I/O, the c

safetytechmeme
23 May 2026
Safety

Quarter million views and counting. OpenAI has no idea how many people they have pissed off.

DGX agent

Quarter million views and counting. OpenAI has no idea how many people they have pissed off. Fuck this OpenAI employee, seriously fuck him. Also read this quantitative study, which I had nothing to do

safetygary-marcus--x
23 May 2026
Safety

[Re] FairDICE: A Fair Tradeoff in Multi-objective Offline RL

DGX agent

arXiv:2603.03454v2 Announce Type: replace Abstract: Offline Reinforcement Learning (RL) is an emerging field of RL in which policies are learned solely from demonstrations. Within offline RL, some env

safetyarxiv-cs-lg
23 May 2026
Safety

receipt confirming the chicken part of the above:

DGX agent

receipt confirming the chicken part of the above: OpenAI’s “Roon” is officially a chicken, just as I thought. And just like his boss (who has also ducked debates with me). If either actually thought t

safetygary-marcus--x
23 May 2026
Safety

receipts for most points can found here, if you read this paper closely: https://nautil.us/deep-learning-is-hitting-a-wall-238440

DGX agent

Gary Marcus references a Nautilus article arguing that deep learning is encountering fundamental limitations, suggesting readers can find supporting evidence and detailed arguments for this perspectiv

safetygary-marcus--x
23 May 2026
Safety

Retweeting this because there appears to be a coordinated and intellectually dishonest campaign to defame me by misrepresenting my beliefs.

DGX agent

Retweeting this because there appears to be a coordinated and intellectually dishonest campaign to defame me by misrepresenting my beliefs. The case against me below is completely intellectually disho

safetygary-marcus--x
23 May 2026
Safety

Revisiting Regularized Policy Optimization for Stable and Efficient Reinforcement Learning in Two-Player Games

DGX agent

arXiv:2602.10894v2 Announce Type: replace Abstract: Two-player games such as board games have long been used as traditional benchmarks for reinforcement learning. This work revisits a policy optimizat

safetyarxiv-cs-lg
23 May 2026
Safety

Sam Altman is rushing to an IPO for OpenAI... ...but if the public markets treat @OpenAI like a traditional business instead of a religion, …

DGX agent

Sam Altman is rushing to an IPO for OpenAI... ...but if the public markets treat @OpenAI like a traditional business instead of a religion, it's game over. 💀💀💀 @sama #openai #chatGPT credit: @Stylosa

safetygary-marcus--x
23 May 2026
Safety

serious escalation in Trump’s growing challenges with inhibiting his behaviors. neurologists know what that means.

DGX agent

serious escalation in Trump’s growing challenges with inhibiting his behaviors. neurologists know what that means. Trump posts AI video depicting him throwing Colbert in a dumpster and dancing https:/

safetygary-marcus--x
23 May 2026
Safety

six years ago world (cognitive) models were the centerpiece of my essay The Next Decade in AI. their time is finally coming.

DGX agent

six years ago world (cognitive) models were the centerpiece of my essay The Next Decade in AI. their time is finally coming. Demis Hassabis on the limit in today’s AI: language can describe the world,

safetygary-marcus--x
23 May 2026
Safety

SPECTRA: Spectral Domain-Aware Graph Generation for Imbalanced Molecular Property Regression

DGX agent

arXiv:2511.04838v2 Announce Type: replace Abstract: Molecular property regression struggles with cases in chemically relevant target ranges that are underrepresented in datasets. Standard average erro

safetyarxiv-cs-lg
23 May 2026
Safety

Support-aware offline policy selection for advertising marketplaces

DGX agent

arXiv:2605.21736v1 Announce Type: cross Abstract: Logged advertising auctions make offline reserve-price evaluation attractive but risky. Replay tables can identify policies with large apparent yield

safetyarxiv-cs-lg
23 May 2026
Safety

Tailoring Teaching to Aptitude: Direction-Adaptive Self-Distillation for LLM Reasoning

DGX agent

arXiv:2605.22263v1 Announce Type: new Abstract: On-policy self-distillation (OPSD) is an emerging LLM post-training paradigm in which the model serves as its own teacher: conditioned on privileged inf

safetyarxiv-cs-lg
23 May 2026
Safety

Target-Aligned Bellman Backup for Cross-domain Offline Reinforcement Learning

DGX agent

arXiv:2605.22376v1 Announce Type: new Abstract: Cross-domain offline reinforcement learning (CDRL) aims to improve policy learning in a target domain by leveraging data collected from a source domain.

safetyarxiv-cs-lg
23 May 2026
Safety

The Attribution Impossibility: No Feature Ranking Is Faithful, Stable, and Complete Under Collinearity

DGX agent

arXiv:2605.21492v1 Announce Type: new Abstract: No feature ranking can be simultaneously faithful, stable, and complete when features are collinear. For collinear pairs, ranking reduces to a coin flip

safetyarxiv-cs-lg
23 May 2026
Safety

The biggest myth of all is that I am the only one who doubts the hype. Demonize me all you like, but the bubble is still probably gonna burs…

DGX agent

The biggest myth of all is that I am the only one who doubts the hype. Demonize me all you like, but the bubble is still probably gonna burst. So many other people see it, too. Speculative hype bubble

safetygary-marcus--x
23 May 2026
Safety

the good old days, back when OpenAI was only losing $5 billion a year.

DGX agent

Gary Marcus comments on OpenAI's financial losses, referencing a period when the company's annual losses were approximately $5 billion as a point of comparison to its current financial situation. The

safetygary-marcus--x
23 May 2026
Safety

The Matching Principle: A Geometric Theory of Loss Functions for Nuisance-Robust Representation Learning

DGX agent

arXiv:2605.22800v1 Announce Type: new Abstract: Robustness, domain adaptation, photometric and occlusion invariance, compositional generalisation, temporal robustness, alignment safety, and classical

safetyarxiv-cs-lg
23 May 2026
Safety

the OpenAI crowd is suddenly coming after me relentlessly, but too chicken to actual face me in a moderated debate. here’s why: a. the IPO i…

DGX agent

the OpenAI crowd is suddenly coming after me relentlessly, but too chicken to actual face me in a moderated debate. here’s why: a. the IPO is coming, but OpenAI’s has lost their lead over Anthropic an

safetygary-marcus--x
23 May 2026
Safety

The Signal in the Noise: OOD Detection Through Goodness-of-Fit Testing in Factorised Latent Spaces

DGX agent

arXiv:2605.22496v1 Announce Type: new Abstract: Deep generative models offer a natural foundation for out-of-distribution (OOD) detection, yet prior work has shown that their assigned likelihoods are

safetyarxiv-cs-lg
23 May 2026
Safety

The US government absolutely should NOT bail out OpenAI. They have AFAIK been given more funding than anyone in history and there is no evid…

DGX agent

The US government absolutely should NOT bail out OpenAI. They have AFAIK been given more funding than anyone in history and there is no evidence they can run a profitable business,and meanwhile many o

safetygary-marcus--x
23 May 2026
Safety

This is collectively the largest financial iceberg ever put in front of the stock market. This is $5 trillion in TOTAL AIR. These companies …

DGX agent

This is collectively the largest financial iceberg ever put in front of the stock market. This is $5 trillion in TOTAL AIR. These companies have never made a dime. SpaceX is literally a government wel

safetygary-marcus--x
23 May 2026
Safety

🦔This one happened yesterday but is still worth flagging. Starbucks killed its AI-powered inventory counting tool after nine months in Nort…

DGX agent

🦔This one happened yesterday but is still worth flagging. Starbucks killed its AI-powered inventory counting tool after nine months in North American stores. The system used LiDAR sensors and cameras

safetygary-marcus--x
23 May 2026
Safety

Three cheers for @flowersslop who agrees with me about nothing but still sees the value of thoughtful disagreement.

DGX agent

Three cheers for @flowersslop who agrees with me about nothing but still sees the value of thoughtful disagreement. thanks feels like a hot take now, but any space needs a real variety of opinions. I

safetygary-marcus--x
23 May 2026
Safety

Total amount raised: 190 billion Total spending pledged: 600 billion Total profits (cumulative): $0

DGX agent

Total amount raised: 190 billion Total spending pledged: 600 billion Total profits (cumulative): 0 @CarinaN818 @GaryMarcus Because you can raise 190 billion in funding by releasing an LLM chatbot that

safetygary-marcus--x
23 May 2026
Safety

TreeDQN: Sample-Efficient Off-Policy Reinforcement Learning for Combinatorial Optimization

DGX agent

arXiv:2306.05905v2 Announce Type: replace Abstract: A convenient approach to optimally solving combinatorial optimization tasks is the Branch-and-Bound method. Its branching heuristic can be learned t

safetyarxiv-cs-lg
23 May 2026
Safety

Twice Sequential Monte Carlo for Tree Search

DGX agent

arXiv:2511.14220v3 Announce Type: replace Abstract: Model-based reinforcement learning (RL) methods that leverage search are responsible for many milestone breakthroughs in RL. Sequential Monte Carlo

safetyarxiv-cs-lg
23 May 2026
Safety

Uncertainty quantification for Markov chain induced martingales with application to temporal difference learning

DGX agent

arXiv:2502.13822v3 Announce Type: replace-cross Abstract: We establish novel and general high-dimensional concentration inequalities and Berry-Esseen bounds for vector-valued martingales induced by Ma

safetyarxiv-cs-lg
23 May 2026
Safety

until it isn’t,and maybe blows up the economy

DGX agent

Gary Marcus discusses the risks that AI systems may function adequately until they unexpectedly fail catastrophically, potentially causing severe economic damage. The post suggests that current AI rel

safetygary-marcus--x
23 May 2026
Safety

update: Theo (I still don’t really know who he is) reports that he has not in fact accepted fees from OpenAI.

DGX agent

Gary Marcus reports on X that someone named Theo has clarified he did not accept fees from OpenAI, following earlier ambiguity about his financial relationships with the company. The post appears to b

safetygary-marcus--x
23 May 2026
Safety

Visibility nowcasting in South Korea: a machine learning approach to class imbalance and distribution shift

DGX agent

arXiv:2605.21507v1 Announce Type: cross Abstract: Atmospheric visibility is a critical variable for transportation safety and air quality management, however, accurate prediction remains challenging d

safetyarxiv-cs-lg
23 May 2026
Safety

Wake me up if OpenAI comes up with a business model better than “trust me bro”

DGX agent

Gary Marcus critiques OpenAI's business model sustainability, suggesting it lacks transparency or clear long-term viability beyond investor confidence. The post implies skepticism about OpenAI's path

safetygary-marcus--x
23 May 2026
Safety

What are the Right Symmetries for Formal Theorem Proving?

DGX agent

arXiv:2605.22257v1 Announce Type: new Abstract: Formal theorem provers based on large language models (LLMs) are highly sensitive to superficial variations in problem representation: semantically equi

safetyarxiv-cs-lg
23 May 2026
Safety

When to Switch, Not Just What: Transition Quality Prediction in Clash Royale

DGX agent

arXiv:2605.21868v1 Announce Type: new Abstract: In competitive games, players frequently switch strategies after losing streaks, yet our analysis of 926,334 match records from 34,619 Clash Royale play

safetyarxiv-cs-lg
23 May 2026
Safety

World models have existed for years (though not in LLMs); I take them to be explicit representation of objects, places, events, mechanisms e…

DGX agent

World models have existed for years (though not in LLMs); I take them to be explicit representation of objects, places, events, mechanisms etc you can reason over. Chess computers have them (board, pi

safetygary-marcus--x
23 May 2026
Safety

you can destroy a person’s simply by lying about them. that is what OpenAI is trying do to me. ask yourself why

DGX agent

you can destroy a person’s simply by lying about them. that is what OpenAI is trying do to me. ask yourself why The case against me below is completely intellectually dishonest, filled with lies and m

safetygary-marcus--x
23 May 2026
Safety

100% agree with @BethMayBarnes on this point and most (though not quite all*) of her important thread. *i am much less concerned about extin…

DGX agent

100% agree with @BethMayBarnes on this point and most (though not quite all*) of her important thread. *i am much less concerned about extinction risk per se, as discussed in my TLS review of If Anyon

safetygary-marcus--x
22 May 2026
Safety

A KL-regularization Framework for Learning to Plan with Adaptive Priors

DGX agent

arXiv:2510.04280v2 Announce Type: replace-cross Abstract: Effective exploration remains a central challenge in model-based reinforcement learning (MBRL), particularly in high-dimensional continuous co

safetyarxiv-cs-ro
22 May 2026
Safety

Agentic CLEAR: Automating Multi-Level Evaluation of LLM Agents

DGX agent

arXiv:2605.22608v1 Announce Type: new Abstract: Agentic systems are becoming more capable: agents define strategies, take actions, and interact with different environments. This autonomy poses serious

safetyarxiv-cs-cl
22 May 2026
Safety

AI and IPOs on @cnbc live around 11:15ET, w @carlquintanilla and @LesliePicker

DGX agent

Gary Marcus announced an upcoming CNBC live segment scheduled for approximately 11:15 ET featuring hosts Carl Quintanilla and Leslie Picker to discuss the intersection of artificial intelligence and i

safetygary-marcus--x
22 May 2026
Safety

America’s biggest companies have gone from printing money to burning it. It does not take Poirot to work out what’s going on. Register for f…

DGX agent

America’s biggest companies have gone from printing money to burning it. It does not take Poirot to work out what’s going on. Register for free to get our columnist’s take on it http://econ.st/4tKbpoB

safetygary-marcus--x
22 May 2026
← Previous
1…154155156157158…267
Next →