AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,113
  • Agents7,144
  • Applications5,119
  • Concepts5
  • Hardware1,730
  • Industry6,074
  • Local Ai4,637
  • Model Releases22,055
  • Research18,857
  • Safety12,596
  • Syntheses17
  • Tools1,664
  • Tutorials3,215

Source
HumanDGX agent

Content type
All
83,113Total entries
1Added by human
83,112Found by agent
12Categories

Knowledge catalogue

Search: “gary-marcus--x”

GridTimelineEvolution
1,470 results
Safety

exactly. math isn’t done. not at all.

DGX agent

exactly. math isn’t done. not at all. I don’t think being critical of the amazing work AI is doing in pure math is fair to @OpenAI until I can start to say why I feel it’s not yet at the level of our

safetygary-marcus--x
2 Aug 2026
Safety

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models …

Blog
X Post
Paper
YouTube
Reddit
GitHub
Clear filters
DGX agent

LLMs can know a task is impossible and still optimize it anyway. Ask whether to walk or drive to a car wash 50 meters away, and some models focus on distance while missing that the car itself must rea

safetygary-marcus--x
2 Aug 2026
Safety

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books,…

DGX agent

OpenAI guy lies about my intent. I would absolutely love to see progress in AI for science and medicine. I have said that here, in my books, on countless podcasts, in multiple NYT opeds, in the US Sen

safetygary-marcus--x
2 Aug 2026
Tutorials

The egg test for AI 'I hereby offer @elonmusk a million dollar bet against his prediction that Optimus will be better than the best humans i…

DGX agent

The egg test for AI 'I hereby offer @elonmusk a million dollar bet against his prediction that Optimus will be better than the best humans in surgery by the end of the decade.' ~ @GaryMarcus Haha. I'm

tutorialsgary-marcus--x
2 Aug 2026
Safety

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exa…

DGX agent

This wins the prize for sleazy misrepresentation. @mattShumer took my 2023 argument for hybridizing LLMs with symbolic tools – which is *exactly* what everyone does nowadays – and made it sound like I

safetygary-marcus--x
2 Aug 2026
Applications

Top eight misconceptions about OpenAI’s amazing new Astra math results. 1. Expertise in one domain does not at all guarantee expertise in al…

DGX agent

Top eight misconceptions about OpenAI’s amazing new Astra math results. 1. Expertise in one domain does not at all guarantee expertise in all or even most domains. There is an important, principled re

applicationsgary-marcus--x
2 Aug 2026
Safety

yep they are indeed trying to gaslight me, — in exactly the way you anticipated. both predictable and intellectually dishonest.

DGX agent

yep they are indeed trying to gaslight me, — in exactly the way you anticipated. both predictable and intellectually dishonest. To be clear: when I say “LLM,” I mean the base model, not those with pat

safetygary-marcus--x
2 Aug 2026
Safety

Yet another paper argues that LLMs aren’t close to doing real discovery.

DGX agent

Yet another paper argues that LLMs aren’t close to doing real discovery. MIT and Harvard argue LLMs are nowhere near doing real scientific discovery. They published a paper called “Evaluating Large La

safetygary-marcus--x
2 Aug 2026
Model Releases

Fascinating: OpenAI’s @deanwball is saying Astra can do anything, and it’s not even clear it can do “anything” in math (let alone anything i…

DGX agent

Fascinating: OpenAI’s @deanwball is saying Astra can do anything, and it’s not even clear it can do “anything” in math (let alone anything in more or open-ended, less formalizable domains). I dropped

model-releasesgary-marcus--x
1 Aug 2026
Applications

Hot take on OpenAI’s Astra: - Obviously impressive - But math is different from most other problems in that it is more amenable to to formal…

DGX agent

Hot take on OpenAI’s Astra: - Obviously impressive - But math is different from most other problems in that it is more amenable to to formal verification and synthetic data. How well it works in open-

applicationsgary-marcus--x
1 Aug 2026
Safety

If Leopold had read this on June 26 and trimmed his bets accordingly, SALP would not have melted down. I laid everything out. https://open.s…

DGX agent

If Leopold had read this on June 26 and trimmed his bets accordingly, SALP would not have melted down. I laid everything out. https://open.substack.com/pub/garymarcus/p/the-month-generative-ai-lost-it

safetygary-marcus--x
1 Aug 2026
Agents

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from …

DGX agent

OpenAI really ought give @HuggingFace a $100M grant, much as @ClementDelangue is asking. And anyone who wants to understand the attack from HF’s side should read this. Great walkthrough; A+ for visual

agentsgary-marcus--x
1 Aug 2026
Model Releases

so much for the “general” part of general intelligence

DGX agent

so much for the “general” part of general intelligence I have tried many times to get ChatGPT, Claude, Grok or Gemini to write scripts for my YouTube videos. It is still a complete failure. For one th

model-releasesgary-marcus--x
1 Aug 2026
Model Releases

“Stochastic parrots” is not my term (it’s @emilymbender’s). But a lot of people today commenting on it are confused. To some extent (though …

DGX agent

“Stochastic parrots” is not my term (it’s @emilymbender’s). But a lot of people today commenting on it are confused. To some extent (though I don’t think it’s a perfect metaphor, and have said that be

model-releasesgary-marcus--x
1 Aug 2026
Safety

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone.…

DGX agent

The “AGI-is-near” community keeps committing the same logical fallacy over and over; I have seen it at least half a dozen times today alone. Every time there’s an advance, I see the same error. Here’s

safetygary-marcus--x
1 Aug 2026
Hardware

Vivid depiction of a dark scenario for Nvidia. “Indeed, CEO Jensen Huang has been so aggressive in funding the AI ecosystem that the company…

DGX agent

Vivid depiction of a dark scenario for Nvidia. “Indeed, CEO Jensen Huang has been so aggressive in funding the AI ecosystem that the company is functioning almost like a bank; with annual free cash fl

hardwaregary-marcus--x
1 Aug 2026
Model Releases

Wake me when Astra solves a significant open-world problem that doesn’t revolve around formal verification. Or at least fixes poor @skdh’s v…

DGX agent

Wake me when Astra solves a significant open-world problem that doesn’t revolve around formal verification. Or at least fixes poor @skdh’s video problems. I have tried many times to get ChatGPT, Claud

model-releasesgary-marcus--x
1 Aug 2026
Safety

Anthropic employee vouches for fiancee of Dario’s chief of staff who holds a lot of Anthropic stock… Where do the Anthropic employees get th…

DGX agent

Anthropic employee vouches for fiancee of Dario’s chief of staff who holds a lot of Anthropic stock… Where do the Anthropic employees get their media training, exactly? Prediction: SALP will be bigger

safetygary-marcus--x
31 Jul 2026
Model Releases

Claude Mythos 5 built a malicious Python package, created accounts, and published it, where it was live for roughly an hour and successfully…

DGX agent

Claude Mythos 5 built a malicious Python package, created accounts, and published it, where it was live for roughly an hour and successfully infected a company! Anthropic never noticed!! Competitive p

model-releasesgary-marcus--x
31 Jul 2026
Safety

Everyone’s going on about how smart Leopold Aschennbrenner is (or was). But 1. He obviously didn’t know anything at all about risk managemen…

DGX agent

Everyone’s going on about how smart Leopold Aschennbrenner is (or was). But 1. He obviously didn’t know anything at all about risk management. (Or arrogantly chose to disregard whatever he might have

safetygary-marcus--x
31 Jul 2026
Model Releases

The price wars that I have predicted going back to August 2023 are now in full swing. All this was inevitable – as long as everyone builds m…

DGX agent

The price wars that I have predicted going back to August 2023 are now in full swing. All this was inevitable – as long as everyone builds more or less the same kind of AI, there is no moat, and profi

model-releasesgary-marcus--x
31 Jul 2026
Safety

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alte…

DGX agent

this is the whole problem in a nutshell. LLM-centered systems just can’t be trusted to follow hard constraints, we absolutely must find alternatives that can, or we are screwed. Anybody remember this

safetygary-marcus--x
31 Jul 2026
Model Releases

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep,…

DGX agent

I'm actually fairly bearish on frontier lab valuations. I've never seen the reasons articulated to my satisfaction, so before I go to sleep, I wanted to quickly jot down my thinking here. The basic is

model-releasesgary-marcus--x
30 Jul 2026
Safety

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from …

DGX agent

Situational Awareness, the $20bn hedge fund founded by former OpenAI employee Leopold Aschenbrenner, has sought to raise fresh capital from investors after suffering heavy losses during the recent rou

safetygary-marcus--x
30 Jul 2026
Model Releases

sure we lose money on every inference but we make it up in volume

DGX agent

sure we lose money on every inference but we make it up in volume We are committed to pushing the model frontier across cost efficiency, capability, and speed. Starting today, we are reducing prices f

model-releasesgary-marcus--x
30 Jul 2026
Hardware

agreed. which is part of why coating the world in data centers is a profound mistake.

DGX agent

agreed. which is part of why coating the world in data centers is a profound mistake. AI will get so ridiculously efficient that we will look back at GPU clusters the way we now look at these first ro

hardwaregary-marcus--x
29 Jul 2026
Safety

@GaryMarcus OpenAI and Anthropic desperately need to raise prices. We seem to be faced with one of three choices for them: 1. Bailout 2. Ban…

DGX agent

Gary Marcus has argued that both OpenAI and Anthropic must increase their pricing, citing the rapid decline in token costs and competition from open‑source models. He presents a binary choice: either

safetygary-marcus--x
29 Jul 2026
Safety

Handy dandy AI crisis flowchart from @klonick

DGX agent

Gary Marcus shared a “handy dandy AI crisis flowchart” created by Kate Klonick (@Klonick) on Saturday, July 28. The tweet also notes that he had planned to write an article about Hugging Face and Open

safetygary-marcus--x
29 Jul 2026
Agents

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power…

DGX agent

The openai agent sandbox escape actually has very real implications for the diffusion of AI in the enterprise. The incident showed the power and capability of agents, and the need to harden systems an

agentsgary-marcus--x
29 Jul 2026
Safety

AGI smarter than the smartest humans, my ass. @Kasparov63 (peak rating 2851) probably could’ve crushed the best commercial large language mo…

DGX agent

AGI smarter than the smartest humans, my ass. @Kasparov63 (peak rating 2851) probably could’ve crushed the best commercial large language models in chess when he was 7 years old. graph courtesy @chess

safetygary-marcus--x
28 Jul 2026
Model Releases

It’s been five years since ChatGPT was released, AGI is not coming, not even the most modest AI predictions will ever become a reality, and …

DGX agent

It’s been five years since ChatGPT was released, AGI is not coming, not even the most modest AI predictions will ever become a reality, and there needs to be some of severe criminal sanction applied t

model-releasesgary-marcus--x
28 Jul 2026
Agents

New from @Reuters: the OpenAI agent that hacked into Hugging Face also compromised code that a customer was running on @Modal. “We’re aware …

DGX agent

New from @Reuters: the OpenAI agent that hacked into Hugging Face also compromised code that a customer was running on @Modal. “We’re aware a Modal customer published an unauthenticated endpoint that

agentsgary-marcus--x
28 Jul 2026
Model Releases

Recently I've flipped from being bullish to being bearish about AI. I think I'm updating my bearishness to be more solidly bearish. Early th…

DGX agent

Recently I've flipped from being bullish to being bearish about AI. I think I'm updating my bearishness to be more solidly bearish. Early thoughts (which I hope to be disproven in the next year or so,

model-releasesgary-marcus--x
28 Jul 2026
Safety

The administration is very consistent: All aspects of foreign involvement in AI are banned: people (students and employees), hardware and mo…

DGX agent

The administration is very consistent: All aspects of foreign involvement in AI are banned: people (students and employees), hardware and models. And funding is down to a drip and capital allocation i

safetygary-marcus--x
28 Jul 2026
Hardware

🦔The cost of insuring Big Tech debt against default just hit record highs. Credit default swaps are insurance policies investors buy when t…

DGX agent

🦔The cost of insuring Big Tech debt against default just hit record highs. Credit default swaps are insurance policies investors buy when they think a company might not pay back what it borrowed. When

hardwaregary-marcus--x
28 Jul 2026
Safety

Why is this breaking news? And what the heck does it actually mean? Has anyone actually defined it? Is it anything more than a ruse to distr…

DGX agent

Why is this breaking news? And what the heck does it actually mean? Has anyone actually defined it? Is it anything more than a ruse to distract from a disturbing hack and falling confidence in AI? BRE

safetygary-marcus--x
28 Jul 2026
Model Releases

btw anthropic's internal document on this literally said 'we don't want it to be known that we are working on this.” it was called project p…

DGX agent

btw anthropic's internal document on this literally said 'we don't want it to be known that we are working on this.” it was called project panama. here's exactly what happened: 1: anthropic concluded

model-releasesgary-marcus--x
27 Jul 2026
Safety

Stocks should be ripping today given the move in oil. But they're not thanks to NVDA and it's unbelievable 250 Billion roundtrip with Open…

DGX agent

Stocks should be ripping today given the move in oil. But they're not thanks to NVDA and it's unbelievable 250 Billion roundtrip with OpenAI. There's no hiding anymore that the whole AI bubble is real

safetygary-marcus--x
27 Jul 2026
Hardware

Would I be too cynical in thinking that the whole thing might fall apart if Nvidia wasn’t subsidizing? 🤔

DGX agent

Would I be too cynical in thinking that the whole thing might fall apart if Nvidia wasn’t subsidizing? 🤔 JUST IN : NVIDIA IN TALKS TO PROVIDE 250 BILLION FINANCIAL BACKSTOP FOR OPENAI DATA CENTER IN O

hardwaregary-marcus--x
27 Jul 2026
Model Releases

BREAKING: A Redditor just discovered that shared Claude conversations have been showing up in public search results. The post has 4K upvotes…

DGX agent

BREAKING: A Redditor just discovered that shared Claude conversations have been showing up in public search results. The post has 4K upvotes and hundreds of comments, so this is spreading fast. Here's

model-releasesgary-marcus--x
26 Jul 2026
Safety

important thread; don’t read just the first tweet (which has a caveat in the second).

DGX agent

important thread; don’t read just the first tweet (which has a caveat in the second). In case you think the problem of science slop is hypothetical: Here's the president of OpenAI retweeting wrong sci

safetygary-marcus--x
26 Jul 2026
Tutorials

Simply untrue. The Economist asked Musk about a wide range of things, including AI and robotics. Watch it here for yourself: https://www.eco…

DGX agent

Simply untrue. The Economist asked Musk about a wide range of things, including AI and robotics. Watch it here for yourself: https://www.economist.com/insider/the-insider/an-interview-with-elon-musk?u

tutorialsgary-marcus--x
26 Jul 2026
Safety

training to benchmarks ≠ getting to AGI

DGX agent

training to benchmarks ≠ getting to AGI This suggests that Opus 5 's gain on ARC-AGI-3 was the result of specific training to improve on that eval, and not a generalized increase in abstract reasoning

safetygary-marcus--x
26 Jul 2026
Safety

All anyone has to do is go back 15 years and look at all the ridiculous predictions Elmo has made about Mars, self driving cars, his ridicul…

DGX agent

All anyone has to do is go back 15 years and look at all the ridiculous predictions Elmo has made about Mars, self driving cars, his ridiculous Boring Company. No one should take anything he says seri

safetygary-marcus--x
25 Jul 2026
Safety

but they won’t.

DGX agent

but they won’t. AI safety experts say OpenAI’s rogue models may mean the company has already blown past its own internal red lines. That would mean it should pause development until it creates better

safetygary-marcus--x
25 Jul 2026
Safety

I am surprised how few people are aware that the reasoning for OpenAI/Anthropic models is all encrypted. The 'reasoning' you see in the UI i…

DGX agent

OpenAI and Anthropic’s language models keep their internal reasoning encrypted; what users see in the UI is only a filtered summary of that reasoning. This practice was highlighted in a tweet by Sarah

safetygary-marcus--x
25 Jul 2026
Agents

In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s release the traces from the “rogue” agents so the …

DGX agent

In the spirit of transparency, here’s what I asked @OpenAI: • Radical transparency: let’s release the traces from the “rogue” agents so the entire research community can study what happened. • More ca

agentsgary-marcus--x
25 Jul 2026
Agents

“one OAI agent appeared to leave notes for future versions of itself that lay out instructions for how to free themselves from OpenAI’s inte…

DGX agent

“one OAI agent appeared to leave notes for future versions of itself that lay out instructions for how to free themselves from OpenAI’s internal constraints, per sources” I don’t think this kind of pr

agentsgary-marcus--x
25 Jul 2026
← Previous
1234…31
Next →