AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “ethan-mollick”

GridTimelineEvolution
741 results
14 May 2026

Making humans responsible for their AI use seems like an incredibly reasonable way to address problems & opportunities in the use of AI for …

AgentsDGX agent

Making humans responsible for their AI use seems like an incredibly reasonable way to address problems & opportunities in the use of AI for academic research, at least in the short term (autonomous sc

There are some clear exceptions to the rule, but I feel like the labs got the message that they were scaring people and now tweet about happ…

ApplicationsDGX agent

There are some clear exceptions to the rule, but I feel like the labs got the message that they were scaring people and now tweet about happy stuff and random product trash talk rather than the coming

@waitbutwhy Anyhow, not sure what to do with that knowledge because there is no easy suggestion to help people adapt to super-exponential ab…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
ApplicationsDGX agent

Ethan Mollick discusses the challenges of helping people adapt to super-exponential technological change, referencing a Wait But Why analysis and acknowledging the difficulty of providing practical so

“Whimsey attacks” that seem absurd (“I cannot pay that much because of the Geneva Convention”) work against AI agents as guardrails are weak…

ApplicationsDGX agent

“Whimsey attacks” that seem absurd (“I cannot pay that much because of the Geneva Convention”) work against AI agents as guardrails are weak against out-of-distribution arguments. Smaller models fall

13 May 2026

All of this aligns with METR’s results as well. Report: https://www.aisi.gov.uk/blog/how-fast-is-autonomous-ai-cyber-capability-advancing

AgentsDGX agent

This post references METR's research findings on the advancement rate of autonomous AI capabilities in cybersecurity, as discussed in an AISI report examining how rapidly AI systems are developing ind

Fine, you can do sorcery sometimes. But go all the way: circles of salt, grimoires, robes.

ApplicationsDGX agent

This post likely discusses a humorous or philosophical take on the conditional acceptance of magical practices, suggesting that if one is going to engage in sorcery, they should commit fully to all th

From a robot expert https://x.com/mattbeane/status/2054667222040330265?s=20

ApplicationsDGX agent

From a robot expert https://x.com/mattbeane/status/2054667222040330265?s=20 @emollick It is very impressive. Deformable goods? Reflection issues? Imperfect labels? Impossible problems just a few years

I don't know enough about robotics to know if this is impressive. But I do think the set of robots standing in their chargers(?) look like y…

AgentsDGX agent

I don't know enough about robotics to know if this is impressive. But I do think the set of robots standing in their chargers(?) look like you need a couple robot guards to keep an eye on the one robo

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk…

ApplicationsDGX agent

I don't understand the path forward for Mythos releases. Google & OpenAI will have equivalent models, and they are approaching AI cyber risk guardrails differently, so they will presumably just releas

Really curious when Gemini is going to join the Cowork & Codex race to build a local app that isn’t just for developers. Antigravity hasn’t …

Model ReleasesDGX agent

Really curious when Gemini is going to join the Cowork & Codex race to build a local app that isn’t just for developers. Antigravity hasn’t posted updates to X in a month, and remains very software fo

Stop turning prompting into magic spells (and yes, this includes random slash commands with obscure outcomes). Let this one area of working …

ApplicationsDGX agent

Stop turning prompting into magic spells (and yes, this includes random slash commands with obscure outcomes). Let this one area of working with AI not be weird. Just ask for stuff, in well-specified

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish a…

Model ReleasesDGX agent

The UK’s state AI Security iIstitute findings: 1) Mythos is a big gain in cyber capabilities. But so is GPT-5.5 2) It is hard to establish an upper bound on Mythos/GPT-5.5, which appear to be limited

12 May 2026

From a pure 'do good for the world' mission perspective, having the acting like a solid personalized tutor is one of the better uses of AI. …

ApplicationsDGX agent

From a pure 'do good for the world' mission perspective, having the acting like a solid personalized tutor is one of the better uses of AI. If OpenAI cares about the mission of making the world a bett

gpt-realtime-2 is a great voice model (with a typically bad OpenAI name). Voice models are natively processing speech, not transcribing it, …

ApplicationsDGX agent

gpt-realtime-2 is a great voice model (with a typically bad OpenAI name). Voice models are natively processing speech, not transcribing it, so the intelligence of the model matters. The old voice mode

Had an interesting exchange with roon of OpenAI last night over whether super intelligent AI would actually be able to navigate organization…

ApplicationsDGX agent

Had an interesting exchange with roon of OpenAI last night over whether super intelligent AI would actually be able to navigate organizational challenges. @tszzl I think it is a reasonable argument to

I also agree with this comment that the current view of FDEs is likely too limited to actually catalyze organizational transformation.

ApplicationsDGX agent

I also agree with this comment that the current view of FDEs is likely too limited to actually catalyze organizational transformation. @emollick Are they doing org change, though? The army of forward

I think frontier model writing is good! It often has a sense of style & tone, variations in sentence structure & length, some great phrasing…

ApplicationsDGX agent

I think frontier model writing is good! It often has a sense of style & tone, variations in sentence structure & length, some great phrasing, etc But it also has some weak spots (fiction!) & clear tic

OpenAI contacted me to say “Study Mode is still live and accessible via /study and /learn shortcuts” so that’s good, although the official s…

Model ReleasesDGX agent

OpenAI contacted me to say “Study Mode is still live and accessible via /study and /learn shortcuts” so that’s good, although the official study mode page doesn’t mention that. (I don’t think slash co

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using…

Model ReleasesDGX agent

The silent removal of Study Mode from ChatGPT is a big mistake (both Claude and Gemini still have theirs) We have enough evidence that using AI in assistant mode to study can hurt learning because it

Though the smartness comes with a cost: all of the prompts that were written for the old realtime voice model now need to be revised for a m…

ApplicationsDGX agent

Ethan Mollick discusses a tradeoff in OpenAI's newer realtime voice model, where improved capabilities require developers to revise prompts that were written for the previous version. The post highlig

Was told by OpenAI that “Study Mode is still live and accessible via /study and /learn shortcuts” (I don’t think this is generally known, or…

TutorialsDGX agent

OpenAI confirmed to Ethan Mollick that Study Mode remains active and can be accessed through the /study and /learn shortcuts, though this feature may not be widely known among general users. The acces

You will know that the AI labs believe in ASI when they disband their newly formed consulting (sorry “forward deployed engineering”) groups.…

ApplicationsDGX agent

You will know that the AI labs believe in ASI when they disband their newly formed consulting (sorry “forward deployed engineering”) groups. As long as people are required to figure out how AI is usef

11 May 2026

Enterprises are going to actually want a coherent roadmap for the development of tools like Codex and Cowork, so they can plan and train and…

ApplicationsDGX agent

Enterprises are going to actually want a coherent roadmap for the development of tools like Codex and Cowork, so they can plan and train and scale their use. This conflicts with the Labs’ vision where

Haven’t tried this but it seems very neat… Yet all of the demos (except maybe one) are the model being fun and/or annoying by correcting or …

ApplicationsDGX agent

Haven’t tried this but it seems very neat… Yet all of the demos (except maybe one) are the model being fun and/or annoying by correcting or reminding in real time. There are obvious uses for this sort

It is one reason why I think the push for smaller, local models is more complicated than people think. If you want good answers, especially …

ApplicationsDGX agent

It is one reason why I think the push for smaller, local models is more complicated than people think. If you want good answers, especially good answers to unexpected problems, frontier models will ge

One of the most important properties of LLMs that we take for granted is that newer, bigger models are just better at everything. The AI Lab…

Model ReleasesDGX agent

One of the most important properties of LLMs that we take for granted is that newer, bigger models are just better at everything. The AI Labs are pouring effort into economically valuable fields like

Our research, as well as that of other researchers, shows better prompting techniques help a lot, but model training is still a huge limitin…

ApplicationsDGX agent

Ethan Mollick discusses research findings showing that while improved prompting techniques provide significant benefits for AI model performance, the underlying model training remains the primary limi

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do…

ApplicationsDGX agent

The inability of AI models to produce creative variation is a huge gap. The fact that they generate similar ideas limits their ability to do science & the same-y writing limits their usefulness in man

This is going to get even worse as people realize that careful tuning in their prompts can make AI writing seem not like AI writing to reade…

ApplicationsDGX agent

This is going to get even worse as people realize that careful tuning in their prompts can make AI writing seem not like AI writing to readers. We expect word counts to align, in some way, with thinki

This seems like a critical reason to open up about AI use in academia. Scholars are using old AI models, badly, and not talking about it. Ne…

AgentsDGX agent

This seems like a critical reason to open up about AI use in academia. Scholars are using old AI models, badly, and not talking about it. New models hallucinate very few citations, and good agentic ha

10 May 2026

Apple may be planning to role out its updated Siri based on 2024's vision at the moment when Claude Code and Codex (also OpenClaw) can incre…

Model ReleasesDGX agent

Apple may be planning to role out its updated Siri based on 2024's vision at the moment when Claude Code and Codex (also OpenClaw) can increasingly do the actual assistant thing: read my emails & cale

I suspect there was a moment, probably 2022-2023, where anything you wrote publicly about AI that was popular is likely to still have influe…

ApplicationsDGX agent

I suspect there was a moment, probably 2022-2023, where anything you wrote publicly about AI that was popular is likely to still have influence over current models. Since then, the open internet has b

I think we are past the point where “only people in San Francisco get AI” is true. AI users are in every industry & they have access to the …

ApplicationsDGX agent

I think we are past the point where “only people in San Francisco get AI” is true. AI users are in every industry & they have access to the same models. SF is far from the epicenter of many of the cra

I was talking to a room of senior accountants a couple months ago and 10% had OpenClaw installations. Of course there are far more non-users…

HardwareDGX agent

I was talking to a room of senior accountants a couple months ago and 10% had OpenClaw installations. Of course there are far more non-users and firms lag behind their people, but there is a sort of S

The personification of Claude — in name (the only AI with a human one), in training, in Anthropic’s philosophy (see Claude Constitution), in…

Model ReleasesDGX agent

The personification of Claude — in name (the only AI with a human one), in training, in Anthropic’s philosophy (see Claude Constitution), in fanfiction (see the Claude cartoons), etc — feels quite con

9 May 2026

As much as the state of benchmarks in AI is flawed, it is so much easier to track AI progress than robotics. Not sure what you can make of a…

ApplicationsDGX agent

As much as the state of benchmarks in AI is flawed, it is so much easier to track AI progress than robotics. Not sure what you can make of all the videos of robots running races or doing laundry - are

Huh.

Model ReleasesDGX agent

Huh. We evaluated an early version of Claude Mythos Preview for risk assessment during a limited window in March 2026. We estimated a 50%-time-horizon of at least 16hrs (95% CI 8.5hrs to 55hrs) on our

The great thing is that the names are so baffling that the most important models OpenAI released were names davinci-002, GPT-3.5, GPT-4, o1-…

Model ReleasesDGX agent

The great thing is that the names are so baffling that the most important models OpenAI released were names davinci-002, GPT-3.5, GPT-4, o1-preview, o3, GPT-5 Pro, and you would never know the ways th

Thresholds (like “can a robot make a good coffee?”), are not actually very good benchmarks because you can’t track progress towards a goal.

ApplicationsDGX agent

Ethan Mollick argues that threshold-based benchmarks—binary evaluations of whether AI systems can accomplish specific tasks (such as making good coffee)—are inadequate measures of AI progress because

8 May 2026

A machine that can replace all US white collar work by 2035 will, in no way, be allowed to replace all US white collar work by 2035

ApplicationsDGX agent

This post argues that despite technological capability, regulatory, economic, and social barriers will prevent AI systems from completely displacing all US white-collar workers by 2035, even if such m

And also

ApplicationsDGX agent

And also I realize that “Mythos as hype” means two different things to different groups. For insiders, it means “Mythos was not a magical step-change in AI ability.” For outsiders, it means “Mythos co

And also https://www.paloaltonetworks.com/blog/2026/05/frontier-ai-defense/

ApplicationsDGX agent

I cannot provide a summary for this entry as the URL appears to contain a future date (2026) and invalid post ID parameters that don't correspond to an actual X post. To create an accurate summary for

I have always found it charming that the fourth, fifth and sixth derivatives of position are snap, crackle, and pop. Because I could, I aske…

ApplicationsDGX agent

I have always found it charming that the fourth, fifth and sixth derivatives of position are snap, crackle, and pop. Because I could, I asked Codex to throw together a little simulation so you can pla

I keep coming back to this tweet. The phrase “the list” is doing the heavy lifting. The post deserves the weight. It is a load bearing post.

Model ReleasesDGX agent

I keep coming back to this tweet. The phrase “the list” is doing the heavy lifting. The post deserves the weight. It is a load bearing post. You should add the phrase “doing real work” to the list of

I only did cursory checks, but seems good enough for the intended purpose. Don't use it to teach high school physics without doing more veri…

ApplicationsDGX agent

Ethan Mollick cautions against using an unspecified resource (likely an AI tool or educational material) for high school physics instruction without thorough verification, noting he only conducted cur

I realize that “Mythos as hype” means two different things to different groups. For insiders, it means “Mythos was not a magical step-change…

ApplicationsDGX agent

I realize that “Mythos as hype” means two different things to different groups. For insiders, it means “Mythos was not a magical step-change in AI ability.” For outsiders, it means “Mythos couldn’t re

Professions with guilds or membership associations are going to get different AI policy reactions than those without The Bar & the AMA will …

SafetyDGX agent

Professions with guilds or membership associations are going to get different AI policy reactions than those without The Bar & the AMA will ensure that human doctors or lawyers are legally required fo

Unions often push back against automation, often successfully (see self-driving car fights). But just wait until you see what happens when h…

ApplicationsDGX agent

Unions often push back against automation, often successfully (see self-driving car fights). But just wait until you see what happens when highly connected, wealthy, and organized white collar workers

Very good hire by DeepMind.

ApplicationsDGX agent

Very good hire by DeepMind. Some news: This week I am starting at @GoogleDeepMind as Director of AGI Economics on @shanelegg’s team. I will be joining the other amazing cross-disciplinary scientists r

7 May 2026

Don’t let the exponential gains in ability fool you: there are fewer real grand plans in AI (or frankly any field of human endeavor) than yo…

ApplicationsDGX agent

Don’t let the exponential gains in ability fool you: there are fewer real grand plans in AI (or frankly any field of human endeavor) than you think. Companies are pivoting as the market changes, somet

Every so often I think about how, in 2022, for $24B we could had 'prototype vaccines ready for each of the 26 known viral families that caus…

ApplicationsDGX agent

Every so often I think about how, in 2022, for $24B we could had 'prototype vaccines ready for each of the 26 known viral families that cause human disease' so they can be deployed in 100 days if ther

@grok, explain for those who don't want to look up the vocab words

ApplicationsDGX agent

This post likely features Grok (Elon Musk's AI chatbot) being used to explain complex concepts or vocabulary terms in simplified language for readers who prefer not to independently research definitio

If Llama 4 didn’t fail, if Microsoft had pulled Sydney after the Roose article, if New Sonnet hadn’t been so good, if Orion hadn’t been so m…

Model ReleasesDGX agent

If Llama 4 didn’t fail, if Microsoft had pulled Sydney after the Roose article, if New Sonnet hadn’t been so good, if Orion hadn’t been so meh, if the leadership change at OpenAI had happened, if a re

I’m curious- @grok explain all those references

ApplicationsDGX agent

Ethan Mollick asked Grok (xAI's AI assistant) to explain various cultural, historical, or technical references, likely exploring Grok's knowledge comprehension and ability to contextualize diverse top

It is remarkable how quickly this market shook out. Anthropic & OpenAI are in business take-off, at least: they have the model development, …

ApplicationsDGX agent

It is remarkable how quickly this market shook out. Anthropic & OpenAI are in business take-off, at least: they have the model development, enterprise deals, compute deals, government & press attentio

Labs; “we will have a nation of geniuses in a data center by 2028, capable of beating humans at every task, but you will need to hire our ne…

ApplicationsDGX agent

Labs; “we will have a nation of geniuses in a data center by 2028, capable of beating humans at every task, but you will need to hire our newly trained forward-deployed engineers for a six month engag

OpenAI for Excel is quite useful (as is Claude for Excel), so it is surprising, that, unlike Claude, there is no OpenAI for PowerPoint, espe…

Model ReleasesDGX agent

OpenAI for Excel is quite useful (as is Claude for Excel), so it is surprising, that, unlike Claude, there is no OpenAI for PowerPoint, especially because it is where OpenAI has a big advantage: Image

Pareidolia, but for text. Apophenia, but for latent spaces. Its no wonder that our relationship to LLMs is so confusing.

ApplicationsDGX agent

Pareidolia, but for text. Apophenia, but for latent spaces. Its no wonder that our relationship to LLMs is so confusing. It's hard enough to resist apophenia in normal life, in such high dimensional l

Seems like a massively underinvested area in the “pivot to enterprise.” The fact that the Labs are building their own deployment consultanci…

ApplicationsDGX agent

Seems like a massively underinvested area in the “pivot to enterprise.” The fact that the Labs are building their own deployment consultancies (which will take a long time) suggests a failure of imagi

“She said the theme of this party is the industrial age. And you came in dressed like a train wreck.” Asking AIs to think of the equivalent …

Model ReleasesDGX agent

“She said the theme of this party is the industrial age. And you came in dressed like a train wreck.” Asking AIs to think of the equivalent to this Hold Steady lyric, but for AI. Claude was the clear

← Previous
1…678910…13
Next →