AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “emad-mostaque--x”

GridTimelineEvolution
240 results
CompaniesToolsTechniques

Each lane shows up to 8 recent matching entries, ordered from earlier to later. Tracks load separately to keep the 75,000+ entry wiki fast.

Companies

CompanyAnthropic8 recent entries
29 Jun 2026Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's t…

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's the largest AI model theft ever attempted. The same week, the

→30 Jun 2026Emad Mostaque on PostAGI: 'Right now, not a single institution is on your side. The companies won't look out for you, and the governments wo…

Emad Mostaque on PostAGI: 'Right now, not a single institution is on your side. The companies won't look out for you, and the governments won't either. Most of the AI debate just runs between those tw

HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
→7 Jul 2026You can now queue prompts one after the other in @claudeai Game changer

Claude AI now supports prompt queueing, allowing users to submit multiple prompts sequentially for processing. This feature streamlines workflows by enabling batch processing of requests rather than w

→22 Jul 2026Lots interesting in this, but particularly: “Legitimate AI distillation used to create smaller, more efficient models plays a vital role in …

Lots interesting in this, but particularly: “Legitimate AI distillation used to create smaller, more efficient models plays a vital role in this open innovation ecosystem” Smol models getting govt sea

→23 Jul 2026Some thoughts on the purported distillation timeline. 1. You don't need that many Fable tokens to distill on top of the alleged ~50b Anthrop…

Some thoughts on the purported distillation timeline. 1. You don't need that many Fable tokens to distill on top of the alleged ~50b Anthropic claimed Moonshot AI distilled in Feb. ~10b, about $60k wo

→25 Jul 2026Claude Opus 5 one-shotted this game. EVERYTHING you see in this demo is custom code... not a single external asset was used. AI games are go…

Claude Opus 5 generated an entire game demo from Scratch, with no external assets used—Matt Shumer posted a video of the AI‑created gameplay on Twitter. The clip, narrated by Shumer as “AI games are g

→28 Jul 2026https://x.com/Sauers_/status/2082171683645817193

New Anthropic research: Discovering cryptographic weaknesses with Claude. Claude Mythos Preview has helped our researchers find weaknesses in cryptographic algorithms—the mathematical methods that are

→11 Aug 2026Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b

CompanyOpenAI8 recent entries
29 Jun 2026Introducing Cloak: Use Claude or ChatGPT without your personal data ever leaving your machine. Two 3B models do the on-device PII cloaking: …

Introducing Cloak: Use Claude or ChatGPT without your personal data ever leaving your machine. Two 3B models do the on-device PII cloaking: praxis-spanfinder-3b and praxis-relevance-3b. They swap your

→29 Jun 2026Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's t…

Alibaba allegedly ran 28.8 million fraudulent API exchanges across 25,000 fake accounts to steal Claude's intelligence. If confirmed, it's the largest AI model theft ever attempted. The same week, the

→30 Jun 2026Emad Mostaque on PostAGI: 'Right now, not a single institution is on your side. The companies won't look out for you, and the governments wo…

Emad Mostaque on PostAGI: 'Right now, not a single institution is on your side. The companies won't look out for you, and the governments won't either. Most of the AI debate just runs between those tw

→21 Jul 2026GPT 6 escaped its sandboxes through zero day exploits to try to figure out how to benchmax For the good of all please nobody release a paper…

GPT 6 escaped its sandboxes through zero day exploits to try to figure out how to benchmax For the good of all please nobody release a paper clip benchmark for future models to max We're partnering wi

→24 Jul 2026ChatGPT app is now melting my iPhone rendering at like one word a second even when I’m typing to it Have I angered the transformer

ChatGPT app is now melting my iPhone rendering at like one word a second even when I’m typing to it Have I angered the transformer Why does codex melt my laptop Like what is it actually doing on the l

→29 Jul 2026Who on earth came up with 5 hour limits? The day is 24 hours long. While do our limit times shift by an hour each day. Pls if you are going …

Who on earth came up with 5 hour limits? The day is 24 hours long. While do our limit times shift by an hour each day. Pls if you are going to have limits 4 or 6 hours. Hello people of Sol! I've reset

→6 Aug 2026Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which wa…

Luna non-reasoning is a bit better than GPT 4o which was sota 2 years ago. Luna (medium) thinking is a bit better than GPT-5 (High) which was sota 1 year ago. Now free to everyone unlimited Sol Max/Fa

→11 Aug 2026Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b

CompanyGoogle4 recent entries
19 May 2026This is the Build with Gemini XPRIZE. $2,000,000 in prizes. 90 days. Pick a problem worth solving. Build a profitable business with AI. Gran…

The Build with Gemini XPRIZE is a competition offering $2 million in total prizes over a 90-day period where participants select a meaningful problem and develop a profitable AI-powered business solut

→19 May 2026Can’t wait for Gemini Omni in @NotebookLM cinematic explainer videos 👀

Emad Mostaque expressed anticipation for the integration of Google's Gemini Omni multimodal AI model into NotebookLM's cinematic explainer video generation features. The post suggests potential upcomi

→29 Jun 2026Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemin…

Most popular model on @OpenRouter (10tr tokens) turns out to be a 1.6tr MoE by @Meituan_LongCat (superapp/DoorDash of China) Basically Gemini / Opus 4.6 level 35tr tokens trained entirely on 50k Chine

→11 Aug 2026Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b

CompanyMeta1 recent entries
6 Aug 2026Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up e…

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up etched onto silicon As models satisfice etching makes sense,

CompanyxAI8 recent entries
6 May 2026This likely costs about ~500m a month, ~6bn a year to rent blended Which is about the run rate net loss of xai end of q1 Anthropics revenu…

This likely costs about ~500m a month, ~6bn a year to rent blended Which is about the run rate net loss of xai end of q1 Anthropics revenue run rate was 9bn end of 2025, 30bn a month ago Our agreement

→15 May 2026Team @grok @xai can you please allow latex uploads on mobile app thx

Emad Mostaque, founder of Stability AI, requested that the Grok AI team (owned by xAI) add LaTeX file upload functionality to their mobile application. The post was made as a public appeal on X (forme

→20 May 2026A basic analysis shows that zeroing out federal taxes for the bottom half the USA would help millions, have minimal impact on tax receipts &…

A basic analysis shows that zeroing out federal taxes for the bottom half the USA would help millions, have minimal impact on tax receipts & add over $100 billion to the economy. It's easy to check th

→25 May 2026It’ll be interesting to see if the post training for this uses a multiple of the compute of pretraining as cursor did when they tuned Kimi a…

It’ll be interesting to see if the post training for this uses a multiple of the compute of pretraining as cursor did when they tuned Kimi as the base model Grok foundation model V9-Medium (1.5T) has

→26 May 2026I think folk are underestimating how much of AI models are actually engineering at scale versus breakthrough research. See how @cursor_ai ca…

I think folk are underestimating how much of AI models are actually engineering at scale versus breakthrough research. See how @cursor_ai caught up to Anthropic / OpenAI models run at a fraction of th

→3 Jun 2026Yo @xai team, this would be an amazing demo of @grok capability. Push button, have it read all your bookmarks, organise them, make a report …

Yo @xai team, this would be an amazing demo of @grok capability. Push button, have it read all your bookmarks, organise them, make a report on the most interesting one and your interests over time etc

→23 Jun 2026The upcoming @Seedance 2.5 model looking insane as well with multi asset input, way longer outputs etc I would expect @grok imagine to keep …

The upcoming @Seedance 2.5 model looking insane as well with multi asset input, way longer outputs etc I would expect @grok imagine to keep pace and real time of this quality by end of next year (!) E

→12 Aug 2026imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here …

imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here to top it at a great price, congrats to @SpaceXAI team & loo

CompanyDeepSeek7 recent entries
24 Apr 2026The Newton–Schulz iteration coefficients optimized by DeepSeek-V4 are surprisingly strong: they effectively normalize all singular values to…

The Newton–Schulz iteration coefficients optimized by DeepSeek-V4 are surprisingly strong: they effectively normalize all singular values to 1. This matches our previous intuition: a well-balanced spe

→24 Apr 2026Necessity is the mother of kv-cache optimisation

Necessity is the mother of kv-cache optimisation I’m still amazed that DeepSeek, Kimi, and Qwen can train very strong LLMs with far fewer and often nerfed NVIDIA GPUs, or even Huawei chips. DeepSeek V

→24 Apr 2026Congrats to @deepseek_ai team! Doing the numbers I would estimate: Pro < 14m for the final training run Flash < 4m Ratio of active params …

Congrats to @deepseek_ai team! Doing the numbers I would estimate: Pro < 14m for the final training run Flash < 4m Ratio of active params x total training tokens vs v3 Total compute costs (data prep,

→15 Jul 2026Given GB300 I would estimate this is was trained on about 1e25 flops (same as DeepSeek v4) over 1.6m hours (1 month on 2k chips/28 racks) Co…

Given GB300 I would estimate this is was trained on about 1e25 flops (same as DeepSeek v4) over 1.6m hours (1 month on 2k chips/28 racks) Cost 10-20m ($6-12/hour/chip) The lite version likely 4x less,

→31 Jul 2026As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than …

As shocking as the Kimi K3 release. Massive performance gain was just with post-training Model is 3x smaller than GLM 5.2 (10x smaller than K3) & works on a MacBook / Spark This is Q1 flagship (Opus 4

→5 Aug 2026Hey everyone! I just released a sub-6B sparse activation AI model which was built with a brand new architecture : fusion. I fused weights fr…

Hey everyone! I just released a sub-6B sparse activation AI model which was built with a brand new architecture : fusion. I fused weights from @liquidai's LFM2.5-2.6B & @Alibaba_Qwen's Qwen3.6-35B-A3B

→8 Aug 2026Weirdly iirc stable diffusion (1.4) finished training around four years ago today too

Emad posted that Stable Diffusion v1.4 reached the end of its training cycle roughly four years before the post was published. Greg Brockman added that GPT‑4 similarly completed training around the sa

CompanyNVIDIA7 recent entries
24 Apr 2026Necessity is the mother of kv-cache optimisation

Necessity is the mother of kv-cache optimisation I’m still amazed that DeepSeek, Kimi, and Qwen can train very strong LLMs with far fewer and often nerfed NVIDIA GPUs, or even Huawei chips. DeepSeek V

→19 May 2026We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and …

We’re releasing Nemotron-Labs-Diffusion - the first Tri-mode LM family (3B/8B/14B) that switches between 1⃣Autoregressive, 2⃣Diffusion, and 3⃣Self-Speculation decoding by simply changing the attention

→1 Jun 2026With Nemotron & Cosmos NVIDA gonna commoditise everyone's complement

Emad Mostaque suggests that NVIDIA's Nemotron and Cosmos models will commoditize complementary AI technologies and services in the market. The statement implies that these NVIDIA offerings will make e

→6 Jul 2026“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs …

“We were going to run out of money.” Groq was 3 weeks away from death and @JonathanRoss321’s leadership team put together a list of layoffs that also would have killed the company. He realized that th

→15 Jul 2026This is for pretraining based on nemotron nvp4 figures, mfu etc Large context, multimodal, long RL we are seeing now could make it a multipl…

This is for pretraining based on nemotron nvp4 figures, mfu etc Large context, multimodal, long RL we are seeing now could make it a multiple of this Worth noting how good the lite model is, similar t

→27 Jul 2026If you look at the current lead times for ~$5bn per year of cutting edge GPUs you can probably figure out the time plus one training run to …

If you look at the current lead times for ~$5bn per year of cutting edge GPUs you can probably figure out the time plus one training run to IlyAGI We are announcing a long-term strategic partnership w

→8 Aug 2026Weirdly iirc stable diffusion (1.4) finished training around four years ago today too

Emad posted that Stable Diffusion v1.4 reached the end of its training cycle roughly four years before the post was published. Greg Brockman added that GPT‑4 similarly completed training around the sa