AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “jeremy-howard--x”

GridTimelineEvolution
212 results
7 May 2026

The most female-led product org in tech right now: Chief Product Officer: Ami Vora Claude Code/Cowork Head of Product: Cat Wu Claude Code/Co…

Model ReleasesDGX agent

The most female-led product org in tech right now: Chief Product Officer: Ami Vora Claude Code/Cowork Head of Product: Cat Wu Claude Code/Cowork Head of Eng: Fiona Fung Claude Platform Head of Product

The only correct answer when a VC asks: 'What's your moat?'

TutorialsDGX agent

# Summary Jeremy Howard explains what founders should actually answer when venture capitalists ask about their competitive moat, addressing a common startup pitching question. The post likely clarifie

Welcome to DS4, a specialized inference engine for DeepSeek v4 Flash. https://github.com/antirez/ds4 This project would have been impossible…


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases
DGX agent

Welcome to DS4, a specialized inference engine for DeepSeek v4 Flash. https://github.com/antirez/ds4 This project would have been impossible without the existence of llama.cpp and GGML and the work of

6 May 2026

ニャッキの伊藤有壱さんにお声掛け頂き、コマ撮りの展覧会に一作家として参加しています。私はコマ撮り分野ではない場所から活動をはじめて、デザインの視点でのコマ撮りに取り組んできましたが、今回初めてコマ撮り界の本丸の方々とご一緒でき嬉しいです。今6年目のマッチ撮影素材等を展示しています

TutorialsDGX agent

A designer shares their experience participating in a stop-motion animation exhibition curated by Iyuichi Itoh from the animation 'Nyacki,' displaying their match photography materials from a six-year

Exciting to work with @googledevs . Dflash is one of the most powerful technique developed here at UCSD by @zhijianliu_ and @jianchen1799 an…

HardwareDGX agent

Exciting to work with @googledevs . Dflash is one of the most powerful technique developed here at UCSD by @zhijianliu_ and @jianchen1799 and glad that our students and collaborators help port them in

Give LLMs 1. A latent space diffusion-like reasoning. 2. A real recurrent state. 3. A world-model pre-pre-training. And we are done.

TutorialsDGX agent

Jeremy Howard proposes three key enhancements for large language models: incorporating latent space diffusion-like reasoning processes, adding a persistent recurrent state mechanism, and implementing

I'm disappointed by repeatedly hearing that my colleagues at Anthropic believe they are the only ones who should be trusted with building AI…

TutorialsDGX agent

I'm disappointed by repeatedly hearing that my colleagues at Anthropic believe they are the only ones who should be trusted with building AI. It is *very good* there are a diversity of people building

Two weeks after release, Hy3 preview is #1 on @OpenRouter's weekly leaderboard with 3.66T tokens processed, up 298% week-over-week. #1 in ov…

Model ReleasesDGX agent

Two weeks after release, Hy3 preview is #1 on @OpenRouter's weekly leaderboard with 3.66T tokens processed, up 298% week-over-week. #1 in overall usage, tool calls, and coding. 15.4% market share acro

5 May 2026

🚀 Day-0 MTP support for Gemma4 now available at vLLM with ready-to-use docker image! ⚡️Enjoy up to 3x faster decoding performance to superc…

Model ReleasesDGX agent

🚀 Day-0 MTP support for Gemma4 now available at vLLM with ready-to-use docker image! ⚡️Enjoy up to 3x faster decoding performance to supercharge your development with zero quality degradation! Check o

Excited to introduce Gemma 4 Multi-Token Prediction Drafters⚡️Accelerated inference right in your pockets - Up to a 3x speedup - Same qualit…

Model ReleasesDGX agent

Excited to introduce Gemma 4 Multi-Token Prediction Drafters⚡️Accelerated inference right in your pockets - Up to a 3x speedup - Same quality guarantees - Available in your favorite open-source tools

i have yet to meet a single person who feels like claude code is getting exponentially better on some kind of fast take off

Model ReleasesDGX agent

i have yet to meet a single person who feels like claude code is getting exponentially better on some kind of fast take off Anthropic pays $750K/ year per senior engineer. The creator of Claude Code j

MiniMax-M2.7 is now available across six inference providers on Artificial Analysis, with significant differentiation in speed and price @Sa…

TutorialsDGX agent

MiniMax-M2.7 is now available across six inference providers on Artificial Analysis, with significant differentiation in speed and price @SambaNovaAI leads on speed at 435 output tokens/s, >3x faster

4 May 2026

Deepseek V4 works more thoroughly than other open source models: It writes its own tests and performs extensive validation. This leads to be…

Model ReleasesDGX agent

Deepseek V4 works more thoroughly than other open source models: It writes its own tests and performs extensive validation. This leads to better performance, but also cases of the model being overconf

hi, i'm a sole proprietor/founder in Austria and i earn many many multiples of what i'd earn as an employee, despite 'predatory income tax'.…

TutorialsDGX agent

hi, i'm a sole proprietor/founder in Austria and i earn many many multiples of what i'd earn as an employee, despite 'predatory income tax'. in fact, i opt out of the many tax optimizations i could us

3 May 2026

i actually don't want this 'but you don't review compiler output either' meme to die. it's the perfect signal for being immediately able to …

AgentsDGX agent

i actually don't want this 'but you don't review compiler output either' meme to die. it's the perfect signal for being immediately able to ignore someone in this space. Interesting article on treatin

30 Apr 2026

Fun fact - if you have a recent commit that mentions OpenClaw in a json blob, Claude Code will either refuse your request or bill you extra …

Model ReleasesDGX agent

Fun fact - if you have a recent commit that mentions OpenClaw in a json blob, Claude Code will either refuse your request or bill you extra money. This is an empty repo, I'm just calling Claude Code d

I can’t believe I stopped using Claude Code max and entirely use DeepSeek and Hermes. It’s so fast, so so fast, 3x faster for the same task.…

Model ReleasesDGX agent

I can’t believe I stopped using Claude Code max and entirely use DeepSeek and Hermes. It’s so fast, so so fast, 3x faster for the same task. So cheap. I spent $5 last week and never need worry about b

The Zig project's rationale for their blanket ban on AI-assisted contributions makes a lot of sense to me - for them, time spent reviewing P…

TutorialsDGX agent

The Zig project's rationale for their blanket ban on AI-assisted contributions makes a lot of sense to me - for them, time spent reviewing PRs isn't about the code, it's about growing new contributors

29 Apr 2026

DeepSeek V4 Pro is crazy good at bug fixing. Costs counted in cents not dollars/tens of dollars. It’s the next level quiet confidence not to…

Model ReleasesDGX agent

DeepSeek V4 Pro is crazy good at bug fixing. Costs counted in cents not dollars/tens of dollars. It’s the next level quiet confidence not to get caught up by benchmarks & rankings, and just let us use

Ernie-5.1 from @ErnieforDevs lands at #13 in Text Arena — now the #1 highest-ranked model from a Chinese lab. Strongest categories: - #9 Mat…

ApplicationsDGX agent

Ernie-5.1 from @ErnieforDevs lands at #13 in Text Arena — now the #1 highest-ranked model from a Chinese lab. Strongest categories: - #9 Math - #1 Legal & Government - #4 Business, Management & Financ

IBM Granite just released two multilingual embedding models with 97M and 311M parameters 🤏🏻 ModernBERT-based, 200+ languages, 32K context,…

Model ReleasesDGX agent

IBM Granite just released two multilingual embedding models with 97M and 311M parameters 🤏🏻 ModernBERT-based, 200+ languages, 32K context, and built for retrieval, search, similarity, and code. And...

The new DeepSeek-V4, like DeepSeek-V3, uses concepts from our 2024 paper on Self-Rewarding LMs -- see screenshots of their tech reports! (Co…

Model ReleasesDGX agent

The new DeepSeek-V4, like DeepSeek-V3, uses concepts from our 2024 paper on Self-Rewarding LMs -- see screenshots of their tech reports! (Congrats!) Classical :) Self-Rewarding LMs from Jan 2024: http

Xiaomi MiMo-V2.5-Pro achieves multiple breakthroughs in the latest Arena rankings (Apr 26, 2026) 🔥 🏆 Text Arena (Expert) — #6 globally | #…

ApplicationsDGX agent

Xiaomi MiMo-V2.5-Pro achieves multiple breakthroughs in the latest Arena rankings (Apr 26, 2026) 🔥 🏆 Text Arena (Expert) — #6 globally | #1 open-source model Also #1 among Chinese models, with Xiaomi

28 Apr 2026

However, older versions have a critical CVE, so you do need to upgrade! If you use uv (thanks to @charliermarsh) you can pop this in your py…

TutorialsDGX agent

However, older versions have a critical CVE, so you do need to upgrade! If you use uv (thanks to @charliermarsh) you can pop this in your pyproject to work around it: ``` [tool.uv] override-dependenci

I know the litellm team have a *lot* going on, & I have no desire to increase their workload even more… But if you're a litellm user, this i…

TutorialsDGX agent

I know the litellm team have a *lot* going on, & I have no desire to increase their workload even more… But if you're a litellm user, this is a critical issue to know about - it currently force downgr

I'm so confused…

Model ReleasesDGX agent

I'm so confused… We're excited to partner with Google to offer Grounding With Exa inside of Gemini models! Using Exa's agent-first search, Gemini models can now access billions of websites, technical

I'm speechless at Google signing a deal to use our AI models for classified tasks. Frankly, it is shameful. For HR, I'm not speaking on beha…

TutorialsDGX agent

I'm speechless at Google signing a deal to use our AI models for classified tasks. Frankly, it is shameful. For HR, I'm not speaking on behalf of Google but in my personal capacity, quoting public inf

This is great - @deepseek_ai V4 supports prefill! :D Most other providers have been dropping support for this critically important capabilit…

Model ReleasesDGX agent

This is great - @deepseek_ai V4 supports prefill! :D Most other providers have been dropping support for this critically important capability, so wonderful to see at least one company stepping up. htt

What search providers are you all using with openclaw/pi/opencode/etc? Brave; serpapi; gemini; ...? Got any favorites?

Model ReleasesDGX agent

This post is a question about search provider preferences for use with various open-source tools and APIs (openclaw, pi, opencode, etc.), with examples of options like Brave, SerpAPI, and Gemini. The

27 Apr 2026

It just takes a little time and care to engage with the work of new folks in the field -- but it can make a big difference :D

TutorialsDGX agent

It just takes a little time and care to engage with the work of new folks in the field -- but it can make a big difference :D On a more personal note, I'm grateful to @jeremyphoward for helping me bre

Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, supporting commercial deployment, continued training, and fine-tuning - no add…

Model ReleasesDGX agent

Xiaomi MiMo-V2.5 is now officially open-sourced! MIT License, supporting commercial deployment, continued training, and fine-tuning - no additional authorization required. Two models, both supporting

26 Apr 2026

🔥DeepSeek Input Cache Price Drop! Effective immediately, the price for input cache hits across the ENTIRE DeepSeek API series is reduced to…

Model ReleasesDGX agent

🔥DeepSeek Input Cache Price Drop! Effective immediately, the price for input cache hits across the ENTIRE DeepSeek API series is reduced to just 1/10th of the original price! Build more efficiently fo

FYI Claude Code is mostly a vibe-coded product (as they say, 100% written by Claude) It's the worst harness for Opus 4.6 among ANY harness o…

Model ReleasesDGX agent

FYI Claude Code is mostly a vibe-coded product (as they say, 100% written by Claude) It's the worst harness for Opus 4.6 among ANY harness on Terminal-Bench 2 I feel sorry for Claude Code I know they'

I am increasingly bullish on open source harnesses (like OpenCode) not because they will be better than SOTA closed harnesses, but because t…

Model ReleasesDGX agent

I am increasingly bullish on open source harnesses (like OpenCode) not because they will be better than SOTA closed harnesses, but because they will never pull shady stuff like what Claude Code and ot

THIS GUY LOST $200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects…

Model ReleasesDGX agent

THIS GUY LOST 200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects. it's a system prompt specification file. not some obscure e

25 Apr 2026

Launching pyptx — a Python DSL for writing NVIDIA PTX kernels. One PTX instruction = one Python call. Write pure PTX in Python. Direct Hoppe…

HardwareDGX agent

Launching pyptx — a Python DSL for writing NVIDIA PTX kernels. One PTX instruction = one Python call. Write pure PTX in Python. Direct Hopper + Blackwell support: wgmma, TMA, tcgen05, mbarriers. JAX +

24 Apr 2026

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T t…

Model ReleasesDGX agent

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top

Sapiens2 is the highest quality ViT backbone that now exists in the public domain. It was pretrained on the equivalent of 1/2 of all human i…

TutorialsDGX agent

Sapiens2 is the highest quality ViT backbone that now exists in the public domain. It was pretrained on the equivalent of 1/2 of all human images on Flickr. First public release by a large lab that is

this is probably the most important piece of software of the decade next to vllm and sglang. i'm not joking.

Model ReleasesDGX agent

this is probably the most important piece of software of the decade next to vllm and sglang. i'm not joking. llama.cpp at 100k stars now that 90% of the code worldwide is being written by AI agents, I

Today I re-iterate: I hate MoEs and we are wasting time on them.... Let's unite and call a global ban on MoEs please. Please 1M+ salary rese…

TutorialsDGX agent

Today I re-iterate: I hate MoEs and we are wasting time on them.... Let's unite and call a global ban on MoEs please. Please 1M+ salary researchers: do better... credits to @IlysMoutawwakil for the gr

22 Apr 2026

AFAIK this is the first time it's been made official from an OpenAI staffer - the `/backend-api/codex/responses` endpoint that Pi and Openco…

TutorialsDGX agent

AFAIK this is the first time it's been made official from an OpenAI staffer - the `/backend-api/codex/responses` endpoint that Pi and Opencode (IIUC) uses is officially supported! :D This is really gr

... and Anthropic reverted this change. Claude Code is now part of Pro, as per the Pricing page. Important note on the growth hack: Anthropi…

Model ReleasesDGX agent

... and Anthropic reverted this change. Claude Code is now part of Pro, as per the Pricing page. Important note on the growth hack: Anthropic advertises safety and integrity as their values. A 'fake d

been working with @Kimi_Moonshot K2.6 for the past 2 hours 'accidentially' and ... it's really good!

TutorialsDGX agent

Jeremy Howard shares positive impressions of working with Kimi Moonshot's K2.6 model for two hours, noting it performed well despite being used somewhat unexpectedly. The post suggests K2.6 demonstrat

clampy clampy clampdown. just waiting for OAI to clamp down as well.

TutorialsDGX agent

This post appears to reference regulatory or policy 'clampdowns' on AI, with the author expressing anticipation that OpenAI will face similar restrictions. Without access to the full context or surrou

For the 'small test' they've modified their docs to remove mention of Claude Code in Claude Pro: https://support.claude.com/en/articles/1114…

Model ReleasesDGX agent

For the 'small test' they've modified their docs to remove mention of Claude Code in Claude Pro: https://support.claude.com/en/articles/11145838-using-claude-code-with-your-max-plan It's been a shock

🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, and Qwen3.6-27B punches way above i…

Model ReleasesDGX agent

🚀 Meet Qwen3.6-27B, our latest dense, open-source model, packing flagship-level coding power! Yes, 27B, and Qwen3.6-27B punches way above its weight. 👇 What's new: 🧠 Outstanding agentic coding — surpa

21 Apr 2026

Anyone know whether OpenAI officially supports the use of the `/backend-api/codex/responses` endpoint that Pi and Opencode (IIUC) uses? It d…

TutorialsDGX agent

Anyone know whether OpenAI officially supports the use of the `/backend-api/codex/responses` endpoint that Pi and Opencode (IIUC) uses? It doesn't seem to be documented, and was reverse engineered by

@nbaschez what test, when it seems to be a full-on update to the pricing page? I could not get it to not show this update, no matter how man…

TutorialsDGX agent

@nbaschez what test, when it seems to be a full-on update to the pricing page? I could not get it to not show this update, no matter how many different devices I checked in (not connected to my accoun

Since this is blowing up on hacker news. Boris said that CLI usage is allowed. Thus we added support for it, only to find out that we are st…

TutorialsDGX agent

Since this is blowing up on hacker news. Boris said that CLI usage is allowed. Thus we added support for it, only to find out that we are still blocked there. It is trival to work around with a few re

This is so confusing. Did Anthropic really just drop Claude Code from their $20/month plan? Why would they do that through a pricing page up…

Model ReleasesDGX agent

This is so confusing. Did Anthropic really just drop Claude Code from their 20/month plan? Why would they do that through a pricing page update without making a proper announcement? Plus, 20/month sti

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the bi…

TutorialsDGX agent

This is what I’ve been cooking in the past 4 months . GPT Image 2 is over a massive 240 elo jump over the second place model, marking the biggest jump bigger than the rest of the leaderboard combined

20 Apr 2026

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up t…

Model ReleasesDGX agent

I upgraded my Claude token counter tool to compare different models and Opus 4.7 does appear to use 1.46x times the tokens for text and up to 3x the tokens for images - it's priced the same as Opus 4.

opus 4.7 seems to have a much better time in claude code if you run without most of the system prompt (claude --system-prompt '.')

Model ReleasesDGX agent

A user reports that Opus 4.7 performs better in Claude Code when executed with a minimal system prompt (using just a period) rather than the full default system prompt, suggesting that reducing system

Philosophy (among other things) grad here. I could write a whole essay about this video, and mostly the reactions to it. People are dunking …

Model ReleasesDGX agent

Philosophy (among other things) grad here. I could write a whole essay about this video, and mostly the reactions to it. People are dunking on her because of what she symbolises more than what she say

19 Apr 2026

in 1982 Titanic survivor Ruth Becker was giving an interview where she stated the ship broke in two. The treasurer of the Titanic Historical…

TutorialsDGX agent

in 1982 Titanic survivor Ruth Becker was giving an interview where she stated the ship broke in two. The treasurer of the Titanic Historical Society actually took the microphone away from her and said

18 Apr 2026

A: Because they're pointless.

TutorialsDGX agent

I cannot provide a reliable summary of this specific post without access to its content, as the URL appears incomplete or potentially inaccurate. To create an accurate knowledge base entry, please pro

Q: Why are stories set in zero dimensional space so uninteresting?

TutorialsDGX agent

This post likely discusses why narratives or thought experiments set in zero-dimensional space—where there is no length, width, height, or any spatial extension—fail to generate compelling storytellin

17 Apr 2026

@AmandaAskell are you the person to thank for this?

TutorialsDGX agent

Amanda Askell is likely a researcher or professional involved in AI safety or alignment work, and Jeremy Howard is publicly crediting or thanking her for a contribution or achievement on social media.

@RoyRogers_HTMS This is the way I like to see the Hurdy-Gurdy played.

TutorialsDGX agent

Roy Rogers demonstrates his preferred style of playing the hurdy-gurdy, showcasing traditional or characteristic technique in this post shared by Jeremy Howard on X. The post reflects appreciation for

So cool to see that open-source, with open experimentation (and with the help of someone posting blog posts about their personal research), …

TutorialsDGX agent

So cool to see that open-source, with open experimentation (and with the help of someone posting blog posts about their personal research), can yield a very robust method for MoE balancing. This metho

← Previous
1234
Next →