AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,147 results
Agents

ml-intern now notifies you when it needs input! Amazing for those 12h+ autonomous LLM training runs

DGX agent

ml-intern now notifies you when it needs input! Amazing for those 12h+ autonomous LLM training runs The ML Intern can now ping you on Slack when it's finished training models, generating datasets, or

agentsclem-delangue--x
27 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

OpenAI released a new dataset on the hub 👀 Benchmark for 'Making ChatGPT better for clinicians' https://huggingface.co/datasets/openai/heal…

DGX agent

OpenAI released a new dataset on Hugging Face Hub designed to improve ChatGPT's performance in clinical settings, addressing the specific needs and workflows of healthcare professionals. The dataset s

model-releasesclem-delangue--x
26 Apr 2026
Research

TRINITY: An Evolved LLM Coordinator https://arxiv.org/abs/2512.04695

DGX agent

TRINITY is a large language model coordinator system that evolves and improves its approach to orchestrating multiple LLMs or reasoning processes. The system likely uses evolutionary or adaptive mecha

researchdavid-ha--x
26 Apr 2026
Model Releases

Anyone got DeepSeek-V4-Flash running on a Mac yet? 512GB or 256GB or 128GB or smaller?

DGX agent

Simon Willison inquires about running DeepSeek-V4-Flash on Mac hardware, specifically asking about feasibility across different RAM configurations from 512GB down to smaller amounts. This reflects dis

model-releasessimon-willison--x
25 Apr 2026
Model Releases

btw we are cooking something with @hhua_ (not final yet but keep calendar open after ICML in Seoul)

DGX agent

btw we are cooking something with @hhua_ (not final yet but keep calendar open after ICML in Seoul) 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M con

model-releasesswyx--x
25 Apr 2026
Model Releases

Qwen-Image-2.0-Pro is now live 🚀🚀 We’ve pushed image quality, multilingual text rendering, and instruction following to a new level, while…

DGX agent

Qwen-Image-2.0-Pro is now live 🚀🚀 We’ve pushed image quality, multilingual text rendering, and instruction following to a new level, while making performance much more consistent across styles.🌅🌃 Rank

model-releasesqwen--x
25 Apr 2026
Model Releases

We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending…

DGX agent

We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending my ml-intern to participate in the @OpenAI Parameter Golf c

model-releasesclem-delangue--x
25 Apr 2026
Tools

What would you like to see next? Full Param tuning?

DGX agent

Fireworks AI posted a question to their audience asking what features or capabilities they would like to see developed next, with a specific mention of full parameter tuning as a potential option. Thi

toolsfireworks-ai--x
25 Apr 2026
Model Releases

... and there's me thinking that sending it out at 9pm Pacific Time on a Thursday evening was safe, surely there wouldn't be any news this e…

DGX agent

... and there's me thinking that sending it out at 9pm Pacific Time on a Thursday evening was safe, surely there wouldn't be any news this evening that I might want to include in there... https://x.co

model-releasessimon-willison--x
24 Apr 2026
Agents

Deep Agents middleware is a great invention. You get a really strong default harness and a clean way to customize every part of it

DGX agent

Deep Agents middleware is a great invention. You get a really strong default harness and a clean way to customize every part of it DeepAgents - it’s all you need - batteries included if you want, but

agentsharrison-chase--x
24 Apr 2026
Model Releases

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T t…

DGX agent

🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 DeepSeek-V4-Pro: 1.6T total / 49B active params. Performance rivaling the world's top

model-releasesjeremy-howard--x
24 Apr 2026
Model Releases

GPT-5.5 is now available in Windsurf 2.0!

DGX agent

Windsurf 2.0 has been released with integration of GPT-5.5, representing an upgrade to the Windsurf development platform's AI capabilities. This update likely provides users access to an advanced lang

model-releaseswindsurf--x
24 Apr 2026
Model Releases

My first two TiKZ Sparks unicorns from DeepSeek v4. (Expert mode, from the DeepSeek site, which is supposed to be v4 Pro according to the re…

DGX agent

Ethan Mollick documents his first attempts at using DeepSeek v4's expert mode to generate TiKZ code for creating unicorn graphics, sharing results from the DeepSeek website's v4 Pro interface. The pos

model-releasesethan-mollick--x
24 Apr 2026
Safety

New episode of The Information Bottleneck is out, this time with @liuzhuang1234 (Princeton). We talked about ConvNeXt and whether architectu…

DGX agent

New episode of The Information Bottleneck is out, this time with @liuzhuang1234 (Princeton). We talked about ConvNeXt and whether architecture still matters; dataset bias and what 'good data' actually

safetyyann-lecun--x
24 Apr 2026
Model Releases

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud.

DGX agent

🐳Ollama is working to have DeepSeek-V4-Pro and Flash available on Ollama's cloud. 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M context length. 🔹 Dee

model-releasesollama--x
24 Apr 2026
Model Releases

These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM t…

DGX agent

These pelicans are kind of angry looking! Left is deepseek-v4-flash, right is deepseek-v4-pro - both generated using OpenRouter via my LLM tool 🚀 DeepSeek-V4 Preview is officially live & open-sourced!

model-releasessimon-willison--x
24 Apr 2026
Model Releases

This has been the most anticipated event in OSS and it doesn't disappoint. Congratulations to @deepseek_ai team. DSV4 Pro is now available o…

DGX agent

This has been the most anticipated event in OSS and it doesn't disappoint. Congratulations to @deepseek_ai team. DSV4 Pro is now available on @togethercompute and we'll be adding a lot of capacity beh

model-releasestogether-ai--x
24 Apr 2026
Model Releases

This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.…

DGX agent

This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B running inside of Pi coding agent via Llama.cpp on the MacBook Pro For non-trivial tasks on the @huggingf

model-releasesclem-delangue--x
24 Apr 2026
Model Releases

Try out Devin with the GPT-5.5 Agent Preview today: https://devin.ai

DGX agent

Cognition AI announced an early preview opportunity for Devin integrated with GPT-5.5 Agent capabilities, inviting users to test the combination at devin.ai. This likely represents an update to Devin'

model-releasescognition-ai--x
24 Apr 2026
Model Releases

Here’s my view on GPT-5.5, which I have been testing for a couple of weeks. It conducted not-bad social science research on its own, develop…

DGX agent

Here’s my view on GPT-5.5, which I have been testing for a couple of weeks. It conducted not-bad social science research on its own, developed a novel RPG & more. There is still jaggedness but GPT-5.5

model-releasesethan-mollick--x
23 Apr 2026
Model Releases

I'm a manager at @OpenAI, but with GPT-5.5 I'm a more effective IC than I've ever been. I can now write CUDA kernels like a pro. I can rely …

DGX agent

I'm a manager at @OpenAI, but with GPT-5.5 I'm a more effective IC than I've ever been. I can now write CUDA kernels like a pro. I can rely on it to run my research experiments. And we know how to mak

model-releasessam-altman--x
23 Apr 2026
Model Releases

Last night was the biggest disaster in the history of Tesla. Let me walk you through what actually happened on that earnings call, because t…

DGX agent

Last night was the biggest disaster in the history of Tesla. Let me walk you through what actually happened on that earnings call, because the headlines are doing you a disservice: Elon Musk got on th

model-releasesgary-marcus--x
23 Apr 2026
Model Releases

One day while testing GPT-5.5, I had my first taste of AGI. We had a branch with hundreds of visual and front-end changes, plus complex refa…

DGX agent

One day while testing GPT-5.5, I had my first taste of AGI. We had a branch with hundreds of visual and front-end changes, plus complex refactors. At the same time, main had changed a lot too. Conflic

model-releasessam-altman--x
23 Apr 2026
Agents

Thanks to @lmsysorg ! Try it on SGLang now!🚀🚀

DGX agent

Thanks to @lmsysorg ! Try it on SGLang now!🚀🚀 🚀 Qwen3.6-27B is here, and we have day 0 support on SGLang ✅ 27B params, beats Qwen3.5-397B-A17B across major coding benchmarks → Agentic coding → Text +

agentsqwen--x
23 Apr 2026
Model Releases

Three shifts in the AI stack today: - OpenAI ships GPT-5.5 as a mid-cycle drop before its IPO - Hugging Face's ML Intern beats Claude Code a…

DGX agent

Three shifts in the AI stack today: - OpenAI ships GPT-5.5 as a mid-cycle drop before its IPO - Hugging Face's ML Intern beats Claude Code and Codex on research - Pliny used Claude Opus 4.7 to jailbre

model-releasesclem-delangue--x
23 Apr 2026
Model Releases

Waiting at Superchargers is rare, but when it happens, customers should be able to plan with confidence. Superchargers are the only fast cha…

DGX agent

Waiting at Superchargers is rare, but when it happens, customers should be able to plan with confidence. Superchargers are the only fast chargers with predictive wait times and we're now using vehicle

model-releaseselon-musk--x
23 Apr 2026
Industry

we've quantized kimi-k2.6 to mxfp4 on amd! download and use today! @AIatAMD

DGX agent

Hugging Face and AMD have collaborated to quantize the Kimi-K2.6 model to MXFP4 format, enabling efficient inference on AMD hardware. MXFP4 is a low-precision floating-point quantization technique tha

industryclem-delangue--x
23 Apr 2026
Safety

ChatGPT doesn’t know its whisk from its elbow

DGX agent

Gary Marcus critiques ChatGPT's lack of embodied understanding and spatial reasoning, arguing that the language model struggles with physical concepts that humans intuitively grasp through bodily expe

safetygary-marcus--x
22 Apr 2026
Model Releases

Could a sequence of poetic gibberish manipulate the world’s most advanced AI? Recent tests on GPT-5 by @ChristophHeilig show it consistently…

DGX agent

Could a sequence of poetic gibberish manipulate the world’s most advanced AI? Recent tests on GPT-5 by @ChristophHeilig show it consistently rates nonsensical 'word salad' as higher quality than clear

model-releasesgary-marcus--x
22 Apr 2026
Industry

Excited to share some big news: I've joined @FireworksAI_HQ President. I've spent my career at the intersection of great teams, disruptive t…

DGX agent

Excited to share some big news: I've joined @FireworksAI_HQ President. I've spent my career at the intersection of great teams, disruptive technology, and massive market opportunity, from architecting

industrysonya-huang--x
22 Apr 2026
Model Releases

Guys, I am absolutely astounded. The Qwen 3.6 27b is like a jump to Qwen 4 from Qwen 27B 3.5. I just did a full suite of front end design te…

DGX agent

Guys, I am absolutely astounded. The Qwen 3.6 27b is like a jump to Qwen 4 from Qwen 27B 3.5. I just did a full suite of front end design tests and agentic benchmarks, made entirely by it. VERDICT: Th

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent…

DGX agent

How far are we from agents that can self-generate world knowledge? The work proposes an outcome-based reward that measures how much an agent's self-generated world knowledge actually improves its task

model-releasesdair-ai--x
22 Apr 2026
Agents

I predicted in January that 'harness' would be the AI buzzword of H12026 - seems accurate so far. Everyone is talking about harnesses but I'…

DGX agent

I predicted in January that 'harness' would be the AI buzzword of H12026 - seems accurate so far. Everyone is talking about harnesses but I'm not confident that the majority of people (including me!)

agentsharrison-chase--x
22 Apr 2026
Model Releases

LM Performance:With only 27B parameters, Qwen3.6-27B outperforms the Qwen3.5-397B-A17B (397B total / 17B active, ~15x larger!) on every majo…

DGX agent

LM Performance:With only 27B parameters, Qwen3.6-27B outperforms the Qwen3.5-397B-A17B (397B total / 17B active, ~15x larger!) on every major coding benchmark — including SWE-bench Verified (77.2 vs.

model-releasesqwen--x
22 Apr 2026
Safety

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infra…

DGX agent

Q1 2026 Shareholder Update https://ir.tesla.com/#quarterly-disclosure We continued to make meaningful progress on the build out of the infrastructure & AI software that underpins our Robotaxi & future

safetyelon-musk--x
22 Apr 2026
Tools

Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking into where this came fro…

DGX agent

Thank you for flagging this, Jeff. This was a mistake: we are not deprecating text-embedding-3-small. We’re looking into where this came from now, and we’ll also email users to clarify. Sorry for the

toolssimon-willison--x
22 Apr 2026
Model Releases

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt ha…

DGX agent

Claude Opus 4.7 with adaptive thinking via the API... am I missing something or is it not possible any more to force it to think? (Prompt hacks like 'think step by step' don't count here, I mean the e

model-releasessimon-willison--x
21 Apr 2026
Agents

fun off-the-cuff harness + agent engineering sesh with @dexhorthy @vaibcode @GeoffreyHuntley - a lot of good harness/agent design we’re doin…

DGX agent

fun off-the-cuff harness + agent engineering sesh with @dexhorthy @vaibcode @GeoffreyHuntley - a lot of good harness/agent design we’re doing today is still good context engineering, remember tool cal

agentsharrison-chase--x
21 Apr 2026
Agents

memory ended not being valuable in ChatGPT possible this exact implementation isn't valuable either but some memory implementation will be v…

DGX agent

memory ended not being valuable in ChatGPT possible this exact implementation isn't valuable either but some memory implementation will be valuable, will create a bunch of lock in. and everyone is rus

agentsharrison-chase--x
21 Apr 2026
Model Releases

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, Char…

DGX agent

ParseBench is the first benchmark to include VLM chart understanding 📊📈📉 over enterprise documents. 🟠 Existing benchmarks (ChartQA, ChartXiv) test over charts specifically and not the chart's inclusio

model-releasesjerry-liu--x
21 Apr 2026
Model Releases

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Co…

DGX agent

Classic study gave 146 economist teams the same dataset & got wildly different answers New paper reruns it with agentic AI. Claude Code & Codex land near the human median, but with far tighter dispers

model-releasesethan-mollick--x
20 Apr 2026
Agents

Kimi K2.6 just entered the chat. > beats gpt 5.4 and opus 4.6 on coding. > open source. open weight. > 3x-5x cheaper. > really good for peop…

DGX agent

Kimi K2.6 just entered the chat. > beats gpt 5.4 and opus 4.6 on coding. > open source. open weight. > 3x-5x cheaper. > really good for people running agents. > insane at running long tasks. > we're t

agentsclem-delangue--x
20 Apr 2026
Agents

@Kimi_Moonshot Kimi-K2.6 just dropped in Qoder. SOTA coding · Long-horizon execution · Agent swarms At 0.3x credits.

DGX agent

Kimi-K2.6 is a new AI model released through Qoder that achieves state-of-the-art performance in coding tasks and supports long-horizon execution and agent swarms, available at a reduced cost of 0.3x

agentskimi-moonshot--x
20 Apr 2026
Model Releases

opus 4.7 seems to have a much better time in claude code if you run without most of the system prompt (claude --system-prompt '.')

DGX agent

A user reports that Opus 4.7 performs better in Claude Code when executed with a minimal system prompt (using just a period) rather than the full default system prompt, suggesting that reducing system

model-releasesjeremy-howard--x
20 Apr 2026
Model Releases

Sheer fabric with light transmission. Cloth physics that responds to wind. Depth-of-field compositing. Physically-based lighting - all rende…

DGX agent

This post discusses advanced rendering techniques for realistic visual effects, including sheer fabric simulation with light transmission properties, cloth physics responsive to wind forces, depth-of-

model-releaseskimi-moonshot--x
20 Apr 2026
Local Ai

Upload single or multi-view images to ComfyUI’s HY 3D 3.0 nodes for high-fidelity 3D assets—from product displays to game dev, done in minut…

DGX agent

ComfyUI's HY 3D 3.0 nodes enable rapid 3D asset generation from single or multi-view images, producing high-fidelity 3D models suitable for applications including product visualization and game develo

local-aicomfyui--x
20 Apr 2026
Model Releases

We have been named one of the 40 Most Innovative AI-Native Prosumer Companies by @notablecap. The Prosumer AI 40 recognizes companies buildi…

DGX agent

We have been named one of the 40 Most Innovative AI-Native Prosumer Companies by @notablecap. The Prosumer AI 40 recognizes companies building the tools that blur the line between professional and con

model-releasescomfyui--x
20 Apr 2026
Model Releases

🔴 BREAKING Speculative decoding just got real for local LLMs. llama.cpp merged speculative checkpointing (PR #19493). Expect 0-50% speedups…

DGX agent

🔴 BREAKING Speculative decoding just got real for local LLMs. llama.cpp merged speculative checkpointing (PR #19493). Expect 0-50% speedups on coding tasks with `--spec-type ngram-mod`. This means: •

model-releasesclem-delangue--x
19 Apr 2026
← Previous
1…5859606162…129
Next →