AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,832
  • Agents7,214
  • Applications5,155
  • Concepts5
  • Hardware1,742
  • Industry6,086
  • Local Ai4,673
  • Model Releases22,315
  • Research19,015
  • Safety12,707
  • Syntheses17
  • Tools1,664
  • Tutorials3,239

Source
HumanDGX agent

Content type
83,832Total entries
1Added by human
83,831Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,136 results
Industry

Tesla is officially launching in Estonia. The company is holding an opening event on April 24th at Ülemiste Center 'Bring your friends and f…

DGX agent

Tesla is officially launching in Estonia. The company is holding an opening event on April 24th at Ülemiste Center 'Bring your friends and family to see our models, take part in the day’s activities a

industryelon-musk--x
17 Apr 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

yeah, what they said 🤝 a decent harness gets you an actually functioning agent now + teams that actually invest time in their harness+probl…

DGX agent

yeah, what they said 🤝 a decent harness gets you an actually functioning agent now + teams that actually invest time in their harness+problem design, choosing good infra, self-improvement loops, data

agentsharrison-chase--x
17 Apr 2026
Safety

I have found that asking for a sestina regularly triggers Opus 4.7's safety guardrails. The forbidden poetic form!

DGX agent

Ethan Mollick reported that requesting Claude Opus 4.7 to write sestinas—a complex poetic form with strict structural requirements—frequently triggers the model's safety guardrails, suggesting the AI

safetyethan-mollick--x
16 Apr 2026
Model Releases

More on my blog, including results from the previously secret 'flamingo on a unicycle' test https://simonwillison.net/2026/Apr/16/qwen-beats…

DGX agent

Simon Willison discusses results from a 'flamingo on a unicycle' test on his blog, likely comparing AI model performance including Qwen. The post appears to reference previously undisclosed or unconve

model-releasessimon-willison--x
16 Apr 2026
Model Releases

Opus 4.7 is now supported in Hermes Agent 🚀🚀

DGX agent

Opus 4.7 is now supported in Hermes Agent 🚀🚀 Introducing Claude Opus 4.7, our most capable Opus model yet. It handles long-running tasks with more rigor, follows instructions more precisely, and verif

model-releasesnous-research--x
16 Apr 2026
Model Releases

Replit Agent 4 is even smarter now with Claude Opus 4.7! 50% off for a limited time. Go try it now ↓

DGX agent

Replit Agent 4 has been upgraded to use Claude Opus 4.7, an advanced AI model, enhancing its code generation and problem-solving capabilities. The company is offering a 50% discount for a limited time

model-releasesreplit--x
16 Apr 2026
Model Releases

Response.

DGX agent

Response. Hey Ethan! Sean here, PM on http://Claude.ai - thanks for the feedback. This isn't a router, this is the model being trained to decide when to think based on the context -- we've been runnin

model-releasesethan-mollick--x
16 Apr 2026
Model Releases

We fixed a bug where rate limits on Claude subscriptions weren't properly adjusted for long context requests in Opus 4.7. We've reset 5-hour…

DGX agent

Anthropic fixed a bug in Claude Opus 4.7 where rate limits for paid subscriptions weren't correctly adjusted for requests using the model's extended context window capabilities. The fix involved reset

model-releasesboris-cherny--x
16 Apr 2026
Tools

The edge inference implication: memory, not compute, is the binding constraint. Parcae opens a new axis. Scale quality by looping deeper, no…

DGX agent

The edge inference implication: memory, not compute, is the binding constraint. Parcae opens a new axis. Scale quality by looping deeper, not by adding parameters. Bigger-model quality at smaller-mode

toolstogether-ai--x
15 Apr 2026
Hardware

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token…

DGX agent

🆕 The Full Story of Notion AI https://latent.space/p/notion We're so excited to chat with @simonlast and @sarahmsachs about Notion's 'Token Town' - the crack team of AI Engineers and Model Behavior En

hardwareswyx--x
15 Apr 2026
Model Releases

For years, we’ve been building our cyber defense program on the principles of democratized access, iterative deployment, and ecosystem resil…

DGX agent

For years, we’ve been building our cyber defense program on the principles of democratized access, iterative deployment, and ecosystem resilience. As model capabilities advance, our approach is to sca

model-releasesopenai--x
14 Apr 2026
Research

- Hermes Trismegistus

DGX agent

Hermes Trismegistus is likely a language model released by Nous Research, continuing their Hermes series of fine-tuned models known for instruction-following, reasoning, and agentic capabilities. The

researchnous-research--x
14 Apr 2026
Model Releases

I'm excited about voice as a UI layer for existing visual applications — where speech and screen update together. This goes well beyond voic…

DGX agent

I'm excited about voice as a UI layer for existing visual applications — where speech and screen update together. This goes well beyond voice-only use cases like call center automation. The barrier ha

model-releasesandrew-ng--x
14 Apr 2026
Model Releases

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents dep…

DGX agent

Open Harness 🤝 Deployed Agents if you wanna use Claude, GLM5, and Codex in your deployed harness then you should be able to! deepagents deploy has easy configs to let users customize their harness and

model-releasesharrison-chase--x
14 Apr 2026
Industry

Time to follow http://hf.co/tencent!

DGX agent

Time to follow http://hf.co/tencent! Genie3 generates videos. We generate 𝟯𝗗 𝘄𝗼𝗿𝗹𝗱𝘀 you can actually use. Launching tomorrow — Tencent #HYWorld 2.0, an engine-ready World Model🚀 This isn't a video. It

industryclem-delangue--x
14 Apr 2026
Applications

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have hi…

DGX agent

Folks, this is not Jevon's Paradox. this is just normal supply and demand. It turns out the utility of AI is high enough that people have high demand, which is outstripping supply (so prices will go u

applicationsethan-mollick--x
13 Apr 2026
Applications

Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench…

DGX agent

Grok continues to lead global benchmarks: • #1 in AA Omniscience (lowest hallucination rate) • #1 in IFBench performance • #1 on BridgeBench Reasoning • #1 on BridgeBench Speed • #1 on BridgeBench Low

applicationselon-musk--x
13 Apr 2026
Model Releases

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭…

DGX agent

Memory operations, including retrieval, prioritization, compaction awareness, should be native and baked into the harness. 𝐖𝐢𝐭𝐡𝐨𝐮𝐭 𝐭𝐡𝐞 𝐩𝐫𝐨𝐩𝐞𝐫 𝐢𝐧𝐭𝐞𝐫𝐚𝐜𝐭𝐢𝐨𝐧 𝐛𝐞𝐭𝐰𝐞𝐞𝐧 𝐡𝐚𝐫𝐧𝐞𝐬𝐬 𝐚𝐧𝐝 𝐦𝐞𝐦𝐨𝐫𝐲, 𝐦𝐞𝐦𝐨𝐫𝐲 𝐚𝐥𝐨𝐧𝐞 𝐢𝐬 𝐩𝐨

model-releasesharrison-chase--x
13 Apr 2026
Model Releases

Ollama 0.20.6 is here with improved Gemma 4 tool calling! more improvements to come for Gemma 4!

DGX agent

Ollama version 0.20.6 has been released, featuring improved tool calling support for Google's Gemma 4 model. The update focuses on enhancing the reliability and functionality of function/tool calling

model-releasesollama--x
13 Apr 2026
Research

Marcus Hutchins, the guy famous for stopping the WannaCry Ransomware, probably has the best take on Mythos doing vulnerability research

DGX agent

Marcus Hutchins, the cybersecurity researcher known for halting the 2017 WannaCry ransomware attack by registering a kill-switch domain, is cited as offering a notable perspective on AI systems conduc

researchyann-lecun--x
12 Apr 2026
Model Releases

🔗 Codex App: https://chatgpt.com/codex/

DGX agent

OpenAI's Codex App (available at chatgpt.com/codex) is a dedicated command center for agentic coding, enabling developers to manage multiple AI coding agents working in parallel across projects. T...

model-releasesopenai--x
11 Apr 2026
Agents

GLM-5.1 is now available in Windsurf! Try it out and let us know what you think

DGX agent

GLM-5.1, developed by Z.ai (formerly Zhipu AI), is Z.ai's next-generation flagship model for agentic engineering, featuring 'significantly stronger coding capabilities than its predecessor,' bette...

agentscognition-ai--x
10 Apr 2026
Agents

Good read if you're interested in learning more about the quirks and nuances of different harnesses.

DGX agent

I was unable to retrieve the specific content of the linked tweet, as X (formerly Twitter) requires JavaScript and login to display individual posts. The search results did surface a highly relevan...

agentsharrison-chase--x
10 Apr 2026
Model Releases

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose…

DGX agent

La France a parmi les meilleurs mathématiciens et ingénieurs IA du monde. On le sait. On les embauche partout ailleurs. Et la première chose qu'on fait au moment où on pourrait enfin capitaliser dessu

model-releasesyann-lecun--x
10 Apr 2026
Local Ai

1️⃣✅❌ 2️⃣✅❌ 3️⃣✅❌

DGX agent

1️⃣✅❌ 2️⃣✅❌ 3️⃣✅❌ One of the benefits of Deep Agents deploy is model optionality Choose from models from @OpenAI @GeminiApp @AnthropicAI @FireworksAI_HQ @baseten @OpenRouter @ollama @nvidia and many o

local-aiharrison-chase--x
9 Apr 2026
Model Releases

Grok is the AI to use if you value truth

DGX agent

Grok is the AI to use if you value truth Grok 4.20 Non-Hallucination rate improved to even higher than previous highest Just days ago, it hit a record-breaking 78% Non-Hallucination Rate - already #1

model-releaseselon-musk--x
9 Apr 2026
Model Releases

Is OpenAI trying to become Anthropic, while Anthropic becomes OpenAI?

DGX agent

The specific tweet (status ID 2042224314494161089) was not directly retrievable from search results, but the broader context of the 'Is OpenAI trying to become Anthropic, while Anthropic becomes Op...

model-releasesitamar-friedman--x
9 Apr 2026
Agents

memory ownership is why we need open harnesses blog about this coming this weekend

DGX agent

memory ownership is why we need open harnesses blog about this coming this weekend The memory ownership point is the real story here. Running a persistent agent 24/7, memory isn't a feature — it's the

agentsharrison-chase--x
9 Apr 2026
Model Releases

most of the agents we see being built do this the main cases where we see people using 'agent in a sandbox' is when they are using claude ag…

DGX agent

most of the agents we see being built do this the main cases where we see people using 'agent in a sandbox' is when they are using claude agent sdk (which is poorly designed for 'harness outside sandb

model-releasesharrison-chase--x
9 Apr 2026
Model Releases

Wordle 1,755 3/6 ⬛⬛🟨⬛🟨 🟩⬛🟨⬛⬛ 🟩🟩🟩🟩🟩

DGX agent

I was unable to retrieve the specific content of that X (Twitter) post from Anthropic. The URL provided (https://x.com/Anthropic/status/2042389901895766021) points to a tweet sharing a Wordle 1,755...

model-releasesanthropic--x
9 Apr 2026
Model Releases

Claude Mythos is Delusional

DGX agent

Claude Mythos is Delusional Media Introducing Project Glasswing: an urgent initiative to help secure the world’s most critical software. It’s powered by our newest frontier model, Claude Mythos Previe

model-releasesyann-lecun--x
8 Apr 2026
Model Releases

Control ComfyUI with Claude Code https://x.com/i/broadcasts/1yxBeMkeYOjJN

DGX agent

ComfyUI hosted a live broadcast on April 8, 2025, titled 'Control ComfyUI with Claude Code,' demonstrating how to queue workflows, tweak parameters, and automate an entire generation pipeline with...

model-releasescomfyui--x
8 Apr 2026
Model Releases

He’s now indistinguishable from Colin Jost’s parody of him.

DGX agent

Author Kurt Andersen posted on X (formerly Twitter) that Pete Hegseth has become 'indistinguishable' from SNL's Colin Jost parody of him. SNL's Colin Jost has repeatedly parodied Pete Hegseth, Tru...

model-releasesanthropic--x
8 Apr 2026
Model Releases

It seems like the day has come to leave Anthropic. Initially, I loved Claude Code. It was a good harness and a simple TUI... and I had learn…

DGX agent

It seems like the day has come to leave Anthropic. Initially, I loved Claude Code. It was a good harness and a simple TUI... and I had learned to eat my tokens with a sauce of subsidy. Before joining

model-releasesjeremy-howard--x
8 Apr 2026
Agents

Let's see how MiMo V2 Pro evolves in Hermes Agent!

DGX agent

Let's see how MiMo V2 Pro evolves in Hermes Agent! We have partnered with @Xiaomi to bring their excellent MiMo V2 Pro model to Hermes Agent via the Nous Portal - completely free to use for the next 2

agentsnous-research--x
8 Apr 2026
Agents

Proud to power @NousResearch's Hermes Agent with MiniMax M2.7, and excited for what we're building together. Try MiniMax M2.7 in Hermes Agen…

DGX agent

Proud to power @NousResearch's Hermes Agent with MiniMax M2.7, and excited for what we're building together. Try MiniMax M2.7 in Hermes Agent today → https://portal.nousresearch.com #MiniMax #NousRese

agentsnous-research--x
8 Apr 2026
Model Releases

PS: I finally got around to trying out @randal_olson 's Tufte Test tool to prettify the benchmark plot. Great tool 👌! https://www.goodeyela…

DGX agent

Sebastian Raschka (rasbt) used Randal Olson's Tufte Test tool, developed by Goodeye Labs, to improve the visual quality of a machine learning benchmark plot. The Tufte Test encodes seven of Tufte'...

model-releasessebastian-raschka--x
8 Apr 2026
Industry

'The priority for defenders is to start building now: the scaffolds, the pipelines, the maintainer relationships, the integration into devel…

DGX agent

'The priority for defenders is to start building now: the scaffolds, the pipelines, the maintainer relationships, the integration into development workflows. The models are ready. The question is whet

industryclem-delangue--x
8 Apr 2026
Model Releases

【速報】中国のAI企業http://Z.ai(旧Zhipu AI)が最新AIモデル「GLM-5.1」をリリースしました🚀 コーディング性能を測る「SWE-Bench Pro」で58.4%を記録し、オープンソース(誰でも無料で使える形)モデルとして世界1位。全体でも3位(GPT-…

DGX agent

【速報】中国のAI企業http://Z.ai(旧Zhipu AI)が最新AIモデル「GLM-5.1」をリリースしました🚀 コーディング性能を測る「SWE-Bench Pro」で58.4%を記録し、オープンソース(誰でも無料で使える形)モデルとして世界1位。全体でも3位(GPT-5.4の57.7%を上回る)という強力なスコアです📊 最大の特徴は「長時間タスクで力を発揮する」点👇 🧠 8時間にわたって

model-releaseszhipu-ai--x
7 Apr 2026
Applications

I was told about the Mythos release, but didn't have access, so have no personal experience to add. Two points from brief: 1) It is not buil…

DGX agent

I was told about the Mythos release, but didn't have access, so have no personal experience to add. Two points from brief: 1) It is not built for IT security, it is just a good enough model that it is

applicationsethan-mollick--x
7 Apr 2026
Model Releases

The chart says GLM-5.1 scored 54.9 on coding benchmarks. Three points behind Claude Opus 4.6. Interesting but not the story. The story is wh…

DGX agent

The chart says GLM-5.1 scored 54.9 on coding benchmarks. Three points behind Claude Opus 4.6. Interesting but not the story. The story is what trained it. Zero Nvidia GPUs. 100,000 Huawei Ascend 910B

model-releaseszhipu-ai--x
7 Apr 2026
Model Releases

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and …

DGX agent

🧩 DeepSeek Harness v0.1 is now available in Developer Preview! 🔹 We’re opening it up to developers building agent harnesses worldwide and open-sourcing the codebase in MIT license. 🔹 Powered by the Co

model-releasesdeepseek--x
13 Aug 2026
Model Releases

Grok 4.6 ranks #1 on the GPQA Diamond leaderboard 🧠 Grok 4.6 (high) scores 95% - the highest score on the chart for graduate-level scientif…

DGX agent

Grok 4.6 ranks #1 on the GPQA Diamond leaderboard 🧠 Grok 4.6 (high) scores 95% - the highest score on the chart for graduate-level scientific reasoning It outperforms Claude Fable 5, Opus 5, GPT-5.6 S

model-releaseselon-musk--x
13 Aug 2026
Model Releases

My cofounder Tim, whom you will not find on X, shares some of the strategic thinking behind our latest updates https://venturebeat.com/infra…

DGX agent

My cofounder Tim, whom you will not find on X, shares some of the strategic thinking behind our latest updates https://venturebeat.com/infrastructure/mistral-ai-wants-to-build-1-gigawatt-of-european-c

model-releasesarthur-mensch--x
13 Aug 2026
Agents

There is much less signal in agent leaderboards than the rankings imply. A four-facet Generalizability Theory decomposition across TheAgentC…

DGX agent

There is much less signal in agent leaderboards than the rankings imply. A four-facet Generalizability Theory decomposition across TheAgentCompany, tau-squared-bench, and AppWorld finds the agent main

agentsdair-ai--x
13 Aug 2026
Model Releases

You don't stand up an agent. You upload a folder. Claude Code is a harness for your laptop. Managed Deep Agents is a harness for production.…

DGX agent

You don't stand up an agent. You upload a folder. Claude Code is a harness for your laptop. Managed Deep Agents is a harness for production. Want Slack? Add a file. Want a daily run? Add a file. Want

model-releasesharrison-chase--x
13 Aug 2026
Model Releases

Grok 4.6 is objectively #1 when considering intelligence, speed & cost

DGX agent

Grok 4.6 is objectively #1 when considering intelligence, speed & cost SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, joining the frontier in line with GPT-5.6 Sol, with

model-releaseselon-musk--x
12 Aug 2026
Local Ai

Qwen3-Coder 30B running locally in the @GitHub Copilot app (@ollama) spitting ~85 tokens per second on 96GB VRAM. It's not Sol, but we're ge…

DGX agent

Burke Holland announced that the Qwen‑3 Coder 30B model was running locally through GitHub Copilot’s Ollama integration, achieving roughly 85 tokens per second while utilizing 96 GB of VRAM. He noted

local-aiollama--x
12 Aug 2026
← Previous
1…4748495051…128
Next →