AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

Search: “agents”

GridTimelineEvolution
4,322 results
Model Releases

NVIDIAとSakana AI、オープンモデルによるイノベーションのため協業拡大 本日、Sakana AIはNVIDIAとのコラボレーションを強化し、日本発の「集合知」の取り組みを次なるフェーズへ進めることを発表します。 私たちのマルチエージェント・オーケストレーションシステム…

DGX agent

NVIDIAとSakana AI、オープンモデルによるイノベーションのため協業拡大 本日、Sakana AIはNVIDIAとのコラボレーションを強化し、日本発の「集合知」の取り組みを次なるフェーズへ進めることを発表します。 私たちのマルチエージェント・オーケストレーションシステム「Sakana Fugu」に、Nemotronファミリーを含むNVIDIAのオープンモデル群を統合します。 Sakana

model-releasesdavid-ha--x
16 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

We’re excited to collaborate with NVIDIA to build the next generation of Fugu orchestration models together, by incorporating leading open-w…

DGX agent

We’re excited to collaborate with NVIDIA to build the next generation of Fugu orchestration models together, by incorporating leading open-weights models. Sakana AI Teams With NVIDIA to Advance Open M

model-releasesdavid-ha--x
16 Jul 2026
Model Releases

🥉 3rd place: CashFromChaos, by David Diaz (@davddiazm) CashFromChaos starts from a single seller input and automates everything up until a …

DGX agent

🥉 3rd place: CashFromChaos, by David Diaz (@davddiazm) CashFromChaos starts from a single seller input and automates everything up until a completed sale. You send a photo and a one-line clue, and Her

model-releasesnous-research--x
15 Jul 2026
Model Releases

BREAKING: Grok 4.5 has climbed to #2 on the FrontierSWE benchmark. The result places Grok 4.5 among the world's top-performing AI models for…

DGX agent

BREAKING: Grok 4.5 has climbed to #2 on the FrontierSWE benchmark. The result places Grok 4.5 among the world's top-performing AI models for software engineering tasks, highlighting its growing streng

model-releaseselon-musk--x
15 Jul 2026
Model Releases

Good reason to try Grok 4.5 with Grok Build. It gets better every day!

DGX agent

Good reason to try Grok 4.5 with Grok Build. It gets better every day! Grok 4.5 just took the #1 spot on the Long-Horizon Terminal-Bench, outperforming Claude Fable 5, Claude Opus 4.8 and GPT-5.6-sol

model-releaseselon-musk--x
15 Jul 2026
Model Releases

I’ve found @angjiang and @katelyn_lesse of @AnthropicAI to be generous, transparent, and thoughtful when it comes to building an ecosystem, …

DGX agent

I’ve found @angjiang and @katelyn_lesse of @AnthropicAI to be generous, transparent, and thoughtful when it comes to building an ecosystem, not a walled garden. Listen and decide for yourself. Just so

model-releasessonya-huang--x
15 Jul 2026
Model Releases

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Cod…

DGX agent

🆕This Year In Claude https://www.youtube.com/watch?v=uU5Gv2h8-9g @simonw chats with @_catwu and @trq212 about the state of: - @claudeai Code - Claude Fable - @anthropicai culture & product strategy -

model-releasesswyx--x
15 Jul 2026
Applications

Training against GPT‑Red makes GPT‑5.6 substantially more resilient. To measure this, we replayed some of GPT‑Red’s strongest attacks—none o…

DGX agent

Training against GPT‑Red makes GPT‑5.6 substantially more resilient. To measure this, we replayed some of GPT‑Red’s strongest attacks—none of which our models had seen during training. GPT‑5.6 Sol pro

applicationsopenai--x
15 Jul 2026
Local Ai

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai fam…

DGX agent

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai family. Based on Qwen3.6 27B, it brings a new capability tier t

local-aiclem-delangue--x
14 Jul 2026
Model Releases

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexi…

DGX agent

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexity Computer. https://research.perplexity.ai/articles/wandr-b

model-releasesperplexity--x
14 Jul 2026
Model Releases

What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • ⁠Qwen 3.6 35b NVFP4 — 256k ctx, 81 …

DGX agent

What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • ⁠Qwen 3.6 35b NVFP4 — 256k ctx, 81 tok/s • ⁠Qwen 3.6 27b NVFP4 — 256k ctx, 33 tok/s 2× DGX Spar

model-releasesclem-delangue--x
14 Jul 2026
Model Releases

Standard RL benchmarks are episodic and stationary, so they don't capture the the characteristics of real-world deployment. Morpheus is a ne…

DGX agent

Standard RL benchmarks are episodic and stationary, so they don't capture the the characteristics of real-world deployment. Morpheus is a new benchmark for continual learning that provides persistent

model-releasesfrancois-chollet--x
13 Jul 2026
Model Releases

I don't think Anthropic realizes how disruptive these changes are to users. I appreciate the extension, but please stop playing games. Eithe…

DGX agent

I don't think Anthropic realizes how disruptive these changes are to users. I appreciate the extension, but please stop playing games. Either keep it under the subscriptions or put it under the API al

model-releasesjeremy-howard--x
12 Jul 2026
Research

VLMは人間のような創造性を持てるか? ケネス・スタンレー教授らの『目標という幻想(Why Greatness Cannot Be Planned)』は、明確な目標を設定することが、かえって真に偉大な発見を遠ざけてしまうという逆説を論じた書籍です。その議論の中核にあったのが「Pi…

DGX agent

VLMは人間のような創造性を持てるか? ケネス・スタンレー教授らの『目標という幻想(Why Greatness Cannot Be Planned)』は、明確な目標を設定することが、かえって真に偉大な発見を遠ざけてしまうという逆説を論じた書籍です。その議論の中核にあったのが「PicBreeder」の実験でした。 PicBreeder では、ユーザーが「面白い」と感じた画像を選び、それを少しずつ進化

researchdavid-ha--x
11 Jul 2026
Model Releases

Another big langchain week

DGX agent

Another big langchain week 🚀langchain launches this week: all about open source models and memory! First: open source models. We partnered with @NVIDIAAI to launch a NemoClaw DeepAgents blueprint. Thi

model-releasesharrison-chase--x
10 Jul 2026
Local Ai

Call the homies, new @UnslothAI NVFP4 just dropped 🔥

DGX agent

Call the homies, new @UnslothAI NVFP4 just dropped 🔥 We’re releasing new Qwen3.6 quants that run 2.5× faster on your GPU. Qwen3.6-27B NVFP4 runs on 24GB VRAM. 35B-A3B can hit 17,561 tok/s (B200). We a

local-aiclem-delangue--x
10 Jul 2026
Model Releases

GROK 4.5 LEADS ON REAL PROFESSIONAL WORK BENCHMARK New data from Snorkel shows Grok 4.5 outperforming other frontier models on real-world pr…

DGX agent

GROK 4.5 LEADS ON REAL PROFESSIONAL WORK BENCHMARK New data from Snorkel shows Grok 4.5 outperforming other frontier models on real-world professional tasks. On their GDPval+ benchmark (expert-created

model-releaseselon-musk--x
10 Jul 2026
Tutorials

'I think it's important for people to understand how code works.' Geoffrey Litt's Design Eng track keynote is live now: https://www.youtube.…

DGX agent

'I think it's important for people to understand how code works.' Geoffrey Litt's Design Eng track keynote is live now: https://www.youtube.com/watch?v=WkBPX-oDMnA Thank you @NotionHQ for supporting h

tutorialsswyx--x
10 Jul 2026
Model Releases

OpenWiki general purpose memory is meant to be complementary to codex/claude code memory: it's proactive & ambient, meaning it'll automatica…

DGX agent

OpenWiki general purpose memory is meant to be complementary to codex/claude code memory: it's proactive & ambient, meaning it'll automatically go out into your world (via connections like gmail, x, n

model-releasesharrison-chase--x
10 Jul 2026
Tutorials

A Visual Introduction to Information Theory (bookmark it) Information Theory is such an beautiful and powerful subject. In the era of AI, it…

DGX agent

A Visual Introduction to Information Theory (bookmark it) Information Theory is such an beautiful and powerful subject. In the era of AI, it's worth spending time learning about it. Here is a highly-r

tutorialsdair-ai--x
9 Jul 2026
Model Releases

ChatGPT Work == Claude Cowork ChatGPT Codex == Claude Code I kinda wish OpenAI created a single unified app surface for all work, coding or …

DGX agent

ChatGPT Work == Claude Cowork ChatGPT Codex == Claude Code I kinda wish OpenAI created a single unified app surface for all work, coding or not, even though I get the UI/UX would be different Introduc

model-releasesjerry-liu--x
9 Jul 2026
Applications

create a really interesting game called 'Don't Discuss Goblins' that should have some story elements but mostly be fun and fast moving - the…

DGX agent

create a really interesting game called 'Don't Discuss Goblins' that should have some story elements but mostly be fun and fast moving - the goal is to avoid mentioning goblins, and this should be cha

applicationsethan-mollick--x
9 Jul 2026
Model Releases

Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it…

DGX agent

Is Meta AI back? Haven't seen Mark post in three years here. Plus, the model is available via API. Not to mention the courage to announce it the same week as the long-awaited GPT-5.6. Great timing if

model-releasesdair-ai--x
9 Jul 2026
Model Releases

OpenWiki Brains 0.1.0 is officially released! We added a general-purpose memory brain to OpenWiki, in addition to the existing code brain. Y…

DGX agent

OpenWiki Brains 0.1.0 is officially released! We added a general-purpose memory brain to OpenWiki, in addition to the existing code brain. You can now use it to seamlessly setup a personal brain to tr

model-releasesharrison-chase--x
9 Jul 2026
Model Releases

SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at …

DGX agent

SpaceXAI's Grok 4.5 takes the #1 spot on AutomationBench-AA with a score of 51%, ahead of Claude Fable 5 (49%) and Claude Opus 4.8 (48%) at roughly a quarter of their cost per task - the first model t

model-releaseselon-musk--x
9 Jul 2026
Model Releases

// The Harness Effect // (bookmark it) Now more that ever pay very close attention to the orchestration harness and its effect on costs and …

DGX agent

// The Harness Effect // (bookmark it) Now more that ever pay very close attention to the orchestration harness and its effect on costs and performance. This study ran 22 evaluation tasks on six found

model-releasesdair-ai--x
9 Jul 2026
Model Releases

Grok 4.5 brings frontier performance across coding and knowledge work

DGX agent

Grok 4.5 brings frontier performance across coding and knowledge work SpaceXAI’s Grok 4.5 scores 54 to place fourth on the Artificial Analysis Intelligence Index following only Fable 5, GPT-5.5, and O

model-releaseselon-musk--x
8 Jul 2026
Model Releases

Grok 4.5 context window will upgrade to 1M probably by next week

DGX agent

Grok 4.5 context window will upgrade to 1M probably by next week SpaceXAI’s Grok 4.5 scores 54 to place fourth on the Artificial Analysis Intelligence Index following only Fable 5, GPT-5.5, and Opus 4

model-releaseselon-musk--x
8 Jul 2026
Model Releases

Not another demo or benchmark. @ShoucongChen is a senior member of our technical staff. A real project, scoped at 1-month. Delivered in 4 da…

DGX agent

Not another demo or benchmark. @ShoucongChen is a senior member of our technical staff. A real project, scoped at 1-month. Delivered in 4 days with GLM5.2 Fast. The best devs deserve >400 t/sec. Take

model-releasesfireworks-ai--x
8 Jul 2026
Tutorials

We are excited to launch 𝗥𝗲𝘀𝘁𝗮𝘁𝗲 𝗕𝗬𝗢𝗖 (Bring-your-own-Cloud) today. It is fully managed @restatedev, with all the features of Res…

DGX agent

We are excited to launch 𝗥𝗲𝘀𝘁𝗮𝘁𝗲 𝗕𝗬𝗢𝗖 (Bring-your-own-Cloud) today. It is fully managed @restatedev, with all the features of Restate Cloud, but in your own account, in a dedicated VPC. Data never lea

tutorialsswyx--x
8 Jul 2026
Applications

We are hiring for @Harvey’s model training team. This team will help Harvey expand from the application layer into the model layer and from …

DGX agent

We are hiring for @Harvey’s model training team. This team will help Harvey expand from the application layer into the model layer and from legal into high end knowledge work more broadly. We are hiri

applicationsharrison-chase--x
8 Jul 2026
Model Releases

We need AI model selection to be MUCH easier ASAP. I want proactive flags from my AI systems suggesting models. I want my AI harness to say …

DGX agent

We need AI model selection to be MUCH easier ASAP. I want proactive flags from my AI systems suggesting models. I want my AI harness to say 'hey allie, my girl, you keep asking for bar recommendations

model-releasesallie-k--miller--x
8 Jul 2026
Model Releases

Xiaomi now processes more AI tokens than OpenAI. On OpenRouter, Chinese models just crossed 45% of all token volume. Anthropic is at 15.3%. …

DGX agent

Xiaomi now processes more AI tokens than OpenAI. On OpenRouter, Chinese models just crossed 45% of all token volume. Anthropic is at 15.3%. OpenAI is at 7.4%. For now, the frontier is American models,

model-releasesallie-k--miller--x
8 Jul 2026
Model Releases

NEW AI paper worth bookmarking. This is something I called early, and this paper confirms it: verification has emerged as a new important sc…

DGX agent

NEW AI paper worth bookmarking. This is something I called early, and this paper confirms it: verification has emerged as a new important scaling axis. Here is the simple explainer and what this paper

model-releasesdair-ai--x
7 Jul 2026
Model Releases

nvidia ai just handed dgx spark owners a real gift. nemotron labs 3 puzzle 75b a9b nvfp4 is basically built for this box. 75b total, 9.3b ac…

DGX agent

nvidia ai just handed dgx spark owners a real gift. nemotron labs 3 puzzle 75b a9b nvfp4 is basically built for this box. 75b total, 9.3b active, nvfp4, mamba plus moe, 256k context in the config, and

model-releasesclem-delangue--x
7 Jul 2026
Model Releases

The AI labs desperately need non-engineer Peters, Borises, and Thariqs. Most demos for 'business users' are about replying to emails or to S…

DGX agent

The AI labs desperately need non-engineer Peters, Borises, and Thariqs. Most demos for 'business users' are about replying to emails or to Slack. And yes, that's helpful to manage the cacophonous hell

model-releasesallie-k--miller--x
7 Jul 2026
Applications

I talk to a lot of companies that still have active efforts to build GPTs. (It remains weird that OpenAI abandoned GPTs after rolling them o…

DGX agent

I talk to a lot of companies that still have active efforts to build GPTs. (It remains weird that OpenAI abandoned GPTs after rolling them out. They were the precursor to Skills & could have been a br

applicationsethan-mollick--x
6 Jul 2026
Tutorials

// In-context Retrieval at Million-token Scale // Great study providing better understanding of retrieval at million-token scale. They run t…

DGX agent

// In-context Retrieval at Million-token Scale // Great study providing better understanding of retrieval at million-token scale. They run the first systematic study of in-context retrieval at the sca

tutorialsdair-ai--x
6 Jul 2026
Tutorials

// ReContext // Models now support 128K context windows and still fail to use evidence that is already in the prompt. Where is the gap? New …

DGX agent

// ReContext // Models now support 128K context windows and still fail to use evidence that is already in the prompt. Where is the gap? New paper introduces ReContext, a training-free inference harnes

tutorialsdair-ai--x
6 Jul 2026
Safety

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release …

DGX agent

Another major Grok Build update just landed, packed with new features, extensive bug fixes, and meaningful performance improvements Release Notes: v0.2.84 — 2026-07-03 Features: • Announcements now up

safetyelon-musk--x
3 Jul 2026
Model Releases

Highly-recommended read from MIT on the part of RL with verifiable rewards that everyone keeps hitting. RLVR only optimizes what you can obj…

DGX agent

Highly-recommended read from MIT on the part of RL with verifiable rewards that everyone keeps hitting. RLVR only optimizes what you can objectively score, so style, structure, and diversity quietly c

model-releasesdair-ai--x
3 Jul 2026
Tutorials

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to rec…

DGX agent

NEW paper worth reading. (bookmark it) The basic idea is to pair a compressive recurrent state with a small exact memory, which helps to recover long-range recall without giving up the efficiency of l

tutorialsdair-ai--x
3 Jul 2026
Model Releases

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as…

DGX agent

This is true… but maybe less important than the fact that people don’t try ambitious things with these systems. Many models are excellent as a Google replacement, for homework “help,” etc. It is someo

model-releasesethan-mollick--x
3 Jul 2026
Model Releases

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their…

DGX agent

Another fascinating paper on LLM Judges. (bookmark it) It's from Amazon, and they show that if you run panels of LLM judges, averaging their scores is a trap. 'Overall, we establish that robust aggreg

model-releasesdair-ai--x
2 Jul 2026
Model Releases

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class …

DGX agent

I just left the final day of the @aiDotEngineer World's Fair Conference in San Francisco. Kudos to @swyx for putting together a world-class lineup of speakers and workshops! It really was an invigorat

model-releasesswyx--x
2 Jul 2026
Applications

My one serious piece of advice having used Fable a bunch before release is that, unless you are careful it develops its own internal bizarre…

DGX agent

My one serious piece of advice having used Fable a bunch before release is that, unless you are careful it develops its own internal bizarre cadence & dialogue over long tasks. If you aren't asking it

applicationsethan-mollick--x
2 Jul 2026
Safety

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming f…

DGX agent

NEW paper from NVIDIA. They discuss robot programming that compounds experience instead of throwing it away. Traditional robot programming forces you to orchestrate perception, contact dynamics, diver

safetydair-ai--x
2 Jul 2026
Tutorials

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes …

DGX agent

New research from Google. LLMs hallucinate with high confidence, miss their own knowledge boundaries, and misreport uncertainty. Most fixes bolt calibration on from the outside. RLMF turns the model o

tutorialsdair-ai--x
2 Jul 2026
← Previous
1…8283848586…91
Next →