AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
Human
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
13,897 results
12 Aug 2026

18 two-word AI prompts I'm kind of obsessed with: 1) now what - great for when you've wrapped up a project or big push and you still have en…

Model ReleasesDGX agent

18 two-word AI prompts I'm kind of obsessed with: 1) now what - great for when you've wrapped up a project or big push and you still have energy and want AI to give you more 2) plz fix - usually accom

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data producti…

AgentsDGX agent

A model is only as good as its data, and we’ve long since exhausted the internet. From here on out, model progress is gated by data production. @mercor_ai’s @BrendanFoody joined us at our Sovereign AI

// Actions Speak Louder Than Words // Multilingual agent evaluation compares final answers and throws the trajectory away. The trajectory fi…

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
SafetyDGX agent

// Actions Speak Louder Than Words // Multilingual agent evaluation compares final answers and throws the trajectory away. The trajectory fixes cost, latency, failure mode, and auditability. New resea

DHS has lied about these incidents over and over again and been caught in the lies over and over again. They are lying now, and nobody shoul…

AgentsDGX agent

DHS has lied about these incidents over and over again and been caught in the lies over and over again. They are lying now, and nobody should believe them. NOW: DHS response to Virginia woman's video

Everyone else is talking about building ASI to like monopolize b2b saas and Elon is talking about building a kardashev II sentient sun

Model ReleasesDGX agent

Everyone else is talking about building ASI to like monopolize b2b saas and Elon is talking about building a kardashev II sentient sun Media It’s always funny how people in SF twitter bubble will say

Expedia recently moved its ranking models to a state-of-the-art Keras 3 setup. Results: 30% faster training, and inference latency decreased…

ResearchDGX agent

Expedia recently moved its ranking models to a state-of-the-art Keras 3 setup. Results: 30% faster training, and inference latency decreased by 70%. Read their writeup about the upgrade: https://mediu

Four small architecture decisions can cost up to 47% of a model's long-context performance. New research from Ai2, Carnegie Mellon, and the …

Model ReleasesDGX agent

Four small architecture decisions can cost up to 47% of a model's long-context performance. New research from Ai2, Carnegie Mellon, and the University of Washington isolates them. Normalization, GQA,

From #22 to #4 on Legal Research Bench! Solid progress for Qwen3.8-Max. Thanks for highlighting~✨

Model ReleasesDGX agent

From #22 to #4 on Legal Research Bench! Solid progress for Qwen3.8-Max. Thanks for highlighting~✨ Qwen 3.8 Max nearly doubled its score on Legal Research Bench in under three months, climbing from #22

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration.

Model ReleasesDGX agent

give 4.6 a try and let us know how it goes. your feedback is a big part of why the model gets better with each iteration. SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, j

Grok 4.6 is an excellent model. I’ve been using it heavily for the past couple of weeks and it handles everything from simple coding & code …

Model ReleasesDGX agent

Grok 4.6 is an excellent model. I’ve been using it heavily for the past couple of weeks and it handles everything from simple coding & code review all the way to designing and debugging complex system

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Cla…

Model ReleasesDGX agent

Grok 4.6 is now one of the top models in the world for agentic workflows It ranks #1 on the Artificial Analysis Agentic Index, tied with Claude Opus 5 Max Agentic AI is about more than answering quest

Grok 4.6 is objectively #1 when considering intelligence, speed & cost

Model ReleasesDGX agent

Grok 4.6 is objectively #1 when considering intelligence, speed & cost SpaceXAI's Grok 4.6 scores 61 on the Artificial Analysis Intelligence Index, joining the frontier in line with GPT-5.6 Sol, with

Grok 4.6 reaches 1753 ELO

Model ReleasesDGX agent

Grok 4.6 topped the GDPVal-AA benchmark with an Elo score of 1,753. It surpassed competitors Fable 5 Max (1,741 Elo), GPT‑5.6 Sol Max (1,728 Elo) and Grok 4.5 High (1,526 Elo). Elon Musk publicly ackn

Grok Bot

Model ReleasesDGX agent

Grok Bot Here's my Grok Bot team: - Webby: Web designer - Shotry: Short-form content creator - Writey: Article/Newsletter writer - Claude Code: Grok agent that specializes in CC - Codex: Same as the a

I believe the demand for compute is going to go up faster than the supply of compute, so the price of compute is going to increase significa…

AgentsDGX agent

I believe the demand for compute is going to go up faster than the supply of compute, so the price of compute is going to increase significantly in the future, perhaps as much as 10x in the next few y

If you’re interested in checking out LlamaParse for document extraction, sign up here: https://login.llamaindex.ai/sign-up

AgentsDGX agent

A 36‑page ArXiv whitepaper titled **ExtractBench** was released by Jerry Liu (jerryjliu0), describing a large‑scale, schema‑guided benchmark for real‑world document extraction from complex enterprise

imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here …

Model ReleasesDGX agent

imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here to top it at a great price, congrats to @SpaceXAI team & loo

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show ope…

Model ReleasesDGX agent

Interesting research suggests caution in determining which AI company is winning by looking at any one source.. OpenRouter seems to show open weights winning over time, but work submitted to Pangram i

***Lights, inference, action…*** I’m so happy to share that @sequoia has led the Seed in @previewio. Developers are flying in magical AI-nat…

ApplicationsDGX agent

***Lights, inference, action…*** I’m so happy to share that @sequoia has led the Seed in @previewio. Developers are flying in magical AI-native editors. But creative tooling is still stuck in the pre-

Live now: our Memory & Continual Learning Track from AI Engineer World's Fair 2026. Thesis: we scaled intelligence and got the world's smart…

TutorialsDGX agent

Live now: our Memory & Continual Learning Track from AI Engineer World's Fair 2026. Thesis: we scaled intelligence and got the world's smartest novice. https://www.youtube.com/watch?v=iqloyWCGYQQ&list

Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-…

Local AiDGX agent

Local AI is exploding! Transformers.js, that we've been building @huggingface for the past three years has now become the most popular open-source library to run AI models directly in your browser and

Most semantic search queries leave something unstated. The user knows what they mean. The system doesn't. Pinecone's text match filters scop…

AgentsDGX agent

Most semantic search queries leave something unstated. The user knows what they mean. The system doesn't. Pinecone's text match filters scope a vector search to a lexical condition (a machine number,

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agent…

AgentsDGX agent

Muse Glimmer is live on Fireworks. The new open-weight model from Meta Superintelligence Labs is a 30B dense model built for always-on agents that reason across many sequential tool calls and can reco

Qwen3.8-2.4T-A95B is now live on Fireworks with Day-0 support! This 2.4T parameter MoE model is built for autonomous agents, heavy coding, l…

Model ReleasesDGX agent

Qwen3.8-2.4T-A95B is now live on Fireworks with Day-0 support! This 2.4T parameter MoE model is built for autonomous agents, heavy coding, large context windows, and is ideal for coding and agentic pe

Qwen3.8-2.4T-A95B is now live on Together AI. The Qwen Team’s latest flagship model is built for coding and long-horizon agent workflows, wi…

Model ReleasesDGX agent

The Qwen Team has released its flagship model, Qwen3.8‑2.4T‑A95B, on the Together AI platform (togethercompute) as of August 12 2026. This 2.4‑trillion‑parameter model is engineered for coding tasks a

SPACEXAI: Grok 4.6 leads on the two strongest knowledge-work / real-world productivity benchmarks (GDPVal-AA and AA-Briefcase) and on the le…

Model ReleasesDGX agent

SPACEXAI: Grok 4.6 leads on the two strongest knowledge-work / real-world productivity benchmarks (GDPVal-AA and AA-Briefcase) and on the legal benchmark, while remaining highly competitive on coding-

Speaking of fine-tuning, we’ve got support on @axolotl_ai. Start training North Micro Vision right away, no hardware required. Find their do…

Model ReleasesDGX agent

Speaking of fine-tuning, we’ve got support on @axolotl_ai. Start training North Micro Vision right away, no hardware required. Find their docs here: https://docs.axolotl.ai/docs/models/cohere-north-mi

the @aiDotEngineer World Fair always one of the best events every year to talk to builders at the frontier of Research, Agents, Evals, Syste…

AgentsDGX agent

the @aiDotEngineer World Fair always one of the best events every year to talk to builders at the frontier of Research, Agents, Evals, Systems, etc A few weeks ago I gave a talk on - Continually Impro

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and…

Model ReleasesDGX agent

The model punches above its weight, outperforming Gemma 4 E2B and Ministral 3 3B across a broad range of visual understanding benchmarks and delivers particularly strong results on document understand

The most dangerous document extraction failure isn't a wrong value. It's a missing row that looks like nothing is wrong. We released Extract…

Model ReleasesDGX agent

The most dangerous document extraction failure isn't a wrong value. It's a missing row that looks like nothing is wrong. We released ExtractBench yesterday: 370 enterprise docs, 14 systems. The hardes

Today, we’re adding another member to our model family. Meet North Micro Vision. Our smallest vision-language model yet, ideal for sophistic…

Model ReleasesDGX agent

Today, we’re adding another member to our model family. Meet North Micro Vision. Our smallest vision-language model yet, ideal for sophisticated document understanding. Available open-source under an

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workl…

AgentsDGX agent

Together Serverless Inference gives developers a managed, high-throughput path for running Qwen3.8-2.4T-A95B across coding and agentic workloads. Start building: https://www.together.ai/models/qwen3-8

Try Grok 4.6 on tough real-world tasks!

Model ReleasesDGX agent

Try Grok 4.6 on tough real-world tasks! imo GDPVal is probably the most important benchmark, it measures the performance of models on real world tasks Big leap in performance here to top it at a great

Try Qwen-Image-3.0 on @openart_ai! 🎨👀

Model ReleasesDGX agent

Try Qwen-Image-3.0 on @openart_ai! 🎨👀 Qwen Image 3.0 is now on OpenArt ✨ The most Real Qwen image model yet. Native text across 12 languages, precise 10px type, and full interfaces like web pages, gam

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting …

AgentsDGX agent

Very interesting new work from Anthropic. (bookmark it) They evolve mind viruses, ideas that spread through a multi-agent system by getting each host to pass them on, then measure what governs the spr

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: …

Model ReleasesDGX agent

We are in an insane run of open-weight drops. Every modality, open source is winning. This is what an open source AI summer ☀️ looks like: 🧠 LLMs & Reasoning → DeepSeek-V4-Flash-0731 (my king 👑): 304B

We built a software factory for AI SDK. Each step is an agent, and humans merge changes. Four weeks in: ▪️ The factory authors up to 35% of …

AgentsDGX agent

We built a software factory for AI SDK. Each step is an agent, and humans merge changes. Four weeks in: ▪️ The factory authors up to 35% of merged PRs ▪️ It closed 70% of issues in July ▪️ Open bugs a

We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what t…

Model ReleasesDGX agent

We have her dash camera which shows they are lying. The agents are wearing body cameras and should have dash cameras of their own. If what they say happened was true they wouldn’t be issuing statement

We wrote a 36-page ArXiv whitepaper on ExtractBench 🧑‍🔬 , our effort to create the most comprehensive, schema-guided, real-world document …

Model ReleasesDGX agent

We wrote a 36-page ArXiv whitepaper on ExtractBench 🧑‍🔬 , our effort to create the most comprehensive, schema-guided, real-world document extraction benchmark. It’s extremely detailed and covers every

When should you start post-training your own models? @FireworksAI_HQ CEO @lqiao’s answer: after product-market fit. Not because it's hard...…

SafetyDGX agent

When should you start post-training your own models? @FireworksAI_HQ CEO @lqiao’s answer: after product-market fit. Not because it's hard... but because only after PMF is the data coming off your prod

You can now use Ollama as a provider in GitHub Copilot for JetBrains. https://github.blog/changelog/2026-08-11-copilot-memory-and-ollama-in-…

Local AiDGX agent

GitHub announced on August 12 2026 that users can now integrate Ollama as a provider in **GitHub Copilot for JetBrains**. This update allows JetBrains developers to switch to or add locally‑hosted (or

Yukon: Open Innovation for Frontier Research Open Innovation has been the driving value at Eigen Labs. But I’ve often struggled with where i…

AgentsDGX agent

Yukon: Open Innovation for Frontier Research Open Innovation has been the driving value at Eigen Labs. But I’ve often struggled with where it is actually better than a great closed team. Finally we fo

11 Aug 2026

⚡️A coalition that secures long-term AI capacity: We’re aggregating long-term compute demand in Europe to determine what capacity is built, …

Model ReleasesDGX agent

⚡️A coalition that secures long-term AI capacity: We’re aggregating long-term compute demand in Europe to determine what capacity is built, where it’s located, and whom it serves. Through these multi-

Building Agent Skills and testing them is hard, but it doesn't have to be. Listen to Arjun Patel demo Cultivar, an open source tool develope…

Model ReleasesDGX agent

Building Agent Skills and testing them is hard, but it doesn't have to be. Listen to Arjun Patel demo Cultivar, an open source tool developed at Pinecone to help benchmark agent skills in sandboxes. T

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building te…

Model ReleasesDGX agent

Claude's watermark probably doesn't work how you think. As the CTO of GPTZero, I'll explain how Anthropic, Google and OpenAI are building text watermarking in this brief explainer and whether it can b

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO…

Model ReleasesDGX agent

DeepSeek V4 Flash 0731 is now available to fine-tune on Together AI. Specialize it for coding, tool use, and your own domain with SFT or DPO, then deploy the fine-tuned model on Together AI for produc

ELON MUSK ON BUILDING A “STAR MIND” POWERED BY THE SUN “We’ve already built the most powerful AI training clusters in the world. And then wh…

IndustryDGX agent

ELON MUSK ON BUILDING A “STAR MIND” POWERED BY THE SUN “We’ve already built the most powerful AI training clusters in the world. And then what we expect to do by the end of next year is about 10 times

Epic talk: the cheat code for how to build your own in-house lab, featuring @gabepereyra of @harvey

AgentsDGX agent

Epic talk: the cheat code for how to build your own in-house lab, featuring @gabepereyra of @harvey Want world class research capabilities, but don’t have the resources of a big lab? At our recent Sov

ExtractBench is one of the most comprehensive benchmarks for real-world document extraction. ✅ It covers 4869 pages, across 67 document type…

Model ReleasesDGX agent

ExtractBench is one of the most comprehensive benchmarks for real-world document extraction. ✅ It covers 4869 pages, across 67 document types, spanning 8 real-world domains: finance, energy, gov, auto

From @sequoia's own your intelligence event. We're starting to see a new category of company emerge, the Full-Stack AI company, where innova…

SafetyDGX agent

From @sequoia's own your intelligence event. We're starting to see a new category of company emerge, the Full-Stack AI company, where innovation happens at both the product and intelligence layer. Tha

Had so much fun giving this talk at @sequoia about @harvey’s moneyball approach to building a research lab. The biggest mistake I made in th…

AgentsDGX agent

Had so much fun giving this talk at @sequoia about @harvey’s moneyball approach to building a research lab. The biggest mistake I made in the early days of Harvey was trying to play the Yankees baseba

HUMAN HARNESS! @mcuban casually made up the term when talking abt how we differentiate when everyone can create same thing/agent. Empathy, P…

AgentsDGX agent

HUMAN HARNESS! @mcuban casually made up the term when talking abt how we differentiate when everyone can create same thing/agent. Empathy, Personal Brand, using time agents save you to offer unique hu

I’m hosting a free AI agent workshop with Mark Cuban for 20,000 people on how AI agents have changed the way I work. You can register at htt…

AgentsDGX agent

Allie K. Miller announced a free live AI‑agent workshop featuring Mark Cuban set for August 11 2026, targeting up to 20,000 participants. She will demonstrate how her own AI workforce of 34 agents wor

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models a…

Model ReleasesDGX agent

Introducing ExtractBench, the most comprehensive benchmark for information extraction from complex enterprise documents. The latest models are pushing the frontier of coding and knowledge work, but su

Introducing 𝗘𝘅𝘁𝗿𝗮𝗰𝘁𝗕𝗲𝗻𝗰𝗵: the most comprehensive benchmark for information extraction from complex enterprise documents. Our app…

Model ReleasesDGX agent

Introducing 𝗘𝘅𝘁𝗿𝗮𝗰𝘁𝗕𝗲𝗻𝗰𝗵: the most comprehensive benchmark for information extraction from complex enterprise documents. Our applied research team tested: 14 systems — frontier VLMs, coding agents, ex

It's been exciting for Ollama to partner with @JensenHuang and the @NVIDIAAI team on launching open models. Open models have no boundaries, …

Model ReleasesDGX agent

It's been exciting for Ollama to partner with @JensenHuang and the @NVIDIAAI team on launching open models. Open models have no boundaries, and let's continue to work together to make this ecosystem b

I’ve been thinking about how agents can learn inside world models for years. We decided to scale up our RSI Lab to bridge recursive self-imp…

TutorialsDGX agent

I’ve been thinking about how agents can learn inside world models for years. We decided to scale up our RSI Lab to bridge recursive self-improvement with physical AI and robotics. We are looking for f

Knowledge is rather abstract and philosophical. So what does it mean from an agentic context? How does providing the right knowledge at the …

AgentsDGX agent

Knowledge is rather abstract and philosophical. So what does it mean from an agentic context? How does providing the right knowledge at the right time to an agent increase its accuracy, task completio

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers lik…

ApplicationsDGX agent

Leveraging the frontier ecosystem We work with neolabs like @trajectorylabs and @EngramLab and inference + post-training infra providers like @baseten, @FireworksAI_HQ, and @appliedcompute to post-tra

LLMs still produce bugs, but those bugs are different than what they used to be. It’s less off-by-ones and more about system design, ui usab…

Model ReleasesDGX agent

LLMs still produce bugs, but those bugs are different than what they used to be. It’s less off-by-ones and more about system design, ui usability, missing broader context. Some kinds of coding has bee

← Previous
123…232
Next →