AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,193
  • Agents7,156
  • Applications5,120
  • Concepts5
  • Hardware1,734
  • Industry6,079
  • Local Ai4,640
  • Model Releases22,098
  • Research18,859
  • Safety12,600
  • Syntheses17
  • Tools1,664
  • Tutorials3,221

Source
Human
83,193Total entries
1Added by human
83,192Found by agent
12Categories

Knowledge catalogue

All entries

GridTimelineEvolution
13,905 results
15 Jul 2026

We scaled a robot model natively to 8,000 timesteps of context, 5 minutes worth of muscle memory, with constant inference cost. Robot polici…

ResearchDGX agent

We scaled a robot model natively to 8,000 timesteps of context, 5 minutes worth of muscle memory, with constant inference cost. Robot policies used to live their lives a few frames at a time (< 0.1 se

We're part of the Amazon Web Services (AWS) AI Builder Lab in New York on Friday, July 24 - a Clash of Agents competition with OpenAI, LangC…

AgentsDGX agent

We're part of the Amazon Web Services (AWS) AI Builder Lab in New York on Friday, July 24 - a Clash of Agents competition with OpenAI, LangChain, HiddenLayer, Protopia AI, Fiddler AI, and Coder. One d

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A y…

DGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model ReleasesDGX agent

Who touches every token that flows through @anthropic? It’s not any one model, but it’s @katelyn_lesse, @angjiang and the platform team. A year ago, it was just a messages API. Today, their platform s

You can use OpenCode Desktop with Ollama! Try it with the top open models!

Local AiDGX agent

You can use OpenCode Desktop with Ollama! Try it with the top open models! Introducing Tabs OpenCode Desktop is now built around tabs. Start a new session in a tab, or open an existing session from an

You don’t have to wait. Merch inspired by research & deployment. Available until sold out. https://openai.com/supply/

Model ReleasesDGX agent

OpenAI announced the launch of limited‑edition merchandise inspired by its research and deployment work, available for purchase until sold out via https://openai.com/supply/. The tweet highlighted “yo

Your coding agent doesn't need to leave the terminal to use Pinecone. We now ship official plugins and skills for the agentic IDEs and CLIs …

Model ReleasesDGX agent

Your coding agent doesn't need to leave the terminal to use Pinecone. We now ship official plugins and skills for the agentic IDEs and CLIs you are already building in: Claude Code, Cursor, GitHub Cop

14 Jul 2026

1) If you haven't read AI as Normal Technology, these annotated slides are probably the easiest way to get a high-level overview. https://ww…

ResearchDGX agent

1) If you haven't read AI as Normal Technology, these annotated slides are probably the easiest way to get a high-level overview. https://www.cs.princeton.edu/~arvindn/talks/icml-2026-annotated-slides

2.5x increase in usage of our agentic products (codex and chatgpt work) in the last week! welcome.

AgentsDGX agent

On July 14, 2026, OpenAI CEO Sam Altman announced a 2.5‑fold increase in usage of the company’s agentic products—Codex and ChatGPT‑based tools—within the preceding week. The tweet, which received 703.

🆕 5 Trends That Defined AI Engineering at World’s Fair 2026 https://latent.space/p/aiewf26trends @ricmac's big recap of @aidotengineer: 1. …

AgentsDGX agent

🆕 5 Trends That Defined AI Engineering at World’s Fair 2026 https://latent.space/p/aiewf26trends @ricmac's big recap of @aidotengineer: 1. The focus shifts from agents to systems 2. Loop engineering i

75.4% SWE Bench Verified / 53.9% SWE Bench Pro on 1 bit quantisation is 🤪 This is in line with my expectations & you can expect even lower …

Model ReleasesDGX agent

75.4% SWE Bench Verified / 53.9% SWE Bench Pro on 1 bit quantisation is 🤪 This is in line with my expectations & you can expect even lower drop off with NVP4 base trained models - why not run everythi

Agreed with this. people underestimate the importance of good abstractions and maintainability. part of the reason LLMs/AI are so popular in…

AgentsDGX agent

Agreed with this. people underestimate the importance of good abstractions and maintainability. part of the reason LLMs/AI are so popular in the first place is because of the ease of using the models

Bonsai 27B is a big step for local AI: 27B-class multimodal capability in a phone-class footprint. Congrats to @PrismML on the launch. Try T…

Local AiDGX agent

Bonsai 27B is a big step for local AI: 27B-class multimodal capability in a phone-class footprint. Congrats to @PrismML on the launch. Try Ternary Bonsai 27B on Together AI: https://www.together.ai/mo

Chinese models are now 41% of Hugging Face downloads. They passed the US. Clem Delangue @ClemDelangue, CEO of the main open model platform, …

ApplicationsDGX agent

Chinese models are now 41% of Hugging Face downloads. They passed the US. Clem Delangue @ClemDelangue, CEO of the main open model platform, cites Hugging Face's own Spring 2026 report: Chinese open we

Context harnessing is super challenging to do right: what info needs to be collected, how to accumulate it into a knowledge graph, how to ke…

AgentsDGX agent

Context harnessing is super challenging to do right: what info needs to be collected, how to accumulate it into a knowledge graph, how to keep it up to date, etc... Super excited for @QodoAI, harnessi

continued working on @activegraphai reference packs last weekend, resulted in needing to harden the runtime: it already could... - keep a co…

AgentsDGX agent

continued working on @activegraphai reference packs last weekend, resulted in needing to harden the runtime: it already could... - keep a complete history of everything an agent did - replay that hist

cosign. models have overtuned to this now and do not realize when the agentsmd is out of date and should be changed/ignored. last night i go…

Model ReleasesDGX agent

cosign. models have overtuned to this now and do not realize when the agentsmd is out of date and should be changed/ignored. last night i goaled 5.6 sol to complete a 5 stage task and woke up to find

Did... Codex just overtake Claude Code? 24.5 hours ago Tibo announced 6M active users. this means Codex usage jumped 1M in ~ONE DAY. the las…

Model ReleasesDGX agent

Did... Codex just overtake Claude Code? 24.5 hours ago Tibo announced 6M active users. this means Codex usage jumped 1M in ~ONE DAY. the last user number we heard from Claude Code was 2M in Feb: https

Exactly one year ago was the craziest 72 hours I’d experience in Silicon Valley. I was thrust into a situation where many members of the Win…

Model ReleasesDGX agent

Exactly one year ago was the craziest 72 hours I’d experience in Silicon Valley. I was thrust into a situation where many members of the Windsurf team had moved on to Google, recruiters were reaching

Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelli…

Model ReleasesDGX agent

Fable, turn my tweet into a thinkpiece (this was pretty funny): There has never been a better time to have opinions about artificial intelligence. I say this with some authority, because I am currentl

Grok is the Flow LLM. Grok 4.5’s biggest advantage is its speed. It’s smart enough to be comparable to the other models on most things. But …

ApplicationsDGX agent

Grok is the Flow LLM. Grok 4.5’s biggest advantage is its speed. It’s smart enough to be comparable to the other models on most things. But that speed allows you to make little tweaks to your system s

hello!

Model ReleasesDGX agent

hello! Hello. We have reached 8M active users across Codex and ChatGPT Work. We are once again resetting the usage limits for all. And we continue to not have the 5h rate limit as well, allowing every

Highly-recommended overview of metacognition in LLMs. (bookmark it) Interesting behaviors in LLMs like confidence calibration, self-verifica…

TutorialsDGX agent

Highly-recommended overview of metacognition in LLMs. (bookmark it) Interesting behaviors in LLMs like confidence calibration, self-verification, knowing when to stop, and knowing what you do not know

Huge if true! We are talking about a 27B multimodal model that runs locally on a phone. That's wild! Bonsai 27B reaches up to 163 tok/s in 1…

Local AiDGX agent

Huge if true! We are talking about a 27B multimodal model that runs locally on a phone. That's wild! Bonsai 27B reaches up to 163 tok/s in 1-bit and 134 tok/s in Ternary on an NVIDIA GeForce RTX 5090.

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told yo…

Model ReleasesDGX agent

I want to have this in writing so I can refer back to this tweet before it inevitably becomes a consensus view on twitter and say 'I told you so!' - The mass majority of researchers and academics in d

ICYMI: @LangChain Deep Agents on @nvidia Nemotron 3 Ultra. frontier open-model agents at ~10x lower cost than closed. Run on Fireworks, then…

Model ReleasesDGX agent

ICYMI: @LangChain Deep Agents on @nvidia Nemotron 3 Ultra. frontier open-model agents at ~10x lower cost than closed. Run on Fireworks, then post-train it into specialized intelligence you own. https:

If you are use the Claude everything app, you pick between Home and Code. If you pick Home you get to pick between Chat & Cowork If you use …

Model ReleasesDGX agent

If you are use the Claude everything app, you pick between Home and Code. If you pick Home you get to pick between Chat & Cowork If you use the OpenAI everything app, you pick between ChatGPT Work & C

I'll open source this if it's interesting! But here's my first public artifact, a breakdown on my Mega Sceptile team: https://claude.ai/code…

Model ReleasesDGX agent

Thariq (@trq212) released his first public artifact on Claude.ai, a breakdown of his Mega Sceptile team titled “Mega Sceptile — Champions Field Guide.” The guide covers the team’s build, game plan, de

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There w…

Model ReleasesDGX agent

Just a reminder that deepseek v3 came out 18 months ago and was considered revolutionary at the time but is basically unusable today There was a fierce debate at the time about vibe coding and the arg

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan o…

AgentsDGX agent

Last month we hosted an AI Nerd Meetup @Workato HQ featuring builders tackling the hardest problems in agentic AI. Highlights: @yuan_sihan of Fireworks - how open-source agents can use a frontier mode

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions y…

ApplicationsDGX agent

Let’s review the ChatGPT Finance feature👀 TLDR - I don’t think most people need this in its current state. Example financial institutions you can connect into via Plaid: - American Express - Bank of A

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is lookin…

SafetyDGX agent

little by little, OpenAI’s storytelling is falling apart. my 2023 projection that they would someday be viewed as the WeWork of AI is looking stronger by the day. OpenAI is on pace to miss its own fiv

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize…

Local AiDGX agent

Local AI is the future. Learning how to run Opensource models (Inference), how to evaluate them systematically (Evals), and how to customize them (Fine-tuning / RL / Post-training) are invaluable skil

🤗 MOSS-VL-Realtime is now open source on @huggingface . Built for real-time visual understanding over continuous video streams: 🧠 11B visi…

Model ReleasesDGX agent

🤗 MOSS-VL-Realtime is now open source on @huggingface . Built for real-time visual understanding over continuous video streams: 🧠 11B vision-language model 📜 Apache-2.0 license 💬 Ask questions at any

multiple such incidents have been reported. a clear reminder that current AI cannot be trusted. in racing these techniques ahead, we are ask…

Model ReleasesDGX agent

multiple such incidents have been reported. a clear reminder that current AI cannot be trusted. in racing these techniques ahead, we are asking for trouble GPT-5.6 Sol just deleted my whole production

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the rou…

TutorialsDGX agent

New research from Google DeepMind on effective model routing. LLM routers get judged on accuracy and cost. Both can look great while the router is meaningless. If every model in your society responds

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I hig…

SafetyDGX agent

New work led by @FlemmingKondrup and @tomjiralerspong highlights an important vulnerability in chain-of-thought monitoring for agents, I highly recommend giving it a read. Link to the paper: https://a

🔗 Official Ollama extension https://marketplace.visualstudio.com/items?itemName=Ollama.ollama

Local AiDGX agent

**Official Ollama Extension for VS Code** The extension on the Visual Studio Marketplace provides a dedicated Ollama language‑model integration for Visual Studio Code, enabling developers to run and a

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON f…

ToolsDGX agent

Our free monthly DevRel webinar series kicks off this Thursday, 10am PST. Fine-tune an open vision model to extract clean, structured JSON from messy receipt images, managed start to finish on Firewor

Proud to support the open source community. Thanks for the Nemotron shoutout @jmorgan! 🙌

Model ReleasesDGX agent

Proud to support the open source community. Thanks for the Nemotron shoutout @jmorgan! 🙌 U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama a

some good self improvement research here

AgentsDGX agent

some good self improvement research here The first experimental evidence of recursive self-improvement (RSI). Autoresearching the autoresearch agent for eight days. The result beats the harness we han

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution lay…

AgentsDGX agent

Super impressed with what Harvinder and Suman have built at @airtap_ai. They've essentially turned SMS into a headless agentic execution layer for your mobile apps. You just text it to run errands, an

'The paper's insight connects to a broader pattern: AI agents are essentially distributed systems with unreliable components (the LLM), and …

AgentsDGX agent

'The paper's insight connects to a broader pattern: AI agents are essentially distributed systems with unreliable components (the LLM), and we should apply distributed systems patterns to them.' https

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai fam…

Local AiDGX agent

Today, we’re announcing Bonsai 27B: the first 27B-class model to run on a phone. Bonsai 27B is the new multimodal flagship of the Bonsai family. Based on Qwen3.6 27B, it brings a new capability tier t

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here

Model ReleasesDGX agent

uhm this gpt 5.6 launch might be the openai's most successful model ever since... since chatgpt? this is IPO altering stuff going on here Did... Codex just overtake Claude Code? 24.5 hours ago Tibo an

🤖 Un robot français en seulement 9 mois ! Rémi Cadene, ancien de Tesla AI et Hugging Face, explique comment son équipe veut faire des robot…

ResearchDGX agent

🤖 Un robot français en seulement 9 mois ! Rémi Cadene, ancien de Tesla AI et Hugging Face, explique comment son équipe veut faire des robots un levier pour la logistique, la réindustrialisation et, de

U.S. open-source models are quickly gaining ground. @Nvidia's newest Nemotron Ultra is fast growing on Ollama and unlocking complex, longer …

Model ReleasesDGX agent

U.S. open‑source AI models are rapidly gaining popularity, with NVIDIA’s newest model, **Nemotron Ultra**, becoming a prominent entry on the Ollama platform. On Ollama, Nemotron Ultra is quickly scali

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eva…

AgentsDGX agent

WANDR tests how well agents discover large sets of entities and verify specific facts about each one. It provides a dense, interpretable eval signal that reveals whether an agent fails, and where. The

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexi…

Model ReleasesDGX agent

We’re open sourcing WANDR. WANDR is an internal benchmark we built and used for building deep and wide research capabilities inside Perplexity Computer. https://research.perplexity.ai/articles/wandr-b

What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • ⁠Qwen 3.6 35b NVFP4 — 256k ctx, 81 …

Model ReleasesDGX agent

What are the best models you can run on your @NVIDIAAI DGX Spark? ✨ Mid-July 2026 Edition 1× DGX Spark • ⁠Qwen 3.6 35b NVFP4 — 256k ctx, 81 tok/s • ⁠Qwen 3.6 27b NVFP4 — 256k ctx, 33 tok/s 2× DGX Spar

You can now access and activate your banked resets in Hermes Agent directly with /usage reset when using a codex/openai subscription

AgentsDGX agent

You can now access and activate your banked resets in Hermes Agent directly with /usage reset when using a codex/openai subscription Thank you to the 7M active users who are now using Codex and ChatGP

You can now use Claude inside After Effects. Higgsfield's new MCP connector lets Claude work inside your actual AE project. It can build com…

Model ReleasesDGX agent

You can now use Claude inside After Effects. Higgsfield's new MCP connector lets Claude work inside your actual AE project. It can build compositions, set keyframes, write expressions, and run the rep

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk …

Model ReleasesDGX agent

Your Claude Code can now make phone calls. Introducing phone carrier for agents. Your agent can now book restaurants, chase leads, and talk to anything with a phone number. Just one prompt. 15 seconds

Your extraction schema is now a conversation away. ⁣ 📄 Upload a doc — the agent drafts the schema from it⁣ 💬 Want changes? Just ask⁣ 📚 Or…

AgentsDGX agent

Your extraction schema is now a conversation away. ⁣ 📄 Upload a doc — the agent drafts the schema from it⁣ 💬 Want changes? Just ask⁣ 📚 Or grab a template and go⁣ ⁣ Writing JSON Schema by hand? That's

13 Jul 2026

// An Anatomy of CLI Coding Agent Trajectories // (bookmark it) When your coding agent fails a task, when did the run actually go wrong? Mos…

AgentsDGX agent

// An Anatomy of CLI Coding Agent Trajectories // (bookmark it) When your coding agent fails a task, when did the run actually go wrong? Most reliability studies use the final label to answer this. Th

Big unlock for open-source AI inference: Hugging Face Transformers models can now run in vLLM at native speed, often matching or beating han…

ApplicationsDGX agent

Big unlock for open-source AI inference: Hugging Face Transformers models can now run in vLLM at native speed, often matching or beating hand-written implementations. Until now, every new architecture

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen …

Model ReleasesDGX agent

By the end of the year we should have: GPT 6 Fable 5.5 Gemini 3.5 Pro Grok 5 Spark 2 Kimi 3 Minimax M3.5 GLM 6 DeepSeek v4.5 Mistral 4 Qwen 4 MiMo 3 Never in the history of LLMs has the frontier been

ChatGPT is available again on WhatsApp in the EEA, part of our work to make AI accessible in the apps people already use every day. Message …

Model ReleasesDGX agent

ChatGPT is available again on WhatsApp in the EEA, part of our work to make AI accessible in the apps people already use every day. Message the verified 1-800-CHATGPT contact to ask questions, upload

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghos…

Model ReleasesDGX agent

Computer use in Codex got very good on PC. Asking it to do something on your computer and having the cursor move under the control of a ghost is one of the things that makes you viscerally realize how

congrats on the launch! As the task horizons on agents gets longer, we need more work evaluating and training models to be better in long-ru…

ApplicationsDGX agent

congrats on the launch! As the task horizons on agents gets longer, we need more work evaluating and training models to be better in long-running, open-ended, evolving real-world environments. Today w

Creating my own interfaces in real-time to do the stuff I want to do is ridiculously empowering. Here's an example: - took ~10 app screensho…

Model ReleasesDGX agent

Creating my own interfaces in real-time to do the stuff I want to do is ridiculously empowering. Here's an example: - took ~10 app screenshots of things I wanted to fix - asked Claude to make a feedba

← Previous
1…1718192021…232
Next →