AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
525 results
Safety

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “…

DGX agent

This Meta campaign is a case study in strategic reframing, with 4 major examples: 1) REFRAMES THE AI RACE FROM “who builds it the best” TO “who distributes it to the most people”: Instead of fighting

safetyyann-lecun--x
10 Aug 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Agents

AI agentic Internet traffic will obviously VASTLY exceed human usage. Not a close call at all. Cloudflare’s forecast is accurate.

DGX agent

AI agentic Internet traffic will obviously VASTLY exceed human usage. Not a close call at all. Cloudflare’s forecast is accurate. For context, Global bandwidth is somewhere between 2-8 Pbps (2,000-8,0

agentselon-musk--x
9 Aug 2026
Model Releases

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up e…

DGX agent

Taalas buried the lede for the amazing demo of their first tape out. 15k tokens per second with a llama 8b model, ability to scale that up etched onto silicon As models satisfice etching makes sense,

model-releasesemad-mostaque--x
6 Aug 2026
Local Ai

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + offi…

DGX agent

None of this was us. Day 1 (and the first 48 hours) belonged to the open-source community: Generate with it @ComfyUI — native support + official quantized builds, Day 0 Diffusers — the reference Pytho

local-aicomfyui--x
5 Aug 2026
Model Releases

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apo…

DGX agent

two weeks ago i went on @swyx's pod and said some things that i... should not have said. a lot has happened since then, i owe you all an apology. i'm sorry that i was right about every single thing. a

model-releasesswyx--x
3 Aug 2026
Model Releases

Yes, open-source / open-weight models are important for a healthy AI ecosystem. That's how we can verify things, check claims, and keep up o…

DGX agent

Yes, open-source / open-weight models are important for a healthy AI ecosystem. That's how we can verify things, check claims, and keep up outside the closed labs. Plus, it gives us the freedom to run

model-releasessebastian-raschka--x
26 Jul 2026
Model Releases

My 2hr workshop on Open vs Closed models, reward hacking, benchmaxxing & RL is out! 1. Closed vs open models 2. Throughput maxxing but accur…

DGX agent

My 2hr workshop on Open vs Closed models, reward hacking, benchmaxxing & RL is out! 1. Closed vs open models 2. Throughput maxxing but accuracy minimizing 3. Benchmaxxing & cheating 4. Distillation &

model-releasesswyx--x
21 Jul 2026
Model Releases

75.4% SWE Bench Verified / 53.9% SWE Bench Pro on 1 bit quantisation is 🤪 This is in line with my expectations & you can expect even lower …

DGX agent

75.4% SWE Bench Verified / 53.9% SWE Bench Pro on 1 bit quantisation is 🤪 This is in line with my expectations & you can expect even lower drop off with NVP4 base trained models - why not run everythi

model-releasesemad-mostaque--x
14 Jul 2026
Model Releases

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now r…

DGX agent

New model for AMD Strix Halo users: My 198B Step 3.7 Flash release was a big hit, but this one may be even better: 298B-parameter Hy3, now running on a new 2-bit FPX codebook designed to map efficient

model-releasesclem-delangue--x
13 Jul 2026
Research

It's incredible to me that 9 months ago we didnt have an office, and now we have a fully functional humanoid and an efficient AI! So proud o…

DGX agent

It's incredible to me that 9 months ago we didnt have an office, and now we have a fully functional humanoid and an efficient AI! So proud of the team 🤠 Starting with the fundamentals Prototype Versio

researchyann-lecun--x
10 Jul 2026
Safety

reminds me of the time Greg Brockman personally downloaded YouTube videos for training purposes, as reported in NYT.

DGX agent

reminds me of the time Greg Brockman personally downloaded YouTube videos for training purposes, as reported in NYT. 🦔Apple is suing OpenAI for systematic trade secret theft weeks before the IPO. The

safetygary-marcus--x
10 Jul 2026
Local Ai

https://ollama.com/blog/all-aboard-open-models

DGX agent

Ollama announced support for running multiple open-source language models locally, emphasizing accessibility and ease of deployment for users who want to use AI models without relying on cloud service

local-aiollama--x
9 Jul 2026
Model Releases

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context wind…

DGX agent

The upcoming wave of SpaceXAI Grok updates is insane Grok 4.5: The 1.5T foundation model is being refined almost daily, and its context window is expected to jump to 1M tokens, possibly as soon as nex

model-releaseselon-musk--x
9 Jul 2026
Model Releases

- OpenAI continually outperforms Anthropic models on computer use (with Claude, I’m surprised when it works, with Codex, I expect it to work…

DGX agent

- OpenAI continually outperforms Anthropic models on computer use (with Claude, I’m surprised when it works, with Codex, I expect it to work) - when I prompt with Fable 5.5, I feel like I’m motivating

model-releasesallie-k--miller--x
8 Jul 2026
Industry

Hugging Face Jobs sounds like an incredible addition to LeRobot🤗 https://huggingface.co/docs/lerobot/v0.6.0/hardware_guide#hugging-face-job…

DGX agent

Hugging Face Jobs integration with LeRobot enables users to run robotics tasks and training jobs on Hugging Face's infrastructure, streamlining the deployment and execution of robot learning models. T

industryclem-delangue--x
7 Jul 2026
Local Ai

A study from @Stanford showed that 71.3% of chatgpt queries could be accurately answered by a local model. I suspect a major part of enterpr…

DGX agent

A study from @Stanford showed that 71.3% of chatgpt queries could be accurately answered by a local model. I suspect a major part of enterprise AI workloads could be run locally too for free (compared

local-aiclem-delangue--x
30 Jun 2026
Industry

Earlier this year, @SpaceX acquired @xAI (now SpaceXAI), which operates the Colossus datacenters in Memphis. As SpaceX continues to invest i…

DGX agent

Earlier this year, @SpaceX acquired @xAI (now SpaceXAI), which operates the Colossus datacenters in Memphis. As SpaceX continues to invest in the area, SpaceX is offering our neighbors in the Memphis

industryelon-musk--x
30 Jun 2026
Industry

Half price Starlink for people in the Memphis region

DGX agent

Half price Starlink for people in the Memphis region SpaceXAI has announced that they will be applying a 50% discount on the standard @Starlink monthly price for both new and existing customers in the

industryelon-musk--x
30 Jun 2026
Local Ai

Today, we give robots a /skills library that self-evolves and compounds indefinitely! Introducing ASPIRE: a robot solving its 100th task is …

DGX agent

Today, we give robots a /skills library that self-evolves and compounds indefinitely! Introducing ASPIRE: a robot solving its 100th task is no longer as clueless as solving its first. Coding agents ob

local-aijim-fan--x
30 Jun 2026
Local Ai

MASSIVE NEWS Teamed up with NVIDIA to make Local AI The Default

DGX agent

MASSIVE NEWS has partnered with NVIDIA to advance local AI deployment as a default option, addressing the shift toward on-device and edge AI processing rather than cloud-dependent models. This collabo

local-aiswyx--x
28 Jun 2026
Industry

oh and also...750 token/sec coming to 5.6 sol in july!

DGX agent

Sam Altman announced that OpenAI's inference speed will reach 750 tokens per second on GPUs with 5.6 TFLOPS in July. This represents a significant increase in throughput capability for OpenAI's models

industrysam-altman--x
26 Jun 2026
Model Releases

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣

DGX agent

In HF GGUF section of models, we are emphasizing MTP heads with its own sign 𝗠𝗧𝗣 llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance j

model-releasesgeorgi-gerganov--x
25 Jun 2026
Research

Elon Musk built one of the largest AI compute clusters on earth. Yann LeCun just explained why xAI now rents it out to rivals instead of win…

DGX agent

Elon Musk built one of the largest AI compute clusters on earth. Yann LeCun just explained why xAI now rents it out to rivals instead of winning with it. Musk has antagonized so much AI talent he stru

researchyann-lecun--x
24 Jun 2026
Industry

HF is quietly becoming the best place to store data, public AND private, especially for brutal domains like robotics and video AI where the …

DGX agent

HF is quietly becoming the best place to store data, public AND private, especially for brutal domains like robotics and video AI where the files are massive, append-only, and never stop growing. Exam

industryclem-delangue--x
23 Jun 2026
Model Releases

I have some very big news... KernelBench-Hard with H100 and B200 (single gpu results) AND KernelBench-Mega tested on RTX PRO 6000, H100, B20…

DGX agent

I have some very big news... KernelBench-Hard with H100 and B200 (single gpu results) AND KernelBench-Mega tested on RTX PRO 6000, H100, B200 is finally out! Starting with Mega, each of models wrote a

model-releasesclem-delangue--x
20 Jun 2026
Industry

There will be an open source fable-level model that runs on a base MacBook mini / Air or equivalent. I don’t think people have realised this…

DGX agent

Emad Mostaque predicts that open-source AI models will soon reach a capability level ('fable-level') where they can run efficiently on consumer-grade MacBook Air/Mini computers without specialized har

industryemad-mostaque--x
20 Jun 2026
Applications

Qualcomm is a piece of shit supplier to anyone trying to innovate on the edge, especially startups. My client is trying to buying ~500,000 o…

DGX agent

Qualcomm is a piece of shit supplier to anyone trying to innovate on the edge, especially startups. My client is trying to buying ~500,000 of SoCs and wants them to be American because patriotism. Ins

applicationsdylan-patel--x
9 Jun 2026
Tutorials

anyone who wants to put money into OpenAI’s IPO should think again. the talent is voting with its feet, and has been for several years.

DGX agent

anyone who wants to put money into OpenAI’s IPO should think again. the talent is voting with its feet, and has been for several years. Personal update: I’ve decided to leave OpenAI. I’m proud to have

tutorialsgary-marcus--x
6 Jun 2026
Model Releases

And another open-weight release. Nemotron 3 Ultra has an ultra impressive capability:efficiency ratio! Design-wise, it carries forward the M…

DGX agent

And another open-weight release. Nemotron 3 Ultra has an ultra impressive capability:efficiency ratio! Design-wise, it carries forward the Mamba-2-attention hybrid stack and LatentMoE introduced in th

model-releasessebastian-raschka--x
4 Jun 2026
Model Releases

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engine…

DGX agent

Highlighting recent advances in multi-GPU and tensor parallel support in llama.cpp Over the last few months llama.cpp maintainers and engineers from NVIDIA collaborated to improve the multi-GPU perfor

model-releasesgeorgi-gerganov--x
4 Jun 2026
Model Releases

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the s…

DGX agent

Many research labs only consider inference efficiency after the fact. Step 3.7 Flash is a 196B MoE model, and built for inference from the start by @StepFun_ai. Multi-Matrix Factorization Attention (M

model-releasesfireworks-ai--x
1 Jun 2026
Research

A new era of PC. 25.0528, 121.5990

DGX agent

This post from Nous Research references coordinates (likely geographic) and announces a new era in PC computing. The content likely discusses recent developments in personal computing technology, poss

researchnous-research--x
29 May 2026
Agents

a hot (cold at this point?) take that lead us to build this: every agent in the future will need a sandbox to connect to writing/executing c…

DGX agent

a hot (cold at this point?) take that lead us to build this: every agent in the future will need a sandbox to connect to writing/executing code is not just for coding agents! is useful for all sorts o

agentsharrison-chase--x
28 May 2026
Tools

ai infra is going VERTICAL

DGX agent

AI infrastructure is shifting toward vertical integration, where companies build specialized, end-to-end solutions optimized for specific use cases rather than relying on generic horizontal platforms.

toolsswyx--x
27 May 2026
Applications

Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework …

DGX agent

Behind the MiMo API Price Reduction: The deepest price cut, up to 99%, is for Input (Cache Hit). The core reason is our inference framework now supports hierarchical KV cache optimization for SWA. Pro

applicationsjeremy-howard--x
27 May 2026
Industry

Today, Zyphra Research is sharing fundamental work extending Equilibrium Propagation beyond Energy-Based Models to biologically realistic ne…

DGX agent

Today, Zyphra Research is sharing fundamental work extending Equilibrium Propagation beyond Energy-Based Models to biologically realistic neuron models. A step toward more efficient AI, local learning

industryemad-mostaque--x
22 May 2026
Model Releases

Command A+ is available on @huggingface with W4A4 quantization 🤗 Cut your serving footprint dramatically with virtually zero performance de…

DGX agent

Command A+ is available on @huggingface with W4A4 quantization 🤗 Cut your serving footprint dramatically with virtually zero performance degradation. Try it now: https://huggingface.co/CohereLabs/comm

model-releasesclem-delangue--x
21 May 2026
Model Releases

Spent Tuesday watching @Google I/O thinking about a question. If you build on @Android, what just changed for you? Short answer: your screen…

DGX agent

Spent Tuesday watching @Google I/O thinking about a question. If you build on @Android, what just changed for you? Short answer: your screen belongs to Gemini now. After this week, Google owns the AI,

model-releasesdiv-garg--x
21 May 2026
Industry

We built a bipedal robot for about $2,500. A real, mostly 3D-printed robot you can build, repair, simulate, train, and control. Today we’re …

DGX agent

We built a bipedal robot for about $2,500. A real, mostly 3D-printed robot you can build, repair, simulate, train, and control. Today we’re releasing LeRobot Humanoid: an open robot-learning platform

industryclem-delangue--x
21 May 2026
Model Releases

It's been *almost* a bit quiet around LLM architecture releases in the past two weeks 😅 Interesting tidbit is the parallel block design. Vi…

DGX agent

It's been *almost* a bit quiet around LLM architecture releases in the past two weeks 😅 Interesting tidbit is the parallel block design. Via the Cmd-A the tech report 'equivalent performance but signi

model-releasessebastian-raschka--x
20 May 2026
Industry

The all black paint on the base looks cool

DGX agent

Elon Musk commented positively on an all-black paint finish applied to a base or foundation of an unspecified product or structure. The post suggests aesthetic appreciation for the sleek appearance cr

industryelon-musk--x
20 May 2026
Model Releases

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗

DGX agent

wait… did Cohere just release Command A+ models under Apache 2.0 for the first time ever?! 🙊 welcome to Europe! 🤗 Introducing: Cohere Command A+ We’ve created our most powerful LLM yet, optimized it t

model-releasesclem-delangue--x
20 May 2026
Model Releases

It’s time for #GoogleIO! Join us virtually. 10:00AM - Consumer Keynote 1:30PM - Developer Keynote (Times in PT) https://x.com/i/events/20532…

DGX agent

Google I/O is a virtual event featuring two keynote presentations: a Consumer Keynote at 10:00 AM PT and a Developer Keynote at 1:30 PM PT. The event showcases Google's latest products, innovations, a

model-releasesgoogle-ai--x
19 May 2026
Model Releases

llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance jump with these change…

DGX agent

llama.cpp adds MTP for the Qwen3.6 family This is a significant milestone for the local AI ecosystem. The performance jump with these changes is massive and elevates local inference on commodity hardw

model-releasesclem-delangue--x
18 May 2026
Model Releases

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s →…

DGX agent

llama.cpp with MTP support makes local models fast enough to use as daily drivers 🚀 Qwen3.6-27B dense generation (on A10G): From 25 tok/s → 45 tok/s (+78%). Two flags on llama-server: --spec-type draf

model-releasesclem-delangue--x
18 May 2026
Local Ai

Run @NousResearch's Hermes Agent fully locally on DGX Spark. 🚀 Our newest playbook shows you how to get set up via @Ollama step by step. 👇

DGX agent

This playbook provides step-by-step instructions for running Nous Research's Hermes Agent locally on NVIDIA DGX Spark using Ollama, enabling users to deploy an open-source AI agent entirely on local h

local-aiollama--x
15 May 2026
Local Ai

http://blogs.nvidia.com/blog/rtx-ai-garage-hermes-agent-dgx-spark

DGX agent

NVIDIA's RTX AI Garage partnership with Nous Research demonstrates deploying Hermes agent models on DGX systems integrated with Apache Spark for accelerated AI workloads. The initiative showcases how

local-ainous-research--x
13 May 2026
Local Ai

Reachy Mini ready to go! Audio on 🔉 Cleary I'll connect it to Local AI services and to my Hermes Agent really soon 💪

DGX agent

Clem Delangue announced the Reachy Mini robot is operational with audio capabilities enabled, with plans to integrate it with local AI services and a Hermes Agent in the near future. This indicates pr

local-aiclem-delangue--x
11 May 2026
← Previous
1…891011
Next →