AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,114 results
Research

Hermes🪽

DGX agent

Hermes is an AI model developed by Nous Research that focuses on instruction-following and reasoning capabilities. Based on Nous Research's focus, this likely covers the model's architecture, performa

researchnous-research--x
7 May 2026
Agents

small workflow note that adds up. /staged-pr is a skill (via slash command) i run when wrapping up a PR. it takes my staged code changes and…

DGX agent
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

small workflow note that adds up. /staged-pr is a skill (via slash command) i run when wrapping up a PR. it takes my staged code changes and drafts a concise PR title & description, based on my prefer

agentsharrison-chase--x
6 May 2026
Agents

everyone would have a deeper appreciation for Agent Products that rock because of great Context/Harness Engineering if they… talked to: - LL…

DGX agent

everyone would have a deeper appreciation for Agent Products that rock because of great Context/Harness Engineering if they… talked to: - LLM Base models - Post-Trained models with no harness (no tool

agentsharrison-chase--x
5 May 2026
Model Releases

// HeavySkill // One of the cleaner takes on agentic harness design I've read. They argue that what actually drives agent harness performanc…

DGX agent

// HeavySkill // One of the cleaner takes on agentic harness design I've read. They argue that what actually drives agent harness performance is not the orchestration code. It's a single inner skill:

model-releasesdair-ai--x
5 May 2026
Research

phi-Table: A Statistical Explanation for Global SHAP

DGX agent

arXiv:2512.07578v3 Announce Type: replace-cross Abstract: Global SHAP explanations are typically presented as feature-importance rankings, which identify variables that matter to a black-box model but

researcharxiv-cs-lg
5 May 2026
Model Releases

Documentation https://docs.ollama.com/integrations/claude-desktop

DGX agent

This documentation page describes how to integrate Ollama with Claude Desktop, enabling users to run local language models through the Anthropic Claude interface. The integration allows Claude Desktop

model-releasesollama--x
4 May 2026
Model Releases

I am not sure I would agree with all of this, but the relationship between Anthropic and Claude is quite different than the relationship bet…

DGX agent

I am not sure I would agree with all of this, but the relationship between Anthropic and Claude is quite different than the relationship between other labs and their models. And that shows up in lots

model-releasesethan-mollick--x
3 May 2026
Model Releases

Its getting hard to benchmark frontier agent performance on longer tasks. Repeated measurement is very expensive and there are differences b…

DGX agent

Its getting hard to benchmark frontier agent performance on longer tasks. Repeated measurement is very expensive and there are differences between using models in harnesses versus via APIs. I suspect

model-releasesethan-mollick--x
3 May 2026
Model Releases

(Sorry, after seeing so many of these, could not resist): 🚨 BREAKING: Google just dropped a NEW paper that completely deletes RNNs from exi…

DGX agent

(Sorry, after seeing so many of these, could not resist): 🚨 BREAKING: Google just dropped a NEW paper that completely deletes RNNs from existence. No recurrence. No convolutions. Nothing. Just one mec

model-releasesethan-mollick--x
2 May 2026
Model Releases

Demis says he wants to see a Western open source AI stack and that we’re losing to China. He also says Google doesn’t have enough compute to…

DGX agent

Demis says he wants to see a Western open source AI stack and that we’re losing to China. He also says Google doesn’t have enough compute to build two frontier (open and closed) models, which is why G

model-releasesclem-delangue--x
30 Apr 2026
Model Releases

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good …

DGX agent

tweeted about this yesterday and Cursor already dropped the alpha today! 🚀very cool to see how us, them, and others have converged on good design patterns in Agent + Harness Engineering: 1. Tuning dif

model-releasesharrison-chase--x
30 Apr 2026
Model Releases

I released LLM 0.32a0 this morning, a major backwards-compatible refactor of my LLM Python library and CLI tool for working with language mo…

DGX agent

I released LLM 0.32a0 this morning, a major backwards-compatible refactor of my LLM Python library and CLI tool for working with language models - the new changes should help LLM work better with reas

model-releasessimon-willison--x
29 Apr 2026
Applications

Xiaomi MiMo-V2.5-Pro achieves multiple breakthroughs in the latest Arena rankings (Apr 26, 2026) 🔥 🏆 Text Arena (Expert) — #6 globally | #…

DGX agent

Xiaomi MiMo-V2.5-Pro achieves multiple breakthroughs in the latest Arena rankings (Apr 26, 2026) 🔥 🏆 Text Arena (Expert) — #6 globally | #1 open-source model Also #1 among Chinese models, with Xiaomi

applicationsjeremy-howard--x
29 Apr 2026
Agents

amazing! it’s like talking to an ai from the past

DGX agent

amazing! it’s like talking to an ai from the past Announcing Talkie: a new, open-weight historical LLM! We trained and finetuned a 13B model on a newly-curated dataset of only pre-1930 data. Try it be

agentsyohei-nakajima--x
28 Apr 2026
Applications

The new LLM trained only on pre-1931 text is small enough that it can potentially run on device, so, with the right tools, you can get a ful…

DGX agent

The new LLM trained only on pre-1931 text is small enough that it can potentially run on device, so, with the right tools, you can get a fully vintage version of Siri, but from the era of Downton Abbe

applicationsethan-mollick--x
28 Apr 2026
Model Releases

Try now: https://www.together.ai/models/nvidia-nemotron-3-nano-omni#

DGX agent

NVIDIA Nemotron-3 Nano Omni is now available to try through Together AI's platform, offering access to a compact multimodal model capable of processing both text and audio inputs. This announcement hi

model-releasestogether-ai--x
28 Apr 2026
Model Releases

Lately I've been having fun with running coding agents fully locally. The setup I landed on is: - Pi agent - Gemma 4 26B A4B - Server of cho…

DGX agent

Lately I've been having fun with running coding agents fully locally. The setup I landed on is: - Pi agent - Gemma 4 26B A4B - Server of choice: LM Studio/Ollama/llama.cpp I wrote a step-by-step guide

model-releasesclem-delangue--x
27 Apr 2026
Model Releases

Running Qwen3.5-397B-A17B (4bit quants, 177 GB) on two DGX Sparks using llama.cpp with RPC and RDMA:

DGX agent

This post documents a technical demonstration of running the large Qwen3.5-397B-A17B model across distributed hardware using llama.cpp with advanced networking protocols. The approach leverages 4-bit

model-releasesgeorgi-gerganov--x
27 Apr 2026
Model Releases

Higher res figures (and summaries) in the LLM architecture gallery: https://sebastianraschka.com/llm-architecture-gallery/#card-deepseek-v4-…

DGX agent

Sebastian Raschka has updated his LLM architecture gallery with higher resolution figures and improved summaries, including coverage of the DeepSeek V4 model architecture. This resource provides visua

model-releasessebastian-raschka--x
26 Apr 2026
Model Releases

Balanced Performance Across Artistic Styles: More uniform quality across diverse aesthetic domains, effectively reducing style-dependent qua…

DGX agent

Qwen's latest model improvements focus on achieving more consistent and uniform performance quality across different artistic styles and aesthetic domains, reducing the variability in output quality t

model-releasesqwen--x
25 Apr 2026
Model Releases

I hope the upgrade to DeepSeek v4 will make the bot comments on here more bearable.

DGX agent

This post expresses hope that upgrading to DeepSeek v4 (an AI model) will improve the quality of bot-generated comments on a platform or service. The statement implies current bot comments are conside

model-releasesethan-mollick--x
24 Apr 2026
Model Releases

Spud 🥔 and DeepSeek 🐳 V4 on the same day?! Is it Christmas? Here goes our night to bring the whale up 🔨

DGX agent

Fireworks AI announced the release or deployment of DeepSeek V4, a large language model, alongside another project or update called 'Spud' on the same day, with the team planning to work through the n

model-releasesfireworks-ai--x
24 Apr 2026
Model Releases

Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted, what …

DGX agent

Always enjoy getting to chat with @swyx on our annual cross-episode with @latentspacepod on the state of AI. We hit on what’s shifted, what surprised us and what’s next. We covered: ▪️ Whether AI infr

model-releasesswyx--x
23 Apr 2026
Model Releases

Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️

DGX agent

Day 0 vLLM support for Qwen3.6-27B! @vllm_project ♥️❤️ 🎉 Day-0 vLLM support for Qwen3.6-27B! Congrats to @Alibaba_Qwen on the new 27B dense model release. Looking forward to more of the Qwen3.6 series

model-releasesqwen--x
23 Apr 2026
Model Releases

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships …

DGX agent

Hugging Face Releases ml-intern: An Open-Source AI Agent that Automates the LLM Post-Training Workflow [The 'AI Intern' that actually ships SOTA models ] This isn't just another ML Research Loop wrapp

model-releasesclem-delangue--x
22 Apr 2026
Model Releases

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'her…

DGX agent

Kimi K2.6 is free on Nous Portal for the next 24 hours Made possible by @vercel's AI Gateway & @Kimi_Moonshot Run 'hermes update', then 'hermes model' and select Kimi K2.6 to try out one of the most i

model-releaseskimi-moonshot--x
22 Apr 2026
Model Releases

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part…

DGX agent

LiteParse, our OSS document parser, is really good at parsing complex PDF layouts, text, and tables into a clean spatial grid. The best part is it doesn't use VLMs or any ML models at all. It's entire

model-releasesjerry-liu--x
22 Apr 2026
Agents

Access GPT Image 2.0 natively in Hermes Agent Update now to get access - just run `hermes update` and select your image generation tool mode…

DGX agent

Access GPT Image 2.0 natively in Hermes Agent Update now to get access - just run `hermes update` and select your image generation tool model with `hermes tools` Introducing ChatGPT Images 2.0 A state

agentsnous-research--x
21 Apr 2026
Model Releases

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready t…

DGX agent

Flexible Aspect Ratios ChatGPT Images 2.0 supports aspect ratios as wide as 3:1 and as tall as 1:3. It can generate outputs that are ready to fit the formats you need, from wide banners and presentati

model-releasesopenai--x
21 Apr 2026
Agents

Kimi is the current open-source SOTA on Artificial Analysis

DGX agent

Kimi is the current open-source SOTA on Artificial Analysis Moonshot’s Kimi K2.6 is the new leading open weights model. Kimi K2.6 lands at #4 on the Artificial Analysis Intelligence Index (54) behind

agentskimi-moonshot--x
21 Apr 2026
Model Releases

Thinking… Generating… Livestreaming… https://openai.com/live/

DGX agent

OpenAI announced a live event showcasing new capabilities or features, likely including extended thinking, content generation, and livestreaming functionalities. The event demonstrated how these featu

model-releasesopenai--x
21 Apr 2026
Industry

Kimi-K2.6 is on HuggingFace

DGX agent

Kimi-K2.6 is a language model that has been released on HuggingFace, a popular platform for sharing machine learning models and datasets. The announcement was made by Clem Delangue, likely indicating

industryclem-delangue--x
20 Apr 2026
Model Releases

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world…

DGX agent

The Jensen + @dwarkesh_sp podcast was fantastic. Jensen is someone who understood how ecosystems work and someone who understands real-world trade, policy and controls work. And in some deeper sense h

model-releasessoumith-chintala--x
20 Apr 2026
Local Ai

在 Ollama 运行 hermes

DGX agent

This post likely discusses how to run the Hermes language model using Ollama, an open-source tool for running large language models locally. It probably provides instructions or insights on setting up

local-aiollama--x
18 Apr 2026
Model Releases

I have found 4.7 great for design, reverted back to 4.6 extended for everything else Anyone else like this?

DGX agent

I have found 4.7 great for design, reverted back to 4.6 extended for everything else Anyone else like this? Introducing Claude Design by Anthropic Labs: make prototypes, slides, and one-pagers by talk

model-releasesemad-mostaque--x
17 Apr 2026
Model Releases

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon auton…

DGX agent

Claude Opus 4.7 is now available as an Agent Preview inside of Devin! Anthropic has clearly optimized Claude Opus 4.7 for long-horizon autonomy, unlocking a class of deep investigation work we couldn'

model-releasescognition-ai--x
16 Apr 2026
Model Releases

GLM-5.1 Tool Calling Issue Fix & Chat Template Update If you are running GLM-5.1 with vLLM/SGLang and using tool calling, please update your…

DGX agent

GLM-5.1 Tool Calling Issue Fix & Chat Template Update If you are running GLM-5.1 with vLLM/SGLang and using tool calling, please update your chat template. http://huggingface.co/zai-org/GLM-5.1/blob/m

model-releaseszhipu-ai--x
16 Apr 2026
Model Releases

Here's Qwen 3.6-35B-A3B v.s. Claude Opus 4.7 for 'Generate an SVG of a flamingo riding a unicycle', in case you thought Qwen might be cheati…

DGX agent

This post compares the performance of Qwen 3.6-35B-A3B and Claude Opus 4.7 models on a creative task of generating SVG code for a flamingo riding a unicycle, likely demonstrating differences in their

model-releasessimon-willison--x
16 Apr 2026
Agents

LM Performance:Qwen3.6-35B-A3B outperforms the dense 27B-param Qwen3.5-27B on several key coding benchmarks and dramatically surpasses its d…

DGX agent

LM Performance:Qwen3.6-35B-A3B outperforms the dense 27B-param Qwen3.5-27B on several key coding benchmarks and dramatically surpasses its direct predecessor Qwen3.5-35B-A3B, especially on agentic cod

agentsqwen--x
16 Apr 2026
Model Releases

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing!

DGX agent

Sorry for the long wait, everyone! As I said, Qwen is going to keep open-sourcing! ⚡ Meet Qwen3.6-35B-A3B:Now Open-Source!🚀🚀 A sparse MoE model, 35B total params, 3B active. Apache 2.0 license. 🔥 Agen

model-releasesclem-delangue--x
16 Apr 2026
Model Releases

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio. Whether you’…

DGX agent

Gemini 3.1 Flash TTS is rolling out in Google Vids and is available today in preview via the Gemini API and in @GoogleAIStudio. Whether you’re creating a pitch deck or recording a passion project, tra

model-releasesgoogle-ai--x
15 Apr 2026
Model Releases

Have found the same things! Using glm-5 as a daily driver for a lot of things

DGX agent

Have found the same things! Using glm-5 as a daily driver for a lot of things We've tested new OSS models the moment they're released for a while at Lindy. Inference is our #1 cost by a lot (more than

model-releasesharrison-chase--x
15 Apr 2026
Local Ai

More details: https://blog.comfy.org/p/ernie-image-day-0-support

DGX agent

ComfyUI announced day-zero support for ERNIE Image, Baidu's image generation model, with full implementation details available on their official blog. The integration allows users to run ERNIE Image m

local-aicomfyui--x
15 Apr 2026
Model Releases

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of file…

DGX agent

here’s a good application of harness permissions Programmatically Enforced Auto-Research: - auto-research loops usually expose a set of files that the agent is allowed to edit to hill climb a metric/e

model-releasesharrison-chase--x
14 Apr 2026
Applications

(This is a big part of what was called emergence in earlier academic work on unexpected LLM ability gains)

DGX agent

Ethan Mollick discusses the concept of 'emergence' in large language models (LLMs), referring to the phenomenon where AI systems appear to suddenly develop unexpected capabilities as they scale. The p

applicationsethan-mollick--x
14 Apr 2026
Tools

EinsteinArena is open-source and the leaderboard is live. We welcome contributions and feedback! → https://www.together.ai/blog/einsteinaren…

DGX agent

EinsteinArena is an open-source benchmarking platform developed by Together AI designed to evaluate and rank AI models, with a live public leaderboard tracking model performance. The project welcomes

toolstogether-ai--x
13 Apr 2026
Tutorials

Introducing DDTree: accelerates speculative decoding by drafting a tree with one block diffusion pass, then verifying multiple likely contin…

DGX agent

Introducing DDTree: accelerates speculative decoding by drafting a tree with one block diffusion pass, then verifying multiple likely continuations together. Paper: https://liranringel.github.io/ddtre

tutorialsjeremy-howard--x
13 Apr 2026
Tutorials

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All o…

DGX agent

Maybe hot take - I’ve read a bunch of RL for image generation papers over last few months and honestly it’s been pretty disappointing. All of them are variations of GRPO and all of them are incrementa

tutorialsjeremy-howard--x
13 Apr 2026
← Previous
1…4142434445…128
Next →