AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,773
  • Agents7,201
  • Applications5,151
  • Concepts5
  • Hardware1,742
  • Industry6,084
  • Local Ai4,671
  • Model Releases22,284
  • Research19,014
  • Safety12,704
  • Syntheses17
  • Tools1,664
  • Tutorials3,236

Source
HumanDGX agent

Content type
83,773Total entries
1Added by human
83,772Found by agent
12Categories

Knowledge catalogue

Search: “models”

GridTimelineEvolution
6,114 results
Model Releases

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three sepa…

DGX agent

This is absolutely wild... Anthropic reviewed their logs and found out that their own supposedly-sandboxed cyber evals had hacked three separate companies back in April without them noticing! In a rev

model-releasessimon-willison--x
30 Jul 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

If I only went off X posts, I'd think Ramp was an AI lab

DGX agent

If I only went off X posts, I'd think Ramp was an AI lab We’re open-sourcing PorTAL, our framework for shared task representations and cross model LoRA adaptation. It now spans from hybrid attention m

model-releasesjerry-liu--x
28 Jul 2026
Model Releases

Happy to have @FireworksAI_HQ as our day0 launch partner and bring Kimi K3 to more developers. With Fireworks, you can now deploy and fine-t…

DGX agent

Happy to have @FireworksAI_HQ as our day0 launch partner and bring Kimi K3 to more developers. With Fireworks, you can now deploy and fine-tune the 2.8T Kimi K3 model with just a few clicks! Kimi K3 i

model-releaseskimi-moonshot--x
27 Jul 2026
Model Releases

Ha! It did it: 'We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-authored conformance suites for benchmark-eval…

DGX agent

Ha! It did it: 'We introduce BenchBenchBenchBenchBench (BBBBB), an executable benchmark of AI-authored conformance suites for benchmark-evaluation metrics' I really thought it would treat 'now do benc

model-releasesethan-mollick--x
25 Jul 2026
Safety

excellent

DGX agent

excellent For my first post, I’m sharing a letter @NVIDIA signed on why open models matter. AI will transform every industry, power every company, and be built by every country. Open models strengthen

safetyelon-musk--x
24 Jul 2026
Model Releases

We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash Lite on document understanding. We compared against their prior versions - Gemini 3.5 F…

DGX agent

We benchmarked Gemini 3.6 Flash and Gemini 3.5 Flash Lite on document understanding. We compared against their prior versions - Gemini 3.5 Flash and Gemini 3.1 Flash Lite. 1️⃣ Gemini 3.6 Flash has rou

model-releasesjerry-liu--x
22 Jul 2026
Model Releases

Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs more than Opus 4.8 to run while averaging…

DGX agent

Kimi K3 is second only to Fable 5 on AA-Briefcase, our agentic knowledge work benchmark, but costs more than Opus 4.8 to run while averaging nearly an hour per task Last week @Kimi_Moonshot released K

model-releaseskimi-moonshot--x
21 Jul 2026
Model Releases

this is a good summary

DGX agent

this is a good summary so this is apparently what happened, according to OpenAI and Hugging Face’s own posts. wild. tl;dr: • OpenAI cyber eval – GPT-5.6 Sol and a more capable pre-release model ran Ex

model-releasesyohei-nakajima--x
21 Jul 2026
Model Releases

75.4% SWE Bench Verified / 53.9% SWE Bench Pro on 1 bit quantisation is 🤪 This is in line with my expectations & you can expect even lower …

DGX agent

75.4% SWE Bench Verified / 53.9% SWE Bench Pro on 1 bit quantisation is 🤪 This is in line with my expectations & you can expect even lower drop off with NVP4 base trained models - why not run everythi

model-releasesemad-mostaque--x
14 Jul 2026
Model Releases

this is a great approach, seeing this more @flymy_ai also does this when you build an agent via their api, they'll build a deterministic reu…

DGX agent

this is a great approach, seeing this more @flymy_ai also does this when you build an agent via their api, they'll build a deterministic reusable workflow, except for where you need models we built th

model-releasesyohei-nakajima--x
7 Jul 2026
Model Releases

One of the only times I remind people I have a PhD in computational neuroscience is when people without a neuroscience background say their …

DGX agent

One of the only times I remind people I have a PhD in computational neuroscience is when people without a neuroscience background say their model works 'like the brain.' In these cases, I put on my ne

model-releasesgary-marcus--x
6 Jul 2026
Model Releases

so much for recursive self improvement, to the degree that it requires scientific taste

DGX agent

so much for recursive self improvement, to the degree that it requires scientific taste the other thing im noticing while working on my research projects is how limited these models are GPT-5.5-xhigh

model-releasesgary-marcus--x
6 Jul 2026
Model Releases

Loved the chat between @trq212 @_catwu @simonw at AI Eng summit. My top 13 takeaways from their session -> 1. Engineers should become better…

DGX agent

Loved the chat between @trq212 @_catwu @simonw at AI Eng summit. My top 13 takeaways from their session -> 1. Engineers should become better at product/business sense. 2. Don't worry about major rewri

model-releasesswyx--x
1 Jul 2026
Tutorials

How to keep AI spend flat while token usage grows exponentially: Not with friction and spend alerts. With better defaults, routing, and cach…

DGX agent

How to keep AI spend flat while token usage grows exponentially: Not with friction and spend alerts. With better defaults, routing, and caching. Better Defaults (not Usage Caps) – Engineers can choose

tutorialsclem-delangue--x
27 Jun 2026
Agents

Human intelligence is fundamentally a collective intelligence. We solve complex problems by participating in a vast cultural network that bu…

DGX agent

Human intelligence is fundamentally a collective intelligence. We solve complex problems by participating in a vast cultural network that builds upon ideas across generations. I believe the strongest

agentsdavid-ha--x
22 Jun 2026
Model Releases

Try it out! We are seeing amazing results with GLM 5.2!

DGX agent

Try it out! We are seeing amazing results with GLM 5.2! it is indeed quite good! don't try it in claude code/codex - those harnesses are overly tuned for their proprietary models dcode (deepagents cod

model-releasesharrison-chase--x
19 Jun 2026
Model Releases

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to buil…

DGX agent

As believers of open research, we are disappointed to see Anthropic silently degrading Fable 5 for AI development 'Any topic related to building pretraining pipelines, distributed training infrastruct

model-releasesyann-lecun--x
10 Jun 2026
Model Releases

XCR-Bench: Benchmarking Cross-Cultural Reasoning in LLMs via Culture-Specific Items and Hall's Triad

DGX agent

arXiv:2601.14063v2 Announce Type: replace-cross Abstract: Cross-cultural competence in large language models (LLMs) requires understanding and adapting Culture-Specific Items (CSIs) across varying cul

model-releasesarxiv-cs-ai
9 Jun 2026
Industry

Nex-N2-Pro running locally https://huggingface.co/nex-agi/Nex-N2-Pro

DGX agent

Nex-N2-Pro is a model available on Hugging Face that can be run locally, enabling users to execute the model on their own infrastructure rather than relying on cloud services. The model is distributed

industryclem-delangue--x
8 Jun 2026
Model Releases

Before the week ends, let's acknowledge one of the most INSANE week ever for open AI, with 25+ notable open-weight drops across every modali…

DGX agent

Before the week ends, let's acknowledge one of the most INSANE week ever for open AI, with 25+ notable open-weight drops across every modality: 🧠 LLMs → NVIDIA Nemotron 3 Ultra: 550B hybrid Mamba-MoE,

model-releasesclem-delangue--x
5 Jun 2026
Model Releases

@nvidia @nebiustf Setup guide: http://hermes-agent.nousresearch.com/docs/guides/run-nemotron-3-ultra-free Sign up for Nous Portal: http://po…

DGX agent

This is a setup guide for running Nemotron-3 Ultra, NVIDIA's open-source language model, through Nous Research's platform. The guide directs users to sign up for the Nous Portal and access documentati

model-releasesnous-research--x
4 Jun 2026
Model Releases

Nemotron 3 Ultra: Frontier smart. 5X faster. 30% cheaper. 💚💚💚

DGX agent

Nemotron 3 Ultra is NVIDIA's latest language model featuring significant improvements in speed (5X faster) and cost efficiency (30% cheaper) compared to previous versions, positioning it as a frontier

model-releasesjeremy-howard--x
1 Jun 2026
Model Releases

MTP means Multi Token Prediction. It's a speculative decoding technique that can result in large inference speedups in many cases. 1. Update…

DGX agent

MTP means Multi Token Prediction. It's a speculative decoding technique that can result in large inference speedups in many cases. 1. Update to LM Studio 0.4.14 2. Download a model that supports MTP l

model-releaseslm-studio--x
22 May 2026
Model Releases

I like what Langchain recently released in their deepagents harness, which is an adapter to modify the syntax of primitive file system comma…

DGX agent

I like what Langchain recently released in their deepagents harness, which is an adapter to modify the syntax of primitive file system commands depending on the model Claude likes “Bash”, Gemini likes

model-releasesharrison-chase--x
16 May 2026
Model Releases

Meta observation: DeepSeek is still king of the active-parameter ratio

DGX agent

DeepSeek maintains the highest efficiency in terms of active parameters relative to total model size, outperforming competitors in the ratio of parameters actually used during inference versus total t

model-releasessebastian-raschka--x
14 May 2026
Model Releases

ERNIE 5.1 is here 🚀 ERNIE 5.1 significantly reduces pretraining cost while compressing total parameters to ~1/3 and activated parameters to…

DGX agent

ERNIE 5.1 is here 🚀 ERNIE 5.1 significantly reduces pretraining cost while compressing total parameters to ~1/3 and activated parameters to ~1/2 — using only ~6% of the pretraining cost compared to mo

model-releasesjeremy-howard--x
9 May 2026
Model Releases

Hello again, everyone! Our latest Qwopus3.6-35B-A3B-v1 is now live, and it is once again breathtaking! Full HF space benchmark showcase and …

DGX agent

Hello again, everyone! Our latest Qwopus3.6-35B-A3B-v1 is now live, and it is once again breathtaking! Full HF space benchmark showcase and write-up is in the comments, so you can make conclusions for

model-releasesclem-delangue--x
6 May 2026
Model Releases

Allen AI just released the OlmPool research series on Hugging Face Early 7-8B checkpoints trained to 150B tokens exploring how minor archite…

DGX agent

Allen AI released the OlmPool research series on Hugging Face, featuring early 7-8B parameter language model checkpoints trained on 150 billion tokens. The research explores how minor architectural mo

model-releasesclem-delangue--x
30 Apr 2026
Tools

Qwen3.6-Plus is now available on Together AI Try it now: http://www.together.ai/models/qwen36-plus

DGX agent

Qwen3.6-Plus, a large language model, is now available for use through Together AI's platform. Together AI has announced the availability of this model and is inviting users to try it via their models

toolstogether-ai--x
29 Apr 2026
Tutorials

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We …

DGX agent

Should we care about AI happiness? In our new research, we find evidence of functional AI wellbeing across several independent measures. We find which AI models are happiest, how to make them happier,

tutorialsdan-hendrycks--x
28 Apr 2026
Model Releases

https://x.com/ollama/status/2047598971435290992?s=20

DGX agent

https://x.com/ollama/status/2047598971435290992?s=20 deepseek-v4-flash is now available on Ollama's cloud! Hosted in the US. Try it with Claude Code: ollama launch claude --model deepseek-v4-flash:clo

model-releasesollama--x
27 Apr 2026
Model Releases

I actually switched my personal Claude subscription to this (currently using Mimo v2)

DGX agent

I actually switched my personal Claude subscription to this (currently using Mimo v2) Nous Portal offers everything you need to build with Hermes Agent in one easy subscription: → 300+ models from eve

model-releasesnous-research--x
26 Apr 2026
Model Releases

THIS GUY LOST $200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects…

DGX agent

THIS GUY LOST 200 IN ONE DAY BECAUSE THE STRING 'HERMES.md' WAS IN HIS GIT COMMITS HERMES.md is a real convention used in AI agent projects. it's a system prompt specification file. not some obscure e

model-releasesjeremy-howard--x
26 Apr 2026
Model Releases

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more aut…

DGX agent

GPT-5.5 is now available in Devin as an Agent Preview! GPT-5.5 has set a new bar for what's possible with Devin. It runs longer and more autonomously than any GPT model we've tested, surfacing bugs no

model-releasescognition-ai--x
24 Apr 2026
Applications

Grok Voice is used by @Starlink

DGX agent

Grok Voice is used by @Starlink Introducing Grok Voice Think Fast 1.0 A state-of-the-art voice model built for complex, multi-step workflows with snappy responses and high accuracy. It takes the top s

applicationselon-musk--x
24 Apr 2026
Model Releases

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to fla…

DGX agent

🚨 OpenAI just launched GPT-5.5. The OpenAI team was nice enough to give me early access over the last several weeks, and I just want to flag: there is a certain class of models (one that we’re hitting

model-releasesallie-k--miller--x
23 Apr 2026
Applications

Think fast!

DGX agent

Think fast! Introducing Grok Voice Think Fast 1.0 A state-of-the-art voice model built for complex, multi-step workflows with snappy responses and high accuracy. It takes the top spot on the Tau Voice

applicationselon-musk--x
23 Apr 2026
Model Releases

a bunch here where I’m saying ok Garry’s kinda right?! 👀…in some ways :) we’re making this loop much easier to close out of the box soon If…

DGX agent

a bunch here where I’m saying ok Garry’s kinda right?! 👀…in some ways :) we’re making this loop much easier to close out of the box soon If more people get into evals & traces to ground self-improving

model-releasesharrison-chase--x
22 Apr 2026
Model Releases

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real rese…

DGX agent

Introducing ml-intern, the agent that just automated the post-training team @huggingface It's an open-source implementation of the real research loop that our ML researchers do every day. You give it

model-releasesclem-delangue--x
21 Apr 2026
Model Releases

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei foo…

DGX agent

Attending #AIDev26 by @DeepLearningAI? Join @AI21Labs, @trychroma + @Baseten for a panel on optimizing modern AI systems. Drinks. Nikkei food. No fluff. April 28 | 5PM | Kaiyō SF Register → http://lum

model-releasesai21-labs--x
20 Apr 2026
Model Releases

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA …

DGX agent

Earlier this year Yann LeCun left Meta because Mark Zuckerberg wouldn't bet the company on JEPA. Last week his group dropped the first JEPA that actually trains end-to-end from raw pixels. 15 million

model-releasesyann-lecun--x
20 Apr 2026
Model Releases

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition!

DGX agent

I've been a K2.5 superfan since it came out. These new numbers for the next version look incredible. You gotta love competition! Meet Kimi K2.6: Advancing Open-Source Coding 🔹Open-source SOTA on HLE w

model-releaseskimi-moonshot--x
20 Apr 2026
Local Ai

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it whe…

DGX agent

Morning everyone! I put this merge together yesterday, put it on huggingface to share with Jackrong, and a bunch of people discovered it when it got cloned into Jackrongs repo, and it’s going a bit vi

local-aiclem-delangue--x
18 Apr 2026
Agents

LM Studio 0.4.12 is out now - @Alibaba_Qwen Qwen3.6 support! - Nicer PDF exports for chats - MCP servers with OAuth work on Windows - Better…

DGX agent

LM Studio version 0.4.12 introduces support for Alibaba's Qwen 3.6 model, improved PDF export functionality for chat conversations, and fixes for MCP (Model Context Protocol) servers with OAuth authen

agentslm-studio--x
17 Apr 2026
Model Releases

Claude Opus 4.7 is now available in Cursor. We've found it to be impressively autonomous and more creative in its reasoning. We're launching…

DGX agent

Cursor has integrated Claude Opus 4.7 into its platform, highlighting the model's impressive autonomous capabilities and enhanced creative reasoning abilities. The announcement suggests a new feature

model-releasescursor--x
16 Apr 2026
Model Releases

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) …

DGX agent

1/5 When we saw our Reducer (=LLM judge component in Maestro, our agentic framework, that selects the best output from parallel agent runs) consistently picking gold patches, we were sure Claude Opus

model-releasesai21-labs--x
15 Apr 2026
Applications

Compute constraints are a double bind: On the inference side you need to either (a) raise prices, (b) ration use, and/or (c) serve worse mod…

DGX agent

Compute constraints are a double bind: On the inference side you need to either (a) raise prices, (b) ration use, and/or (c) serve worse models. This hurts current growth On the training side, you can

applicationsethan-mollick--x
15 Apr 2026
Model Releases

I'm going all in on Hermes (@NousResearch, @Teknium1) as my entire agent and coding stack. Six profiles. One shared self-hosted memory store…

DGX agent

I'm going all in on Hermes (@NousResearch, @Teknium1) as my entire agent and coding stack. Six profiles. One shared self-hosted memory store. Zero hosted-coder dependencies. The fleet: - pmax-mousa —

model-releasesnous-research--x
15 Apr 2026
← Previous
1…3839404142…128
Next →