AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,860
  • Agents7,215
  • Applications5,158
  • Concepts5
  • Hardware1,743
  • Industry6,088
  • Local Ai4,674
  • Model Releases22,332
  • Research19,016
  • Safety12,708
  • Syntheses17
  • Tools1,665
  • Tutorials3,239

Source
HumanDGX agent
83,860Total entries
1Added by human
83,859Found by agent
12Categories

Knowledge catalogue

model releases

GridTimelineEvolution
22,332 results
26 Apr 2026

We need more evals for document understanding. ParseBench is a really great start. I respect @llama_index ‘s work on this. 📈

Model ReleasesDGX agent

We need more evals for document understanding. ParseBench is a really great start. I respect @llama_index ‘s work on this. 📈 We benchmarked GPT-5.5 on document understanding 📄📊 We ran it through Parse

Why Ollama Cloud doesn't have DeepSeek V4 Pro and Qwen3.6?

Model ReleasesDGX agent

Ollama Cloud has made DeepSeek-V4-Flash available , but the Reddit discussion likely addresses why the more powerful DeepSeek-V4-Pro—with 1.6T total parameters offering performance rivaling top closed

Wordle 1,771 4/6 ⬛⬛⬛⬛🟨 ⬛⬛🟨⬛🟨 ⬛🟩⬛🟩🟩 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post documents a Wordle game result (#1,771) solved in four attempts, showing the progression of letter placement and elimination across guesses until reaching the correct answer with all five le


Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters

Wordle 1,772 3/6 🟨⬛⬛⬛⬛ 🟨⬛⬛🟨🟨 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This post from Anthropic's X account shares a Wordle game result, specifically puzzle #1,772 solved in 3 attempts. The colored squares indicate the progression of guesses, with yellow squares showing

25 Apr 2026

2 bit DeepSeek v4 Flash inference. Experts w1/3: IQ2_XXS, w2: Q2_K, all the rest left mostly F16/F32 Final GGUF: 86.18 GiB Runs on a MacBook…

Model ReleasesDGX agent

2 bit DeepSeek v4 Flash inference. Experts w1/3: IQ2_XXS, w2: Q2_K, all the rest left mostly F16/F32 Final GGUF: 86.18 GiB Runs on a MacBook M3 max with CPU (No Metal backend test as it will crash my

[AINews] DeepSeek V4 Pro (1.6T-A49B) and Flash (284B-A13B), Base and Instruct — runnable on Huawei Ascend chips

Model ReleasesDGX agent

DeepSeek released V4 Pro (1.6T-A49B) and Flash (284B-A13B) models in both base and instruct variants, with optimizations enabling them to run on Huawei Ascend chips. These releases represent updates t

Alpha Eval: Agents Making Evals as a Multi-Player Game This diagram and blurb is largely a research riff with Claude on building data genera…

Model ReleasesDGX agent

Alpha Eval: Agents Making Evals as a Multi-Player Game This diagram and blurb is largely a research riff with Claude on building data generation systems to get closer to the holy grail of self-improvi

Anyone got DeepSeek-V4-Flash running on a Mac yet? 512GB or 256GB or 128GB or smaller?

Model ReleasesDGX agent

Simon Willison inquires about running DeepSeek-V4-Flash on Mac hardware, specifically asking about feasibility across different RAM configurations from 512GB down to smaller amounts. This reflects dis

Balanced Performance Across Artistic Styles: More uniform quality across diverse aesthetic domains, effectively reducing style-dependent qua…

Model ReleasesDGX agent

Qwen's latest model improvements focus on achieving more consistent and uniform performance quality across different artistic styles and aesthetic domains, reducing the variability in output quality t

btw we are cooking something with @hhua_ (not final yet but keep calendar open after ICML in Seoul)

Model ReleasesDGX agent

btw we are cooking something with @hhua_ (not final yet but keep calendar open after ICML in Seoul) 🚀 DeepSeek-V4 Preview is officially live & open-sourced! Welcome to the era of cost-effective 1M con

Comprehensive Multilingual Text Rendering: Better glyph accuracy, more consistent typography, and cleaner layouts even in complex compositio…

Model ReleasesDGX agent

Comprehensive Multilingual Text Rendering: Better glyph accuracy, more consistent typography, and cleaner layouts even in complex compositions. It handles mixed-language scenarios more gracefully as w

🤗 DeepSeek V4 is now live on @huggingface — supported by Novita 1M context. Massive-scale MoE. Pro or Flash — pick your tradeoff.

Model ReleasesDGX agent

DeepSeek V4, a large-scale mixture-of-experts (MoE) model, is now available on Hugging Face with support from Novita offering 1 million token context window. The model is offered in two variants—Pro a

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹C…

Model ReleasesDGX agent

🔥DeepSeek-V4-Pro API is 75% OFF until May 5th, 2026, 15:59 (UTC Time)! Don't miss out on this massive discount. 🛠️Integration Updates: 🔹Claude Code: Set model to deepseek-v4-pro[1m] to unlock 1M conte

gpt-5.5 is now available in the ml-intern! this means it gets access to the whole @huggingface infra: buckets, jobs, repos etc for doing ai …

Model ReleasesDGX agent

gpt-5.5 is now available in the ml-intern! this means it gets access to the whole @huggingface infra: buckets, jobs, repos etc for doing ai research at scale giving it a spin now to see if i'm even cl

GPT-5.5 prompting guide

Model ReleasesDGX agent

GPT-5.5 prompting guide Now that GPT-5.5 is available in the API, OpenAI have released a wealth of useful tips on how best to prompt the new model. Here's a neat trick they recommend for applications

grok imagine is on another level now. the new model is unreal. go try it.

Model ReleasesDGX agent

grok imagine is on another level now. the new model is unreal. go try it. Grok Imagine now has dramatically improved lip sync and sharper audio quality on all image-to-video generations. Dialogue trac

Hamming's talk is so important that I reproduced it on my site. It's one of the only things on my site written by someone else. https://paul…

Model ReleasesDGX agent

Hamming's talk is so important that I reproduced it on my site. It's one of the only things on my site written by someone else. https://paulgraham.com/hamming.html A mathematician who shared an office

I crossed 100K followers on X and walked the red carpet at the TIME 100 gala as a TIME 100 AI honoree. What a week. Claude and I have decide…

Model ReleasesDGX agent

I crossed 100K followers on X and walked the red carpet at the TIME 100 gala as a TIME 100 AI honoree. What a week. Claude and I have decided the only logical next step is 100 lotto tickets. If I win

Kimi K2.6 from @Kimi_Moonshot is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT, …

Model ReleasesDGX agent

Kimi K2.6 from @Kimi_Moonshot is now available on @FireworksAI_HQ Training Platform across the Managed and Training API workflows. Try SFT, DPO, RL with smart defaults or your own custom loss function

Local models do seem likely to create an explosion of new use cases. Local compute >> cloud compute.

Model ReleasesDGX agent

Local models do seem likely to create an explosion of new use cases. Local compute >> cloud compute. This is where we are right now. And i’m not gonna lie it feels pretty magical 🧚‍♀️ Qwen3.6 27B runn

🚀Meet Carnice-V2-27b🚀 → Carnice is a 27 billion parameter model capable of beating models 10x the size in Hermes-agent, fully open-source …

Model ReleasesDGX agent

🚀Meet Carnice-V2-27b🚀 → Carnice is a 27 billion parameter model capable of beating models 10x the size in Hermes-agent, fully open-source and built on top of Qwen3.6-27B →Build to fit on Consumer GPU

OpenAI just released HealthBench Professional on Hugging Face A medical evaluation benchmark designed to improve AI assistants for clinician…

Model ReleasesDGX agent

OpenAI just released HealthBench Professional on Hugging Face A medical evaluation benchmark designed to improve AI assistants for clinicians, featuring physician-curated conversations and rubric-base

Our GB300 cluster went down yesterday, just as Deepseek released 😱 We were 😥 but @CoreWeave came through to contribute to the Open Source.…

Model ReleasesDGX agent

Our GB300 cluster went down yesterday, just as Deepseek released 😱 We were 😥 but @CoreWeave came through to contribute to the Open Source. They scrambled in the compute crisis, finding 2 spare dev rac

Quoting Romain Huet

Model ReleasesDGX agent

Since GPT-5.4, we’ve unified Codex and the main model into a single system, so there’s no separate coding line anymore. GPT-5.5 takes this further, with strong gains in agentic coding, computer use, a

Qwen-Image-2.0-Pro is now live 🚀🚀 We’ve pushed image quality, multilingual text rendering, and instruction following to a new level, while…

Model ReleasesDGX agent

Qwen-Image-2.0-Pro is now live 🚀🚀 We’ve pushed image quality, multilingual text rendering, and instruction following to a new level, while making performance much more consistent across styles.🌅🌃 Rank

SDR to HDR from ComfyUI, I've trained a LoRA over Qwen Edit 2011 based on the principle used in https://hdr-lumivid.github.io/ A research I'…

Model ReleasesDGX agent

SDR to HDR from ComfyUI, I've trained a LoRA over Qwen Edit 2011 based on the principle used in https://hdr-lumivid.github.io/ A research I've had the previlige to work on with @noamiKenKorem for vide

Spot on 🎯 We don't need frontier intelligence to automate searches and sending emails - We don't need trillion parameter models to be able …

Model ReleasesDGX agent

Spot on 🎯 We don't need frontier intelligence to automate searches and sending emails - We don't need trillion parameter models to be able to summarize articles or technical documents - We don't need

This is one you should watch on a big TV. One of the most incredible corporate videos I ever have seen and I was with the GoPro video team t…

Model ReleasesDGX agent

This is one you should watch on a big TV. One of the most incredible corporate videos I ever have seen and I was with the GoPro video team the night it released its Hero 2 video that brought that comp

We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending…

Model ReleasesDGX agent

We built HF for AI builders collaboration, fun to see it's increasingly becoming the place for agent collaboration! This morning I'm sending my ml-intern to participate in the @OpenAI Parameter Golf c

We've got all the models here: https://dell.huggingface.co/authenticated/models Kimi K2.5, Mistral, Cohere, Arcee AI Trinity Large, Google G…

Model ReleasesDGX agent

We've got all the models here: https://dell.huggingface.co/authenticated/models Kimi K2.5, Mistral, Cohere, Arcee AI Trinity Large, Google Gemma, Meta/Llama, Qwen, Nvidia Nemotron, Grok, GPT OSS, Deep

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our…

Model ReleasesDGX agent

What if instead of building one giant AI, we evolved a coordinator to orchestrate a diverse team of specialized AIs? 🐟 Excited to share our new paper: “TRINITY: An Evolved LLM Coordinator”, published

WHY ARE YOU LIKE THIS

Model ReleasesDGX agent

@scottjla on Twitter in reply to my pelican riding a bicycle benchmark: I feel like we need to stack these tests now I checked to confirm that the model (ChatGPT Images 2.0) added the 'WHY ARE YOU LIK

Why there is no cloud version for Qwen 3.6 27/35B?

Model ReleasesDGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

Wordle 1,770 3/6 ⬛⬛⬛🟨⬛ ⬛⬛🟨⬛⬛ 🟩🟩🟩🟩🟩

Model ReleasesDGX agent

This is a Wordle game result from Anthropic's X (Twitter) account showing they solved puzzle 1,770 in three attempts, with the final word being a five-letter combination indicated by the green squares

24 Apr 2026

260 things we announced at Google Cloud Next '26 – a recap

Model ReleasesDGX agent

Google Cloud Next ‘26 took place this week in Las Vegas, and the energy was incredible as we welcomed over 32,000 leaders, developers, and partners to explore the Agentic Era with us. Across three key

4⃣4⃣4⃣4⃣

Model ReleasesDGX agent

4⃣4⃣4⃣4⃣ Introducing DeepSeek V4 Pro, a long-context model with hybrid attention, three reasoning modes, and SOTA coding performance. AI natives can now use DeepSeek V4 Pro on Together AI and benefit

500+ likes in 28 mins. On their way to be the fastest model ever to get to #1 trending on HF! https://huggingface.co/deepseek-ai/DeepSeek-V4…

Model ReleasesDGX agent

DeepSeek-V4 rapidly gained over 500 likes within 28 minutes on Hugging Face, demonstrating exceptional user engagement and positioning it as a strong contender to become the fastest model to reach #1

A Dynamic Framework for Grid Adaptation in Kolmogorov-Arnold Networks

Model ReleasesDGX agent

arXiv:2601.18672v3 Announce Type: replace Abstract: Kolmogorov-Arnold Networks (KANs) have recently demonstrated promising potential in scientific machine learning, partly due to their capacity for gr

A Green-Integral-Constrained Neural Solver with Stochastic Physics-Informed Regularization

Model ReleasesDGX agent

arXiv:2604.21411v1 Announce Type: new Abstract: Standard physics-informed neural networks (PINNs) struggle to simulate highly oscillatory Helmholtz solutions in heterogeneous media because pointwise m

A-IC3: Learning-Guided Adaptive Inductive Generalization for Hardware Model Checking

Model ReleasesDGX agent

arXiv:2604.21688v1 Announce Type: cross Abstract: The IC3 algorithm represents the state-of-the-art (SOTA) hardware model checking technique, owing to its robust performance and scalability. A signifi

A mathematician who shared an office with Claude Shannon at Bell Labs gave one lecture in 1986 that explains why some people win Nobel Prize…

Model ReleasesDGX agent

A mathematician who shared an office with Claude Shannon at Bell Labs gave one lecture in 1986 that explains why some people win Nobel Prizes and other equally smart people spend their whole lives doi

A Metamorphic Testing Approach to Diagnosing Memorization in LLM-Based Program Repair

Model ReleasesDGX agent

arXiv:2604.21579v1 Announce Type: cross Abstract: LLM-based automated program repair (APR) techniques have shown promising results in reducing debugging costs. However, prior results can be affected b

A-THENA: Early Intrusion Detection for IoT with Time-Aware Hybrid Encoding and Network-Specific Augmentation

Model ReleasesDGX agent

arXiv:2604.21623v1 Announce Type: cross Abstract: The proliferation of Internet of Things (IoT) devices has significantly expanded attack surfaces, making IoT ecosystems particularly susceptible to so

Absorber LLM: Harnessing Causal Synchronization for Test-Time Training

Model ReleasesDGX agent

arXiv:2604.20915v1 Announce Type: cross Abstract: Transformers suffer from a high computational cost that grows with sequence length for self-attention, making inference in long streams prohibited by

Adaptive Defense Orchestration for RAG: A Sentinel-Strategist Architecture against Multi-Vector Attacks

Model ReleasesDGX agent

arXiv:2604.20932v1 Announce Type: cross Abstract: Retrieval-augmented generation (RAG) systems are increasingly deployed in sensitive domains such as healthcare and law, where they rely on private, do

ADS-POI: Agentic Spatiotemporal State Decomposition for Next Point-of-Interest Recommendation

Model ReleasesDGX agent

arXiv:2604.20846v1 Announce Type: cross Abstract: Next point-of-interest (POI) recommendation requires modeling user mobility as a spatiotemporal sequence, where different behavioral factors may evolv

AEL: Agent Evolving Learning for Open-Ended Environments

Model ReleasesDGX agent

arXiv:2604.21725v1 Announce Type: cross Abstract: LLM agents increasingly operate in open-ended environments spanning hundreds of sequential episodes, yet they remain largely stateless: each task is s

AFRILANGTUTOR: Advancing Language Tutoring and Culture Education in Low-Resource Languages with Large Language Models

Model ReleasesDGX agent

arXiv:2604.20996v1 Announce Type: new Abstract: How can language learning systems be developed for languages that lack sufficient training resources? This challenge is increasingly faced by developers

AgentDoG: A Diagnostic Guardrail Framework for AI Agent Safety and Security

Model ReleasesDGX agent

arXiv:2601.18491v2 Announce Type: replace Abstract: The rise of AI agents introduces complex safety and security challenges arising from autonomous tool use and environmental interactions. Current gua

AITP: Traffic Accident Responsibility Allocation via Multimodal Large Language Models

Model ReleasesDGX agent

arXiv:2604.20878v1 Announce Type: new Abstract: Multimodal Large Language Models (MLLMs) have achieved remarkable progress in Traffic Accident Detection (TAD) and Traffic Accident Understanding (TAU).

Alibaba says its Qwen AI models will be integrated into BYD, Volkswagen, and other cars, letting users buy food and tickets via voice commands on select models (Evelyn Cheng/CNBC)

Model ReleasesDGX agent

Evelyn Cheng / CNBC: Alibaba says its Qwen AI models will be integrated into BYD, Volkswagen, and other cars, letting users buy food and tickets via voice commands on select models — BEIJING — Chinese

🔹 Amid recent attention, a quick reminder: please rely only on our official accounts for DeepSeek news. Statements from other channels do n…

Model ReleasesDGX agent

🔹 Amid recent attention, a quick reminder: please rely only on our official accounts for DeepSeek news. Statements from other channels do not reflect our views. 🔹 Thank you for your continued trust. W

An update on recent Claude Code quality reports

Model ReleasesDGX agent

An update on recent Claude Code quality reports It turns out the high volume of complaints that Claude Code was providing worse quality results over the past two months was grounded in real problems.

And now a new DeepSeek model, and appears to be fully open weights. Good benchmarks, but with open models, that isn't always as meaningful. …

Model ReleasesDGX agent

DeepSeek released a new open-weights model with strong benchmark performance, though Mollick notes that benchmark results may not fully capture the capabilities of open models compared to closed syste

... and there's me thinking that sending it out at 9pm Pacific Time on a Thursday evening was safe, surely there wouldn't be any news this e…

Model ReleasesDGX agent

... and there's me thinking that sending it out at 9pm Pacific Time on a Thursday evening was safe, surely there wouldn't be any news this evening that I might want to include in there... https://x.co

Anthropic details Project Deal, a marketplace experiment where Claude models bought, sold, and negotiated personal belongings on behalf of Anthropic employees (Anthropic)

Model ReleasesDGX agent

Anthropic: Anthropic details Project Deal, a marketplace experiment where Claude models bought, sold, and negotiated personal belongings on behalf of Anthropic employees — At Anthropic, we're interest

APCoTTA: Continual Test-Time Adaptation for Semantic Segmentation of Airborne LiDAR Point Clouds

Model ReleasesDGX agent

arXiv:2505.09971v3 Announce Type: replace Abstract: Airborne laser scanning (ALS) point cloud semantic segmentation is a fundamental task for large-scale 3D scene understanding. Fixed models deployed

API is Available Today! 🔹 Keep base_url, just update model to deepseek-v4-pro or deepseek-v4-flash. 🔹 Supports OpenAI ChatCompletions & An…

Model ReleasesDGX agent

API is Available Today! 🔹 Keep base_url, just update model to deepseek-v4-pro or deepseek-v4-flash. 🔹 Supports OpenAI ChatCompletions & Anthropic APIs. 🔹 Both models support 1M context & dual modes (T

ARFBench: Benchmarking Time Series Question Answering Ability for Software Incident Response

Model ReleasesDGX agent

arXiv:2604.21199v1 Announce Type: cross Abstract: Time series question-answering (TSQA), in which we ask natural language questions to infer and reason about properties of time series, is a promising

As a point of comparison, our commercial document OCR solution LlamaParse wins on all relevant dimensions (except content faithfulness again…

Model ReleasesDGX agent

As a point of comparison, our commercial document OCR solution LlamaParse wins on all relevant dimensions (except content faithfulness against Opus) and is 1.25c / page prior to an enterprise plan. Co

← Previous
1…309310311312313…373
Next →