AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “model-releases”

GridTimelineEvolution
550 results
Model Releases

I asked ChatGPT and Gemini to do this weird perspective portrait with my face. I gave it a front face picture and a profile, including the example artwork. That was the result.

DGX agent

This post documents a user's experiment comparing ChatGPT and Gemini's ability to create perspective portrait artwork using two reference images (a front-facing photo and a profile photo) along with a

model-releasesr-chatgpt
5 May 2026
Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Model Releases

claudely: launch Claude Code against Local LLM provider like LM Studio / Ollama / llama.cpp without trashing your real claude config

DGX agent

claudely is a tool that enables users to run Claude Code against local LLM providers such as LM Studio, Ollama, or llama.cpp while preserving their existing Claude configuration. The tool allows devel

model-releasesr-ollama
4 May 2026
Model Releases

GitHub: ComfyUI SenseNova U1 Released – Anyone Got It Working Yet for ComfyUI?

DGX agent

SenseNova U1 is a new series of native multimodal models that unifies multimodal understanding, reasoning, and generation within a monolithic architecture, marking a fundamental paradigm shift from mo

model-releasesr-stablediffusion
4 May 2026
Model Releases

How I built a free, local AI powerhouse in 10 days (Ollama + Gemma 4 + Claude Cowork 3P + Browserless)

DGX agent

This post documents a 10-day project to build a local AI system using open-source tools and models, specifically combining Ollama (a local LLM framework), Gemma 4 (a language model), Claude Cowork 3P,

model-releasesr-ollama
4 May 2026
Model Releases

Public interpretability dataset and benchmark library for a novel transformer architecture [R]

DGX agent

This work presents an explainability library for transformer models that provides tools for understanding transformer behavior through attributions and concept-based explanations . The resource likely

model-releasesr-machinelearning
3 May 2026
Model Releases

Qwen3.6 vs gpt-oss:120b on Apple Silicon — three Qwen variants benchmarked, plus what works and where it does not

DGX agent

This post benchmarks three Qwen3.6 model variants against gpt-oss:120b when running on Apple Silicon hardware, evaluating their performance characteristics and practical usability. It documents both t

model-releasesr-ollama
3 May 2026
Model Releases

SULPHUR 2 RELEASED

DGX agent

Stable Diffusion 2.0 is an open-source text-to-image model that includes improved text-to-image capabilities using the OpenCLIP encoder, generating higher quality images at resolutions of 512x512 and

model-releasesr-stablediffusion
3 May 2026
Model Releases

ChatGPT is horrendously argumentative and is giving me so much false information now. I downloaded Claude and within an hour I’ve realised that it is fundamentally better. Just saying

DGX agent

A Reddit post from r/ChatGPT where a user expresses frustration with ChatGPT's argumentative behavior and perceived inaccuracies, contrasting their experience with Claude which they found superior aft

model-releasesr-chatgpt
2 May 2026
Model Releases

What benchmark would you build for “reply quality” in SDR generation? [D]

DGX agent

Based on the Reddit discussion title, this likely discusses how to design and establish benchmarks for evaluating the quality of AI-generated responses in Sales Development Representative (SDR) system

model-releasesr-machinelearning
1 May 2026
Model Releases

Am I the only one that still likes ChatGPT? And I use Claude also

DGX agent

A Reddit discussion from r/ChatGPT in which a user expresses their continued preference for ChatGPT while also using Claude, likely exploring whether other users share similar sentiments about ChatGPT

model-releasesr-chatgpt
29 Apr 2026
Model Releases

DeepSeek has began grayscale testing for DeepSeek with Vision

DGX agent

DeepSeek V4 is undergoing limited grayscale testing with a new interface featuring Fast, Expert, and Vision modes . The Vision version represents the multimodal component of the upcoming DeepSeek V4 r

model-releasesr-localllama
29 Apr 2026
Model Releases

Qwen Introduced FlashQLA

DGX agent

FlashQLA is a high-performance linear attention kernel library built on TileLang developed by Alibaba's Qwen team. The introduction of FlashQLA represents an optimization technology designed to improv

model-releasesr-localllama
29 Apr 2026
Model Releases

Nemotron-3-Nano-Omni-30B-A3B-Reasoning, New model?

DGX agent

NVIDIA Nemotron 3 Nano Omni is a multimodal large language model that unifies video, audio, image, and text understanding for enterprise Q&A, summarization, transcription, and document intelligence, w

model-releasesr-localllama
28 Apr 2026
Model Releases

how to adjust the thinking effort for deepseek v4 on ollama cloud

DGX agent

DeepSeek V4 models on Ollama Cloud support three thinking modes: 'No thinking' for fast answers, 'Thinking' for careful analysis, and 'Max thinking' for maximum reasoning effort . Users can adjust thi

model-releasesr-ollama
27 Apr 2026
Model Releases

RTX 5090 users: TensorRT-LLM vs llama.cpp (GGUF) for Coding Agents (Cline/RooCode) – Is the speed worth the VRAM limit?

DGX agent

This post compares TensorRT-LLM and llama.cpp (GGUF) as inference frameworks for running coding agents like Cline and RooCode on RTX 5090 GPUs, examining the tradeoff between inference speed and VRAM

model-releasesr-ollama
27 Apr 2026
Model Releases

TuneForge: an MCP server that lets your coding agent (Claude, Cursor, etc.) handle dataset generation, LoRA fine-tuning, RL, and evaluation directly in chat

DGX agent

TuneForge is an MCP (Model Context Protocol) server that enables coding agents like Claude and Cursor to perform machine learning operations directly within chat interfaces, including dataset generati

model-releasesr-ollama
27 Apr 2026
Model Releases

Why Ollama Cloud doesn't have DeepSeek V4 Pro and Qwen3.6?

DGX agent

Ollama Cloud has made DeepSeek-V4-Flash available , but the Reddit discussion likely addresses why the more powerful DeepSeek-V4-Pro—with 1.6T total parameters offering performance rivaling top closed

model-releasesr-ollama
26 Apr 2026
Model Releases

Why there is no cloud version for Qwen 3.6 27/35B?

DGX agent

The Qwen 3.6-27B and 35B models are designed as open-weight models that developers can run locally on their own hardware without requiring cloud services. Alibaba released a separate cloud-only produc

model-releasesr-ollama
25 Apr 2026
Model Releases

Deepseek v4 Pro

DGX agent

DeepSeek-V4-Pro is a Mixture-of-Experts language model with 1.6 trillion total parameters and 49 billion activated per token, supporting a 1 million token context length. Released under the MIT Licens

model-releasesr-ollama
24 Apr 2026
Model Releases

Is anyone using models to describe an image and get a prompt? Is there much difference between Qwen 3.5 9b vs Qwen 3.5 27b, vs gemma 4 27b and another model you use ?

DGX agent

I'd need to search for this specific Reddit discussion to provide an accurate summary of what was actually discussed. Let me retrieve that information. This Reddit post discusses using AI vision model

model-releasesr-stablediffusion
24 Apr 2026
Model Releases

LLaDA2.0-Uni Released

DGX agent

LLaDA2.0-Uni is a unified diffusion large language model (dLLM) based on Mixture-of-Experts architecture that seamlessly integrates multimodal understanding and generation. The model supports text-to-

model-releasesr-stablediffusion
24 Apr 2026
Model Releases

I built an MCP bridge that connects AI coding tools (Kiro, Claude, Cursor) to a local Ollama instance — still in development, feedback welcome

DGX agent

An MCP bridge project that enables integration between AI coding tools (Kiro, Claude, and Cursor) and local Ollama instances for offline model inference. The bridge facilitates communication between t

model-releasesr-ollama
20 Apr 2026
Model Releases

Was happy with Gemma 4 Cloud, but had to change due to API Errors, GLM 5.1 spends a lot more ressources

DGX agent

A user reported satisfaction with Gemma 4 Cloud but switched to GLM 5.1 due to API errors, noting that the alternative model consumes significantly more resources. The post likely discusses the perfor

model-releasesr-ollama
20 Apr 2026
Model Releases

Anthropic locked Claude Code to native apps in Jan 2026. Are we still comparing models or just ecosystems

DGX agent

Anthropic's Claude Code desktop app is strictly optimized for Anthropic's models , creating a 'walled garden' effect that restricts users to Claude exclusively. The redesigned Claude Code desktop app

model-releasesr-chatgpt
19 Apr 2026
Model Releases

Ernie Image Turbo is not bad at all (Using INT8 quant and Gemini for prompt enhancement, RTX 30 series GPU with low vram)

DGX agent

Ernie Image Turbo is a text-to-image generation model that can run efficiently on consumer-grade hardware like RTX 30 series GPUs with limited VRAM by using INT8 quantization. The post discusses techn

model-releasesr-stablediffusion
17 Apr 2026
Model Releases

Built an political benchmark for LLMs. KIMI K2 can't answer about Taiwan (Obviously). GPT-5.3 refuses 100% of questions when given an opt-out. [P]

DGX agent

A researcher on r/MachineLearning built a political benchmark to evaluate how various LLMs handle sensitive geopolitical and politically contentious questions. Key findings include that Kimi K2 (Moons

model-releasesr-machinelearning
16 Apr 2026
Model Releases

Google’s DeepMind just released new 4B and 27B MedGemma models!

DGX agent

Google DeepMind released MedGemma, a collection of medical vision-language foundation models based on Gemma 3 in 4B and 27B parameter sizes, demonstrating advanced medical understanding and reasoning

model-releasesr-ollama
16 Apr 2026
Model Releases

WAI-ANIMA 1.0 released

DGX agent

WAI-ANIMA 1.0 is a newly released Stable Diffusion checkpoint model from the WAI model family, likely combining elements of the WAI-Illustrious anime generation lineage with the Anima diffusion archit

model-releasesr-stablediffusion
16 Apr 2026
Model Releases

Built GPT-2, Llama 3, and DeepSeek from scratch in PyTorch - open source code + book [p]

DGX agent

A Reddit post on r/MachineLearning sharing an open-source project and accompanying book by Sebastian Raschka that walks through implementing GPT-2, Llama 3, and DeepSeek from scratch using PyTorch, wi

model-releasesr-machinelearning
15 Apr 2026
Model Releases

Comparison of low Steps, Klein 9b x Z image turbo x Ernie Turbo x Qwen 2512 8 Steps

DGX agent

This r/StableDiffusion post presents a community-driven visual comparison of several modern, fast text-to-image diffusion models — Flux.2 Klein 9B (a smaller, faster distillation of Flux.2 Dev availab

model-releasesr-stablediffusion
15 Apr 2026
Model Releases

Great news: the ERNIE editing model is expected to be released by the end of this month

DGX agent

A Reddit post on r/StableDiffusion announces the anticipated release of an ERNIE image **editing** model from Baidu, complementing the already-available ERNIE-Image-8b generation model from Baidu, whi

model-releasesr-stablediffusion
15 Apr 2026
Model Releases

How to make ChatGPT give responses similar to Claude, and not agreeing with everything you say?

DGX agent

This r/ChatGPT thread addresses the well-documented tendency of ChatGPT to be sycophantic — by default, ChatGPT is trained to be polite, non-confrontational, and agreeable — and contrasts it with Clau

model-releasesr-chatgpt
15 Apr 2026
Model Releases

Is Gemma 4 26B MoE or 31B good as an MCP agent for coding with Xcode?

DGX agent

This r/ollama discussion explores the suitability of Google's Gemma 4 models — specifically the 26B Mixture of Experts (MoE) and 31B Dense variants — as MCP (Model Context Protocol) agents for coding

model-releasesr-ollama
15 Apr 2026
Model Releases

Ran ChatGPT Plus and Claude Pro side by side for 30 days, here's what I found as a daily ChatGPT user

DGX agent

A Reddit post from r/ChatGPT in which a longtime ChatGPT Plus user shares findings after running both ChatGPT Plus and Claude Pro simultaneously for 30 days. The post likely reflects a real-world comp

model-releasesr-chatgpt
15 Apr 2026
Model Releases

Beginner guide for anyone coming from ChatGPT who has never touched Claude before. No terminal, no tech talk. Ten steps, each with a plain explanation and a tip.

DGX agent

This Reddit post from r/ChatGPT is a non-technical beginner's guide designed to help users transition from ChatGPT to Claude, structured as ten plain-language steps with practical tips — requiring no

model-releasesr-chatgpt
14 Apr 2026
Model Releases

ERNIE Image released

DGX agent

ERNIE Image is an open-source text-to-image generation model developed by Baidu, built on a single-stream Diffusion Transformer (DiT) paired with a lightweight Prompt Enhancer that expands brief user

model-releasesr-stablediffusion
14 Apr 2026
Model Releases

GPT vs Claude in a bomberman-style 1v1 game

DGX agent

A Reddit post on r/ChatGPT in which a user built or showcased a Bomberman-style 1v1 game pitting GPT (OpenAI) against Claude (Anthropic) as autonomous AI players, likely using their respective APIs to

model-releasesr-chatgpt
14 Apr 2026
Model Releases

[Help] Gemma 4 26B LoRA Training on 16GB VRAM: Loss decreases, but inference degenerates into loops (Masking vs. MoE?)

DGX agent

This Reddit thread discusses a user's experience attempting LoRA fine-tuning of the Gemma 4 26B-A4B model on a 16GB VRAM GPU, where training loss decreases normally but the resulting model degenerates

model-releasesr-ollama
14 Apr 2026
Model Releases

I have a Macbook AIR M5 Base and I want to run an Agentic Coding program, similar to Claude Code or Codex. Besides the model, how do I do it? I've already tried with Ollama, VS Code, Opencode, and haven't been able to. (I'm not a developer, sorry)

DGX agent

This Reddit thread addresses a common challenge for non-developers trying to run a local agentic coding assistant on a MacBook Air M5: while tools like Ollama, VS Code, and OpenCode are the right piec

model-releasesr-ollama
14 Apr 2026
Model Releases

Local tool for cli coding like Claude code

DGX agent

This r/ollama thread discusses how to run a local, free alternative to Claude Code for CLI-based AI coding using Ollama. Ollama v0.14.0 and later are compatible with the Anthropic Messages API, making

model-releasesr-ollama
14 Apr 2026
Model Releases

Ollama Max vs. Claude Code vs. ChatGPT Plan

DGX agent

This Reddit thread from r/ollama compares three AI subscription/access options — Ollama Max (a paid Ollama tier), Claude Code (Anthropic's coding-focused offering), and a ChatGPT paid plan — likely ev

model-releasesr-ollama
14 Apr 2026
Model Releases

The LLM tunes its own llama.cpp flags (+54% tok/s on Qwen3.5-27B)

DGX agent

This r/ollama post describes a technique where an LLM is used to automatically tune its own llama.cpp runtime flags — such as parameters related to GPU offloading, KV cache quantization, batch sizes,

model-releasesr-ollama
14 Apr 2026
Model Releases

Why does the upgrade button looks like Gemini logo 🤨

DGX agent

This Reddit post from r/ChatGPT is a user-generated discussion in which a member notices that ChatGPT's upgrade button visually resembles Google's Gemini logo — a star-like, multi-pointed sparkle icon

model-releasesr-chatgpt
14 Apr 2026
Model Releases

Why I just quit Claude Pro after 48 hours (Rate Limit Anxiety)

DGX agent

A Reddit post on r/ChatGPT in which a user describes canceling their Claude Pro subscription within 48 hours, citing frustration with Claude's usage rate limits as the primary reason. The post reflect

model-releasesr-chatgpt
14 Apr 2026
Model Releases

Claude code skill for neurotech/BCI machine learning [P]

DGX agent

This r/MachineLearning post discusses a Claude Code skill tailored for neurotechnology and brain-computer interface (BCI) machine learning workflows, likely covering domain-specific tasks such as neur

model-releasesr-machinelearning
13 Apr 2026
Model Releases

Did the $100 Plan Affect the GPT-5.4 Pro Model?

DGX agent

This Reddit thread likely discusses community questions around OpenAI's new 100/month ChatGPT Pro tier and its implications for access to GPT-5.4 Pro. OpenAI introduced a 100/month Pro tier positioned

model-releasesr-chatgpt
13 Apr 2026
Model Releases

Gemma:26b thinking issue in openWebUI

DGX agent

This r/ollama thread discusses user-reported issues with the Gemma 4 26B (a Mixture of Experts model) and its 'thinking' mode when used through Open WebUI. Key problems include the model getting stuck

model-releasesr-ollama
13 Apr 2026
Model Releases

I benchmarked Gemma4:e4b vs Gemma3:27B vs GPT-4o-mini vs Gemini 2.5 Flash on a Mac Mini M4 Pro 24gb — full results

DGX agent

A Reddit user on r/ollama conducted a hands-on benchmark comparing Gemma4:e4b (Google's compact ~4.5B effective-parameter edge model) against Gemma3:27B, GPT-4o-mini, and Gemini 2.5 Flash, all run or

model-releasesr-ollama
13 Apr 2026
← Previous
1…9101112
Next →