CompanyAnthropic4 recent entries10 Apr 2026Advanced inpaint/edit Klein/Qwen workflowsA Reddit post on r/StableDiffusion discussing advanced ComfyUI workflows that combine the FLUX Klein and Qwen Image Edit models for precision inpainting and image editing tasks. FLUX Klein offers ...→10 Apr 2026Ace Step 1.5 XL ComfyUI automation workflow without lama for generating random tags using qwen, generate song and then give it a rating by using waveform analysisThis is a ComfyUI automation workflow for the ACE-Step 1.5 XL music generation model that uses Qwen (a language model text encoder) to randomly generate music tags/captions without requiring LAMA, ...
CompanyOpenAI8 recent entries8 Jun 2026Ideogram 4Ideogram 4 is Ideogram's first open-weight text-to-image model trained from scratch, introducing structured JSON prompting with multilingual text rendering and explicit layout controls. Released on Ju→10 Jun 2026Ideogram4 vs Flux.2 Dev vs GPT Image 2 vs Nano Banana ProA practical comparison of Ideogram 4.0, Nano Banana Pro, and GPT Image 2 across text rendering, design, photorealism, and real creator workflows. Ideogram 4.0 is the first open-weight challenger to ra→23 Jul 2026TRELLIS.2 can now generate a high-quality 3D asset in under 7 minutes on a 6 GB VRAM CUDA GPU. No ComfyUI Node Nightmare.Not Self Promotion: Just sharing an open-source tool I built to democratize image to 3D creations. For OpenAI Build Week Hackathon I built a free, open-source local Image-to-3D Studio that makes TRELL→25 Jul 2026Launching ComfyUI with a Blank Canvas (StabilityMatrix)Hey everyone, I’m posting this question in the subreddit because I haven’t been able to figure it out with the help of AI. I’ve asked ChatGPT and Gemini, but their answers are all over the place. So, →1 Aug 2026Trying to understand VRAM usage and find the sweet spot for Wan/SCAIL-2 (or other models) on a GPUSo upfront I'll admit that this is a ChatGPT summary of my chat with it about this idea i had, but this post wouldn't exist any other way, so... I’m trying to get a better understanding of how VRAM is→1 Aug 2026Possible to create accurate medieval woodcut style art?Wondering if it's possible to actually produce ai art works that are indistinguishable from authentic medieval woodcut illustrations like the one attached. All the AI attempts I've seen at re creating→5 Aug 2026Stable Diffusion might actually be remembered in the history books, and I don’t think that’s an overstatementHear me out before you roll your eyes. We tend to only recognize turning points in hindsight. Nobody in 1993 thought the Mosaic browser would be a history book moment, but the web is. I think Stable D→9 Aug 2026It took two years, but we finally have a 'local Sora'Who remembers when OpenAI previewed Sora two years ago and the quality felt unreal? We had never seen anything like it. Back then, Sora 1 didn't even generate audio and was heavily censored. Prompt: i
CompanyGoogle7 recent entries14 Apr 2026AI tool to analyze a video and generate a prompt?This Reddit thread from r/StableDiffusion discusses the community's search for AI tools capable of analyzing an existing video and automatically generating a descriptive text prompt from it — essentia→17 Apr 2026Ernie Image Turbo is not bad at all (Using INT8 quant and Gemini for prompt enhancement, RTX 30 series GPU with low vram)Ernie Image Turbo is a text-to-image generation model that can run efficiently on consumer-grade hardware like RTX 30 series GPUs with limited VRAM by using INT8 quantization. The post discusses techn→10 Jun 2026Ideogram4 vs Flux.2 Dev vs GPT Image 2 vs Nano Banana ProA practical comparison of Ideogram 4.0, Nano Banana Pro, and GPT Image 2 across text rendering, design, photorealism, and real creator workflows. Ideogram 4.0 is the first open-weight challenger to ra→23 Jul 2026LTX Desktop v1.1.0 is out: local generation on Apple Silicon, a built-in LoRA library, video extend, and moreLTX Desktop v1.1.0 just shipped with some big updates. Apple Silicon Macs can generate video locally now, there's a built-in LoRA/IC-LoRA library you can browse and apply from inside the app with per-→25 Jul 2026Launching ComfyUI with a Blank Canvas (StabilityMatrix)Hey everyone, I’m posting this question in the subreddit because I haven’t been able to figure it out with the help of AI. I’ve asked ChatGPT and Gemini, but their answers are all over the place. So, →1 Aug 2026Possible to create accurate medieval woodcut style art?Wondering if it's possible to actually produce ai art works that are indistinguishable from authentic medieval woodcut illustrations like the one attached. All the AI attempts I've seen at re creating→5 Aug 2026Stable Diffusion might actually be remembered in the history books, and I don’t think that’s an overstatementHear me out before you roll your eyes. We tend to only recognize turning points in hindsight. Nobody in 1993 thought the Mosaic browser would be a history book moment, but the web is. I think Stable D
CompanyMeta3 recent entries28 Apr 2026Meta is about to release a pixel space model (Tuna-2)Tuna-2 is a unified multimodal model that performs visual understanding and generation directly based on pixel embeddings, employing simple patch embedding layers to encode visual input without a VAE →3 May 2026Best Local Vision-Language Models?This discussion thread explores lightweight vision-language models that can run locally, including options like Llama 3.2 Vision, Qwen2.5-VL, and SmolVLM2, optimized for tasks like OCR and visual ques→31 Jul 2026Using an AMD V620 workstation card for ComfyUI - successA few weeks ago I posted about if it was worth using a V620 for Comfyui, and was told it likely wouldn't work, at least in Windows 11. And if it did, it would be far too slow and unusable. I decided t
CompanyMistral2 recent entries4 May 2026OneTrainer now supports Ernie LoRAOneTrainer is a one-stop solution for all diffusion training needs. The tool now supports the Ernie Image model, which can be trained using LoRA (Low-Rank Adaptation) methods. This adds support for tr→23 Jul 2026Trained a 32B FLUX.2 LoRA on a 24GB AMD 7900 XTX, native ROCm on Windows — full guide + patchesTL;DR: Everyone says QLoRA past ~13B is dead on a 24GB card. I got the full 32B FLUX.2 dev transformer QLoRA-training resident on the GPU on a 7900 XTX under native ROCm on Windows (no ZLUDA, no CUDA
CompanyxAI6 recent entries10 Apr 2026kugel-2 model (VibeVoice finetune) repo is gone. Does anyone know why?The kugel-2 model is a community fine-tune of Microsoft's VibeVoice, a text-to-speech model — and its disappearance is rooted in Microsoft's own removal of the base VibeVoice repository. Microsoft...→10 Apr 2026HappyHorse is from Alibaba ATH, not Grok / Veo 3.2 / Wan 2.7 / Seedance 2HappyHorse-1.0 is an AI video generation model that anonymously appeared on the Artificial Analysis Video Arena leaderboard in early April 2026, where both its V1 and V2 versions rapidly climbed to...→10 Apr 2026Got early access to a real-time interactive video model, here's what I foundI was unable to retrieve the specific Reddit post at the provided URL through my search. The post (reddit.com/r/StableDiffusion/comments/1shxmfk) did not surface in the search results, and I cannot...→10 Apr 2026Bad news on Happy Horse from twitterHappyHorse-1.0 is a pseudonymous AI video generation model that appeared on April 7, 2026, topping the Artificial Analysis Video Arena leaderboard in both text-to-video and image-to-video (no audio...→15 Apr 2026need a grok/neno-banana like img2img generation colab cellA r/StableDiffusion thread where a user seeks a Google Colab notebook cell that replicates the img2img generation style or workflow of tools like 'grok' or 'neno-banana' — likely referring to specific→16 Apr 2026Motif-Video-2BMotif-Video-2B is a 2-billion-parameter open-source video generation diffusion transformer released by Motif Technologies in April 2026, capable of both text-to-video and image-to-video generation fro
CompanyNVIDIA8 recent entries24 Jul 2026Nvidia releases Qwen-Image-Flash'The NVIDIA Qwen-Image-Flash model generates images from text prompts using a four-step, DMD2-distilled version of Qwen/Qwen-Image. The distillation used DMD2 from NVIDIA FastGen, NVIDIA Model Optimiz→25 Jul 2026SVDQuant + native INT8/W4A4 for Krea 2 on ComfyUI — up to 2x faster, works on any modern NVIDIA GPUQuantized Krea 2 Turbo checkpoints for ComfyUI, up to 2x faster and about a third smaller than the usual FP8 version — no calibration dataset, no quality cliff. How to use it (short version): clone th→25 Jul 2026I tried making a cinematic action trailer using Krea 2 + LTX 2.3I wanted to challenge myself and see how far I could push Krea 2 and LTX 2.3, so I decided to create a short cinematic action trailer. It ended up being one of the most enjoyable AI projects I've work→28 Jul 2026Manga Coloring Tool 2Hey everyone! 👋 I'm excited to announce the official release of Manga Coloring Tool 2.0, a completely free, local, open-source web application designed to colorize manga pages and chapters effortlessl→1 Aug 2026Trying to understand VRAM usage and find the sweet spot for Wan/SCAIL-2 (or other models) on a GPUSo upfront I'll admit that this is a ChatGPT summary of my chat with it about this idea i had, but this post wouldn't exist any other way, so... I’m trying to get a better understanding of how VRAM is→2 Aug 2026Comfyui VRAM trackerHello! VRAM tracker is a node that track the full memory lifecycle of a comfyui run: when each weight is reserved, paged into VRAM, computed on, evicted, and freed. It renders it as an interactive HTM→5 Aug 2026RTX 3090 MiniMax H3 Speed Comparison: FP8 Scaled vs INT8 ConvRot (W8A8)Setup: GPU: RTX 3090 24GB RAM: 32GB ComfyUI 0.30.0 PyTorch 2.13.0+cu130 CUDA 13.0 SageAttention enabled (sageattention-2.2.0+cu130torch2.10.0andhigher.post6-cp310-abi3-win_amd64) Spectrum node + Euler→7 Aug 2026~45% lower MiniMax H3 sampler time with new Spectrum settings — degree 1 works surprisingly well (v0.1.8)Follow-up to my original Spectrum MiniMax H3 post: https://www.reddit.com/r/StableDiffusion/comments/1vf1ze3/spectrum_acceleration_for_minimax_h3_in_comfyui/ In that first post, I released the MiniMax