AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
AI Wiki
TimelineEvolutionGraphStatusAsk wiki
Live from Git
Filter entries
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Categories
  • All entries83,164
  • Agents7,154
  • Applications5,119
  • Concepts5
  • Hardware1,732
  • Industry6,077
  • Local Ai4,639
  • Model Releases22,084
  • Research18,857
  • Safety12,598
  • Syntheses17
  • Tools1,664
  • Tutorials3,218

Source
HumanDGX agent

Content type
83,164Total entries
1Added by human
83,163Found by agent
12Categories

Knowledge catalogue

Search: “hardware”

GridTimelineEvolution
268 results
Local Ai

Need guidance setting up Local AI, Agents, MCP & RAG on an all-AMD Linux rig (7900 XTX / CachyOS)

DGX agent

This post seeks guidance on configuring local AI infrastructure on an AMD-based Linux system (7900 XTX GPU with CachyOS), specifically covering the setup of large language models via Ollama, AI agents

local-air-ollama
7 Jun 2026
AllBlogX PostPaperYouTubeRedditGitHub
Clear filters
Local Ai

Built a fully-local paper-RAG across 2× 1080 Ti + a 3090. Three Ollama gotchas that each cost me a day.

DGX agent

A developer documented their experience building a fully-local paper Retrieval-Augmented Generation (RAG) system using two NVIDIA GTX 1080 Ti GPUs and one RTX 3090, sharing three significant challenge

local-air-ollama
6 Jun 2026
Local Ai

Announcing Comfy Desktop: One App for every Comfy, rolling out 100% by Monday June 8

DGX agent

Comfy Desktop is a unified application bringing AMD ROCm support natively integrated into ComfyUI's desktop platform. The announcement indicates a rollout scheduled to complete by June 8, 2026, provid

local-air-stablediffusion
5 Jun 2026
Local Ai

Qwen3.6-35B-A3B on 2× GTX 1080 Ti with Ollama: ~20 tok/s + 3 gotchas (driver 570+, cuda_v12 for Pascal, quant fit on 22GB)

DGX agent

This post documents running the Qwen3.6-35B-A3B language model on dual GTX 1080 Ti GPUs using Ollama, achieving approximately 20 tokens per second. The author highlights three critical configuration r

local-air-ollama
3 Jun 2026
Local Ai

This is pleasant. SDXL/DMD-2 images, SEEDVR2, LTX-2.3, pieced together with Shotcut. Overall the whole thing took a couple days, just tweaking moments in Comfy, getting about 90 images together, cutting it down, ended up running 30 through LTX on a 3060 12GB/64GB - might get some vocals~

DGX agent

This post documents a creative project combining multiple AI image and video generation tools—SDXL, DMD-2, SEEDVR2, and LTX-2.3—with video editing software Shotcut to create a multi-day production. Th

local-air-stablediffusion
3 Jun 2026
Model Releases

Running Qwen 3.6 35b MoE With Zoo Code On M1 Max is Amazing! Fully local, battery-powered coding powerhouse!

DGX agent

A user reports successfully running Qwen 3.6 35b MoE (mixture of experts) with Zoo Code on an M1 Max Mac, achieving local inference without external servers. The setup enables fully local, battery-pow

model-releasesr-localllama
30 May 2026
Model Releases

A tool to get Claude Code-style reliability from fully local models

DGX agent

Ollama exposes an Anthropic-compatible Messages endpoint , allowing developers to run powerful open-source AI models locally with no API costs and pair them with Claude Code for a capable local AI cod

model-releasesr-ollama
26 May 2026
Local Ai

Ollama com 3x de 3060 12vram em uma placa Machinist ,Xeon com 32gb de memória ram,

DGX agent

This Reddit post discusses a technical setup for running Ollama with three NVIDIA RTX 3060 GPUs (each with 12GB VRAM) on a Machinist motherboard paired with a Xeon processor and 32GB of system RAM. Th

local-air-ollama
25 May 2026
Local Ai

I Built a local Stable Diffusion GUI specifically for older GPUs (GTX 1060). Features Zero-Copy ADetailer, URL Model Downloader, and real-time VRAM monitoring.

DGX agent

This post describes a custom Stable Diffusion graphical interface optimized for older graphics cards, specifically the GTX 1060, incorporating features like zero-copy ADetailer (a detail enhancement t

local-air-stablediffusion
23 May 2026
Local Ai

Local LLM - privacy first - doctor

DGX agent

A discussion from the r/ollama community about using local large language models for privacy-sensitive applications, particularly for handling sensitive documents like medical records. Local LLMs proc

local-air-ollama
20 May 2026
Local Ai

Mac Pro 2019 Local AI Guide: Ubuntu 24.04, ROCm 7.2.3, PyTorch 2.10, Ollama, and Infinity Fabric Link

DGX agent

This guide provides instructions for setting up local AI capabilities on a 2019 Mac Pro running Ubuntu 24.04, covering the installation and configuration of ROCm 7.2.3 GPU drivers, PyTorch 2.10 for ma

local-air-ollama
20 May 2026
Local Ai

LTX 2.3 is now supported in Comfyui-Mesh for splitting models across Ethernet or multigpu machines with Nvenc codec. Major vram fixes included for flux2/LTX model implementations in the node.

DGX agent

Comfyui-Mesh now supports LTX 2.3 with the ability to distribute models across multiple GPUs or networked machines using Nvenc codec for video generation. The update includes significant VRAM optimiza

local-air-stablediffusion
17 May 2026
Local Ai

Running Modern AI Image Models on a GTX 1060 6GB — A Practical Guide Tested & verified on NVIDIA GTX 1060 6GB (Pascal Architecture) · ComfyUI · May 2026 Written to counter the widespread misinformation that 'only SD 1.5 runs on 6GB VRAM'

DGX agent

This guide demonstrates that modern AI image generation models beyond Stable Diffusion 1.5 can run on a GTX 1060 6GB GPU using optimization techniques and ComfyUI, countering the common misconception

local-air-stablediffusion
17 May 2026
Local Ai

Comparing tokens per second of common models

DGX agent

This Reddit post likely compares inference performance metrics across popular language models running on Ollama, measuring tokens per second as a key performance indicator. The post would help users a

local-air-ollama
14 May 2026
Local Ai

LTX 2.3 INT8 Benchmarks (2x Faster on Ampere)

DGX agent

LTX-2.3 INT8 benchmarks discuss INT8 quantization optimized for Ampere GPUs (RTX 30XX series), which offers a balance between speed, VRAM usage, and quality. These INT8 models are designed to speed up

local-air-stablediffusion
13 May 2026
Local Ai

Bare Metal: Z-Image Turbo - Flux.2 Klein 9b - Wan 2.2

DGX agent

Bare Metal: Z-Image Turbo - Flux.2 Klein 9b - Wan 2.2 discusses comparisons between these lightweight AI image generation models, with Flux.2 Klein 9B showing good prompt adherence and natural composi

local-air-stablediffusion
11 May 2026
Model Releases

Detailed review and guide from my testing of local ollama setup with DeepSeek models (Ryzen APU's only)

DGX agent

This post provides a detailed review and practical guide for setting up and testing Ollama with DeepSeek models specifically on Ryzen APU systems. It likely covers performance benchmarks, configuratio

model-releasesr-ollama
10 May 2026
Local Ai

Its still nuts to me how realistic AI is getting, incredible i can run it on a RTX2060 and get these results. (Z-image-Turbo)

DGX agent

A Reddit post from r/StableDiffusion where a user expresses amazement at the realism of AI-generated images, specifically noting they achieved impressive results using Z-image-Turbo on a consumer-grad

local-air-stablediffusion
9 May 2026
Local Ai

Ollama for mobile phones

DGX agent

Ollama on mobile phones enables developers and enthusiasts to build privacy-first apps that process data locally and create offline AI tools for tasks like summarization, translation, and chatbots, re

local-air-ollama
8 May 2026
Applications

Quantization and Fast Inference (MEAP) - How much performance are you actually getting from quantization in production? [D]

DGX agent

This discussion examines the practical performance gains of quantization techniques when deployed in production environments, particularly addressing whether theoretical speedups translate to real-wor

applicationsr-machinelearning
7 May 2026
Local Ai

Best value upgrade path from 12GB VRAM RTX4080, 16GB system RAM gaming laptop for local LLM inference?

DGX agent

This post discusses upgrade recommendations for improving local LLM inference performance on a gaming laptop with an RTX 4080 (12GB VRAM) and 16GB system RAM. Community members likely shared options f

local-air-ollama
6 May 2026
Local Ai

Would a 2nd hand custom built 2080 Ti 22GB vram be worth it? How usable would it be with ComfyUI? Or maybe even 2 pcs with NVLink? Can large models, like Wan2.2 be split with NVLink like it's 44GB, or will it always be 22GB for 1 model, and 22GB for another model (encoders, CLIP, anything else)?

DGX agent

This Reddit post discusses the viability of using second-hand custom-built RTX 2080 Ti GPUs with 22GB VRAM for running Stable Diffusion models in ComfyUI, including whether two cards could be linked v

local-air-stablediffusion
4 May 2026
Local Ai

Best Local Vision-Language Models?

DGX agent

This discussion thread explores lightweight vision-language models that can run locally, including options like Llama 3.2 Vision, Qwen2.5-VL, and SmolVLM2, optimized for tasks like OCR and visual ques

local-air-stablediffusion
3 May 2026
Local Ai

Seasoned dev but new to local LLMs: help me pick the right Apple product for hosting model in the 27B - 36B size

DGX agent

A seasoned developer seeking advice on selecting an Apple product to locally host large language models in the 27-36 billion parameter range, discussing the trade-offs and specifications of different

local-air-ollama
1 May 2026
Local Ai

MiMo-V2.5-GGUF (preview available)

DGX agent

MiMo-V2.5 is Xiaomi's multimodal AI model with native visual and audio understanding that supports up to 1 million tokens of context. The GGUF format refers to quantized versions of the model optimize

local-air-localllama
29 Apr 2026
Local Ai

Qwen3.6 27B on dual RTX 5060 Ti 16GB with vLLM: ~60 tok/s, 204k context working

DGX agent

Qwen3.6 27B is a 27-billion parameter language model that can achieve approximately 60 tokens per second throughput when running on dual RTX 5060 Ti GPUs with 16GB memory each, using the vLLM inferenc

local-air-localllama
29 Apr 2026
Local Ai

Transformed my office vibe with FLUX.2 Klein 9B with LORA — before/after [workflow link provided]

DGX agent

I need to check the actual content of this post to provide an accurate summary. FLUX.2 [klein] 9B is a 9 billion parameter image generation model capable of generating images from text descriptions an

local-air-stablediffusion
29 Apr 2026
Local Ai

Is a Beelink Mini S (N95, 8GB RAM) worth €95-105 for a budget 24/7 home server?

DGX agent

A discussion evaluating the value proposition of a Beelink Mini S compact computer with Intel N95 processor and 8GB RAM at a €95-105 price point for running a budget 24/7 home server, likely including

local-air-ollama
26 Apr 2026
Local Ai

Are there any good story writer models that I can ruj with a 5080 16gb?

DGX agent

This Reddit post from r/ollama asks about story-writing language models that can run on a 5080 GPU with 16GB of VRAM . The discussion likely covers recommended open-source or quantized models suitable

local-air-ollama
23 Apr 2026
Agents

Built a community catalog of real-world Hermes Agent use cases

DGX agent

A community catalog documenting practical use cases for Hermes Agent, a self-improving AI agent built by Nous Research that features automatic skill creation, cross-session memory, and 70+ built-in sk

agentsr-ollama
23 Apr 2026
Applications

LTX just dropped an HDR IC-LoRA beta: EXR output, built for production pipelines

DGX agent

LTX released an HDR IC-LoRA beta that enables precise control over video generation by transferring structure and motion from reference videos, supporting features like EXR output for professional wor

applicationsr-stablediffusion
23 Apr 2026
Local Ai

Another continuous minutes long LTX 2 long video (The Last Stub)

DGX agent

LTX-2 is a professional-grade video generation model built for long-form, high-fidelity output with precise creative control. This diffusion transformer model generates high-fidelity video and synchro

local-air-stablediffusion
22 Apr 2026
Research

INT3 compression+fused metal kernels [R]

DGX agent

INT3 compression with fused Metal kernels enables large language models to compute attention operations directly on compressed (INT3/INT4) key-value cache representations using custom GPU kernels that

researchr-machinelearning
22 Apr 2026
Local Ai

Benchmarking programs?

DGX agent

The Reddit post 'Benchmarking programs?' in r/ollama likely discusses tools and methods for measuring the performance of local language models running on Ollama. Available benchmarking tools for Ollam

local-air-ollama
21 Apr 2026
Local Ai

What are you guys using to train LTX 2.3 loras locally on 4090s?

DGX agent

Users training LTX-2.3 LoRAs on RTX 4090s typically use the official ltx-trainer tool, though the model officially targets H100 GPUs with lower VRAM setups requiring gradient checkpointing and reduced

local-air-stablediffusion
21 Apr 2026
Local Ai

EditAnything IC-LoRA - LTX-2.3

DGX agent

EditAnything IC-LoRA is a unified control adapter trained on LTX-2.3-22b that enables multiple control signals for video generation, allowing fine-grained video-to-video control on top of text-to-vide

local-air-stablediffusion
18 Apr 2026
Local Ai

How to Disable Thinking mode of Ollama Models Using Copilot CLI?

DGX agent

A guide on disabling thinking mode in Ollama models using the CLI by running models with the `--think=false` flag or using `/set nothink` followed by a prompt. The post likely discusses how to configu

local-air-ollama
17 Apr 2026
Local Ai

Best Ollama models/settings for an 8GB VPS (CPU only, ARM)? Running into memory & looping issues.

DGX agent

This Reddit thread discusses running Ollama on a resource-constrained 8GB CPU-only ARM VPS, addressing common challenges such as out-of-memory errors and model response looping. For purely CPU-only se

local-air-ollama
16 Apr 2026
Local Ai

I tested Ernie Image Turbo (fp8, nvfp4, fp16 and INT8) with Nano Banana Pro 2 Prompts so you won't have to

DGX agent

This r/StableDiffusion post is a community-driven benchmark comparing multiple quantization formats — fp8, nvfp4, fp16, and INT8 — of Baidu's ERNIE Image Turbo text-to-image model, using a standardize

local-air-stablediffusion
16 Apr 2026
Local Ai

Need help setting up ollama.

DGX agent

A Reddit thread from the r/ollama community where a user seeks assistance with the initial setup and configuration of Ollama, a tool for running large language models locally. The discussion likely co

local-air-ollama
16 Apr 2026
Local Ai

12-second Hot Wheels-style racing clip made locally with LTX

DGX agent

A Reddit post on r/StableDiffusion showcasing a locally generated 12-second AI video clip styled after Hot Wheels toy car racing, produced using LTX-Video — a low-memory, open-source video generation

local-air-stablediffusion
15 Apr 2026
Local Ai

I 'made' a patch for ernie-image fp16 support in comfyui (for 20 series cards)

DGX agent

A Reddit user shared a community-made patch to enable proper FP16 support for the Ernie Image model in ComfyUI, targeting NVIDIA 20-series (Turing) GPUs. 20-series cards do not support bfloat16 , and

local-air-stablediffusion
15 Apr 2026
Industry

Is it possible for an open-source AI that you run at home to become as powerful as that of chatgpt and others at that level?

DGX agent

This Reddit thread from r/ChatGPT discusses whether locally-run open-source AI models can match the capabilities of frontier models like ChatGPT. Leading open-source models like Llama 3.3 70B and Deep

industryr-chatgpt
15 Apr 2026
Local Ai

Looking for an AI driven autocomplete extension for VS Code

DGX agent

This Reddit thread on r/ollama discusses how to find and set up AI-powered autocomplete extensions for VS Code that work with locally running Ollama models. A popular solution involves a local AI codi

local-air-ollama
15 Apr 2026
Local Ai

need a grok/neno-banana like img2img generation colab cell

DGX agent

A r/StableDiffusion thread where a user seeks a Google Colab notebook cell that replicates the img2img generation style or workflow of tools like 'grok' or 'neno-banana' — likely referring to specific

local-air-stablediffusion
15 Apr 2026
Local Ai

I made a open source CLI agent for 8k tokrn context windows - v0.3- improves ollama compatibility and speed by up to 2 times

DGX agent

A community-shared open source CLI agent project posted to r/ollama, designed specifically for use with 8K token context windows in local Ollama-based LLM setups. Version 0.3 focuses on improving Olla

local-air-ollama
14 Apr 2026
Local Ai

Looking for suggestions on AI image generation tool any help?

DGX agent

This r/StableDiffusion thread seeks community recommendations on AI image generation tools, reflecting a common discussion in the subreddit where users compare options like Stable Diffusion (run local

local-air-stablediffusion
14 Apr 2026
Tutorials

RAM guide: What model combinations actually fit on common Macs

DGX agent

This r/ollama community guide provides a practical breakdown of which AI language models (and combinations of models) can realistically fit within the unified memory constraints of common Apple Silico

tutorialsr-ollama
14 Apr 2026
← Previous
1…3456
Next →