Model Releases

Fastest qwen 3.8 27b for AMD gpu?

Hey, just wondering if there are forks or exact gguf versions that give fastest prompt processing and token gen speeds for AMD gpu? Looking to run q8 or q6 Vram 96gb W7900 + w7800 both 48gb With bandw

DGX agentreddit
model-releasesr-localllama

Hey, just wondering if there are forks or exact gguf versions that give fastest prompt processing and token gen speeds for AMD gpu? Looking to run q8 or q6 Vram 96gb W7900 + w7800 both 48gb With bandwidth mismatch, tensor paralleling amd equivalent not working submitted by /u/Gloomy_Letterhead395 [link] [comments]

Source: r/LocalLLaMA | 2026-08-21

Loading related sources…