u/Gloomy_Letterhead395

▲ 15 r/ROCm+1 crossposts

Fastest qwen 3.8 27b for AMD gpu?

Hey, just wondering if there are forks or exact gguf versions that give fastest prompt processing and token gen speeds for AMD gpu?
Looking to run q8 or q6
Vram 96gb
W7900 + w7800 both 48gb
With bandwidth mismatch, tensor paralleling amd equivalent not working

reddit.com
u/Gloomy_Letterhead395 — 13 hours ago

Huge conundrum regarding W7900 and W7800

I am building a local token machine
W7900 and w7800 variant have both same 48gb
But a price difference of 1000$
Has anyone used them multiple of these
How is the performance which one works for local lmstudio

reddit.com
u/Gloomy_Letterhead395 — 3 months ago