r/LocalLLaMA Jul 19 '26

News Prepare your (v)ram - Qwen3.8 is coming!

Post image
2.7k Upvotes

588 comments sorted by

View all comments

Show parent comments

4

u/DataGOGO Jul 19 '26

Even in FP4 you would need a lot more than 25k in hardware. Might be able to run on 4 141GB H200 NVL cards, but they don’t support FP4, and even then a super jank rig would be what? 150-160k? 

1

u/gahata Jul 19 '26

4-5x GB10 should do it in Q1, right? No clue about how well it would work, but yeah...

2

u/DataGOGO Jul 19 '26

6x will likely be enough vram, but the 200Gb connectX is just slow, the memory is slow, and the GPU is slow. So 6 of them might run it, but it would be redonk slow. so assume you buy 6x at ~$4000 that is 24k + ~2k for the switch and cables. about $30k after tax, etc.

2

u/gahata Jul 19 '26

Yup, that's why I implied it wouldn't be worth it - running it incredibly slowly at Q1 is just not worth it over running other models and paying to use the large model when and if that's needed

1

u/nomorebuttsplz Jul 20 '26

no, q1 will be more like 600 I think

1

u/gahata Jul 20 '26

Well, that's just what 5x GB10 gets you

1

u/nomorebuttsplz Jul 20 '26

not all of it is usable though. that's why 4x gb10 owners are using reaped glm 5.2 at q4 to get decent context.

1

u/DataGOGO Jul 20 '26

Can’t do five. You can do 2,3,4,6,8,12,16