Very large amounts of gaming gpus vs AI gpus

TheMightyCat@ani.social · 8 days ago

Well that’s annoying

TheMightyCat@ani.social · 1 month ago

I am really curious how the T/s scales with network speed, as said in the video 5G is quite slow.

TheMightyCat@ani.social · 2 months ago

Yeah i should have specified for at home when saying its a scam, i honestly doubt the companies that are buying thousands of B200s for datacenters are even looking at their pricetags lmao.

Anyway the end goal is to run something like Qwen3-235B at fp8, with some very rough napkin math 300GB vram with the cheapest option the 9060XT comes down at €7126 with 18 cards, which is very affordable. But ofcourse that this is theoretically possible does not mean it will actually work in practice, which is what im curious about.

The inference engine im using vLLM supports ROCm so CUDA should not be strictly required.

TheMightyCat@ani.social · 2 months ago

Very large amounts of gaming gpus vs AI gpus