r/LocalLLaMA πŸ€– Ai πŸ‘ 0

4x RTX 3090 PCIe 4.0 x16 - advice? Qwen 3.8 Next Flash?

We're upgrading our server (Threadripper Pro 5955WX, 128GB 8-channel DDR4) from two RTX 3090 to four cards. Currently we're running Qwen 3.8 27b Q8 with vLLM for a few users. I'm wondering what we should do next: Keep ru

πŸ“„

This source provides headlines only. Use the button below to read the complete article on the original site.

πŸ“° Read the original article on r/LocalLLaMA

Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes β€” full credit and traffic to the original publisher.