r/LocalLLaMA 🤖 Ai 👁 0

Kimi K3 full model running on 16x GB10 cluster at 20+tps

Kimi K3 full model running on 16x GB10 cluster at 20+tps average (llama-benchy coherent corpus) 38tps peak, 750tps prefill. This is the first run of full k3 with dspark on my cluster. I will be doing some tests and try t

Kimi K3 full model running on 16x GB10 cluster at 20+tps
📄

This source provides headlines only. Use the button below to read the complete article on the original site.

📰 Read the original article on r/LocalLLaMA

Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.