r/LocalLLaMA 🤖 Ai 👁 0

DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395

Hey fellow llamas. we have something new for Strix Halo owners we thought would be useful to share. i'll keep it short: We were able to fit DeepSeek V4 Flash plus its speculative draft on a single Ryzen AI MAX+ 395 with

DeepSeek V4 Flash, up to 32 tok/s on AMD Ryzen AI MAX+ 395
📄

This source provides headlines only. Use the button below to read the complete article on the original site.

📰 Read the original article on r/LocalLLaMA

Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.