r/LocalLLaMA 🤖 Ai 👁 0

Luce Spark: a 35B MoE on a 16 GB GPU, without the offload tax

Hey fellow Llamas, your time is precious, so I'll keep it short. TL;DR: 33-35B MoE on a 16 GB GPU. Qwen3.6 35B-A3B: 13.3 GiB (was ~20.5). Laguna XS.2 33B-A3B: 14.6 GiB (was 18.8). Both measured on an RTX 3090, both unde

Luce Spark: a 35B MoE on a 16 GB GPU, without the offload tax
📄

This source provides headlines only. Use the button below to read the complete article on the original site.

📰 Read the original article on r/LocalLLaMA

Originally published by r/LocalLLaMA. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.