I just realized how good MoE models are for consumer hardware
I've been tinkering around with LLM for a while now, started with LM Studio like probably all of us β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
I just realized how good MoE models are for consumer hardware
I've been tinkering around with LLM for a while now, started with LM Studio like probably all of us β¦
Suggestion - this sub should have post flairs that mention the amount of vram/unified ram
The amount of fast ram is the single most important factor for llm use. There are lots of people thaβ¦
Intel B70 vs AMD R9700: Has anyone actually tested the noise levels (dB) at full load?
Both 32GB GDDR6. Intel somewhat slower but lower TDP (230W) and a little cheaper. I wish AMD did offβ¦
438 USD for a 3080 20GB isnβt bad
submitted by /u/xw1y [link] [comments]β¦
Microsoft should've released something like Qwen3.6-27B / Gemma-4-31B already. They released MAI models now
Did they abandon Phi series? I remember that few were expecting for Phi-5. I see that they came withβ¦
[NEW MODEL] SupraLabs just released a new model! - Supra-50M-Reasoning
SupraLabs just released a new model! - Supra-50M-Reasoning Hello again r/LocalLLaMA! Supra-50M-Reasoβ¦
Bringing Gemma 4 12B to your Laptop: Unlocking Local, Agentic Workflows with Google AI Edge
submitted by /u/zxyzyxz [link] [comments]β¦
What is your current go-to stack for running a fully local AI agent?
Curious to know what quantization level (GGUF/EXL2) you find balances speed and smarts for daily useβ¦
Benchmark & Reality Check on Gemma 4 12B: Great model, but your local settings are probably breaking it (Fix inside)
I completed a Python bug hunting benchmark with Gemma 4 12B. I used the Unsloth Dynamic Q5 GGUF modeβ¦
RTX Pro 4500 Blackwell Performance Numbers
RTX Pro 4500 Blackwell About one month ago I asked the fine people of Reddit for some upgrade adviceβ¦
Is there a quant of Granite 30b I can run in 12gb of VRAM/32gb of RAM?
I am hoping there is submitted by /u/MrMrsPotts [link] [comments]β¦
[Opinion] Gemma4-12B means that Google is going hard after the market of IoT and mobile and we're helping them
I know it might be a no-brainer in retrospect, but hear me out, y'all, it's not the whole story. [tiβ¦