LocalLLaMA post tier list
Since there is much (justified) whining about post quality, I thought it would be helpful to get a sβ¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
LocalLLaMA post tier list
Since there is much (justified) whining about post quality, I thought it would be helpful to get a sβ¦
When every other post is an AI generated benchmark report, a question about the best model, or a slop-coded application or engine that pretends to be groundbreaking
submitted by /u/Honest-Kangaroo-1830 [link] [comments]β¦
How-to guide to create audiobooks?
There are a number of projects posted in this sub aiming to convert ePub or RTF files to MP3, or jusβ¦
Friends from the localllama community, if you love local llm, don't participate in the IPO (spaceX, OpenAI, Anthropic)
I'm not going to. And you shouldn't either. The frontier labs are the ones who are harming our commβ¦
An Implementation of NanoQuant: A flexible binary quantization method
https://github.com/pitbox46/NanoQuant TLDR: NanoQuant is a quantization method to create 2 bit/weighβ¦
Here are some tips on hitting nearly 200 tok/s for DeepSeek v4 Flash on Hopper
I needed a smarter model for my local Hermes Agent setup, so I moved to DeepSeek v4 Flash. First thiβ¦
Why is the MLX version of the Gemma 4 QAT so big??
the MLX version of the QAT 4bit is like 27gb but the none QAT version is 17gb and the regular 4bit Mβ¦
I bundled a fully local LLM inside my Unity game. No internet, no cloud, no API key. The conversation is the gameplay.
I am making a game that is bundled with a local LLM and every conversation is unique. The game, 'Simβ¦
Levi: Run AlphaEvolve on your local QWEN 30B
Hi r/LocalLLaMA, Wanted to share something I'm excited about. I've been fascinated by AlphaEvolve anβ¦
Xiaomi just claimed 1,000+ tps on a 1T model using a standard 8-GPU server
Just saw Xiaomi MiMo announce MiMo-V2.5-Pro UltraSpeed, claiming they broke the 1,000 tokens/sec outβ¦
Nex N2 has a funny "few words do trick" reasoning
I've been playing with Nex N2 Pro (Qwen 3.5 397B finetune) locally today. I noticed straight away thβ¦
Luce Spark: a 35B MoE on a 16 GB GPU, without the offload tax
Hey fellow Llamas, your time is precious, so I'll keep it short. TL;DR: 33-35B MoE on a 16 GB GPU. β¦