Whatβs your most unusual non-LLM AI you actually use daily?
Whatβs your most unusual or underrated non-LLM AI tool you actually use daily (weird, niche, or non-β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
Whatβs your most unusual non-LLM AI you actually use daily?
Whatβs your most unusual or underrated non-LLM AI tool you actually use daily (weird, niche, or non-β¦
llama.cpp Gemma4 MTP support merged!
submitted by /u/pinkyellowneon [link] [comments]β¦
llama : add Gemma4 MTP by am17an Β· Pull Request #23398 Β· ggml-org/llama.cpp
now your personal local Gemma will be faster than ever submitted by /u/jacek2023 [link] [cβ¦
Need some guidance toying with local models
Hi, so I have a pretty low-end laptop regarding running LLMs locally (NVIDIA GeForce RTX 3050 with 4β¦
Qwen 3.6 27B KV cache quant benchmarks: 75 pairs, q8/q6/q5/q4, KVarN, Turbo/TCQ
Full benchmark results and in-depth analysis are available in the articles: KV Cache Quantization Beβ¦
Dockerized Nemotron 3.5 ASR β Switched from Parakeet, better multilingual support + streaming (4.5x realtime speed on cpu)
I was originally using Parakeet for my speech recognition pipeline but decided to give Nemotron 3.5 β¦
How do you increase prompt processing speed ?
I am rocking Qwen like we all know, at 24GB 7900XTX 230k context, but it starts at 850t/s and then lβ¦
Clustering 3x Jetson Nano Orin Supers
Hey everyone! Recently, I released a blog on how to setup a cluster out of your Raspberry Pi 4bs andβ¦
Gemma 4 31B QAT GGUF loads with MTP branch, but outputs repeated <unused49> - any working recipe?
Iβm trying to run: unsloth/gemma-4-31B-it-qat-GGUF gemma-4-31B-it-qat-UD-Q4_K_XL.gguf on an RTX 5090β¦
Open models to win β
submitted by /u/pmttyji [link] [comments]β¦
How to compare Original vs QAT Gemma 4 31B Q4 quants
I just came across the following post, where a user found some confusing divergence results between β¦
You don't need a GPU to run gemma-4-26B-A4B
I've been running LLMs on my old potato i5-8500 with 32GB of RAM and *no GPU* for awhile now, runninβ¦