NVFP4 GGUF vs Q4_K / Q6_K GGUF for precision
Hey all Mostly a curious question. I've done a bit of research in this sub and other sites, and the β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
NVFP4 GGUF vs Q4_K / Q6_K GGUF for precision
Hey all Mostly a curious question. I've done a bit of research in this sub and other sites, and the β¦
I'm brand new to running LLMs and the sheer number of tools is overwhelming
Hey everyone. I'm brand new to running LLMs in general, even more new to running them locally, and tβ¦
gemma4 QATs vs higher-bit regular quantizations?
I have enough RAM+VRAM to use gemma4 26b a4b up to q6_k quantizations w/ decent performance. Does anβ¦
Without open llm competition, closed source LLM companies will become insatiable.
I can't imagine how arrogant one must be to make such a decision. People pay $200 a month for Anthroβ¦
Releasing Apodex-1.0 Smol Models (0.8B, 2B, 4B Open-Weights) optimized for Agentic Verification + AgentHarness Evals
Hey r/LocalLLaMA, We just released Apodex 1.0, and alongside our flagship API, we are releasing the β¦
How I got inspired to build a version manager for llama.cpp
Hey everyone, I wanted to share a little side project I cooked up over the last week. So, long storyβ¦
Fine-tuned Qwen2.5-7B to 96% of Claude Haiku on a domain-specific task using ~$3 of API calls and zero human labelers
Built a decision-reasoning engine (Orlog) and wanted to fine-tune a local model for it instead of paβ¦
Furiosa AI selling inference chip to consumer market will be a game changer to local llm
β This is south Korean start up all-in on inference chip: https://furiosa.ai/renegade-spec Tsmc 5nmβ¦
hot take (or really not so hot take): WE ARE USING "VIBECODING" FOR TWO DIFFERENT THINGS AND IT CAUSES UNNECESSARY FRICTION IN COMMUNICATION
vibe coding meaning 1: Thrown together without care, by dumping it all on the AI, without deeper undβ¦
Newer Qwen models are worse at summarization?
We have summaries annotated by real humans that we benchmark various models, using an LLM as a judgeβ¦
Since when the RTX 6000 PRO is priced at 13250USD on the official NVIDIA Page?
https://marketplace.nvidia.com/en-us/enterprise/laptops-workstations/nvidia-rtx-pro-6000-blackwell-wβ¦
[PSA] 5070ti 16GB is as low as $500.99 at Best Buy.
The Best Buy 5070ti clearance continues. Some stores have marked it down to $500.99 this last weekenβ¦