RTX 3090 EBay Pricing is Crazy!!
Couple of years ago, before Local LLMs were in vogue, I bought 8 RTX 3090 @ $700 each to build a AI β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
RTX 3090 EBay Pricing is Crazy!!
Couple of years ago, before Local LLMs were in vogue, I bought 8 RTX 3090 @ $700 each to build a AI β¦
Another 1-click admin account takeover in pewdiepie's AI tool (language in video nsfw)
submitted by /u/theonejvo [link] [comments]β¦
Best Coding Harness for Qwen3.6 35B?
I've been happily using GitHub Copilot for 7-8 months, primarily in Visual Studio and VS Code, mostlβ¦
Z.ai, we need Air! GLM GGUF wen?
First we never saw an upgraded Air model after 4.5. Then GLM 4.7 Turbo was great, but quickly surpasβ¦
What are you running on 16Gb VRAM + 64Gb Ram?
I know this gets asked a lot, but I can only find threads that are at least a couple of months old, β¦
AMD MI50 on Debian Testing is doing great and getting better.
There is probably some relevant information to other cards here but my benchmarks are on dual MI50 cβ¦
120 tok/s on 12GB VRAM with Gemma 4 12B QAT MTP
Google just released the QAT (Quantization-Aware Training) variant of their Gemma 4 models, includinβ¦
KV cache quant benchmarks: KVarN 6-bit matches q8_0, 4-bit matches q5_0. Massive!
TL;DR Based on long context KLD benchmarks, KVarN appears to be just better than usual llama.cpp KV β¦
Fuck, sucessfully ran minecraft server on GLM AI's Agent lol.
I just told it, make a minecraft server and let me play and it worked lol. I just asked "host a minβ¦
It felt good to return my Asus Spark
It's an incredible little package but too expensive of a price to pay for the performance and I simpβ¦
Gemma 4 QAT Unquantized Heretic is here
26B MoE is being built. Now someone needs to quantize them to 4bit, also I have intentionally kept tβ¦
Open WebUI vs Kobold for isolated document review
I wanted to know which is the best option for the task of reviewing documents. I was originally goinβ¦