People are making single-slot, half height pcie v100 with nvlink in China
https://preview.redd.it/yu3jolmnm96h1.png?width=899&format=png&auto=webp&s=4a938ad5675015cc6b5e4bb9dβ¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
People are making single-slot, half height pcie v100 with nvlink in China
https://preview.redd.it/yu3jolmnm96h1.png?width=899&format=png&auto=webp&s=4a938ad5675015cc6b5e4bb9dβ¦
Rick & Morty
nobody expected HF there submitted by /u/jacek2023 [link] [comments]β¦
PSA: Throttle GPU power limits, with minor performance deficits
I just feel i need to post this here again so more people see: Test around with throttling the powerβ¦
Apple announced new on device inference engine for Apple Silicon
This news seem to have flown under the radar. Apple announced CoreAI on WWDC which is basically a fuβ¦
I put together a Rust-native, CPU-only implementation of LFM2.5-8B-A1B
This is still a work in progress, but since recording the video, I added callbacks for tool use, morβ¦
Semantic distance as routing layer: an on-device, serverless alternative to the central-index model
Premise: For ~30 years, discovery (of information or of people) has been mediated by a central indexβ¦
Jetson Orin NX Build for Hermes Agent + Benchmarking
I had a huge LLM server, and now I have a tiny one! I had a Jetson Orin NX gathering dust from a lonβ¦
Still a VERY lightweight open web-search tool for smaller local LLMs - now with SearXNG support
Hey everyone, TinySearch v0.2.0 (first stable beta) is out. The first version used DuckDuckGo directβ¦
Gemma 4 31B's competence surprised me
I'm just getting started using local LLMs for code. I'm not interested vibe coding, but I am hoping β¦
Cheapest setup for >10 tok/sec for 120B dense LLM
Hi all, I'm trying to wrap my head around hardware variables when it comes to LLM, and I have anotheβ¦
Have we reached the point where open-source LLMs are βjust good enoughβ?
The question Iβm asking myself is whether open-source LLMs are now βjust good enoughβ to meet 95% ofβ¦
benchmark idea: political compass for finetuned/abliterated models
There are political compass benchmarks for cloud models, like this one:https://trackingai.org/politiβ¦