Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second.
I needed a relatively simple but acceptable level of AI for working on one project. I didn't have an…
AI tools, cybersecurity and development news aggregated from top sources — saved permanently with unique URLs.
Built a home server from an old PC with GPU upgrade. Qwen3.8 27B runs at ~30 tokens per second.
I needed a relatively simple but acceptable level of AI for working on one project. I didn't have an…
OpenAI's Sam Altman to brief UN Security Council next week
submitted by /u/johnnyApplePRNG [link] [comments]…
I enjoyed the daily HF papers today
Top 3 papers on HF Daily Paper are all unusually delightful and interesting reads for anyone on the …
DiffusionGemma: How It Generates Text in Parallel (From Scratch in PyTorch) [P]
submitted by /u/Winter_Mistake_3185 [link] [comments]…
Steer LLMs and Agents at the Token Level: An interactive tool for token visualization & control, model inspection and data annotation.
onPanda is designed for geeks, power users, curious minds, and engineers. Its UI is built for deep e…
Stepfun new model "Step 5 Preview" just leaked
https://artificialanalysis.ai/models/step-5 https://preview.redd.it/tt7z192zkeqh1.png?width=1035&for…
Training a Neural Network on AMD MI50s Using Vulkan: Proof of ConceptOr: Why I Stopped Listening and Just Did It
Note: This writeup was put together with the help of AI. So I've got dual AMD MI50 32GB cards. If yo…
Qwen3.8-27B at 144 tok/s on an M5 Max MacBook Pro
Meet Inco Splash, open-source inference engine, built around the model and around Apple silicon. Up …
JMLR submission experience [D]
Hi just wondering has anyone submitted anything to JMLR before? Especially in the last 2 years? What…
Alibaba open-sources medical AI model that can detect cancer and nearly 150 conditions
Hopefully things like this let people understand there is good things that can come out of AI. su…
I truly think every major AI lab is purposefully making fear-mongering headlines to get regulations that hurt open-source models
submitted by /u/Fusseldieb [link] [comments]…
Tuning Qwen 3.8 27B and OMP as a coding agent on 2× 3090s
Oh My Pi + vLLM on two 3090s. Average wait per turn went from 28s to 7s, mostly from changing omp se…