Having some fun with LMX-Omni-52B-Halo in Open WebUI
submitted by /u/jfowers_amd [link] [comments]β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
Having some fun with LMX-Omni-52B-Halo in Open WebUI
submitted by /u/jfowers_amd [link] [comments]β¦
New models released: Nex-N2 Pro 397B and Nex-N2 Mini 35B
They are FTs of Qwen3.5 and the benchmarks look pretty good https://huggingface.co/nex-agi/Nex-N2-miβ¦
Where are we with computer-control harnesses?
Seems like local vision language models models are getting smart enough so that it would be useful tβ¦
Straight angle vs 90 degree angle PCI-E Riser cables?
Can I use 90 degree angle PCI-E cables in AI rig without major headache? Here is a sample: https://β¦
advice for dual-gpu asymmetric
Hello everyone, i had a 3080ti 12gb and added a 3080 20gb, so it has a bit less speed but more memorβ¦
xdna-top: unified NPU+iGPU terminal monitor for Strix Halo (Ryzen AI Max) β finally see the NPU work
If you're running local models on a Ryzen AI Max / Strix Halo box, you've probably noticed it's hardβ¦
DiffusionGemma made me rethink what memory bandwidth means for local agent inference
Been testing DiffusionGemma 26B A4B for the last few days and the bottleneck profile is completely dβ¦
I built a graph-memory layer on top of turbovec for local/constrained RAG β looking for feedback
Disclosure: I built this. I like turbovec for compact local vector search, but in real RAG apps my bβ¦
Reviewing speed optimizations on llamacpp for large MoE models on multiGPU rigs? (fitparams vs -ngl/-ncmoe vs other flags, P2P, overclocking)
In anticipation of MiniMax reported upcoming open-weight release of M3, wanted to do comprehensive rβ¦
I tried the same prompt people are talking about in the vibecoding subreddit on my local setup
In reference to this: https://www.reddit.com/r/vibecoding/comments/1u26r5z/we_gave_the_same_exact_prβ¦
DifussionGemma 4 on 4x7900xtx
Just got 100 tps on generation, but in total time it around 45-60 t/s in case of prompt processing wβ¦
Any chances for a 12B diffusion Gemma?
Currently recompiling my llama.cpp with support for diffusion Gemma, but I know on my hardware it woβ¦