How LLM-driven NPCs work in Ultima Online (ServUO)
submitted by /u/Zolty [link] [comments]β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
How LLM-driven NPCs work in Ultima Online (ServUO)
submitted by /u/Zolty [link] [comments]β¦
Hi Reddit, I posted my Build Your Own LLM workshop to Youtube (GPT2 & Qwen3.6 style)
Hi internet friends, I recorded a workshop about building your own LLM without any math / ML prerequβ¦
RTX Spark Ads: DJT Edition
"Weβre going to have the most beautiful laptops, theyβll be the slimmest laptops ever. A total masteβ¦
finally
submitted by /u/KvAk_AKPlaysYT [link] [comments]β¦
RTX 3090, Xid 79: 'GPU has fallen off the bus' fixed by cleaning dust out of PCIe riser
Hopefully my experience helps out someone else: I bought a used prebuilt RTX 3090 system (ROG Strix β¦
Higgs Audio v3 TTS 4B. Built for voice chat. Support 100 languages and inline control.
submitted by /u/FerretLegitimate6929 [link] [comments]β¦
Any one still use gpt-oss-120b?
Is anyone still using GPT-OSS-120B? How has it been for tool calling, summarization, coding assistanβ¦
BeeLlama v0.3.1 β latest llama.cpp with extras! DFlash, MTP, q6_0 cache, TurboQuant. Single RTX 3090: Qwen 3.6 27B & Gemma 4 31B up to 177.8 tps (4.93x over baseline)
BeeLlama v0.3.0 and v0.3.1 are here! Big architectural update to align the fork with upstream llama.β¦
cyankiwi AWQ 4-bit β 26.05 update, NVFP4 + FP8 Dynamic quantization and benchmarks across Qwen3.6 4-bit quants
We are happy to share cyankiwi AWQ update: better AWQ implementation, now with NVFP4 and FP8 Dynamicβ¦
You guys were right - Qwen 3.6 35B IS good...and KV Cache DOES matter.
WARNING: I'm speed typing this, no time to organizea/format, so if short paragraph chunks bother youβ¦
Qwen3.6 27B collapse in performance for agentic coding
Hi everyone, I've been trying to optimize my setup to use OpenCode with Qwen 3.6 27B (Unsloth quant β¦
Dynamic KV Cache Quantization and Load-on-demand mmproj/MTP: my llama.cpp wishlist
We all know the struggle of optimizing your VRAM usage: quantized model, quantized kvcache, mmproj oβ¦