How important is it for Chinese LLMs to reach the Opus 4.8 level?
In mid-August, Ramp published spending data collected from 70,000 U.S. companies: Fable 5 ,the most β¦
AI tools, cybersecurity and development news aggregated from top sources β saved permanently with unique URLs.
How important is it for Chinese LLMs to reach the Opus 4.8 level?
In mid-August, Ramp published spending data collected from 70,000 U.S. companies: Fable 5 ,the most β¦
If your t/s is low enough, you can see speculative decoding with your own eyes
The other day I was trying out a distillation of DS4 Pro, and it came with MTP. It was slow as hell β¦
I always wonder how much more speed and/or context they'd be getting..
Nothing personal. I just have too much time on my hands. Probably because I spend none of it inspectβ¦
Terminal Bench 4.0 just dropped, GLM-5.3 is at the same level as Fable 5, accounting for margin of error
Announcement: https://www.tbench.ai/news/terminal-bench-4-0 Leaderboard: https://www.tbench.ai/ Imo β¦
Weβre the Team Behind Apodex 1.1 β Ask Us Anything!
Hi r/LocalLLaMA ! Weβre Apodex, the team behind Apodex 1.1, our new model family built to scale agenβ¦
[Megathread] GLM-5.3-Flash - former ox-alpha
Megathread for discussing the release of GLM-5.3-Flash. Quants Fine-Tunes & Abliterations Chat Tempβ¦
Best Local Vision Language Models - August 2026
Share what your favorite models are right now and why. Given the nature of the beast in evaluating Vβ¦
Qwen 3.8 27b with DSH(DeepSeek Harness) is Amazing!! Experiences so far and perfomance.
https://preview.redd.it/wkg27e152qjh1.png?width=853&format=png&auto=webp&s=2e3f8b11ea6393041f501e95cβ¦
Genie-style playable world model running 720p at 16 FPS on a single 5090 in 19GB VRAM
submitted by /u/juanviera23 [link] [comments]β¦
Single 3090 homies, whats your config for Qwen 3.8 ?
I sadly was not able to follow as much as I wanted the new advancements. Care to share your optimiseβ¦
Paper claims RL for reasoning only changes 1-3% of tokens, and they replicate the gains without RL at ~1000x less compute
submitted by /u/juanviera23 [link] [comments]β¦
Qwen3.8 27B reasoning effort low/medium/xhigh comparison
I did a short test of the different reasoning efforts, since on default xhigh the model thinks a lotβ¦