r/MachineLearning 🤖 Ai 👁 0

LessThink-Qwen3-4B: the same model, with far less thinking [P]

I post-trained Qwen3-4B to spend 44% fewer tokens on reasoning, keeping its knowledge and answer style. The whole pipeline ran on one GPU. folks, you can check it out on : https://5ivatej.com/lessthink/ submitted by

📄

This source provides headlines only. Use the button below to read the complete article on the original site.

📰 Read the original article on r/MachineLearning

Originally published by r/MachineLearning. Aggregated on AIWithGhost for educational purposes — full credit and traffic to the original publisher.