Get answers from a 27B model with less thinking-out-loud
Made by UkisAI
View profile on Hugging Face (opens in a new tab)A modified version of Qwen3.8-27B trained to cut down how much it reasons to itself before answering.
The creator reports 58.3% fewer thinking tokens and close to the same scores on most of its benchmarks.
Can I use this?
You'll need A high-end GPU running vLLM · Setup needed
Worth knowing The licence is listed only as "other", so read the terms on the model page before any use; the creator also points to separate enterprise licensing. All the speed and accuracy comparisons are the creator’s own, and its own table shows some scores fall, most noticeably on the maths benchmarks. BuildTube has not run or verified this model.
Why it's here
Not trending this week
Last seen on the Hugging Face trending list on 20 September 2026. It stays on BuildTube so you can still find it.
- 615 people have liked it on Hugging Face.
- It was downloaded 22,561 times in the last 30 days.
Numbers checked on 6 October 2026; not refreshed since.
Behind it
See the code on Hugging FaceBuildTube has not run or verified this project. Everything above is written from what the creator published.
Made this?