BuildTube Preview About
Back to the feed

Get answers from a 27B model with less thinking-out-loud

A modified version of Qwen3.8-27B trained to cut down how much it reasons to itself before answering.

The creator reports 58.3% fewer thinking tokens and close to the same scores on most of its benchmarks.

Can I use this?

Needs a model runtime on your computer

You'll need A high-end GPU running vLLM · Setup needed

Worth knowing The licence is listed only as "other", so read the terms on the model page before any use; the creator also points to separate enterprise licensing. All the speed and accuracy comparisons are the creator’s own, and its own table shows some scores fall, most noticeably on the maths benchmarks. BuildTube has not run or verified this model.

Why it's here

Not trending this week

Last seen on the Hugging Face trending list on 20 September 2026. It stays on BuildTube so you can still find it.

  • 615 people have liked it on Hugging Face.
  • It was downloaded 22,561 times in the last 30 days.

Numbers checked on 6 October 2026; not refreshed since.

Behind it

See the code on Hugging Face

BuildTube has not run or verified this project. Everything above is written from what the creator published.

Made this?

Back to the feed