
Now you can set "Thinking Effort" with Qwen 3.8 in TurboLLM
If you have been using Qwen 3.8 27B lately, you must have observed that it thinks a lot, that is because it supports thinking effort embedded in its chat template and by default it is "xHigh". So I added a reasoning effort slider just like claude in TurboLLM so you can control it. For all other models it stays the old "Thinking Budget" where you can control number of tokens allowed for thinking. Go give it a try.
npx turbollm
u/Agitated_Problem5320 — 4 days ago