u/Agitated_Problem5320

Now you can set "Thinking Effort" with Qwen 3.8 in TurboLLM
▲ 8 r/LLMStudio+2 crossposts

Now you can set "Thinking Effort" with Qwen 3.8 in TurboLLM

If you have been using Qwen 3.8 27B lately, you must have observed that it thinks a lot, that is because it supports thinking effort embedded in its chat template and by default it is "xHigh". So I added a reasoning effort slider just like claude in TurboLLM so you can control it. For all other models it stays the old "Thinking Budget" where you can control number of tokens allowed for thinking. Go give it a try.

npx turbollm

u/Agitated_Problem5320 — 4 days ago
▲ 4 r/LLMStudio+2 crossposts

Claude like Routines but for your Local LLM

I have been working on TurboLLM so we local LLM runners can get claude like experience with a model that can run on consumer GPU. And for that I added “Routines”. If you have already used claude, that is self explanatory. If not you can check out this video

youtu.be
u/Agitated_Problem5320 — 14 days ago