▲ 3 r/LocalLLM+1 crossposts

Reasoning Effort Toggle

With Qwen3.8-27B out, there has been a lot of discussion about the `reasoning_effort` settings. For those who use Open WebUI, I thought this might be helpful for anyone interested. I made a little plugin for that gives you a toggle and drop-down to set the reasoning level for each message:

https://openwebui.com/posts/reasoning_effort_selector_ee572967

I hope others find this useful!

reddit.com
u/theminor — 4 days ago

Image Generation

Possibly a dumb question. I've been honing-in my home lab and it is working quite well with llama.cpp running the Qwen3.6 models (and several others I've been experimenting with). I run dual Tesla V100s (32GB each) which is pretty fast even though the hardware is dated. Anyway, I'm looking to add capabilities to my rig and started looking at image generation; but I'm clueless about how these work. Seems not quite as simple as loading up a .gguf file in llama.cpp.

What are the current "best" text-to-image models out there right now and where can I find a primer on how to get them running? Basically I'm looking for the "Qwen3.6 of Image Generation" if such a thing exists.

reddit.com
u/theminor — 22 days ago

Late to the party - what is the issue with RD that everyone keeps talking about?

I have not seen any issues with RD but I keep seeing posts alluding to something going on with it. I guess I'm just missing something - what is the issue?

reddit.com
u/theminor — 2 months ago