u/A_Moist_Towe1

▲ 0 r/oMLX

RANT / WARNING V 0.6.1 CAN RENDER MODELS USELESS

I have been running qwen 3.6 35bMOE and qwen 3.8 27b at q4 with a q6 turboquant kv cache, with lighntning MTP Enabled for the past few days. After the latest update, I cannot use lightning MRP with turboquant, or the model will prompt process repeatedly in a loop or output gibberish. The only way to fix was to downgrade to v 0.6.

I don't know wtf the devs were thinking pushing 0.6.1 to "Stable" but holy fuck is it anything but!

reddit.com
u/A_Moist_Towe1 — 3 days ago

Rant: This program NEEDS A stable build and beta build

The fact that everytime I update Hermes its a question of will it add features or just break core features is really hard to work around. Every other app in my AI stack has the option to select a tested and verified stable build or an untested beta build.

Holy fuck it’s really simple but in terms of quality of life it’s essential. Half the time when I update, I have to spend half a day fixing something or waiting for them to fix the things that were broken. For christs sake the desktop app can’t even properly handle chats right now after the latest update.

Sorry for the rant, holy fuck it’s frustrating.

reddit.com
u/A_Moist_Towe1 — 1 month ago