
Ornith-1.5-35B-A3B-MLX-8bit on Apple M5 Max — 92.6 tok/s — llm-bench.io
Good speed, decent quality for some usecases.
u/DerTomsn — 8 hours ago

Good speed, decent quality for some usecases.
Good, but thinking budget needs to be set on oMLX, otherwise it buuuurns tokens.
Decent, but can't keep up with the hype.
Faster than qwen3.6:27b and muse-glimmer (if thinking is off!).