u/fridder

▲ 7 r/oMLX

Trying to understand the ANE tuning

I have a m1max with 64GB of RAM. I'm not sure how to tune it effectively and when the auto tuner runs it errors out with

Prefill would require ~61.47 GB peak (current 45.55 GB + KV+SDPA 15.91 GB) but metal_cap ceiling is 54.62 GB. Raise kernel iogpu.wired_limit_mb in Terminal (currently caps Metal at 58.00 GB), or reduce context length.

reddit.com
u/fridder — 1 day ago
▲ 52 r/oMLX

v0.4.4 has made Qwen-3.6-27B usable for me, finally

Just an appreciation post and heads up. I had gotten some use out of this model before but the prompt prefill performance was terrible. It still isn't blistering but on my m1max 64GB I am finally seeing triple digit prompt prefill stats!

reddit.com
u/fridder — 2 months ago
▲ 1 r/oMLX

"Phantom" Model showing up

After I upgrade to 0.4 I noticed this model, TheCluster--amoral-gemma-3-12B-v2-mlx-4bit, one I didn't download. It doesn't show up in the web ui but does in the new settings app. Anyone know what the heck this is?

u/fridder — 3 months ago

I miss New Jack Swing (Bel Biv DeVoe: Poison)

Very late 80's and early 90's but I really liked BBD, Boys II Men, Soul For Real , Heavy D and the Boyz. Is it just me or is it one of the genres that haven't really come back up again?

youtube.com
u/fridder — 3 months ago