u/pharrt

Qwen-3.8-35B-A3B? Maybe not... cryptic reply direct from Qwen co-author.
▲ 168 r/LocalLLM

Qwen-3.8-35B-A3B? Maybe not... cryptic reply direct from Qwen co-author.

I asked Shuai Bai, co-author and prominent AI developer for Qwen, about this model. Not the answer I was hoping for, but let's see what comes next. In the meantime, I guess all we can do is speculate!

X-link

u/pharrt — 3 days ago
▲ 138 r/LocalLLM

Don't laugh - it works!

A 10yo server was busy collecting dust, but it has 32Gb RAM (2x 16Gb DDR4 @ 2133 MHz )... No GPU.

Now it does some amazing work running heavy tasks with qwen3.6-35b-a3b (IQ4_XS). Running in the background, generating quality output at 5-10tok/s. Even with 128k context! It just chugs along for hours, but does such great, high-context work.

Uses ~26 GB of the 32 GB RAM, uses CPU i-7-6700 @ about 60% (not the bottleneck, of course)

Never would have believed that would be possible until a few months ago, but this ol' gal has a new lease on life.

reddit.com
u/pharrt — 25 days ago

New Discover UI Layout sucks

More clicks, more scrolling, more data consumption. And what is Perplexity's obsession with summaries? We are not all trying to digest one-liners at 100 miles an hour.

reddit.com
u/pharrt — 2 months ago