
▲ 152 r/LocalLLaMA
Meituan just dropped LongCat-Flash-Lite-Sparse
It’s an MoE with ~3B active params and a 30B n-gram lookup table offloaded to RAM for fast 256k context on a 24GB GPU. Reminds me of Gemma 4’s PLE trick.
Initial analysis suggest it wont be replacing my Qwen 3.6 27b.
u/Gohab2001 — 20 days ago