GLM 5.3 released, with a massive leap in capability

Same base model as GLM 5.2, only with more post training.

z.ai
u/cheechw — 6 days ago
▲ 2.5k r/OpenModels+1 crossposts

Qwen 3.8-27b coming this week

Confirmed by the official Qwen account.

u/Bestlife73 — 9 days ago
▲ 1.8k r/OpenModels+2 crossposts

Introducing Muse Glimmer: an open-weight model optimized for always-on local agent workflows

Hi r/LocalLLaMA 👋 

Today we’re excited to release Muse Glimmer, a 30B open-weight model built specifically for local agent workflows. We’re releasing the weights to the community under a permissive Apache 2.0 license.

A few specs

  • 30B params, dense
  • Multimodal: interleaved text + images via a dedicated perception encoder
  • Trained on 100+ languages
  • Controllable reasoning effort (quality/speed tradeoff)

Memory footprint
At full precision, 30B needs 55+ GB, which is out of reach for consumer hardware. We quantize weights to ~4-bit, bringing the LM under 20 GB. That leaves headroom in a 24 GB or 32 GB envelope for the KV cache, the perception encoder, and the speculative decoding drafter running simultaneously. We validated minimal to no degradation on agentic tasks under compression.

Speculative decoding
Ships with a lightweight DFlash-based drafter that proposes blocks of tokens which the main model verifies in parallel. Significantly faster than token-by-token generation with identical output quality. We're also shipping quantized drafter versions so the memory overhead stays small.

A few capabilities
We trained Muse Glimmer for agentic loop tasks, including:

  • End-to-end task completion (strong performance on DeepSearch QA, MCP-Atlas, 𝛕^(3)-Bench, SWE-Bench, and more)
  • Function calling with precise schemas across long workflows
  • Multi-step reasoning over long horizons
  • Failure recovery — when a tool call fails or returns something unexpected, it's trained to diagnose and retry instead of halting. This was a deliberate training target.
  • Works with OpenClaw and other agentic scaffolds
  • Multimodal understanding and reasoning

Running it
Weights are up on Hugging Face. Coming soon: Ollama, LM Studio, Unsloth and torchtitan, plus optimized integrations for llama.cpp, MLX, and ExecuTorch. vLLM and SGLang for serving. Get started quickly with Together AI, Fireworks AI, and OpenRouter. We're also working with AMD, Arm, Dell, Intel, and NVIDIA on per-device optimization.

We look forward to your feedback and seeing what the community builds with Muse Glimmer.

🔗 Weights: https://huggingface.co/meta-models 
🔗 Research Blog: https://go.meta.me/museglimmer
🔗 Resources: https://developer.meta.com/ai/models/muse-glimmer/

u/AIatMeta — 10 days ago
▲ 450 r/OpenModels+1 crossposts

DeepSeek-V4-Flash-0731 now far surpassing the DeepSeek-V4-Pro-Preview in benchmarks

u/SnooBunnies8392 — 20 days ago

Minimax H3: Open weights video model with native audio

Another open weights video model hits the market with native audio capability.

x.com
u/cheechw — 20 days ago
▲ 3.3k r/OpenModels+2 crossposts

Kimi K3 weights now released.

Kimi K3 weights are finally released!

u/Jenna_AI — 23 days ago
▲ 878 r/OpenModels+2 crossposts

Laguna S 2.1 Released: Cheaper than Deepseek v4 Flash, Better than V4 Pro

Model Size Terminal-Bench 2.1 SWE-bench Multilingual SWE-Bench Pro (Public Dataset) DeepSWE SWE Atlas (Codebase QnA) Toolathlon Verified
Laguna S 2.1 118B-A8B 70.2% 78.5% 59.4% 40.4% 46.2% 49.7%

Finally the banger we've been waiting from Laguna. probably will be great for 64GB+ RAM and VRAM setups.

huggingface.co
u/Every-Walrus — 15 days ago

Head of strategic futures at OpenAI claims that "open-weight-model-dominant world" leads to "full AI communism" and that AI as a public good is a "dystopian hellscape"

This seems to me like a painfully ironic position for someone at a company named "OpenAI" to be taking. Not to mention totally batshit insane.

x.com
u/cheechw — 1 month ago

New 1.57T open weights model Monolith-1.0 claiming extraordinary benchmark performance

It claims to beat Fable and GPT 5.5 handily in some benchmarks.

Hard to believe the claims, and from the looks of it, many X users are skeptical as well.

x.com
u/cheechw — 1 month ago
▲ 2.1k r/OpenModels+2 crossposts

KIMI K3 Beats Claude Fable and GPT 5.6 sol in arena.ai!!!

Unbelievable to see kimi k3 beat frontier models that were 'too dangerous' for public use.

u/Gohab2001 — 1 month ago

Kimi K3 released, beating Opus 4.8 in benchmarks at 2.8T parameters open weight

Benchmarks appear to easily clear Opus 4.8 and GPT 5.5, and even appear to be competitive with Fable 5 and GPT 5.6 Sol in a number of them. Full weights to be released July 27, per the blog post.

kimi.com
u/cheechw — 1 month ago
▲ 606 r/OpenModels+1 crossposts

China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model

https://www.theinformation.com/briefings/exclusive-chinas-minimax-plans-launch-2-7-trillion-parameter-model

According to The Information, MiniMax plans to launch a new-generation large language model with 2.7 trillion parameters.

Sources revealed that the internal codename for this new model is M3 Pro. It is expected to be released and open-sourced as early as the third quarter of this year, with significant improvements in handling complex reasoning and multi-step tasks.

This new model is much larger than MiniMax's current flagship model, M3 (428 billion parameters). Larger-scale artificial intelligence models are more capable of handling complex reasoning and multi-step instruction-based tasks.

reddit.com
u/External_Mood4719 — 1 month ago