
u/External_Mood4719

Startup founders urge Trump not to shut off Chinese open weight AI
reddit.combasaltlabsai/monolith-1.0 • HuggingFace
internlm/Intern-S2-Preview-397B • HuggingFace
Kimi K3 released on web and app
# Key Features:
- Kimi K3: 2.8T parameters. 1M context. It leads the field in coding, agentic tasks, lonhorizon reasoning, visual understanding, and agent swarm capabilities.
Kimi K3 maybe coming very soon
Source: From kimi test page .They down the page now
China’s MiniMax Plans to Launch 2.7-Trillion Parameter Model
According to The Information, MiniMax plans to launch a new-generation large language model with 2.7 trillion parameters.
Sources revealed that the internal codename for this new model is M3 Pro. It is expected to be released and open-sourced as early as the third quarter of this year, with significant improvements in handling complex reasoning and multi-step tasks.
This new model is much larger than MiniMax's current flagship model, M3 (428 billion parameters). Larger-scale artificial intelligence models are more capable of handling complex reasoning and multi-step instruction-based tasks.
ascend-tribe/openPangu-2.0-Flash (They haven't uploaded it to Huggingface yet)
https://ai.gitcode.com/ascend-tribe/openPangu-2.0-Flash
openPangu-2.0-Flash is an MoE model trained on Ascend. The model has 92B total parameters and 6B activated parameters. Its context length is 512k. The total pretraining data contains 34T tokens. During Post-training, openPangu-2.0-Flash is trained through unified SFT with slow and fast thinking capability, multiple specialist RL traning, on-policy distillation combining multiple RL specialists.
DeepSeek V4 official version will be launch on mid-July
Source: Email sent from deepseek
used gpt image 2 translate image into english
CEOs of Anthropic and Google DeepMind call for U.S.-led AI coalition in meeting at G7
Anthropic forced to abruptly disable Fable 5 & Mythos 5 globally by US Gov over a jailbreak. This is exactly why we need local models.
I just saw this statement regarding Anthropic being hit with an emergency export control directive from the US government. They were forced to pull the plug on Fable 5 and Mythos 5 for all customers globally. The tl;dr is that the government got spooked by a narrow jailbreak (which basically just sounds like asking the model to fix vulnerabilities in a specific codebase), and forced a complete shutdown without a transparent process. Anthropic is pushing back, but the API access is completely gone for now.
A centralized API can be nuked globally at a moment's notice by a single government decree over something as trivial as a prompt lol.
Banning a model for hundreds of millions of users because someone figured out how to make it fix software flaws is insane. Anthropic admits this standard would halt all new frontier models.
Huawei Released openPangu 2.0 (Will open source on June 30)
At the Huawei Developer Conference (HDC 2026) held on June 12, Richard Yu, Executive Director of Huawei, officially launched the brand-new, open-source Pangu large model—openPangu 2.0. The model is fully adapted to the HarmonyOS ecosystem and has achieved deep optimization and performance breakthroughs on Ascend computing power.
openPangu 2.0 features a 512K context processing capability and comes in two versions tailored for different application scenarios. It sets a record for the largest sparsity ratio in the hundred-billion-parameter category at 28:1:
- openPangu 2.0 Pro: Total parameters: 505B ; Activated parameters: 18B.
- openPangu 2.0 Flash: Total parameters: 92B ; Activated parameters: 6B.
According to the conference presentations and live demonstrations, openPangu 2.0 has been comprehensively upgraded in throughput, latency, and task processing:
- Highly optimized for Ascend computing power, its single-card user throughput is up to 2x that of mainstream open-source models in the industry.
- Built on Ascend-native training, hyper-node optimized training efficiency has improved by 30%, 512K long-sequence training throughput has increased by 50%, and training consistency exceeds 99%.
- Utilizes a high-precision architecture (mHC | Muon | ModAttn) and pioneers the DSA+SWA independent layered hybrid architecture (ultra-sparse attention) for more precise computing power allocation.
Huawei announced plans to progressively open-source the core components of openPangu 2.0 starting June 30, fully empowering developers:
Basic Components: Model architecture, model weights, technical reports, and inference code.
Newly Open-Sourced Components: Pre-training code, post-training code, and training operators.
Addressing the public attention surrounding the 505B total parameter count of the 2.0 Pro version, Richard Yu explained at the conference that this design is due to Huawei allocating a vast amount of its computing power to support the needs of other china enterprises, leaving limited computing power for itself. Furthermore, considering the exorbitant costs of AI computing, Huawei's current strategy ocuses more heavily on achieving substantial improvements in latency and throughput rate.
(Image used Nano banana 2 to translate the image to English)
MooreThreads/MusaCoder-27B • Huggingface
silx-ai/Quasar-Preview • Huggingface (5M context length)
mindlab-research/Macaron-V1-Preview-749B • Huggingface
reddit.comnex-agi/Nex-N2-mini • Huggingface
Keye-VL-2.0-30B-A3B -- Introducing DSA attention into multimodality for the first time
Meet Keye-VL-2.0-30B-A3B — the latest 30B-class flagship base model in the Keye series, purpose-built to push the frontier of long-video understanding and to unlock the first generation of Agent capabilities in the Keye family.