11 tracked signals on mixture-of-experts.
Open-weight AI just hit 2.8 trillion parameters…
Fireship · Jul 22, 2026
Moonshot AI's Kimi K3 is a 2.8T-parameter open-weight model matching frontier closed models on coding benchmarks.
“it has OpenAI and Anthropic terrified because its Trust Me Bro benchmark performance is on par with and in some cases beating Claude Fable and GPT 5.6 Soul”
The mystery is solved... and the answer is 40x cheaper than Claude
Fireship · Sep 01, 2026
Zhipu's GLM 5.3 Flash anonymously dominated Open Router and costs 40x less than Claude
“at its peak, OX Alpha accounted for nearly a third of Open Router's entire weekly traffic”
This $12 billion startup finally shipped something...
Fireship · Jul 20, 2026
Thinking Machines releases Inkling, a 970B-parameter Apache-licensed open-weights multimodal model.
“the weights are Apache licensed and sitting on hugging face right now”
[AINews] Reflection Beam - 501B-A23B American Open Model
Latent Space Blog · Oct 06, 2026
Reflection AI launches Beam, a 501B/23B-active US-trained open MoE model for coding and agentic work
Experiment with Qwen3.8-Flash-Next 176B Model on NVIDIA GB300 NVL72 for Agentic Coding
NVIDIA Developer Blog · Aug 26, 2026
Alibaba previews Qwen4 architecture via 176B MoE model with 1M token context
Introducing GLM 5.3 on Amazon Bedrock
AWS Machine Learning Blog · Oct 05, 2026
Z.ai's 753B-parameter GLM 5.3 MoE model is now available on Amazon Bedrock for enterprise use
Efficient MoE Training for Biological Foundation Models
NVIDIA Developer Blog · Sep 24, 2026
NVIDIA details Mixture-of-experts training methods to efficiently scale biological foundation models.
“Mixture-of-experts (MoE) architectures take a different approach to scaling by using many subnetworks, or experts, while activating only a small subset for each token.”
This Free AI Just Caught The Billion Dollar Giants
Two Minute Papers · Aug 28, 2026
Qwen 3.8 Flash Next, a free open-weights MoE model, rivals paid closed AI systems.
“it seems that we can switch out paid closed systems to open weights AI that we can download and run ourselves forever”
Introducing Gemma 4 models on Amazon Bedrock
AWS Machine Learning Blog · Jun 15, 2026
Google DeepMind's open-weight Gemma 4 model family is now available on Amazon Bedrock.
“Artificial Analysis reports an Intelligence Index of 39 for Gemma 4 31B, well above the median of 15 in the 4B–40B open-weights class.”
Boosting MoE Training Throughput with Advanced Fusion Kernels
NVIDIA Developer Blog · Jun 15, 2026
NVIDIA details advanced fusion kernels that boost training throughput for mixture-of-experts (MoE) models.
“Mixture-of-experts (MoE) models have quickly become a foundational component of modern, large-scale AI systems.”
Introducing Mellum2: A 12B Mixture-of-Experts Model by JetBrains
Hugging Face · Hugging Face Blog · Jun 01, 2026
JetBrains released Mellum2, a 12B mixture-of-experts model, on Hugging Face.