The Hallway Track

mixture-of-experts

11 tracked signals on mixture-of-experts.

Open-weight AI just hit 2.8 trillion parameters…

Fireship · Jul 22, 2026

Moonshot AI's Kimi K3 is a 2.8T-parameter open-weight model matching frontier closed models on coding benchmarks.

“it has OpenAI and Anthropic terrified because its Trust Me Bro benchmark performance is on par with and in some cases beating Claude Fable and GPT 5.6 Soul”
Efficient MoE Training for Biological Foundation Models

NVIDIA Developer Blog · Sep 24, 2026

NVIDIA details Mixture-of-experts training methods to efficiently scale biological foundation models.

“Mixture-of-experts (MoE) architectures take a different approach to scaling by using many subnetworks, or experts, while activating only a small subset for each token.”
This Free AI Just Caught The Billion Dollar Giants

Two Minute Papers · Aug 28, 2026

Qwen 3.8 Flash Next, a free open-weights MoE model, rivals paid closed AI systems.

“it seems that we can switch out paid closed systems to open weights AI that we can download and run ourselves forever”
Introducing Gemma 4 models on Amazon Bedrock

AWS Machine Learning Blog · Jun 15, 2026

Google DeepMind's open-weight Gemma 4 model family is now available on Amazon Bedrock.

“Artificial Analysis reports an Intelligence Index of 39 for Gemma 4 31B, well above the median of 15 in the 4B–40B open-weights class.”
Boosting MoE Training Throughput with Advanced Fusion Kernels

NVIDIA Developer Blog · Jun 15, 2026

NVIDIA details advanced fusion kernels that boost training throughput for mixture-of-experts (MoE) models.

“Mixture-of-experts (MoE) models have quickly become a foundational component of modern, large-scale AI systems.”