Qwen 3.8 27B matches GPT-5.6 Luna and nearly ties 1.7T-parameter DeepSeek on benchmarks
“Qwen 3.8 27B is a truly astonishing model.”
11 tracked signals on qwen.
Qwen 3.8 27B matches GPT-5.6 Luna and nearly ties 1.7T-parameter DeepSeek on benchmarks
“Qwen 3.8 27B is a truly astonishing model.”
Qwen releases 2.4T open-weight model with autonomous 10-day coding and AI research capabilities
China commits to AI openness as Xi's WAIC speech and Qwen's open-weight pivot signal strategic shift
“Xi gave his speech where he directly committed to openness and open source as a strategy.”
Ben Thompson proposes US law legalizing AI training data collection and distillation to compete with China
“The U.S. should pass a law that (1) makes explicit that collecting data for training models is fair use, and (2) bars terms of service that forbid distillation, for U.S. companies at a minimum.”
Alibaba previews Qwen4 architecture via 176B MoE model with 1M token context
Qwen 3.8B runs on consumer laptops while matching frontier model performance
“if we wait a bit, we might get frontier level systems running on our laptops”
RL-trained Qwen 27B model Faraday outperforms Claude and GPT-5 on scientific replication tasks
“they can get this so-called AI scientist agent which can outperform you know much larger models such as Claude and GBD5”
Qwen 3.8 Max challenges OpenAI and Anthropic at 5-10x lower API pricing
“it sat there thinking for 16 days, starting from an empty folder, writing, testing, and repairing its own code”
Qwen releases 125B MoE model with only 6B active parameters as Qwen4 architecture preview
“a multimodal MoE model that also serves as an early preview of the architecture used in Qwen4”
Qwen 3.8 27B defaults to extreme overthinking, consuming 22K tokens for simple tasks
“This is a hilarious default. It's absolutely not a good way to run the model, especially on consumer hardware.”
Georgi Gerganov confirms Qwen3.6-27B is a capable local model for daily coding tasks on consumer hardware.
“I can 100% attest to the fact that Qwen3.6-27B is a very capable local model for coding tasks.”