#mixture-of-experts

Mixture-Of-Experts: 13 AI articles covering mixture-of-experts news, analysis, and research

Articles

GLM-5.3-Flash vs Qwen3.8-Flash-Next: Two Chinese AI Labs Independently Converge on the Same Model Architecture
From Mixtral to Kimi K3: The Evolution of Mixture-of-Experts Models
Z.ai Unveils GLM-5.3-Flash: A 320B-A18B Multimodal MoE Model with 1M-Token Context
NVIDIA Unveils Nemotron 3.5 Lightning: A 30B Open MoE with 3B
TEXAS: Task-Expert-Aware Supervision for Downstream
Cursor Open-Sources Mixture-of-Kittens (MoK): A Deterministic
Alibaba Qwen Unveils Qwen3.8-Max: A 2.4-Trillion-Parameter MoE
Topology-Aware Data Movement for Disaggregated GPU Inference
Thinking Machines Lab Releases Inkling-Small: A 276B Total, 12B
AMD Releases Instella-MoE-16B-A3B: A Fully Open