← Back to Payloads
llm-releases2026-08-06

Ant Group Just Open-Weighted Ling-3.0-Flash. The Hybrid-Linear Architecture From Day Zero Is the Real Story

Ant Ling released Ling-3.0-flash on August 4 as a 124B-total / 5.1B-active MoE with native hybrid-linear attention (KDA + Gated MLA, 5:1 alternating) trained from day zero, MIT-licensed weights, and an SGLang HiCache + Mooncake serving stack that cuts TTFT 60–80% on long inputs. Active parameters drop ~8x vs the 1T-class predecessor while matching or beating it on SWE-Bench Pro, MCP-Atlas, and SkillsBench. The real story is not the benchmark chart. It is that hybrid-linear is now the default architecture for open-weight frontier models, and Ant Ling shipped the first one trained that way from scratch.
Quick Access
Install command
$ mrt install llm-releases
Browse related skills
Ant Group Just Open-Weighted Ling-3.0-Flash. The Hybrid-Linear Architecture From Day Zero Is the Real Story
Related Dispatches