Ant Group’s Ling 3.0 Flash packs 124B parameters into a model built for speed, not size

1 hour ago 2



Ant Group, the fintech giant behind Alipay, just made a notable move in the AI efficiency race. Its inclusionAI research lab released Ling-3.0-Flash on July 23, 2026, an open-weights model that challenges a core assumption baked into the AI industry’s growth story: that bigger always means better. The model carries 124 billion total parameters but activates only around 5.1 billion of them per token during inference. Think of it like a massive library where the librarian only ever needs to pull five books off the shelf at once, no matter how complex the question. Outperforming its trillion-parameter predecessor The headline achievement here is benchmark performance. Ling-3.0-Flash matches or beats Ling-2.6-1T, Ant’s previous flagship model, across core reasoning and instruction-following tasks. That predecessor had a full trillion parameters, making the new model roughly eight times smaller in total parameter count yet competitive on the metrics that matter for production deployments. On the Artificial Analysis Intelligence Index, which measures agentic and reasoning capabilities across a standardized battery of tests, Ling-3.0-Flash scores 38. The architecture behind this is a hybr...

Read Entire Article