Blog
Explore the latest research results and technical insights from Ant Ling Foundation Model
Ling-3.0-flash: More Useful Work per Token
A 124B hybrid-linear MoE built for token efficiency and sustainable intelligence. AntLing's Ling-3.0-flash is a 124B-parameter Mixture-of-Experts model with only 5.1B active parameters per token, designed to increase the amount of useful work each token can carry.
Ling-2.6-flash Release: Faster Response, Stronger Execution, Higher Token Efficiency
Ant Ling has officially launched Ling-2.6-flash—a instruction model with a total parameter count of 10.4 billion and 7.4 billion activated parameters. Faced with ever-increasing token demands, Ling-2.6-flash has chosen a different technical path: rather than simply relying on longer outputs to achieve higher scores, it adopts a systematic optimization approach centered on inference efficiency, token efficiency, and agent scenario performance. While maintaining a competitive level of intelligence, this model strives to be faster, more resource-efficient, and better suited to real-world business scenarios.