Requesty
Ant Group officially announced the release of Ling-2.6-flash, a large language model built on a sparse Mixture-of-Experts architecture with 104 billion total parameters and 7.4 billion active parameters. The press release, published from Hangzhou on April 22, 2026, positions the model as prioritizing efficiency and rea According to Ant Group's release, Ling-2.6-flash achieved an Artificial Analysis Intelligence Index of 26 while consuming only 15 million output tokens, compared to over 110 million tokens for comparable models like Nemotron-3-Super, representing an 86% reduction in inference cost. The model achieves inference speeds o
