Today, we're releasing Ling-3.0-flash-a hybrid-reasoning MoE model built for production-scale agents.
124B parameters. Just 5.1B active per token.
With 1/8 of the total and 1/12 of the active parameters, it matches or beats our 1T flagship model on most benchmarks shown.