There's an underserved market for tiny MoEs like this. Could really take off with how much smarter tiny models are.
Today, we're releasing Ling-3.0-tiny: 7.9B total parameters, with only 1.3B active per token. A native hybrid reasoning model built for real-world tasks, math, ...