*Ling-3.0-flash* is a *124B-parameter Mixture-of-Experts (MoE) model*, with approximately *5.1B parameters activated per token*. The model is designed with *token efficiency and production-scale agentic inference* as key priorities, enabling developers...
Rank
#64
No change
Tokens
88.3B
2026-08-14
Requests
12.8M
12,800,046
Tool calls
-
Input / M
$0.02
Output / M
$0.06
Cache / M
$0.0042
Context
262K
262,144 tokens