Ling-2.6-flash is an instant (instruct) model from inclusionAI with 104B total parameters and 7.4B active parameters, designed for real-world agents that require fast responses, strong execution, and high token efficiency....
Rank
#40
Down 8% week over week
Tokens
213.9B
2026-08-14
Requests
70.6M
70,557,299
Tool calls
-
Input / M
$0.01
Output / M
$0.03
Cache / M
$0.002
Context
262K
262,144 tokens