Qwen3-VL-30B-A3B-Instruct is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Instruct variant optimizes instruction-following for general multimodal tasks. It excels in perception...
Rank
#149
Up 13% week over week
Tokens
8.1B
2026-08-14
Requests
5.4M
5,411,491
Tool calls
-
Input / M
$0.13
Output / M
$0.52
Cache / M
Free
Context
262K
262,144 tokens