Qwen3 VL 8B Instruct
by AlibabaRank #313Score 40.0
Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
Performance Overview
Score
40.0
Rank
#313
24h Change
0
7d Change
0
State
stableConfidence
highSignal Scores
| Signal | Normalized | Weight | Contribution | Freshness |
|---|---|---|---|---|
Capabilities capability | 66.7 | 30% | 20.0 | 2026-08-08T12:24:25.873Z |
Pricing pricing_tier | 99.5 | 25% | 24.9 | 2026-08-08T12:24:25.873Z |
Context Window context_window | 86.0 | 15% | 12.9 | 2026-08-08T12:24:25.873Z |
Recency recency | 78.6 | 15% | 11.8 | 2026-08-08T12:24:25.873Z |
Output Capacity output_capacity | 75.0 | 15% | 11.3 | 2026-08-08T12:24:25.873Z |
Top Drivers
positive
Pricing
$0.45/M output tokens
$0.45
positive
Capabilities
Supports vision, tools, JSON mode, streaming
4/7
positive
Context Window
262K token context window
262K
positive
Recency
Released over 1 year ago
10mo ago
Capabilities
| Capability | Supported |
|---|---|
| Vision | Yes |
| Reasoning | No |
| JSON Mode | Yes |
| Streaming | Yes |
| Function Calling | Yes |
| Web Search | No |
Pricing
Input / 1M tokens
$0.12
Output / 1M tokens
$0.45
Context Window
262K
Max Output
33K