Skip to content

Qwen3 VL 8B Instruct

by AlibabaRank #358Score 40.0

Qwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...

Performance Overview

Score
40.0
Rank
#358
24h Change
0
7d Change
0
State
stable
Confidence
high

Signal Scores

SignalNormalizedWeightContributionFreshness
Capabilities
capability
66.730%20.02026-09-22T14:24:58.308Z
Pricing
pricing_tier
99.525%24.92026-09-22T14:24:58.308Z
Context Window
context_window
86.015%12.92026-09-22T14:24:58.308Z
Recency
recency
70.415%10.62026-09-22T14:24:58.308Z
Output Capacity
output_capacity
72.215%10.82026-09-22T14:24:58.308Z

Top Drivers

positive
Pricing
$0.45/M output tokens
$0.45
positive
Capabilities
Supports vision, tools, JSON mode, streaming
4/7
positive
Context Window
262K token context window
262K
positive
Output Capacity
Up to 33K output tokens per request
33K

Capabilities

CapabilitySupported
VisionYes
ReasoningNo
JSON ModeYes
StreamingYes
Function CallingYes
Web SearchNo

Pricing

Input / 1M tokens
$0.12
Output / 1M tokens
$0.45
Context Window
262K
Max Output
33K

Related Models

Related

Qwen3 VL 8B Instruct Tracker - Performance History | LM Market Cap