Models / Qwen
Qwen3 VL 8B Instruct
by Qwen ·
qwen/qwen3-vl-8b-instructQwen3-VL-8B-Instruct is a multimodal vision-language model from the Qwen3-VL series, built for high-fidelity understanding and reasoning across text, images, and video. It features improved multimodal fusion with Interleaved-MRoPE for long-horizon...
tool-usevisionreleased 2025-10-14
Legal score · raw
-
not yet benchmarked
Context
262K
max output 236K
Input
$0.31
per 1M tokens
Output
$0.94
per 1M tokens
Suite cost
-
no run recorded
Benchmark results
This model has not been run through the suite yet. See Humanity's Last Lawsuit for the benchmark models are ranked on.