DataLearner logo
QW

Qwen3-VL-8B-Instruct

Multimodal modelQwen3

Qwen3-VL-8B-Instruct

Release date: 2025-10-15Updated: 2026-07-17 23:18:35.8701,560
Parameters
8.8B
Context length
256K
Chinese support
Not supported
Reasoning ability

Qwen3-VL-8B-Instruct is a multimodal model from Alibaba, released on 2025-10-15. It accepts text, image, and video input and returns text output. The cataloged parameter count is 8.8B. The recorded context window is 256K. The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Qwen3-VL-8B-Instruct

Model basics

Reasoning traces
Not supported
Thinking modes
Thinking modes not supported
Context length
256K tokens
Max output length
No data
Model type
Multimodal model
Modality (in / out)
Text, Image, Video → Text
Release date
2025-10-15
Model file size
No data
MoE architecture
No
Total params / Active params
8.8B / N/A
Knowledge cutoff
No data
Qwen3-VL-8B-Instruct

Open source & experience

Code license
Weights license
Apache 2.0- Commercial use permitted
Live demo
No live demo
Qwen3-VL-8B-Instruct

Official resources

Paper
DataLearnerAI blog
No blog post yet
Qwen3-VL-8B-Instruct

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-¥0.0015/ 1K¥0.0060/ 1K
Qwen3-VL-8B-Instruct

Benchmark Results

Qwen3-VL-8B-Instruct currently shows benchmark results led by DocVQA (2 / 5, score 96.10), MMMU (24 / 29, score 69.60). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking

Multimodal Understanding

2 evaluations
Benchmark / mode
Score
Rank/total
96.10
2 / 5
69.60
24 / 29

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

Qwen3-VL-8B-Instruct

Publisher

Qwen3-VL-8B-Instruct

Model Overview

Qwen3-VL-8B-Instruct is a multimodal model from Alibaba, released on 2025-10-15.

It accepts text, image, and video input and produces text output. The cataloged parameter count is 8.8B. The recorded context window is 256K.

The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. The page records 2 API pricing rules from the listed provider; current provider pricing and conditions should be checked before deployment. The evaluation section contains 2 cataloged benchmark results with their recorded modes and scores. The page links 3 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

Qwen3-VL-8B-Instruct

FAQ

What is Qwen3-VL-8B-Instruct?

Qwen3-VL-8B-Instruct is a multimodal model from Alibaba, released on 2025-10-15. It accepts text, image, and video input and returns text output. The cataloged parameter count is 8.8B. The recorded context window is 256K. The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does Qwen3-VL-8B-Instruct support?

The current model record lists text, image, and video as input and text as output.

What are the main recorded specifications for Qwen3-VL-8B-Instruct?

The cataloged parameter count is 8.8B. The recorded context window is 256K. Fields without a source-backed value remain undisclosed.

Does Qwen3-VL-8B-Instruct have API pricing?

The page records 2 API pricing rules from the listed provider; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for Qwen3-VL-8B-Instruct?

The evaluation section contains 2 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is Qwen3-VL-8B-Instruct open source?

The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code