Qwen3 VL 8B Instruct

qwen/qwen3-vl-8b-instruct

Qwen3 VL 8B Instruct is a model from Qwen that accepts image and text inputs and generates text output. It provides a context window of 256 000 tokens, supports tool usage but not reasoning, and its weights are open source. Input tokens cost $0.117 per M tokens, output tokens $0.455 per M tokens, and the model was released on 2025‑10‑15.

  • tool use
  • structured outputs
  • open-weight

At a glance

providerqwen
releasedOct 15, 2025
context262K
input / M$0.12

Specifications

What the API gives you

Context window
262K tokens
Max output
33K tokens
Input price
$0.12 / 1M tokens
Output price
$0.45 / 1M tokens
Cache read
Modalities
image, text → text
Knowledge cutoff
Released
Oct 15, 2025

Recommendation coverage

Where this model appears

Not currently included in a published recommendation. This means the available task-specific evidence did not support a ranking yet—not that the model was reviewed and rejected.

Benchmarks

No independent benchmark results are linked to this model yet. We keep it visible while withholding quality claims until usable results arrive.

Open-weight model on Hugging Face.

Catalog data via OpenRouter and models.dev. Last checked within the past 6 hours.