Qwen / qwen/qwen3-vl-8b-thinking

Qwen3-Vl-8B-Thinking - access through LLMTR

Qwen3-Vl-8B-Thinking is a vision-language model for understanding text, images, and video. It is well suited to OCR, screenshot analysis, document understanding, and multimodal assistant workflows. The thinking variant is better suited to harder visual problems and more deliberate step-by-step analysis.

Technical specifications

Canonical IDqwen/qwen3-vl-8b-thinking
ProviderQwen
Context window256,000 tokens
OperationsCHAT_COMPLETIONS
Modalitiestext

Pricing

An 8% platform margin applies to credit top-ups; model usage prices are not separately marked up.

OperationMetricUnitPrice
CHAT_COMPLETIONSINPUT_TEXTPER_1M_TOKENS$0.180000
CHAT_COMPLETIONSOUTPUT_TEXTPER_1M_TOKENS$2.10

Example usage

With existing OpenAI SDK flows, change only the base URL and model identifier.

curl https://llmtr.com/v1/chat/completions \
  -H "Authorization: Bearer llmtr-your_key" \
  -H "Content-Type: application/json" \
  -d '{"model":"qwen/qwen3-vl-8b-thinking","messages":[{"role":"user","content":"Hello"}]}'

Related models