Qwen/Qwen3-VL-32B-Instruct
Qwen3-VL-32B-Instruct is a large-scale multimodal vision-language model designed for high-precision understanding and reasoning across text, images, and video. With 32 billion parameters, it combines deep visual perception with advanced text...
The most important technical and pricing details in one place.
Gateway-compatible routes are tried from the lowest cost; other sources are shown for catalog reference.

Current request, token, and latency data across TamgaStudio.
Call Qwen3-VL-32B-Instruct with a single request through the OpenAI-compatible TamgaStudio API. Create your API key from the API Keys section in your dashboard.
curl https://api.tamga.studio/v1/chat/completions \
-H "Content-Type: application/json" \
-H "Authorization: Bearer tl-sk-XXXXXX" \
-d '{"model": "Qwen/Qwen3-VL-32B-Instruct", "messages": [{"role": "user", "content": "Hello!"}]}'
Common questions about this model.
Explore other models from the same provider or in a similar price range as Qwen3-VL-32B-Instruct.