Full Model Pricing
Model profile

Qwen: Qwen3 VL 30B A3B Thinking

qwen chat
qwen/qwen3-vl-30b-a3b-thinking

Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels...

At a glance

Model specifications

The most important technical and pricing details in one place.

Context window
1M tokens
Model type
Chat
Input / 1M
$0.20
Output / 1M
$2.40

Providers and routing order

Gateway-compatible routes are tried from the lowest cost; other sources are shown for catalog reference.

2 sources
1
SiliconFlowGateway activeQwen/Qwen3-VL-30B-A3B-Thinking · Updated 05.10.2026 20:17
Input$0.0000 / 1M Output$0.0000 / 1M Cache$0.0000 / 1M Media-
2
OpenRouterGateway activeqwen/qwen3-vl-30b-a3b-thinking · Updated 05.10.2026 20:17
Input$0.2000 / 1M Output$2.4000 / 1M Cache$0.0000 / 1M Media-
Live telemetry

Usage and performance

Current request, token, and latency data across TamgaStudio.

Usage Example

Call Qwen: Qwen3 VL 30B A3B Thinking with a single request through the OpenAI-compatible TamgaStudio API. Create your API key from the API Keys section in your dashboard.

curl https://api.tamga.studio/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer tl-sk-XXXXXX" \
  -d '{"model": "qwen/qwen3-vl-30b-a3b-thinking", "messages": [{"role": "user", "content": "Hello!"}]}'
import requests

resp = requests.post(
    "https://api.tamga.studio/v1/chat/completions",
    headers={"Authorization": "Bearer tl-sk-XXXXXX"},
    json={
        "model": "qwen/qwen3-vl-30b-a3b-thinking",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(resp.json()["choices"][0]["message"]["content"])
<?php
$ch = curl_init('https://api.tamga.studio/v1/chat/completions');
curl_setopt_array($ch, [
    CURLOPT_POST => true,
    CURLOPT_HTTPHEADER => [
        'Authorization: Bearer tl-sk-XXXXXX',
        'Content-Type: application/json',
    ],
    CURLOPT_POSTFIELDS => json_encode([
        'model' => 'qwen/qwen3-vl-30b-a3b-thinking',
        'messages' => [['role' => 'user', 'content' => 'Hello!']],
    ]),
    CURLOPT_RETURNTRANSFER => true,
]);
echo curl_exec($ch);

Frequently Asked Questions

Common questions about this model.

What is Qwen: Qwen3 VL 30B A3B Thinking?
Qwen3-VL-30B-A3B-Thinking is a multimodal model that unifies strong text generation with visual understanding for images and videos. Its Thinking variant enhances reasoning in STEM, math, and complex tasks. It excels... It is accessible through a single OpenAI-compatible API via TamgaStudio.
How much does Qwen: Qwen3 VL 30B A3B Thinking cost?
Pricing is $0.20 per 1M input tokens and $2.40 per 1M output tokens. Billing is credit-based: you only pay for what you use.
What is the context length of Qwen: Qwen3 VL 30B A3B Thinking?
Qwen: Qwen3 VL 30B A3B Thinking supports a context of up to 1,048,576 tokens in one request.
Can I use Qwen: Qwen3 VL 30B A3B Thinking with the OpenAI SDK?
Yes. TamgaStudio exposes all models through an OpenAI-compatible endpoint: https://api.tamga.studio/v1/chat/completions. Just swap the base URL and API key in your existing OpenAI code.
How do I get an API key?
After signing in, create a key from the "API Keys" section of your dashboard. The key is shown only once and is stored as a SHA-256 hash.