Full Model Pricing
Model profile

Qwen2.5 72B Instruct

qwen chat
qwen/qwen-2.5-72b-instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

At a glance

Model specifications

The most important technical and pricing details in one place.

Context window
32.8K tokens
Model type
Chat
Input / 1M
$0.36
Output / 1M
$0.40

Providers and routing order

Gateway-compatible routes are tried from the lowest cost; other sources are shown for catalog reference.

2 sources
1
OpenRouterGateway activeqwen/qwen-2.5-72b-instruct · Updated 05.10.2026 20:17
Input$0.3600 / 1M Output$0.4000 / 1M Cache$0.0000 / 1M Media-
2
Novita AICatalog sourceqwen/qwen-2.5-72b-instruct · Updated 05.10.2026 20:17
Input$0.3800 / 1M Output$0.4000 / 1M Cache$0.0000 / 1M Media-
Live telemetry

Usage and performance

Current request, token, and latency data across TamgaStudio.

Usage Example

Call Qwen2.5 72B Instruct with a single request through the OpenAI-compatible TamgaStudio API. Create your API key from the API Keys section in your dashboard.

curl https://api.tamga.studio/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer tl-sk-XXXXXX" \
  -d '{"model": "qwen/qwen-2.5-72b-instruct", "messages": [{"role": "user", "content": "Hello!"}]}'
import requests

resp = requests.post(
    "https://api.tamga.studio/v1/chat/completions",
    headers={"Authorization": "Bearer tl-sk-XXXXXX"},
    json={
        "model": "qwen/qwen-2.5-72b-instruct",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(resp.json()["choices"][0]["message"]["content"])
<?php
$ch = curl_init('https://api.tamga.studio/v1/chat/completions');
curl_setopt_array($ch, [
    CURLOPT_POST => true,
    CURLOPT_HTTPHEADER => [
        'Authorization: Bearer tl-sk-XXXXXX',
        'Content-Type: application/json',
    ],
    CURLOPT_POSTFIELDS => json_encode([
        'model' => 'qwen/qwen-2.5-72b-instruct',
        'messages' => [['role' => 'user', 'content' => 'Hello!']],
    ]),
    CURLOPT_RETURNTRANSFER => true,
]);
echo curl_exec($ch);

Frequently Asked Questions

Common questions about this model.

What is Qwen2.5 72B Instruct?
Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and... It is accessible through a single OpenAI-compatible API via TamgaStudio.
How much does Qwen2.5 72B Instruct cost?
Pricing is $0.36 per 1M input tokens and $0.40 per 1M output tokens. Billing is credit-based: you only pay for what you use.
What is the context length of Qwen2.5 72B Instruct?
Qwen2.5 72B Instruct supports a context of up to 32,768 tokens in one request.
Can I use Qwen2.5 72B Instruct with the OpenAI SDK?
Yes. TamgaStudio exposes all models through an OpenAI-compatible endpoint: https://api.tamga.studio/v1/chat/completions. Just swap the base URL and API key in your existing OpenAI code.
How do I get an API key?
After signing in, create a key from the "API Keys" section of your dashboard. The key is shown only once and is stored as a SHA-256 hash.