Full Model Pricing
Model profile

OpenAI GPT-OSS 120B

OpenAI chat
openai/gpt-oss-120b

gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized...

At a glance

Model specifications

The most important technical and pricing details in one place.

Context window
1M tokens
Model type
Chat
Input / 1M
$0.04
Output / 1M
$0.17

Providers and routing order

Gateway-compatible routes are tried from the lowest cost; other sources are shown for catalog reference.

8 sources
1
Fireworks AIGateway activeaccounts/fireworks/models/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.0000 / 1M Output$0.0000 / 1M Cache$0.0000 / 1M Media-
2
SiliconFlowGateway activeopenai/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.0000 / 1M Output$0.0000 / 1M Cache$0.0000 / 1M Media-
3
OpenRouterGateway activeopenai/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.0370 / 1M Output$0.1700 / 1M Cache$0.0000 / 1M Media-
4
DeepInfraCatalog sourceopenai/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.0370 / 1M Output$0.1700 / 1M Cache$0.0000 / 1M Media-
5
Novita AICatalog sourceopenai/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.0500 / 1M Output$0.2500 / 1M Cache$0.0000 / 1M Media-
6
GroqGateway activeopenai/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.1500 / 1M Output$0.6000 / 1M Cache$0.0750 / 1M Media-
7
Together AIGateway activeopenai/gpt-oss-120b · Updated 05.10.2026 20:17
Input$0.1500 / 1M Output$0.6000 / 1M Cache$0.0000 / 1M Media-
8
CerebrasGateway activegpt-oss-120b · Updated 05.10.2026 20:17
Input$0.3500 / 1M Output$0.7500 / 1M Cache$0.0000 / 1M Media-
Live telemetry

Usage and performance

Current request, token, and latency data across TamgaStudio.

Usage Example

Call OpenAI GPT-OSS 120B with a single request through the OpenAI-compatible TamgaStudio API. Create your API key from the API Keys section in your dashboard.

curl https://api.tamga.studio/v1/chat/completions \
  -H "Content-Type: application/json" \
  -H "Authorization: Bearer tl-sk-XXXXXX" \
  -d '{"model": "openai/gpt-oss-120b", "messages": [{"role": "user", "content": "Hello!"}]}'
import requests

resp = requests.post(
    "https://api.tamga.studio/v1/chat/completions",
    headers={"Authorization": "Bearer tl-sk-XXXXXX"},
    json={
        "model": "openai/gpt-oss-120b",
        "messages": [{"role": "user", "content": "Hello!"}],
    },
)
print(resp.json()["choices"][0]["message"]["content"])
<?php
$ch = curl_init('https://api.tamga.studio/v1/chat/completions');
curl_setopt_array($ch, [
    CURLOPT_POST => true,
    CURLOPT_HTTPHEADER => [
        'Authorization: Bearer tl-sk-XXXXXX',
        'Content-Type: application/json',
    ],
    CURLOPT_POSTFIELDS => json_encode([
        'model' => 'openai/gpt-oss-120b',
        'messages' => [['role' => 'user', 'content' => 'Hello!']],
    ]),
    CURLOPT_RETURNTRANSFER => true,
]);
echo curl_exec($ch);

Frequently Asked Questions

Common questions about this model.

What is OpenAI GPT-OSS 120B?
gpt-oss-120b is an open-weight, 117B-parameter Mixture-of-Experts (MoE) language model from OpenAI designed for high-reasoning, agentic, and general-purpose production use cases. It activates 5.1B parameters per forward pass and is optimized... It is accessible through a single OpenAI-compatible API via TamgaStudio.
How much does OpenAI GPT-OSS 120B cost?
Pricing is $0.04 per 1M input tokens and $0.17 per 1M output tokens. Billing is credit-based: you only pay for what you use.
What is the context length of OpenAI GPT-OSS 120B?
OpenAI GPT-OSS 120B supports a context of up to 1,048,576 tokens in one request.
Can I use OpenAI GPT-OSS 120B with the OpenAI SDK?
Yes. TamgaStudio exposes all models through an OpenAI-compatible endpoint: https://api.tamga.studio/v1/chat/completions. Just swap the base URL and API key in your existing OpenAI code.
How do I get an API key?
After signing in, create a key from the "API Keys" section of your dashboard. The key is shown only once and is stored as a SHA-256 hash.