OpenAI: gpt-oss-20b
gpt-oss-20b- Input / Output
- $0.075 / $0.30
- Context
- 128K
- Max output
- 32K
- Modalities
- Chat
- Capabilities
- Tool use
Getting started
Create API key
Create a key in Console, then use it with every model on the platform.
Send your first request
Copy the example for your language and run it against the endpoint.
chat_completionOpenAI-compatiblePOST/v1/chat/completionscurl https://api.tokenlab.sh/v1/chat/completions \ -H "Content-Type: application/json" \ -H "Authorization: Bearer sk-xxx" \ -d '{ "model": "gpt-oss-20b", "messages": [ {"role": "user", "content": "Hello!"} ] }'
Pricing
Official price is the model maker's public baseline. TokenLab price is what you pay for this model on TokenLab.
Pricing
Per 1M tokens- TokenLab price
- Input$0.075Output$0.30Per 1M tokens
- Official price
- Input$0.075Output$0.30Per 1M tokens
| Official pricePer 1M tokens | TokenLab pricePer 1M tokens | Discount | |
|---|---|---|---|
| Input | $0.075 | $0.075 | - |
| Output | $0.30 | $0.30 | - |
Usage & activity
Success rate is the share of requests that completed. Latency is how long a full response takes; P95 means 95% of requests finished within that time.
Usage & availability
Last 24 hoursNo data yet
Data is based on aggregate user requests, excluding status checks.
Open in Console
Open gpt-oss-20b in Console with a prompt ready to edit or send.
Help me try gpt-oss-20b with a short message at /v1/chat/completions. Show the reply, latency, and cost.
Use cases
Agents and tools
Handle reasoning, support triage, tool calls, and multi-step tasks.
Coding
Generate, review, or debug code in the tools you already use.
Knowledge assistants
Build chat, search, and retrieval with a clear price and capability profile.
Side-by-side test
Compare response quality, latency, and price side by side.
Prompt examples
Write a concise support reply and list the assumptions behind it.
Review this API design and call out the top three integration risks.
Turn a long changelog into release notes a non-engineer would read.
Cost calculator
FAQ
How much does gpt-oss-20b cost?
On TokenLab, gpt-oss-20b costs Input $0.075 / Output $0.30 Per 1M tokens. The pricing table above shows the full breakdown.
What is gpt-oss-20b best for?
gpt-oss-20b supports JSON mode, Tool use. You can open it directly in Create.
How do I test gpt-oss-20b?
Open gpt-oss-20b in Create. A sample for /v1/chat/completions will be ready to try.
Which endpoint should gpt-oss-20b use?
Use https://api.tokenlab.sh/v1/chat/completions for gpt-oss-20b. The request example below shows the matching code shape.
Can I test gpt-oss-20b before integrating it?
Yes. Open Console starts a ready draft for gpt-oss-20b and keeps your prompt after sign-in, so you don’t lose context.
Which operations does gpt-oss-20b support?
gpt-oss-20b supports chat_completion. Select an operation above to see its endpoint and request example.