GPT OSS 20B
Model details
- Model
openai/gpt-oss-20b- Provider
nvidia- API
openai-completions- Base URL
https://integrate.api.nvidia.com/v1- Input
- text
- Reasoning
- Yes
- Context window
- 131,072
- Max tokens
- 32,768
Show configuration
{
"providers": {
"nvidia": {
"apiKey": "YOUR_API_KEY",
"models": [
{
"id": "openai/gpt-oss-20b",
"name": "GPT OSS 20B",
"reasoning": true,
"input": [
"text"
],
"contextWindow": 131072,
"maxTokens": 32768,
"cost": {
"input": 0,
"output": 0,
"cacheRead": 0,
"cacheWrite": 0
},
"compat": {
"supportsStore": false,
"supportsDeveloperRole": false,
"supportsReasoningEffort": false,
"maxTokensField": "max_tokens",
"supportsStrictMode": false,
"supportsLongCacheRetention": false
}
}
],
"api": "openai-completions",
"baseUrl": "https://integrate.api.nvidia.com/v1"
}
}
}Pricing
USD per million tokens. A tier is selected from the total input tokens in each request and applies to that entire request.
| Request input | Input | Output | Cache read | Cache write |
|---|---|---|---|---|
| All requests | $0 | $0 | $0 | $0 |
Session cost calculator
Estimate the requests made during an agent session. A user turn can make several model calls while using tools, so costs are calculated per model request.
Estimated session$0.00
- Uncached input
- —
- Cache reads
- —
- Uncached prefixes
- —
- Cache writes
- —
- Output
- —
- Without caching
- —
The first request starts cold unless marked otherwise. Output is billed separately from context growth because reasoning tokens are not always retained. This remains a directional estimate: providers differ in eligibility, rounding, and retention.
Compatibility flags
Effective values after applying Pi's API defaults and model overrides.
| Feature | Value |
|---|---|
supportsStore | No |
supportsDeveloperRole | No |
supportsReasoningEffort | No |
supportsUsageInStreaming | Yes |
maxTokensField | max_tokens |
requiresToolResultName | No |
requiresAssistantAfterToolResult | No |
requiresThinkingAsText | No |
requiresReasoningContentOnAssistantMessages | No |
thinkingFormat | openai |
openRouterRouting | Empty |
vercelGatewayRouting | Empty |
chatTemplateKwargs | Empty |
zaiToolStream | No |
supportsStrictMode | No |
cacheControlFormat | None |
sendSessionAffinityHeaders | No |
supportsLongCacheRetention | No |