GPT OSS 120B

Open-source 120B parameter model with reasoning capabilities via Groq inference.

gpt-oss-120b
STABLEGet StartedView uptime
131,072 context
Starting at $0.03/M input tokens
Starting at $0.14/M output tokens
Streaming
Tools
Reasoning
JSON Output
No ratings yetSign in to rate

All Providers for GPT OSS 120B

LLM Gateway routes requests to the best providers that are able to handle your prompt size and parameters.

SCX.aiUp to 4x faster
Context: 131.1kQuant: fp8
Input
$0.17
/M tokens
Cache Read
/M tokens
Output
$0.55
/M tokens
Get Started
Groq
Context: 131.1k
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.75
/M tokens
Get Started
Cerebras
Context: 131.1k
Input
$0.35
/M tokens
Cache Read
/M tokens
Output
$0.75
/M tokens
Get Started
NanoGPT
Context: 131.1k
Input
$0.05
/M tokens
Cache Read
/M tokens
Output
$0.25
/M tokens
Get Started
ByteDance
Context: 128k
Input
$0.1
/M tokens
Cache Read
$0.02
/M tokens
Output
$0.5
/M tokens
Get Started
Nebius AI
Context: 131.1kQuant: fp4
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.6
/M tokens
Get Started
Together AI
Context: 131.1k
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.6
/M tokens
Get Started
Azure
Context: 131.1k
Input
$0.15
/M tokens
Cache Read
/M tokens
Output
$0.6
/M tokens
Get Started
Runware
Context: 131.1k
Input
$0.032
/M tokens
Cache Read
$0.032
/M tokens
Output
$0.14
/M tokens
Get Started