gpt-oss-120b by Cerebras
Free tier, no credit card. Ultra-fast inference (~2,600 tok/s). 1M tokens/day cap. 8K context cap on free tier.
Access
Free API tier
Free limits
30 RPM, 14,400 RPD, 1M TPD
Modality
text
Credit card
Not required
Commercial use
Unclear, verify
Context window
128,000 tokens
Rate limit
30 RPM · 14400 RPD · 1,000,000 TPD
Model ID
gpt-oss-120b
Base URL
https://api.cerebras.ai/v1
Last verified
June 2026
Quickstart
curl https://api.cerebras.ai/v1/chat/completions \
-H "Authorization: Bearer $KEY" \
-d '{"model":"gpt-oss-120b","messages":[{"role":"user","content":"hi"}]}'Frequently asked
Is gpt-oss-120b free?
Yes. Cerebras offers it as Free API tier with these limits: 30 RPM, 14,400 RPD, 1M TPD. No credit card is required.
Can I use gpt-oss-120b commercially?
Commercial terms are not clearly documented. Check the provider's current terms before shipping.
How do I start using gpt-oss-120b?
Free tier, no credit card. Ultra-fast inference (~2,600 tok/s). 1M tokens/day cap. 8K context cap on free tier.
Building with this model?
We wire free-tier and open-weight models into production stacks: evals against your real inputs, failover, cost caps.
Free monthly update
The monthly free-models update
What is newly free, what got rate-limited, and what to switch to. One email a month. Unsubscribe anytime.
Related free models
- Kimi K2.6 (Ollama Cloud) · Free tier
- GLM 4.6 (Z.ai) · Self-host free
- Qwen3.7 Max (Qwen) · Free tier
- Gemini 2.5 Flash (Google AI Studio) · Generous · no card
- Inference API (Hugging Face) · Rate-limited · 1000s of models
- Gemini 3.5 Flash (Google Gemini) · 15 RPM, 1,500 RPD