llama-3.3-70b-versatile by Groq
Free tier, no credit card. Ultra-fast LPU inference.
Access
Free API tier
Free limits
30 RPM, 1,000 RPD
Modality
text
Credit card
Not required
Commercial use
Unclear, verify
Context window
131,000 tokens
Rate limit
30 RPM · 1000 RPD
Model ID
llama-3.3-70b-versatile
Base URL
https://api.groq.com/openai/v1
Last verified
June 2026
Quickstart
curl https://api.groq.com/openai/v1/chat/completions \
-H "Authorization: Bearer $KEY" \
-d '{"model":"llama-3.3-70b-versatile","messages":[{"role":"user","content":"hi"}]}'Frequently asked
Is llama-3.3-70b-versatile free?
Yes. Groq offers it as Free API tier with these limits: 30 RPM, 1,000 RPD. No credit card is required.
Can I use llama-3.3-70b-versatile commercially?
Commercial terms are not clearly documented. Check the provider's current terms before shipping.
How do I start using llama-3.3-70b-versatile?
Free tier, no credit card. Ultra-fast LPU inference.
Building with this model?
We wire free-tier and open-weight models into production stacks: evals against your real inputs, failover, cost caps.
Free monthly update
The monthly free-models update
What is newly free, what got rate-limited, and what to switch to. One email a month. Unsubscribe anytime.
Related free models
- Kimi K2.6 (Ollama Cloud) · Free tier
- GLM 4.6 (Z.ai) · Self-host free
- Gemini 2.5 Flash (Google AI Studio) · Generous · no card
- Qwen3.7 Max (Qwen) · Free tier
- Inference API (Hugging Face) · Rate-limited · 1000s of models
- Gemini 3.5 Flash (Google Gemini) · 15 RPM, 1,500 RPD