Groq API pricing
LPU-based inference host delivering 300-1000+ tokens/sec on open models, with a no-credit-card free dev tier.
Free tier or trialBilled per token·Verified Jun 27, 2026
Free tier
Free dev tier, no card: 30 RPM / 6K TPM / 14,400 req/day
Plans and prices
| Plan | Price | What it covers |
|---|---|---|
| Free (GroqCloud) | $0 | Get started free with low rate limits (e.g. ~30 RPM, capped requests-per-day); shared per-organization limits. |
| Developer / Pay-as-you-go (On-Demand) | Per-token, usage-based | Linear per-token pricing with substantially higher rate limits; no idle infrastructure fees. |
| Batch API | 50% off standard per-token pricing | Asynchronous processing at half the on-demand token cost. |
| Enterprise | Custom (contact sales) | Private/co-cloud instances, SSO/SCIM/MFA, enterprise-only models (e.g. Minimax M2.5, Qwen3-VL 32B), higher capacity. |
What to watch for on the bill
- Persistent rate-limit and over-capacity (429) complaints; free tier is tight (~30 RPM)
- Flex/best-effort service tier can return over-capacity errors under load
- SRAM-only LPU architecture (few hundred MB per chip) raises questions about cost-efficiency at very large model sizes
The pricing page, as we saw it

Source: https://groq.com/pricing. Prices change. Confirm on the vendor page before you commit.
Other LLM APIs
OpenAI API · Anthropic Claude API · Google Gemini API · Mistral La Plateforme · xAI Grok API · DeepSeek API