Groq

Groq API Β· Ultra-fast LLM inference on custom LPU hardware.

Groq logo

Groq

βœ“ Verified Pricing LLM API
Groq API

Ultra-fast LLM inference on custom LPU hardware.

All Models & Pricing
Model Input / 1M tokCached / 1M tokOutput / 1M tok
GPT OSS 120B
OpenAI open-weight 120B on LPU
$0.150 - $0.600
GPT OSS 20B
OpenAI open-weight 20B on LPU
$0.075 - $0.300
Llama 4 Scout
17Bx16E MoE, 128K context
$0.110 - $0.340
Llama 3.3 70B Versatile
128K context
$0.590 - $0.790
Llama 3.1 8B Instant
Fastest, cheapest tier
$0.050 - $0.080
Qwen3 32B
131K context
$0.290 - $0.590
Price History

Tracked automatically β€” a marker appears only when a price actually changed.

Groq
Get started with Groq

Ultra-fast LLM inference on custom LPU hardware.

View official pricing β†’
Prices verified on official pricing page. Always confirm before purchase.
Details
Category
LLM API
Last checked
2026-08-12 21:02 EST
Last price change
2026-03-14
Pricing status
βœ“ Verified
Compare with alternatives
See how Groq stacks up against other providers.
Browse API providers β†’
✦
CloudMart AI Compute
Answers from live pricing data
✦
What are you building?
Compare providers, estimate costs, or get a recommendation - answers come from CloudMart's live dataset, never made-up prices.
Enter to send Β· Shift+Enter for a new line
Compare:
Report a problem