Groq
Groq is the premier neocloud for fast inference. One fully integrated platform for infrastructure, inference, and control. Millions of developers run trillions of tokens on Groq every week. Inference platform · OpenAI-compatible API · Fast Inference · Low Latency · Lpu · Open Weight
Intelligence vs Price
Best value among Groq models on this chart: Qwen3.8 27B · GPT OSS 120B · GPT OSS 20B. Prices use each model's lowest available Groq price across regions. Hover any dot for full pricing, or click a creator in the legend to isolate.
Groq models
11 models in Global, 7 with pricing| Select | Model | Creator | Input Price, $ | Output Price, $ | Context | Max Output | Inference Providers | Intelligence | Coding |
|---|---|---|---|---|---|---|---|---|---|
| Qwen3.8 27B | 0.8 | 4.00 | 1.0M | 236K | #1 | #1 | |||
| GPT OSS 120B | 0.15 | 0.6 | 131K | 131K | #2 | #2 | |||
| GPT OSS 20B | 0.075 | 0.3 | 131K | 131K | #3 | #3 | |||
| GPT OSS 20B Safeguard | 0.075 | 0.3 | 131K | 66K | N/A | N/A | |||
| Llama Prompt Guard 2 22M | 0.03 | 0.03 | 512 | 512 | N/A | N/A | |||
| Llama Prompt Guard 2 86M | 0.04 | 0.04 | 512 | 512 | N/A | N/A | |||
| LlamaGuard 3 8B | 0.2 | 0.2 | 131K | 16K | N/A | N/A | |||
| Orpheus 1 English | Canopylabs | N/A | N/A | 4K | N/A | N/A | N/A | ||
| Orpheus Arabic Saudi | Canopylabs | N/A | N/A | 4K | N/A | N/A | N/A | ||
| Whisper 3 Large | N/A | N/A | N/A | N/A | N/A | N/A | |||
| Whisper 3 Large Turbo | N/A | N/A | N/A | N/A | N/A | N/A |