Groq is the premier neocloud for fast inference. One fully integrated platform for infrastructure, inference, and control. Millions of developers run trillions of tokens on Groq every week. Inference platform · OpenAI-compatible API · Fast Inference · Low Latency · Lpu · Open Weight

Intelligence vs Price

Best value among Groq models on this chart: Qwen3.8 27B · GPT OSS 120B · GPT OSS 20B. Prices use each model's lowest available Groq price across regions. Hover any dot for full pricing, or click a creator in the legend to isolate.

Language Models
Intelligence
Blended Price, $
Log X

Groq models

11 models in Global, 7 with pricing
All Model Types
All Creators
US Dollar ($)
Per 1M tokens
Input/1M
to
Output/1M
to
Select
Model
Creator
Input Price, $
Output Price, $
Context
Max Output
Inference Providers
Intelligence
Coding
Qwen3.8 27BAlibaba logoAlibaba0.84.001.0M236K#1#1
GPT OSS 120BOpenAI logoOpenAI0.150.6131K131K#2#2
GPT OSS 20BOpenAI logoOpenAI0.0750.3131K131K#3#3
GPT OSS 20B SafeguardOpenAI logoOpenAI0.0750.3131K66KN/AN/A
Llama Prompt Guard 2 22MMeta logoMeta0.030.03512512N/AN/A
Llama Prompt Guard 2 86MMeta logoMeta0.040.04512512N/AN/A
LlamaGuard 3 8BMeta logoMeta0.20.2131K16KN/AN/A
Orpheus 1 EnglishCanopylabsN/AN/A4KN/AN/AN/A
Orpheus Arabic SaudiCanopylabsN/AN/A4KN/AN/AN/A
Whisper 3 LargeOpenAI logoOpenAIN/AN/AN/AN/AN/AN/A
Whisper 3 Large TurboOpenAI logoOpenAIN/AN/AN/AN/AN/AN/A