Groq is the premier neocloud for fast inference. One fully integrated platform for infrastructure, inference, and control. Millions of developers run trillions of tokens on Groq every week. Inference platform · OpenAI-compatible API · Fast Inference · Low Latency · Lpu · Open Weight

Intelligence vs Price

Best value among Groq models on this chart: Qwen3.8 27B · Qwen3.6 27B · GPT OSS 120B (and 2 more on the dashed frontier). Prices use each model's lowest available Groq price across regions. Hover any dot for full pricing, or click a creator in the legend to isolate.

Language Models
Intelligence
Blended Price, $
Log X

Groq models

20 models in Global, 15 with pricing
All Model Types
All Creators
US Dollar ($)
Per 1M tokens
Input/1M
to
Output/1M
to
Select
Model
Creator
Input Price, $
Output Price, $
Context
Max Output
Inference Providers
Intelligence
Coding
Qwen3.8 27BAlibaba logoAlibaba0.84.001.0M131K#1#1
Qwen3.6 27BAlibaba logoAlibaba0.63.00262K236K#2#2
GPT OSS 120BOpenAI logoOpenAI0.150.6131K131K#3#3
GPT OSS 20BOpenAI logoOpenAI0.0750.3131K131K#4#4
Qwen3 32BAlibaba logoAlibaba0.290.59131K41K#5#5
Llama 3.3 70B InstructMeta logoMeta0.590.79131K120K#6#6
Llama 3.1 8B InstructMeta logoMeta0.050.08200K128K#7#7
Gemma 7B ITGoogle logoGoogle0.050.088K8KN/AN/A
GPT OSS 20B SafeguardOpenAI logoOpenAI0.0750.3131K66KN/AN/A
Kimi K2 InstructMoonshot AI (Kimi) logoMoonshot AI (Kimi)1.003.00262K100KN/AN/A
Llama 4 17B Maverick InstructMeta logoMeta0.20.61.0M16KN/AN/A
Llama 4 17B Scout InstructMeta logoMeta0.110.3410.0M16KN/AN/A
Llama Prompt Guard 2 22MMeta logoMeta0.030.03512512N/AN/A
Llama Prompt Guard 2 86MMeta logoMeta0.040.04512512N/AN/A
LlamaGuard 4 12BMeta logoMeta0.20.21.0M16KN/AN/A
Orpheus 1 EnglishCanopylabsN/AN/A4KN/AN/AN/A
Orpheus Arabic SaudiCanopylabsN/AN/A4KN/AN/AN/A
PlayAI TTSPlayAI logoPlayAIN/AN/A10K10KN/AN/A
Whisper 3 LargeOpenAI logoOpenAIN/AN/AN/AN/AN/AN/A
Whisper 3 Large TurboOpenAI logoOpenAIN/AN/AN/AN/AN/AN/A