Gemma 4 31B is Google's language model with a 256K context window and up to 41K output tokens, available from 3 providers, starting at $0.14 / 1M input and $0.4 / 1M output. Google DeepMind's 31B-parameter Gemma 4 open multimodal LLM, handling text and image inputs with text output.
Specifications
Canonical IDgoogle-gemma-4-31b
TypeLanguage
StatusActive
CreatorGoogleGoogle
Providers
Context Window256K tokens
Max Output41K tokens
Input ModalitiesImageText
Output ModalitiesText
Reasoning Effortsdefault
Parameters31B
Benchmarks
Intelligence Index
#176
Coding Index
#90
GPQA
#92
HLE
#116
IFBench
#30
Time to First Token
#419
SciCode
#87
LCR
#141
TerminalBench Hard
#57
TAU2
#163
Output TPS
#193

Capabilities

Input2/5
Text✓
Image✓
Audio·
Video·
PDF·
Output1/5
Text✓
Image·
Audio·
Video·
Embedding·
Capabilities4/13
Reasoning✓
Adaptive Reasoning·
Function Calling✓
Parallel Function Calling✓
Structured Outputs✓
Native JSON Schema·
Web Search·
URL Context·
Computer Use·
Code Execution·
File Search·
Prompt Caching·
Assistant Prefill·

Pricing by Provider

US Dollar ($)
Per 1M tokens
ProviderStandardBatch
Input
$ / 1M
Output
$ / 1M
Input
$ / 1M
Output
$ / 1M
Amazon Bedrock logo
Amazon Bedrock
bedrock_mantle/google.gemma-4-31b
$0.14$0.4——
Amazon Bedrock logo
Amazon Bedrock
bedrock_mantle/us-gov-west-1/google.gemma-4-31b
$0.168$0.48——
Cerebras logo
Cerebras
cerebras/gemma-4-31b
$0.99$1.49——
Snowflake logo
Snowflake
gemma-4-31b
$0.168$0.48$0.084$0.24

Cost Calculator

US Dollar ($)
Preset:

Cheapest Instances to Run It

Cloud GPU instances that can host Gemma 4 31B, ranked by cheapest on-demand price. The model needs about 74 GB of GPU memory at FP16 precision (estimated from its parameter count), so treat the fit as guidance rather than a guarantee.

All clouds
FP16 (full precision)
US Dollar ($)
Instance
Cloud
GPU
VRAM
Price
Cheapest region
Standard_NP20sAzure2× AMD Alveo U250 FPGA (64GB)128 GB$3.30/hr
g7e.2xlargeAWSRTX PRO Server 600096 GB$3.36/hr
Standard_NC24ads_A100_v4AzureNVIDIA A10080 GB$3.67/hr
7 more instances can run Gemma 4 31B
Unlock the full ranked list and FP8 / INT4 quantization with a CloudPrice subscription.

Versions

VersionReleasedContextInput / 1MOutput / 1MStatus
Gemma 4 31B—256K$0.140$0.400Current
Gemma 4 31B————Available
Gemma 4 26B A4B————Available
Gemma 4 26B A4B—256K$0.130$0.400Available
Gemma 4 12B————Available
Gemma 4 E4B————Available
Gemma 4 E4B————Available
Gemma 4 E2B—128K$0.040$0.080Available
Gemma 4————Available
Gemma 4 E2B————Available
Gemma 4 12B Instruct—16K$0.300$2.00Available

Model IDs

bedrock_mantle/google.gemma-4-31b
bedrock_mantle/us-gov-west-1/google.gemma-4-31b
cerebras/gemma-4-31b
gemma-4-31b
gemma-4-31b-non-reasoning
google-gemma-4-31b