Poly Logo

Polylabs

Free ToolsBlog
Gemini
Google

Gemma 4 31B

Updated: August 2026

Gemma 4 31B Instruct is Google DeepMind's 30.7B dense multimodal model supporting text and image input with text output. Features a 256K token context window, configurable thinking/reasoning mode, native function...

Specifications

Context
262K
Input
$0.1/M
Output
$0.34/M

Capabilities

VISIONTEXTWEBCODINGTHINKINGWRITING

Similarly Priced Models

ModelProviderContextInput PriceOutput Price
GPT-5.6 Luna Pro
OpenAIOpenAI
1M$0.1/M$0.6/M
GPT-5.6 Luna
OpenAIOpenAI
1M$0.1/M$0.6/M
Qwen3.5-9B
QwenQwen
262K$0.1/M$0.15/M
ByteDance Seed: Seed-2.0-Mini
ByteDanceByteDance
262K$0.1/M$0.4/M
Qwen3.5 Flash
QwenQwen
1M$0.1/M$0.4/M

Performance Metrics

Intelligence Index

01

29.4

> 57% OF MODELS

Coding Index

02

43.4

> 54% OF MODELS

Agentic Index

03

14.4

> 43% OF MODELS

Average Response Performance

Output Speed
35.2 tok/s
Time To First Token
1.09s
Time To First Answer Token
50.36s
End To End Response Time
64.55s

DATA SOURCE: Artificial Analysis

Curious about Gemma 4 31B?