GooglePaid
Gemma 4 31b
API model name: gemma-4-31b
Gemma 4 31b is Google's chat model, served on the Api.Airforce unified API. It is priced at $0.14 per million input tokens and $0.40 per million output tokens. Access it through the OpenAI-compatible API with one key, alongside 100+ other models on Api.Airforce.
Pricing
Api.Airforce price vs. the provider's official rate.
Specifications
- Provider
- Type
- chat model
Benchmarks
Independent evaluations and measured speed from Artificial Analysis.
Measured atSame model, different reasoning effort.
Intelligence Index
29.7/100
Coding Index
43.4/100
GPQA Diamond86%
Humanity's Last Exam24%
SciCode43%
τ²-bench60%
τ-bench Banking15%
Terminal-Bench Hard36%
Output speed35.6 tok/s
Time to first token0.95 s
Gemma 4 31B (Reasoning) · 22.3–29.7
Source: Benchmark data by Artificial Analysis (artificialanalysis.ai)
What is Gemma 4 31b used for?
- Chatbots & assistants — conversational AI, drafting, summarizing and Q&A.
Gemma 4 31b vs. similar models
| Model | Intelligence | Context | Input / 1M | Output / 1M |
|---|---|---|---|---|
| Gemma 4 31b | 29.7 | — | $0.14 | $0.40 |
| Aqa | — | — | $57.41 | $57.41 |
| Gemini 2.0 Flash Lite | 8.4 | — | $0.08 | $0.29 |
| Gemini 2.5 Flash | 20.3 | 1M | $0.40 | $2.50 |
Prices are Api.Airforce pay-as-you-go rates per 1M tokens. Context is the maximum input length.
Related models
AqaGoogle · $57.41 / 1MGemini 2.0 Flash LiteGoogle · $0.08 / 1MGemini 2.5 FlashGoogle · $0.40 / 1MGemini 2.5 Flash Image TokenGoogle · $0.24 / 1MGemini 2.5 Flash LiteGoogle · $0.08 / 1MGemini 2.5 Flash Lite Preview 09Google · $0.11 / 1MGemini 2.5 Flash Lite Preview 09Google · $0.08 / 1MGemini 2.5 Flash Native Audio Preview 09Google · $57.41 / 1MGemini 2.5 Flash Preview 04.17Google · $0.21 / 1MGemini 2.5 Flash Preview 09Google · $0.32 / 1MGemini 2.5 Flash Preview TTSGoogle · $0.28 / 1MGemini 2.5 ProGoogle · $1.19 / 1M
Gemma 4 31b — frequently asked questions
- How much does Gemma 4 31b cost?
- Gemma 4 31b is billed pay-as-you-go at $0.14 per 1M input tokens and $0.40 per 1M output tokens. There is no subscription — you only pay for what you use.
- What can Gemma 4 31b do?
- Gemma 4 31b is Google's chat model, accessible through the Api.Airforce API.
- Is Gemma 4 31b free to use?
- Gemma 4 31b is a paid, pay-as-you-go model — no subscription, you are only charged for usage.
- How do I use Gemma 4 31b via the API?
- Gemma 4 31b is OpenAI-compatible. Point any OpenAI SDK at https://api.airforce/v1 and pass the model ID gemma-4-31b with your Api.Airforce API key.
- Who makes Gemma 4 31b?
- Gemma 4 31b is Google's chat model, served through the unified Api.Airforce gateway alongside 100+ other models.
Loading live metrics…
Loading live metrics…
Loading live metrics…
Use Gemma 4 31b via the API
OpenAI-compatible — point any OpenAI SDK at https://api.airforce/v1 and pass gemma-4-31b as the model.
cURL
curl https://api.airforce/v1/chat/completions \
-H "Authorization: Bearer $AIRFORCE_API_KEY" \
-H "Content-Type: application/json" \
-d '{
"model": "gemma-4-31b",
"messages": [{ "role": "user", "content": "Hello!" }]
}'Python
from openai import OpenAI
client = OpenAI(base_url="https://api.airforce/v1", api_key="$AIRFORCE_API_KEY")
r = client.chat.completions.create(
model="gemma-4-31b",
messages=[{"role": "user", "content": "Hello!"}],
)
print(r.choices[0].message.content)JavaScript
import OpenAI from "openai";
const client = new OpenAI({ baseURL: "https://api.airforce/v1", apiKey: process.env.AIRFORCE_API_KEY });
const r = await client.chat.completions.create({
model: "gemma-4-31b",
messages: [{ role: "user", content: "Hello!" }],
});
console.log(r.choices[0].message.content);