GooglePaidOperational

Gemini 3.1 Flash Lite Preview

API model name: gemini-3.1-flash-lite-preview

Gemini 3.1 Flash Lite Preview is Google's chat model, served on the Api.Airforce unified API. It has a 1M-token context window. Beyond text, it accepts image, audio, video, document as input. Capabilities include Vision, Tool calling, Reasoning, Documents, Prompt caching. It is priced at $0.33 per million input tokens and $1.95 per million output tokens. Knowledge cutoff: 2026-03. Access it through the OpenAI-compatible API with one key, alongside 100+ other models on Api.Airforce.

Pricing

Input / 1M tokens
$0.33
Output / 1M tokens
$1.95
Cache read / 1M tokens
$0.04

Specifications

Provider
Google
Type
chat model
Context window
1M tokens
Max output
33K tokens
Knowledge cutoff
2026-03
Input
text, image, audio, video, document
Output
text
Prompt caching
Supported

Capabilities

VisionTool callingReasoningDocumentsPrompt cachingStreaming

What is Gemini 3.1 Flash Lite Preview used for?

  • Chatbots & assistants — conversational AI, drafting, summarizing and Q&A.
  • Image understanding — analyze photos, screenshots, charts and scanned documents.
  • Agents & automation — function calling and tool use for multi-step workflows.
  • Complex reasoning — math, coding and step-by-step problem solving.
  • Document analysis — summarize and answer questions across long files.
  • Long-context tasks — process entire documents or codebases in a single prompt.
  • Real-time experiences — stream tokens for responsive chat and apps.

Gemini 3.1 Flash Lite Preview vs. similar models

ModelIntelligenceContextInput / 1MOutput / 1M
Gemini 3.1 Flash Lite Preview1M$0.33$1.95
[SP]gemini 2.5 Pro$3.97$0.02
[SP]gemini 3.1 Pro Preview$3.97$0.02
[SP]gemini 3.5 Flash$13.27$0.02

Prices are Api.Airforce pay-as-you-go rates per 1M tokens. Context is the maximum input length.

Related models

Gemini 3.1 Flash Lite Preview — frequently asked questions

How much does Gemini 3.1 Flash Lite Preview cost?
Gemini 3.1 Flash Lite Preview is billed pay-as-you-go at $0.33 per 1M input tokens and $1.95 per 1M output tokens. There is no subscription — you only pay for what you use.
What is the context window of Gemini 3.1 Flash Lite Preview?
Gemini 3.1 Flash Lite Preview supports a context window of up to 1M tokens. It can return up to 33K tokens in a single response.
What can Gemini 3.1 Flash Lite Preview do?
Gemini 3.1 Flash Lite Preview supports Vision, Tool calling, Reasoning, Documents, Prompt caching.
Is Gemini 3.1 Flash Lite Preview free to use?
Gemini 3.1 Flash Lite Preview is a paid, pay-as-you-go model — no subscription, you are only charged for usage.
How do I use Gemini 3.1 Flash Lite Preview via the API?
Gemini 3.1 Flash Lite Preview is OpenAI-compatible. Point any OpenAI SDK at https://api.airforce/v1 and pass the model ID gemini-3.1-flash-lite-preview with your Api.Airforce API key.
Who makes Gemini 3.1 Flash Lite Preview?
Gemini 3.1 Flash Lite Preview is Google's chat model, served through the unified Api.Airforce gateway alongside 100+ other models.