GooglePaidOperational

Gemini 2.5 Flash Lite

API model name: gemini-2.5-flash-lite

Gemini 2.5 Flash Lite is Google's chat model, served on the Api.Airforce unified API. Capabilities include Vision, Tool calling, Reasoning, Prompt caching. It is priced at $0.08 per million input tokens and $0.32 per million output tokens. That is below the provider's $0.10 official input rate. Access it through the OpenAI-compatible API with one key, alongside 100+ other models on Api.Airforce.

Pricing

Input / 1M tokens
$0.08
Output / 1M tokens
$0.32
Cache read / 1M tokens
$0.70
Cache write / 1M tokens
$8.75
Official input rate
$0.10
Official output rate
$0.40

Api.Airforce price vs. the provider's official rate.-20%

Specifications

Provider
Google
Type
chat model
Prompt caching
Supported

Capabilities

VisionTool callingReasoningPrompt caching

Benchmarks

Independent evaluations and measured speed from Artificial Analysis.

Measured atSame model, different reasoning effort.
Intelligence Index
11.4/100
Math Index
53.3/100
MMLU-Pro76%
GPQA Diamond63%
Humanity's Last Exam7%
LiveCodeBench59%
AIME 202553%
MATH-50097%

Gemini 2.5 Flash-Lite (Reasoning) · 6.7–11.4

Source: Benchmark data by Artificial Analysis (artificialanalysis.ai)

What is Gemini 2.5 Flash Lite used for?

  • Chatbots & assistants — conversational AI, drafting, summarizing and Q&A.
  • Image understanding — analyze photos, screenshots, charts and scanned documents.
  • Agents & automation — function calling and tool use for multi-step workflows.
  • Complex reasoning — math, coding and step-by-step problem solving.

Gemini 2.5 Flash Lite vs. similar models

ModelIntelligenceContextInput / 1MOutput / 1M
Gemini 2.5 Flash Lite11.4$0.08$0.32
Aqa$57.41$57.41
Gemini 2.0 Flash Lite8.4$0.08$0.32
Gemini 2.5 Flash20.31M$0.40$2.50

Prices are Api.Airforce pay-as-you-go rates per 1M tokens. Context is the maximum input length.

Related models

Gemini 2.5 Flash Lite — frequently asked questions

How much does Gemini 2.5 Flash Lite cost?
Gemini 2.5 Flash Lite is billed pay-as-you-go at $0.08 per 1M input tokens and $0.32 per 1M output tokens. There is no subscription — you only pay for what you use.
What can Gemini 2.5 Flash Lite do?
Gemini 2.5 Flash Lite supports Vision, Tool calling, Reasoning, Prompt caching.
Is Gemini 2.5 Flash Lite free to use?
Gemini 2.5 Flash Lite is a paid, pay-as-you-go model — no subscription, you are only charged for usage.
How do I use Gemini 2.5 Flash Lite via the API?
Gemini 2.5 Flash Lite is OpenAI-compatible. Point any OpenAI SDK at https://api.airforce/v1 and pass the model ID gemini-2.5-flash-lite with your Api.Airforce API key.
Who makes Gemini 2.5 Flash Lite?
Gemini 2.5 Flash Lite is Google's chat model, served through the unified Api.Airforce gateway alongside 100+ other models.