OpenAIPaidOperational

GPT Oss 120b

API model name: gpt-oss-120b

GPT Oss 120b is OpenAI's chat model, served on the Api.Airforce unified API. It has a 131K-token context window. Capabilities include Tool calling, Reasoning. It is priced at $0.09 per million input tokens and $0.37 per million output tokens. That is below the provider's $0.15 official input rate. Knowledge cutoff: 2024-06. Access it through the OpenAI-compatible API with one key, alongside 100+ other models on Api.Airforce.

Pricing

Input / 1M tokens
$0.09
Output / 1M tokens
$0.37
Cache read / 1M tokens
$0.05
Official input rate
$0.15
Official output rate
$0.60

Api.Airforce price vs. the provider's official rate.-38%

Specifications

Provider
OpenAI
Type
chat model
Context window
131K tokens
Max output
33K tokens
Knowledge cutoff
2024-06
Input
text
Output
text

Capabilities

Tool callingReasoningStreaming

Benchmarks

Independent evaluations and measured speed from Artificial Analysis.

Measured atSame model, different reasoning effort.
Intelligence Index
11.6/100
Coding Index
30.4/100
Math Index
93.4/100
MMLU-Pro81%
GPQA Diamond78%
Humanity's Last Exam20%
LiveCodeBench88%
AIME 202593%
SciCode34%
Output speed219.8 tok/s
Time to first token0.50 s

gpt-oss-120b (high) · 10.2–11.6

Source: Benchmark data by Artificial Analysis (artificialanalysis.ai)

What is GPT Oss 120b used for?

  • Chatbots & assistants — conversational AI, drafting, summarizing and Q&A.
  • Agents & automation — function calling and tool use for multi-step workflows.
  • Complex reasoning — math, coding and step-by-step problem solving.
  • Real-time experiences — stream tokens for responsive chat and apps.

GPT Oss 120b vs. similar models

ModelIntelligenceContextInput / 1MOutput / 1M
GPT Oss 120b11.6131K$0.09$0.37
Babbage 002——$0.34$0.34
Chatgpt 4o Latest——$2.75$14.50
Codex Auto Review——$0.09$0.56

Prices are Api.Airforce pay-as-you-go rates per 1M tokens. Context is the maximum input length.

Related models

GPT Oss 120b — frequently asked questions

How much does GPT Oss 120b cost?
GPT Oss 120b is billed pay-as-you-go at $0.09 per 1M input tokens and $0.37 per 1M output tokens. There is no subscription — you only pay for what you use.
What is the context window of GPT Oss 120b?
GPT Oss 120b supports a context window of up to 131K tokens. It can return up to 33K tokens in a single response.
What can GPT Oss 120b do?
GPT Oss 120b supports Tool calling, Reasoning.
Is GPT Oss 120b free to use?
GPT Oss 120b is a paid, pay-as-you-go model — no subscription, you are only charged for usage.
How do I use GPT Oss 120b via the API?
GPT Oss 120b is OpenAI-compatible. Point any OpenAI SDK at https://api.airforce/v1 and pass the model ID gpt-oss-120b with your Api.Airforce API key.
Who makes GPT Oss 120b?
GPT Oss 120b is OpenAI's chat model, served through the unified Api.Airforce gateway alongside 100+ other models.