Skip to content

GPT-5.1 Mini

openai/gpt-5.1-mini
Language

The volume workhorse of the GPT line, most of the capability at a fifth of the price.

Model overview

Context window

400K

tokens

Input price

quoted on request

Output price

quoted on request

Weekly volume

tokens / week

Open weights

No

API only

Capabilities

5

of 11 tags

Description

GPT-5.1 Mini is a language model from OpenAI. Webparam does not stock this model yet and can source it on request — tell us what you need it for and we will come back with availability and a price. OpenAI’s stated focus is the reference point for frontier capability. The weights are not published, so it is API-only.

Capabilities

Tool calling
Calls functions you define and returns their arguments as structured data, the basis of agents.
JSON mode
Constrains the response to valid JSON matching a schema you supply.
Streaming
Emits tokens as they are generated, so answers appear progressively rather than all at once.
Vision input
Accepts images alongside text in the prompt.
Caching
Reuses already-processed prompt prefixes across requests, cutting input cost for repeated context.

Strengths & weaknesses

Strengths

  • Excellent capability-per-dollar
  • Same 400K context as the flagship
  • Fast

Weaknesses

  • Hard problems still need the flagship
  • Closed weights

Pricing

Free
Pricing for GPT-5.1 Mini, per 1M tokens
RatePriceUnit
Input$0.00per 1M tokens
Output not charged

Prompt caching is supported, so repeated prefixes bill below the listed input rate. Free at the point of use; fair-use rate limits apply to the free tier. Illustrative placeholder pricing.

Context window

400Ktokens

covers prompt and response together, so a long input leaves less room for the answer.

400K tokens against a peer range of 32K to 10M, median 256K. Compared across 24 language models. Logarithmic scale.

Supported features

Feature support for GPT-5.1 Mini
FeatureSupport
Tool callingSupported
JSON modeSupported
StreamingSupported
Vision inputSupported
Audio inputNot supported
Long contextNot supported
Open weightsNot supported
ReasoningNot supported
Fine-tunableNot supported
BatchNot supported
CachingSupported

Example use cases

  • Production default

    Excellent capability-per-dollar

  • High-volume features

    Same 400K context as the flagship

  • Cost-tiered routing

    Fast

Comparison

GPT-5.1 Mini compared with GPT-5.1 and Doubao 1.5 Pro
AttributeGPT-5.1 MiniGPT-5.1Doubao 1.5 Pro
Context window400K400K256K
Input price$0.00$0.00$0.00
Output price not charged not charged not charged
Tool callingSupportedSupportedSupported
JSON modeSupportedSupportedSupported
Vision inputSupportedSupportedNot supported
Open weightsNot supportedNot supportedNot supported
Prompt cachingSupportedSupportedSupported

Best value in each row is highlighted. Illustrative placeholder data.

Provider information

View OpenAI
OpenAI4 models

OpenAI · San Francisco, United States · est. 2015

OpenAI set the template the whole industry follows, from GPT's scaling curves to the API-first business model every provider now imitates. Its GPT-5 generation and o-series reasoning models remain the benchmark others measure against, and its developer platform is the most imitated interface in AI.

Recent releases

    Documentation

    Frequently asked questions

    Input is billed at $0.00 per 1M tokens, and there is no separate output charge. Prompt caching bills repeated prefixes below the listed input rate. Every figure on this page is an illustrative placeholder.