Openai model page

GPT-4o-mini

openai/gpt-4o-mini

GPT-4o mini is OpenAI's newest model after [GPT-4 Omni](/models/openai/gpt-4o), supporting both text and image inputs with text outputs.\n\nAs their most advanced small model, it is many multiples more affordable than other recent frontier models, and more than 60% cheaper than [GPT-3.5 Turbo](/models/openai/gpt-3.5-turbo). It maintains SOTA intelligence, while being significantly more cost-effective.\n\nGPT-4o mini achieves an 82% score on MMLU and presently ranks higher than GPT-4 on chat preferences [common leaderboards](https://arena.lmsys.org/).\n\nCheck out the [launch announcement](https://openai.com/index/gpt-4o-mini-advancing-cost-efficient-intelligence/) to learn more.\n\n#multimodal

TextVisionStructured outputFast responsesLong context
Provider In / 1M
$0.15
Provider Out / 1M
$0.60
Context
128K
UAI Status
Pending sync

Availability on UAI

A public fallback landing page for a model that is not in the live UAI catalog yet.

Provider
Openai

Pulled from the provider's public model page

Context window
128K tokens

Public model metadata

UAI status
Pending sync

This route exists now, but the model is not active in /v1/models yet

Pricing source
Provider public page

UAI pricing will appear automatically once the catalog row is synced

This page is available for direct linking and discovery, but the row is not active in the live UAI catalog yet. Until it is synced, /v1/models may not list it and requests for openai/gpt-4o-mini can fail validation.

Why use this model through UAI

The same model route, but with one account, one balance, and one API shape.

Long-context inputs

GPT-4o-mini exposes 128K tokens, which makes it a good fit for large documents, transcripts, and retrieval-heavy chat flows.

Multimodal prompts

The synced capabilities for this row indicate image or vision support, so you can route multimodal requests through the same UAI account and billing flow.

Interactive latency

GPT-4o-mini is positioned well for chat UIs, demos, and agent loops where faster turn time matters.

Pricing and source data

Public metadata until the live UAI catalog row is synced.

UAI catalog status
Pending sync
Openai input / 1M
$0.15
Openai output / 1M
$0.60
Context window
128,000
Supported params
max_completion_tokens, temperature, top_p, stop, frequency_penalty, presence_penalty

Once this row is synced into UAI, this page will automatically switch from public fallback metadata to live UAI pricing, request validation, and app links.

How to use it on UAI

This public page exists already, but the callable catalog row is still pending sync.

Create a UAI key in API Keys, then use /v1/models or the public models catalog to see which rows are active right now.

Requests for openai/gpt-4o-mini may fail until the live UAI catalog sync includes this row. This page stays up now so the URL can still be shared, indexed, and linked from rankings.

Related models

Other active rows to compare without leaving the public catalog.

openai
OpenAI: GPT-5.6 Luna (batch)
openai/gpt-5.6-luna:batch
Input / 1M$0.13
Context1.1M tokens
openai
OpenAI: GPT-5.6 Luna Pro (batch)
openai/gpt-5.6-luna-pro:batch
Input / 1M$0.13
Context1.1M tokens
openai
OpenAI: GPT-5.6 Luna
openai/gpt-5.6-luna
Input / 1M$0.26
Context1.1M tokens