ColinBuilds.com

ColinBuilds.com · Current official OpenAI API pricing

GPT-4.1 mini

Input / 1M
$0.40
Output / 1M
$1.60
OpenAI-published context window
1,047,576 tokens

GPT-4.1 mini / gpt-4.1-mini; OpenAI pricing row lists standard pricing at $0.40 input, $0.10 cached input, and $1.60 output per 1M tokens · as of July 9, 2026

Compare models

Same per-1M input/output rates as above. Enter your monthly token volumes to compare bills.

GPT-4.1 mini

Input: 1M × $0.40 = $0.40
Output: 1M × $1.60 = $1.60
Total
$2.00

GPT-4o

Input: 1M × $2.50 = $2.50
Output: 1M × $10.00 = $10.00
Total
$12.50

Difference

$10.50

GPT-4o costs $10.50 more than GPT-4.1 mini.

Popular comparisons

30 total featuring GPT-4.1 mini.

Open compare hub →

GPT-4.1 mini is OpenAI’s smaller, faster GPT-4.1-family API model for instruction-following, tool calling, and long-context text or image-input work at a lower price than GPT-4.1.

What this model is

OpenAI’s model page lists GPT-4.1 mini as the smaller, faster version of GPT-4.1 and identifies the model alias as gpt-4.1-mini. The same page lists the snapshot gpt-4.1-mini-2025-04-14.

GPT-4.1 mini is a separate model from GPT-4.1 and GPT-4o. OpenAI’s pricing page lists separate rows for gpt-4.1-mini, gpt-4.1, gpt-4o, and gpt-4o-mini, with different prices.

Pricing notes

The calculator on this page uses OpenAI’s official standard pricing row for gpt-4.1-mini: $0.40 per 1M input tokens and $1.60 per 1M output tokens.

OpenAI’s pricing row also lists cached input at $0.10 per 1M tokens for gpt-4.1-mini. The top-level calculator fields here use the standard input and output rates, not the cached-input, Batch, Flex, or Priority rates.

Benchmarks and specs

OpenAI’s model page lists GPT-4.1 mini with a 1,047,576-token context window, a 32,768-token maximum output, and a June 01, 2024 knowledge cutoff. It lists text and image as inputs and text as output.

No benchmark score is shown until an exact benchmark source and model identity are matched to GPT-4.1 mini.

Best fit

GPT-4.1 mini is best for API tasks that need GPT-4.1-style instruction following, tool calling, image input, and a very large context window at lower cost and higher speed than GPT-4.1.

For more complex tasks where OpenAI recommends a newer mini route, compare GPT-4.1 mini against the current GPT-5 mini family rather than treating it as the default current mini model.

Sources