OpenAI: GPT-4.1 Nano

GPT-4.1 Nano is OpenAI's smallest and fastest model in the GPT-4.1 series, released 2025-04. It is built for low-latency tasks such as classification and autocompletion, and scores 80.1% on MMLU, 50.3% on GPQA, and 9.8% on Aider polyglot coding.

Modalities
inImageinTextinFileoutText
Price / 1M−80%
$0.02 / $0.08$0.10 / $0.40
Context
1.048M
Knowledge cutoff
2024-06-30

Capabilities

VisionTool useLong contextPrompt cachingStructured output

Channels

Same weights on every channel. Requests route to the cheapest healthy one; the channel used is printed on each billing line.

Provider
Discount
Input /M
Output /M
Cache Read /M
Cache Create 5m /M
Cache Create 1h /M
Codexcheap
−80%
$0.02
$0.08
$0.006
Defaultstable
−52%
$0.06
$0.18
$0.01

Cost Calculator

Pick a channel and a usage scenario, then drag to your workload — each scenario assumes a different prompt-cache hit rate.

Channel
Scenario
Unit $0.02 / $0.08 /1MDaily chat: 40% cache hit assumed · cache read $0.006 /1M
List price $25.92On MirAPI$5.36/ moYou save $20.56Create an API key

Benchmarks

bar = score (0–100) · ▏ median

Scores from vendor reports and public leaderboards. No private evals.

Intelligence Index
General ability
9.6
#79
Coding Index
Code
11.1
#86
Agentic Index
Agentic
1.2
#81
GPQA Diamond
Science QA
51.9
#89
τ-bench (airline)
Agentic tool use
10.7
#99

Showcase

Three fixed prompts, identical for every chat model. Outputs are archived verbatim — copy the prompt to reproduce.

Task prompt · identical for all models
Build a single-file HTML page using three.js from a CDN: a low-poly planet with a tilted ring of ~200 orbiting particles, a soft key light plus rim light, slow auto-rotation, and drag-to-orbit controls. No build tools — one file only.
Archived output
Archive not published yet

Every model will run this exact prompt; the output lands here verbatim, with tokens and per-run cost.

Quick Start

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.mirapi.ai/v1",   # changed
    api_key=os.environ["MIRAPI_API_KEY"],   # changed
)

resp = client.chat.completions.create(
    model="openai/gpt-4.1-nano",
    messages=[{"role": "user", "content": "Hello"}],
)
Wire GPT-4.1 Nano into your product

1M free tokens to start · no card · OpenAI-compatible, one base_url change

Start free