OpenAI: GPT-3.5 Turbo 16k

GPT-3.5 Turbo 16k is an OpenAI chat model released 2023-08, extending GPT-3.5 Turbo with a 16K token context window for longer documents. It supports tool use and structured output, and its training data runs through September 2021.

Modalities
inTextoutText
Price / 1M−61%
$0.90 / $1.80$3.00 / $4.00
Context
16K
Knowledge cutoff
2021-09-30

Capabilities

Tool useStructured output

Channels

Same weights on every channel. Requests route to the cheapest healthy one; the channel used is printed on each billing line.

Provider
Discount
Input /M
Output /M
Cache Read /M
Cache Create 5m /M
Cache Create 1h /M
Codexcheap
−61%
$0.90
$1.80
Defaultstable
−10%
$2.10
$4.20

Cost Calculator

Pick a channel and a usage scenario, then drag to your workload — each scenario assumes a different prompt-cache hit rate.

Channel
Scenario
Unit $0.90 / $1.80 /1MDaily chat: 40% cache hit assumed · cache read $0.09 /1M
List price $376.32On MirAPI$141.70/ moYou save $234.62Create an API key

Showcase

Three fixed prompts, identical for every chat model. Outputs are archived verbatim — copy the prompt to reproduce.

Task prompt · identical for all models
Build a single-file HTML page using three.js from a CDN: a low-poly planet with a tilted ring of ~200 orbiting particles, a soft key light plus rim light, slow auto-rotation, and drag-to-orbit controls. No build tools — one file only.
Archived output
Archive not published yet

Every model will run this exact prompt; the output lands here verbatim, with tokens and per-run cost.

Quick Start

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.mirapi.ai/v1",   # changed
    api_key=os.environ["MIRAPI_API_KEY"],   # changed
)

resp = client.chat.completions.create(
    model="openai/gpt-3.5-turbo-16k",
    messages=[{"role": "user", "content": "Hello"}],
)
Wire GPT-3.5 Turbo 16k into your product

1M free tokens to start · no card · OpenAI-compatible, one base_url change

Start free