Google: Gemini 3.1 Flash Lite Preview

Coming soonQuick Start

Gemini 3.1 Flash Lite Preview is Google's high-efficiency model, released 2026-03, designed for high-volume workloads. It outperforms Gemini 2.5 Flash Lite on overall quality and approaches Gemini 2.5 Flash across key capabilities, with improvements in audio input, RAG snippet ranking, translation, data extraction, and code completion.

This model is not routed on MirAPI yet — prices shown are the official list prices. Create an account to get notified when it goes live.

Modalities
inTextinImageinVideoinFileinAudiooutText
Price / 1M
$0.25 / $1.50
Context
1.049M
Knowledge cutoff

Capabilities

ReasoningVisionTool useLong contextPrompt cachingStructured output

Benchmarks

bar = score (0–100) · ▏ median

Scores from vendor reports and public leaderboards. No private evals.

Intelligence Index
General ability
25.6
#59
Coding Index
Code
34.7
#61
Agentic Index
Agentic
6.5
#63

Showcase

Three fixed prompts, identical for every chat model. Outputs are archived verbatim — copy the prompt to reproduce.

Task prompt · identical for all models
Build a single-file HTML page using three.js from a CDN: a low-poly planet with a tilted ring of ~200 orbiting particles, a soft key light plus rim light, slow auto-rotation, and drag-to-orbit controls. No build tools — one file only.
Archived output
Archive not published yet

Every model will run this exact prompt; the output lands here verbatim, with tokens and per-run cost.

Quick Start

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.mirapi.ai/v1",   # changed
    api_key=os.environ["MIRAPI_API_KEY"],   # changed
)

resp = client.chat.completions.create(
    model="google/gemini-3.1-flash-lite-preview",
    messages=[{"role": "user", "content": "Hello"}],
)
Wire Gemini 3.1 Flash Lite Preview into your product

1M free tokens to start · no card · OpenAI-compatible, one base_url change

Start free