Google: Gemma 3n 4B

Coming soonQuick Start

Gemma 3n 4B is Google's compact on-device model, released 2025-05, designed for efficient execution on phones, laptops, and tablets. It uses the MatFormer architecture with selective parameter activation to manage memory and computation, and supports over 140 languages.

This model is not routed on MirAPI yet — prices shown are the official list prices. Create an account to get notified when it goes live.

Modalities
inTextoutText
Price / 1M
$0.06 / $0.12
Context
33K
Knowledge cutoff
2024-08-31

Capabilities

Structured output

Benchmarks

bar = score (0–100) · ▏ median

Scores from vendor reports and public leaderboards. No private evals.

Coding Index
Code
3.2
#93

Showcase

Three fixed prompts, identical for every chat model. Outputs are archived verbatim — copy the prompt to reproduce.

Task prompt · identical for all models
Build a single-file HTML page using three.js from a CDN: a low-poly planet with a tilted ring of ~200 orbiting particles, a soft key light plus rim light, slow auto-rotation, and drag-to-orbit controls. No build tools — one file only.
Archived output
Archive not published yet

Every model will run this exact prompt; the output lands here verbatim, with tokens and per-run cost.

Quick Start

import os
from openai import OpenAI

client = OpenAI(
    base_url="https://api.mirapi.ai/v1",   # changed
    api_key=os.environ["MIRAPI_API_KEY"],   # changed
)

resp = client.chat.completions.create(
    model="google/gemma-3n-e4b-it",
    messages=[{"role": "user", "content": "Hello"}],
)
Wire Gemma 3n 4B into your product

1M free tokens to start · no card · OpenAI-compatible, one base_url change

Start free