Skip to main content
The next-gen Flash workhorse, with a big jump in coding and agents: DeepSWE v1.1 at 65.3%, three thinking tiers, 1M context.

Specifications

Pricing

Prices in USD per 1M tokens ($/1M).
The table shows list prices. Recharge promotions and group discounts stack; the console reflects the actual charge in real time. See Pricing and Recharge promotions.

Endpoints

Billing groups

Some groups carry extra discounts, stackable with recharge bonuses. See Tokens and groups.

Supported features

Example request

The example below calls gemini-3.7-flash through OpenAI Chat Completions (/v1/chat/completions). Point base_url at https://api.apiyi.com/v1 — everything else matches the official API.
Read the key from an environment variable, never hard-coded. In production, issue separate tokens per use case so you can revoke and attribute usage individually.

Launch announcement

Background, benchmarks and migration notes for Gemini 3.7 Flash

Native Calls

Parameters, usage and best practices

Gemini 3.6 Flash

Details for a model in the same family

Gemini 3.5 Flash

Details for a model in the same family

Model pricing directory

Live pricing, endpoints and groups for all 296 models
Specs on this page are maintained by hand in models/data/model-details.json; pricing and endpoints come from the live pricing API, updated 2026-08-17 13:47 (UTC+8), specs last verified 2026-08-17.