Key Takeaways
- Two models at once: Google’s July 2026
gemini-3.6-flash(multimodal flagship Flash) andgemini-3.5-flash-lite(high-frequency lightweight) are live on APIYI in thedefault/svipgroups - Fully tested on both endpoints: 30+ cases per model across the native Gemini format and the OpenAI-compatible format — Search grounding, Maps grounding, URL context, code execution, and Computer Use (3.6) all verified working, not just “listed”
- Official pricing: 3.6 Flash at $1.50 in / $7.50 out, 3.5 Flash-Lite at $0.30 in / $2.50 out (per 1M tokens, output includes thinking); discounts come from top-up bonuses (up to 20%, ≈17% off)
- Opposite thinking defaults: 3.6 thinks by default (four controllable tiers, 0–837 tokens measured); Lite ships with zero thinking and ~2s responses — this one line decides which to pick
- Same 1M context: both offer 1,048,576 input / 65,536 output with text/image/video/audio/PDF input
Background
Gemini 3.6 Flash and Gemini 3.5 Flash-Lite are Google’s two stable text models updated in July 2026: the former takes over the Flash mainline and brings native tools including Computer Use (Preview) to the Flash tier; the latter continues the Flash-Lite “fast and cheap” line, with audio input priced the same as text. Before launch, APIYI completed a full dual-endpoint test pass (26 main cases + 5 follow-ups per model) covering every testable item in the official Capabilities table. All numbers in this article come from our July 22, 2026 test records.Deep Dive
Verified capability matrix
Advanced tools (Search/Maps/URL/code execution/Computer Use) are native-format exclusive; the OpenAI-compatible endpoint covers the standard set — chat, streaming, function calling, JSON Schema, vision. The native endpoint takes your APIYI token directly (
x-goog-api-key: sk-...), no Google API Key needed.
Which one to pick?
- 3.6 Flash: deep reasoning, complex planning, heavy native-tool use, Computer Use
- 3.5 Flash-Lite: throughput-first workloads — high-frequency Q&A, classification, extraction, translation, support bots (about twice as fast, down to a fifth of the cost)
In Practice
Pricing & Availability
Identical to Google’s official pricing; top-up bonuses go up to 20% (+10% on $100), roughly 17% off overall. Google Search grounding bills at $14 / 1K queries.