Skip to content

Google

Gemini 3.5 Flash-Lite

API model IDs
gemini-3.5-flash-lite
Context window
1,048,576 tokens
Max output
65,536 tokens
Released
2026-07-21
Lifecycle
active
License
proprietary

Lifecycle verified 2026-09-08 (0 days ago). Source

Data handling

Trains on API data by default
No
Retention
Not published
Zero data retention
on-request
Residency
global
HIPAA BAA
Not available
SOC 2 / FedRAMP
SOC 2 not published, FedRAMP Not published

The Gemini API Additional Terms of Service state that for Paid Services, Google does not use prompts or responses to improve its products, and logs them for a limited period solely to detect Prohibited Use Policy violations without stating a specific number of days. Zero data retention is available on request per project. Unpaid use, including Google AI Studio and the free quota, is used to develop Google products and may be reviewed by human annotators. HIPAA BAA, SOC 2, and FedRAMP coverage are published for Vertex AI and Gemini Enterprise, not for this first-party Developer API surface, so they are not claimed here.

Verified 2026-09-08 (0 days ago). Source 1, Source 2

Pricing snapshot

$0.3 input, $2.5 output, per million tokens. Cached input $0.03 per million tokens. Batch requests are 50% off.

This is a snapshot verified on 2026-09-08, 0 days ago, not a live price. Confirm on the provider’s pricing page before you budget against it.

Where you can call it

  • direct · gemini-3.5-flash-lite First-party Gemini Developer API (ai.google.dev / Google AI Studio).