minimal · low · medium · high · default minimal
Only provider-documented reasoning controls are shown; missing controls remain unpublished rather than inferred.
MODEL REFERENCE / GOOGLE
Google stable low-latency Flash-Lite model optimized for high-volume agentic tasks, document parsing, extraction, and cost-sensitive multimodal workloads.
PRIMARY-SOURCE VERIFIED
Google documents a 1,048,576-token context window. Input modalities: text + image + video + audio + PDF. Output: text. Pricing status: Official paid API.
CAPABILITY CONTRACT
Only provider-documented reasoning controls are shown; missing controls remain unpublished rather than inferred.
Input modalities come from the cited official model documentation.
Output modality and output-limit claims remain separate so an undocumented token cap is not guessed.
PRICING STATUS
$0.30 input · $0.03 cached · $2.50 output per 1 million tokens.
SXF preserves the provider-published native billing basis. Only calculator-eligible generative token models enter token-cost arithmetic; specialist units and input-only embeddings remain visible without forced conversion.
CAVEATS
Google lists gemini-3.5-flash-lite as stable with no shutdown date announced.
Standard paid pricing is $0.30 input, $0.03 cached input, and $2.50 output per 1M tokens.
Knowledge cutoff remains null because the cited current model documentation does not publish one.
EXPLORE
OFFICIAL SOURCES
WHAT CHANGED
No post-baseline factual changes recorded for Gemini 3.5 Flash-Lite since 2026-10-06.
VERIFIED CHANGE HISTORY
Baseline verified: 1,048,576 context · 65,536 max output · $0.30 input / $0.03 cached / $2.50 output per 1 million tokens
Official evidence ↗Append-only SXF ledger · 1 verified event · chain head a81376cb9fe2…