Reported Gemini 3.6 Flash release claims lower inference costs and a modest output‑token price cut
Coverage says Gemini 3.6 Flash targets lower inference cost for agentic and production workloads and that Google introduced a modest output-token price reduction for Flash models.
In this brief: 3 sections 1 min read
Marketed as faster and more token‑efficient than Gemini 3.5 Flash.
Claimed improvements in coding, multimodal reasoning and built‑in capabilities (code exec, file search, structured outputs).
Said to be available across Gemini app, AI Studio, API, Gemini Enterprise and Antigravity.
Reports output price fell from $9 to $7.50 per million output tokens.
Input pricing reportedly unchanged at $1.50 per million tokens.
Publisher frames cut as modest but meaningful for high‑volume enterprise workloads.
Article references Flash‑Lite and cybersecurity‑focused models (e.g., Gemini 3.5 Flash Cyber) in limited pilots.
Cyber model described as available to governments/trusted partners under stricter controls.