source: google deepmind: introducing gemini 3.6 flash, 3.5 flash-lite, and 3.5 flash cyber

level: technical

google deepmind introduced gemini 3.6 flash, an update to its flash series that improves coding, knowledge work, and multimodal tasks while using 17% fewer output tokens than 3.5 flash. it also lowers the cost per output token, priced at $1.50 per million input tokens and $7.50 per million output tokens. the model shows gains on benchmarks like deepswe (49% vs. 37%) and mle bench (63.9% vs. 49.7%), and includes built-in computer use tools via the api.

gemini 3.5 flash-lite is a faster, cheaper model designed for high-throughput agentic tasks like search and document processing. it reaches 350 output tokens per second and costs $0.30 per million input tokens and $2.50 per million output tokens. it outperforms 3.1 flash-lite on coding and agentic benchmarks, including swe-bench pro (54.2% vs. 49.6%) and osworld-verified (74.0% vs. 65.1%), and supports configurable thinking levels for latency or quality trade-offs.

google also announced gemini 3.5 flash cyber, a specialized model fine-tuned for finding and fixing security vulnerabilities, available only to governments and trusted partners through the codemender agent. additionally, gemini 3.5 pro is in testing with partners, and the team has started pre-training for gemini 4. both 3.6 flash and 3.5 flash-lite are available now via the gemini api, google ai studio, and enterprise platforms.

why it matters: these models offer lower cost and higher speed for building ai agents, making it easier to scale automated coding, data analysis, and security tasks.


source: google deepmind: introducing gemini 3.6 flash, 3.5 flash-lite, and 3.5 flash cyber