Why it matters
Google introduced Gemini 3.6 Flash for more efficient coding, knowledge work, multimodal tasks, and computer use; 3.5 Flash-Lite for high-throughput, low-latency agent workflows; and 3.5 Flash Cyber for vulnerability research inside CodeMender. Google reports lower token use for 3.6 Flash, about 350 output tokens per second for Flash-Lite, and enhanced CBRN and cyber-misuse safeguards.
My takeaway: The three variants optimize different constraints, so selection should be workload-specific. Reproduce quality, latency, token, tool-call, refusal, and safety results on your own tasks; define rollback criteria; and apply tighter access controls to computer-use and cyber-capable deployments than headline benchmark scores imply.