Google’s Gemini 3.6 Flash targets enterprise agent token costs
artificialintelligence-news.com Jul 21, 2026

Google’s Gemini 3.6 Flash targets enterprise agent token costs

AI-summarised brief · reviewed before publication

Google unveiled Gemini 3.6 Flash, Gemini 3.5 Flash‑Lite, and a restricted Gemini 3.5 Flash Cyber aimed at reducing latency and token costs for enterprise AI agents. Gemini 3.6 Flash cuts output tokens by 17 % versus 3.5 Flash and shows up to 65 % reduction in synthetic benchmarks, pricing at $1.50 per million input and $7.50 per million output tokens. It improves success rates on DeepSWE (49 % vs 37 %) and MLE Bench (63.9 % vs 49.7 %). Figma, Hebbia and Harvey have integrated the model for faster design iteration and multimodal document processing. Gemini 3.5 Flash‑Lite delivers 350 output tokens per second, costs $0.3/$2.5 per million tokens, and doubles its GDPval‑AA v2 score. Flash Cyber, limited to vetted partners, targets automated vulnerability remediation today.

💡 Why It Matters

  • · Enterprises can now run continuous AI agents at a fraction of previous cost, unlocking real‑time automation at scale.