As AI based workflows evolve, success relies heavily on three factors: lower latency, higher token efficiency, and predictable ROI. Google’s latest updates to the Gemini ecosystem directly address these demands, delivering purpose-built performance for enterprise scaling.

Here is a summary at the newly introduced models according to Artificial Analysis Index:

  • Gemini 3.6 Flash: by streamlining tool execution, it cuts output token consumption by up to 17% compared to its predecessor – reducing operational overhead while maintaining accuracy at $1.50/1M input and $7.50/1M output tokens.
  • Gemini 3.5 Flash-Lite: designed for speed and tasks where high throughput is critical for developers workflows – running at an impressive 350 output tokens per second at just $0.30/1M input tokens.
  • Gemini 3.5 Flash Cyber: engineered specifically for enterprise security, it empowers frontline defenders to automatically detect, validate, and patch codebase vulnerabilities.

Beyond model-level performance, Google is streamlining the overall experience: LLM Notebook is now officially Gemini Notebook. While retaining its power as a flexible research tool, it now features deeper integration across the Google ecosystem – including seamless connection with the Gemini app – for smoother workflow orchestration.

At Perceptiva, we help enterprises connect, orchestrate, and innovate with AI platforms that drive real business impact. Explore our customer success stories to see how we transform complex AI capabilities into measurable growth.