Google Expands the Gemini Ecosystem with New Cost-Efficient Models and Rebranding
TL;DR
- Google DeepMind introduced three new proprietary models prioritizing token efficiency and performance.
- The new Gemini 3.6 Flash model slashes token costs for long-horizon engineering tasks by up to 65%.
- Popular research assistant NotebookLM has been rebranded to Gemini Notebook alongside new native coding features.
- Updated usage quotas and rate changes mean developers and power users will need to closely monitor their daily consumption.
Google is doubling down on enterprise efficiency and developer accessibility with a sweeping update to its artificial intelligence ecosystem. Leading the charge is the debut of three new proprietary models from DeepMind: Gemini 3.6 Flash, Gemini 3.5 Flash-Lite, and Gemini 3.5 Flash Cyber. Designed explicitly to make autonomous AI agents faster and smarter at scale, these releases focus heavily on optimizing token usage without sacrificing capability. Notably, the Gemini 3.6 Flash model slashes token costs by up to 65 percent on complex, long-horizon software engineering tasks, while maintaining a competitive pricing structure for API integration.
Alongside the hardware-adjacent infrastructure updates, Google is streamlining its software branding. The popular research and note-taking assistant formerly known as NotebookLM has officially been renamed Gemini Notebook, bringing it deeper into the core AI family while expanding access to native code-writing features. Meanwhile, users navigating the new ecosystem will need to adjust to updated usage quotas and tracking metrics, as Google has revised how rate limits are tallied across its tiers, potentially changing response volumes for heavy users.
Even as these mid-cycle releases roll out across the developer community, anticipation is already building for the future. Google has begun teasing its next-generation Gemini 4 architecture, signaling that the rapid pace of model iteration shows no signs of slowing down. For organizations building scalable AI solutions, the latest batch of updates offers an immediate path to lower operational costs while setting the stage for even more powerful capabilities on the horizon.
Sources
Google Releases 3 New Gemini Models, 3.5 Pro Still Not Available (cnet.com) – Cnet reports that Google launched three new Gemini models focusing on performance and token efficiency while keeping Gemini 3.5 Pro under wraps.
Google Renames NotebookLM to Gemini Notebook (cnet.com) – Cnet also details how NotebookLM was rebranded to Gemini Notebook alongside the rollout of native code-writing capabilities.- Google Just Teased Its Huge Gemini 4 Release (droid-life.com) – Droid Life notes that Google has already started generating buzz by teasing the upcoming release of Gemini 4.
- Google’s Gemini 3.6 Flash model cuts AI agent token costs by up to 65% on long horizon engineering tasks —and 3.5 Pro is on the way (venturebeat.com) – VentureBeat breaks down the technical specs and pricing of the three new DeepMind models, highlighting how Gemini 3.6 Flash drastically reduces engineering token costs.
How Google’s New Gemini Rates Work and How to Track Your Usage (wired.com) – Wired explains how Google’s restructuring of usage quotas and tracking mechanisms will impact daily AI response limits for users.

Powered by News Ranker