Google Gemini 3.6 Flash launch cuts token use and undercuts 3.5 Flash
Google’s Gemini 3.6 Flash is the main update here, and the company says it uses 17% fewer output tokens than Gemini 3.5 Flash. That is the part that matters most for teams running AI agent workflows at scale, where every tool call and response can affect API spend and throughput planning.
Gemini 3.5 Flash-Lite hits 350 tokens per second for bulk jobs
Google is also pairing the launch with Gemini 3.5 Flash-Lite, a model it describes as built for high-throughput use rather than heavier reasoning. At 350 tokens per second, it is aimed at low-latency bulk usage in the Gemini API pricing update cycle, where speed can matter more than model depth.
For teams handling large volumes of routine tasks, that kind of speed is usually the bigger draw. It points to a model family that is being split more clearly by workload type, with Flash-Lite handling the faster, lighter end of the range. In practice, that can help developers choose a model based on throughput needs instead of using one general option for everything.
- Gemini 3.5 Flash-Lite is positioned for high-throughput work.
- Google says it can reach 350 tokens per second.
- The focus is low latency bulk usage rather than deeper reasoning.
- The use case appears tied to efficiency in API-heavy workflows.
CodeMender keeps Gemini 3.5 Flash Cyber on security duty
Gemini 3.5 Flash Cyber is staying inside CodeMender security agent workflows instead of getting a broad public push. Google is limiting access to governments and trusted partners, which suggests the model is being treated as a specialist tool for vulnerability hunting rather than a general-purpose chat model.
That distinction matters because it shows how Google is separating consumer-facing models from more controlled security use cases. Rather than opening the model broadly, the company appears to be keeping it focused on a narrower operational role. For organizations working in security, that usually means tighter access, clearer deployment boundaries, and more specific expectations around what the model is meant to do.
The update also fits with the broader pattern in Google’s Gemini 3.5 line. One model is being pushed for lower token use, another for speed, and another for controlled security work. Taken together, that gives the lineup a more practical shape, with each model aimed at a different kind of workload.
- Gemini 3.5 Flash Cyber remains tied to CodeMender workflows.
- Access is limited to governments and trusted partners.
- The model is being treated as a specialist security tool.
- Google is not positioning it as a broad public chat model.
For developers and enterprise teams, the key takeaway is simple. Google is making the Gemini 3.5 family more task-specific, and the biggest efficiency gain so far is the 17% reduction in output tokens on Gemini 3.6 Flash. That could matter most where scale, latency, and cost control all sit in the same decision.
Mahi Gupta
author
✉ mahigupta708076@gmail.comHi, I'm Mahi Gupta the Tech Writer at JhatpatLo. I write about smartphones, Android, Apple, AI, gadgets, software updates, and consumer technology. My goal is to make technology easy to understand by publishing accurate, well-researched, and reader-friendly content.Through JhatpatLo, I help readers stay updated with the latest tech news, buying guides, comparisons, and practical tips.
Related Products
Trending News
Xiaomi 18 Pro 3C certification reveals 100W charging
Xiaomi 18 Pro and Xiaomi 18 Pro Max have surfaced in China’s 3C certification reveals database, with M154FF and M311AD both linked to a 100W charger listing. That is the clearest early sign yet that Xiaomi’s next flagship pair is moving toward launch, and it also lines up with earlier leak...
Vivo Y-Series Price Hike Raises Y51 Pro Y31 5G and Y21 in India
Vivo has raised prices for five Y-series smartphones in India, with increases going up to Rs 4,000 on its official online store. The biggest jump is on the Y51 Pro 8GB+256GB model, while the Y31 5G and Y21 have also moved higher across multiple variants. Vivo Y51 Pro price update lea...
Poco F9 Ultra battery leak points to 8,050mAh and 100W charging
Poco F9 Ultra battery leak details surfaced in HyperOS code, pointing to an 8,050mAh cell for the global model and a 100W charging setup. The Poco F9 Pro is also tipped to use a 6,330mAh battery with the same 100W limit, while both phones are still expected to keep the Snapdragon 8 Elite G...
Realme 16x 5G price leak points to ₹25,999 India launch
Realme 16x 5G price leak details suggest the base 4GB+128GB model could land at ₹25,999 when it launches in India on August 12. Leaked live images also point to a 7,000mAh battery, a 144Hz display, and Dimensity 6300, making this one of the more battery-focused 5G launches in the segment....
Moto Pad 70 launched in India with 12.1 inch 2.5K display
Moto Pad 70 has launched in India with a 12.1-inch 2.5K display, 5G support, and a Moto Pen in the box. Motorola is aiming this tablet at buyers who want more than basic entertainment, and the early positioning makes it clear that the device is being pushed as a productivity-focused option...
ColorOS 17 Closed Beta Testing Opens for OnePlus 15 and Find X9 Pro
Oppo has opened ColorOS 17 Closed Beta Testing for a limited set of flagship phones, with 300 seats per device and applications running only from August 6 to August 7. The first wave includes the OnePlus 15, OnePlus 15R, Find X9 Pro, Reno 15 Pro, and Realme GT 8 Pro, giving early testers a...