Gemini 2.5 Flash

Google has announced the release of updated preview versions of Gemini 2.5 Flash and Gemini 2.5 Flash-Lite, now available on Google AI Studio and Vertex AI. These updates focus on delivering higher-quality responses while reducing latency and cost—making them especially attractive for developers building high-throughput and agentic applications.

Key Highlights

Performance and Efficiency

Gemini 2.5 Flash-Lite Updates

Gemini 2.5 Flash Updates

Early testers are already seeing strong results. Yichao "Peak" Ji, Co-Founder & Chief Scientist at Manus, highlighted:

“The new Gemini 2.5 Flash model offers a remarkable blend of speed and intelligence. Our evaluation on internal benchmarks revealed a 15% leap in performance for long-horizon agentic tasks. Its outstanding cost-efficiency enables Manus to scale to unprecedented levels—advancing our mission to Extend Human Reach.”

Easier Access with -latest Aliases

To simplify adoption, Google has introduced -latest aliases for each model family:

These aliases always point to the newest preview versions, letting developers experiment with new features without constantly updating model strings. A 2-week notice will be provided before updates or deprecations.

For production stability, developers should continue using: