Gemini 4 Argon

blog.google
gemini-4-argon

Enabling coding and enterprise workflows across domains Gemini 4 Argon’s capabilities across coding, reasoning, and multimodality and its ability to sustain long, multi-step tasks enable it to excel across a range of enterprise workflows. Google engineers have been using Argon … Read more

Gemini 3.8 text-to-speech says hello

blog.google
gemini-3.8-text-to-speech-says-hello

Get expressive high-quality speech generation built for global scale Gemini 3.8 Flash TTS delivers leading voice customization capabilities, securing the #1 overall spot on Hume AI’s Voice Design Benchmark (71.4) and also leading in accent modeling (60.8). Gemini 3.8 Flash … Read more

Gemini 3.8 Live and 3.8 Live Extended Thinking

blog.google
gemini-38-live-and-3.8-live-extended-thinking

Today, we’re introducing two new models that bring advancements in near real-time reasoning to more effectively enable voice agents and make conversing with AI feel more intuitive and intelligent. Gemini 3.8 Live: Built for scale and cost efficiency, combining conversational … Read more

The Gemini app is now available for Windows

blog.google
the-gemini-app-is-now-available-for-windows

Today, we’re launching the Gemini app for Windows. Designed to work seamlessly alongside your favorite tools and daily applications, the new desktop app gives you instant assistance without breaking your flow. Here are three ways you can use the Gemini … Read more

Gemini 3.8 Flash

deepmind.google
gemini-3.8-flash

Benefit and Intended Usage Gemini 3.8 Flash is well-suited for users, developers, and enterprises, designed for cost-effective scaling of general-purpose, production-ready agents. Some use cases include: software engineering, agent tasks, and complex knowledge workflows. Known Limitations Gemini 3.8 Flash may … Read more

Gemini-3.5-Transcribe

blog.google
gemini-3.5-transcribe

Today, we’re introducing Gemini 3.5 Transcribe, our most precise speech-to-text model yet, designed for intelligent voice interactions. Unlike conventional speech recognition models that struggle with background noise, complex jargon, and disfluency cleanup, Gemini 3.5 Transcribe converts raw audio directly into … Read more

Gemini Omni 1.1 Flash

blog.google
gemini-omni-1.1-flash

Today, we’re introducing Gemini Omni 1.1 Flash, a new suite of creative controls and generative video capabilities to support developers. Gemini Omni brought real-world reasoning to generative creation, and today’s updates make Omni 1.1 production-ready for professional use via the … Read more

Gemini 3.7 Flash

ai.google.dev
gemini-3.7-flash

The Interactions API is now generally available. We recommend using this API for access to all the latest features and models. Send feedback Gemini 3.7 Flash is the next iteration in the Gemini 3 series of highly-capable, natively multimodal, reasoning … Read more