Models
29 posts
Gemini 4 Argon: our next era of frontier intelligence
Announcing Gemini 4 Argon, our frontier model for real-world coding, enterprise knowledge work, and cyber defense, rolling out soon.
See what 4 builders are making with Gemini 3.8 Flash
See what developers have built with Gemini 3.8 Flash, our most intelligent workhorse model, launched in early September.
Introducing Gemini 3.8 Live with Live Avatar
Introducing Gemini 3.8 Live with Live Avatar, which brings near real-time visual presence to Gemini’s conversational AI.
Gemini 3.8 text-to-speech says hello
Gemini 3.8 Flash-Lite TTS and Gemini 3.8 Flash TTS are our most expressive audio models yet.
Introducing Gemini 3.8 Live and 3.8 Live Extended Thinking
Gemini 3.8 Live and Gemini 3.8 Live Extended Thinking are our most advanced live dialogue models yet, built for natural conversation.
Introducing Gemini 3.8 Flash and 3.8 Flash Cyber
Gemini 3.8 Flash and 3.8 Flash Cyber deliver next-generation intelligence for agentic workflows and cybersecurity.
Introducing agentic video understanding with Gemini
We’re launching agentic video understanding across our latest Gemini models for improved accuracy and lower costs and token usage.
Intelligent transcription with Gemini 3.5 Transcribe
Now you can get more intelligent speech-to-text transcription with Gemini 3.5 Transcribe.
What does “full-stack” AI actually mean?
A Google DeepMind engineer breaks full-stack development into five simple layers and explains how it affects everyday users.
Omni experts share what excites them most about the model.
We interviewed some of the experts behind Gemini Omni to hear what excites them most about the model.
Introducing Gemini 3.7 Flash
Gemini 3.7 Flash is our most intelligent workhorse model yet for coding and agents.
See what 5 builders are making with Gemini Omni
Gemini Omni makes creating videos as easy as having a conversation. Here’s how five people use it to edit videos and visualize ideas.
Simplify your morning with this vibe-coded schedule app.
Tired of starting your morning staring at bright screens and stressful notifications? Raph, a creative technologist at Google, built a custom app called Glan…
How Gemini Flash agents are helping a Michigan dairy farmer
See how Paul Windemuller, a Michigan dairy farmer, is changing the way he works by using AI agents built with Gemini 3.6 Flash.
Introducing Gemini 3.6 Flash, 3.5 Flash-Lite, and 3.5 Flash Cyber
We’re introducing new Gemini models, including Gemini 3.6 Flash, 3.5 Flash-Lite and 3.5 Flash Cyber.
Start building with Nano Banana 2 Lite and Gemini Omni Flash
Scale your ideas with Nano Banana 2 Lite, our fastest, most cost-efficient Gemini Image model, and Gemini Omni Flash for high-quality video and conversationa…
Introducing computer use in Gemini 3.5 Flash
A look at the built-in computer use tool in Gemini 3.5 Flash.
Fluid, natural voice translation with Gemini 3.5 Live Translate
Gemini 3.5 Live Translate brings near real-time, natural speech translation to Google AI Studio, Google Translate and Google Meet.
9 demos of Gemini Omni and Gemini 3.5 in action
Watch 9 videos showing the capabilities of Gemini Omni and Gemini 3.5, announced at Google I/O 2026.
Gemini 3.5: frontier intelligence with action
At Google I/O we released Gemini 3.5, our latest series of models combining frontier intelligence with action.
Introducing Gemini Omni
Introducing Gemini Omni, which allows you to create anything from any input and edit naturally using conversational language.
Gemini Embedding 2 is now generally available.
We’re announcing the general availability of Gemini Embedding 2 via the Gemini API and Vertex AI.
Deep Research Max: a step change for autonomous research agents
Introducing Deep Research and Deep Research Max, the next generation of Google’s autonomous research agents.
Gemini 3.1 Flash TTS: the next generation of expressive AI speech
Gemini 3.1 Flash TTS is now available across Google products.
Gemini 3.1 Flash Live: Making audio AI more natural and reliable
Gemini 3.1 Flash Live is now available across Google products.
Gemini Embedding 2: Our first natively multimodal embedding model
An overview of Gemini Embedding 2, our first fully multimodal embedding model that maps text, images, video, audio and documents into a single space.
Gemini 3.1 Flash-Lite: Built for intelligence at scale
Gemini 3.1 Flash-Lite is our fastest and most cost-efficient Gemini 3 series model yet.
Gemini 3.1 Pro: A smarter model for your most complex tasks
3.1 Pro is designed for tasks where a simple answer isn’t enough.
Gemini 3 Deep Think: Advancing science, research and engineering
We’re releasing a major upgrade to Gemini 3 Deep Think, our specialized reasoning mode.