🎉 1 billion Gemma downloads! Thank you to the global developer community for reaching this incredible milestone with us. Google DeepMind's Paige Bailey catches up with developers and partners like Unsloth AI and Qualcomm to hear how they build local coding agents, optimize chipsets, and deploy platforms on device:
Google for Developers
Technology, Information and Internet
Mountain View, CA 4,281,459 followers
Join a community of creative developers and learn how to use the latest in technology—from AI and cloud, to mobile & web
About us
Discover the latest technologies, resources, events, and announcements to help you build smarter and ship faster. Explore more at developers.google.com
- Website
-
http://developers.google.com
External link for Google for Developers
- Industry
- Technology, Information and Internet
- Company size
- 10,001+ employees
- Headquarters
- Mountain View, CA
- Specialties
- coding, engineering, firebase, android, cloud, web development, and mobile development
Updates
-
Build richer music experiences into your apps with Lyria 3.5, our cleanest, most controllable music generation model yet 🎵✨ With the Gemini API, developers can ship production-ready music featuring significantly reduced noise, distortion, and artifacts. Now devs have more granular generation control with authentic regional accents across multiple languages, strict lyric adherence, and native handling for bracketed cues like [guitar solo]. Try Lyria 3.5 in Google AI Studio or explore the developer documentation → https://goo.gle/4gMk7OC
-
Introducing Gemini 3.8 Flash ⚡️ Built to be your sharpest coding partner, 3.8 Flash brings major upgrades over 3.7 Flash across SWE and agentic tasks. You get dynamic thinking controls to balance reasoning depth with your budget— all at the same introductory price of $0.75/1M input and $3.75/1M output tokens through the end of the year. To see it in action, watch 3.8 Flash build this app from scratch in Antigravity. We gave it a single goal to create a dynamic UI with the Google Maps API and the model tackled the complex API integrations and wrote all of the code completely autonomously. Try it today via Google Antigravity and the Gemini API via Google AI Studio and Android Studio. More details in the blog: https://goo.gle/46B6Fs3
-
📽️⚡ Process long-form video more with significantly fewer tokens without sacrificing quality. Agentic video understanding unlocks new ways to process long-form video to deliver massive token reductions (up to 88%), better accuracy, and lower costs. With this feature, Gemini actively scans visual frames, audio, and transcripts instead of passively ingesting video at 1 FPS. Now you can tackle complex video workflows like: ▸ Sub-second retrieval: Catch split-second cuts missed at 1 FPS. ▸ Long-form search: Query multi-hour videos without overflowing context. ▸ Anomaly detection: Resample interesting windows at higher FPS. ▸ Counting: Accurately track repeated actions and objects over time. And more. Learn here: https://goo.gle/4gFY2kK Agentic video understanding is supported by Gemini 3.7 Flash, 3.6 Flash and 3.5 Flash-Lite and is available today for video uploads and YouTube videos via the Gemini API on Google AI Studio and Gemini Enterprise Agent Platform. See how agentic video understanding improves token efficiency in Gemini 3.7 Flash when compared with static video processing:
-
Last week we introduced Gemini 3.5 Transcribe, our latest text-to-speech model. But, what does this actually mean for your projects? We built this app in Google AI Studio to demonstrate just how much smarter 3.5 Transcribe is. When streaming live audio simultaneously through two parallel transcription pipeline modes, you can see how smart transcription automatically strips out disfluencies (ums, ahs) and condenses long, rambling thoughts while verbatim keeps exact match transcription live:
-
🗣️✨ Speak your vibes into existence with Gemini 3.5 Transcribe The latest audio model is built for apps that need live transcription and quick visual prototyping. Watch Geneviève Huskens turn spoken ideas into a dynamic mood board using Gemini 3.5 Transcribe on the Live API. The model maps speech to user intent to create prompts that trigger live image generation and editing with Nano Banana 2 Lite: Try the voice-powered app yourself: https://goo.gle/4wYBsdo
-
🎥 Gemini Omni 1.1 Flash brings a new suite of creative controls and generative video capabilities to developers. Here’s what’s new in Omni 1.1 Flash: ☑️ Scene Extension & Interpolation: Build longer videos by extending videos in 10-second increments, up to 40 seconds. Omni 1.1 can take up to 10 seconds of past context, an improvement over previous models which only took 1 second of context. You can also lock in your first and last frames to achieve complex camera moves and smooth transitions. ☑️ Video References: Drop in 3-second videos to keep your characters and visual style consistent across generations. ☑️ Fast, Low-Cost Prototyping: Draft ideas faster and at a third of the cost of standard 720p by generating lightweight 360p previews first. ☑️ Studio-Quality Upscaling: Ready to ship? Turn your favorite drafts into high-resolution 1080p or crisp 4K video ready for professional production. Gemini Omni 1.1 Flash is available now via the Gemini API in Google AI Studio for production-ready control, and on the Gemini Enterprise Agent Platform for full enterprise security and compliance. Read the blog for more details: https://goo.gle/4hXVOQa See how Omni 1.1 Flash turns static photos into cinematic 4K home tours in this property visualization app:
-
Introducing Gemini 3.5 Transcribe, our latest speech-to-text model built for intelligent audio understanding and transcription. 🎙️ Whether you’re building real-time, voice-first interfaces, or transcribing multi-speaker recordings, Gemini 3.5 Transcribe adapts to your needs by: > Transcribing complex speech, emails, and phone numbers with lower WER (Word Error Rates) > Adhering to provided custom vocabulary and specialized jargon > Auto-detecting 85+ languages, regional accents, and live language switches Available in public preview for developers via the Gemini API in Google AI Studio and Google Antigravity, and for businesses via the Gemini Enterprise Agent Platform. Explore more in the blog: https://goo.gle/4qMie9m 🛠️ Learn how you can build with 3.5 Transcribe as Ammaar Reshi breaks down the latest capabilities:
-
Syncing to and from GitHub is now supported in Google AI Studio Build → https://goo.gle/4qxA9QO Start from an existing repo, pull changes into Build, push updates back to GitHub, and continue developing across environments while keeping your codebase in sync.