Gemini 3.
Google's Gemini 3.5 Live API lacks speaker diarization, forcing developers to choose between real-time transcription and speaker identification, according to a developer's testing. The gemini-3.5-tran…
Google's Gemini 3.5 Live API lacks speaker diarization, forcing developers to choose between real-time transcription and speaker identification, according to a developer's testing. The gemini-3.5-tran…
A developer detailed the addition of real-time transcription and speaker diarization to a macOS meeting translation app using Google's Gemini 3.5 models. The project, gemini-live-translate-macos, leve…
Google DeepMind released Gemini Robotics 2.0, a trio of AI sub-models that improve robot dexterity and safety, with the embodied reasoning model Gemini Robotics ER 2 now publicly available for develop…
Google DeepMind released Gemini Robotics ER 2, an AI robotics model that enables multi-robot collaboration and real-time planning through temporal intelligence, achieving 91.3% accuracy for moment fin…
Google's Gemini API now supports streaming text-to-speech, enabling developers to reduce latency in AI voice applications by generating audio from text chunks before the full response is complete. Thi…
A developer created an animation dashboard for the Open Duck Mini robot that integrates with the Gemini Live API, enabling real-time human-robot interaction. The project, shared on Hacks, provides cod…
A developer built a custom animation tool for the Open Duck Mini robot, pairing it with the Gemini Live API to enable human-robot interaction through gestures like nodding and shaking its head. The pr…
Google released Gemini 3.5 Live Translate, a streaming speech-to-speech audio model that translates over 70 languages in real time. The model processes audio continuously rather than waiting for a spe…
Google released Gemini 3.5 Live Translate, a new audio model for live speech-to-speech translation that automatically detects over 70 languages and generates continuous, natural-sounding translated sp…
Google Cloud Storage has introduced Model Context Protocol (MCP) servers that enable AI agents to directly access and process unstructured data stored in GCS, turning passive objects into actionable c…
The Gemini Live Agent Challenge, organized by Google, attracted 11,878 participants from 151 countries who submitted 1,536 projects using the Gemini Live API and Google Cloud tools. The competition fe…