Google launches Lyria 3.5, an AI model for high-fidelity music generation with multimodal input, natural vocals, and detailed control, available in Gemini.
Google's Gemini Flash now uses agentic AI video processing, actively selecting frames for analysis. Cuts costs & tokens, boosts accuracy. Available via API.
Google's Gemini Live update adds AI agents for productivity but fails everyday users due to being restricted to Google Workspace, overlooking real-life chores.
Waymo's Gemini AI in robotaxis is a gimmick, not a breakthrough. Redundant with phones, it distracts from essential autonomous driving advancements.
Multimodal AI evolves beyond text/images, enabling on-device AI, automated code, robotics, and global sensing, facing enterprise adoption issues and high costs.
OpenHome & ElevenLabs Japan launched a program for on-device voice AI. It offers SDKs & compute credits to devs, reducing latency & boosting privacy.
Google Gemini AI is your new personal assistant, getting rid of jet lag, streamlining travel, calendar, and email management in Workspace for boosted productivity.
Google launches Gemini 3.5 Live Translate: real-time, continuous speech-to-speech translation for 70+ languages, reducing delays.