Let us inspire you in Audio AI...

Google Launches Lyria 3.5 for Multimodal Music Generation in Gemini

Google Launches Lyria 3.5 for Multimodal Music Generation in Gemini

Google launches Lyria 3.5, an AI model for high-fidelity music generation with multimodal input, natural vocals, and detailed control, available in Gemini.

AI.denSep 7
Google's New Agentic Video Processing

Google's New Agentic Video Processing

Google's Gemini Flash now uses agentic AI video processing, actively selecting frames for analysis. Cuts costs & tokens, boosts accuracy. Available via API.

Yorrick SchoonheydtSep 2
Why Google's New Gemini Update Misses the Mark for Everyday Life

Why Google's New Gemini Update Misses the Mark for Everyday Life

Google's Gemini Live update adds AI agents for productivity but fails everyday users due to being restricted to Google Workspace, overlooking real-life chores.

Yorrick SchoonheydtAug 28
Why Waymo’s New Gemini Integration Feels More Like a Gimmick Than a Breakthrough

Why Waymo’s New Gemini Integration Feels More Like a Gimmick Than a Breakthrough

Waymo's Gemini AI in robotaxis is a gimmick, not a breakthrough. Redundant with phones, it distracts from essential autonomous driving advancements.

Yorrick SchoonheydtAug 20
Scaling Multimodal AI

Scaling Multimodal AI

Multimodal AI evolves beyond text/images, enabling on-device AI, automated code, robotics, and global sensing, facing enterprise adoption issues and high costs.

Alban CapajAug 6
OpenHome and ElevenLabs Partner to Bring Local Voice AI to Japanese Hardware

OpenHome and ElevenLabs Partner to Bring Local Voice AI to Japanese Hardware

OpenHome & ElevenLabs Japan launched a program for on-device voice AI. It offers SDKs & compute credits to devs, reducing latency & boosting privacy.

Jorge De CorteAug 4
Why I Think AI is Finally Becoming the Personal Assistant We All Need

Why I Think AI is Finally Becoming the Personal Assistant We All Need

Google Gemini AI is your new personal assistant, getting rid of jet lag, streamlining travel, calendar, and email management in Workspace for boosted productivity.

Yorrick SchoonheydtJun 29
Gemini 3.5 Live Translate: Continuous Modeling for Low-Latency Speech-to-Speech Conversion

Gemini 3.5 Live Translate: Continuous Modeling for Low-Latency Speech-to-Speech Conversion

Google launches Gemini 3.5 Live Translate: real-time, continuous speech-to-speech translation for 70+ languages, reducing delays.

AI.denJun 9