Google Just Dropped LYRIA 3: New AI Feature No One Expected
By AI Revolution
Google’s Recent AI Updates: Lyria 3, Pomelli, & Stitch – A Detailed Breakdown
Key Concepts:
- Lyria 3: Google’s latest text-to-music generation model, integrated into Gemini and YouTube’s Dream Track.
- Synth ID: Google’s imperceptible watermark technology for audio attribution and copyright protection.
- Pomelli: Google’s AI-powered marketing platform for SMBs, now featuring AI-generated product photoshoots.
- Stitch: Google’s AI design tool, evolving with new agents for app design and direct code integration.
- Model Context Protocol (MCP): A protocol enabling seamless integration between design tools like Stitch and coding environments.
- Multimodal AI: AI systems capable of processing and generating content across multiple modalities (text, image, audio, video).
I. Lyria 3: Revolutionizing Music Generation
Google has officially launched Lyria 3, a significant advancement in music generation technology. Unlike previous iterations, Lyria 3 is now directly integrated into consumer-facing products like the Gemini app and YouTube’s Dream Track, moving beyond research demos.
- Functionality: Lyria 3 generates 30-second music tracks based on natural language prompts specifying genre, mood, tempo, and even lyrics (which are now automatically generated). It also accepts image or video input to create music matching visual content.
- Technical Specifications: The model outputs audio at a professional-grade 48 kHz sample rate with 16-bit PCM stereo. This signifies a move towards production-quality audio, not just demo-level output.
- Underlying Technology: Lyria 3 generates music from scratch, avoiding the limitations of assembling pre-made components. While the exact architecture isn’t fully disclosed, it operates within a newer generation of models capable of handling high-fidelity audio while maintaining structural coherence over time. Google contrasts this with models that generate spectrograms or work with compressed audio tokens.
- Synth ID Watermarking: A crucial feature is the inclusion of an imperceptible watermark (Synth ID) embedded directly into the audio waveform. This watermark remains detectable even after compression, slowing, or recording, providing a robust solution for attribution and copyright protection. Users can verify Synth ID presence via the Gemini app. As stated, “Instead of relying on metadata that can be stripped out, they're embedding a digital signature into the sound itself.”
- Real-Time Capabilities (Lyria Realtime): Google DeepMind has introduced Lyria Realtime, which generates audio in two-second chunks via a bidirectional websocket connection. This allows for live steering of the music using weighted prompts, with control latency under 2 seconds.
- Music AI Sandbox: A hands-on environment for musicians, allowing them to transform hums or piano lines into orchestral arrangements, generate vocal choirs from MIDI chords, and change instruments via text prompts.
II. Higsfield’s Cinema Studio 2: Multimodal Production Pipeline
Sponsored segment highlighting Higsfield’s Cinema Studio 2, a platform designed for multimodal production.
- Workflow: Cinema Studio 2 emphasizes a studio pipeline approach, allowing users to plan shots, control camera movements, and define lens behavior before generation. This ensures consistency in lighting, composition, and subjects during motion.
- Features: The platform supports sequencing multiple shots into cohesive scenes and hosts a variety of AI video models. The upcoming Cedance 2 engine will focus on long-form, cohesive video generation.
- Current Promotion: Higsfield is offering a 50% discount on access to Cling 3, a high-performing video model.
III. Pomelli: AI-Powered Marketing for SMBs – Introducing Photoshoot
Pomelli, Google’s AI marketing experiment for small and medium-sized businesses, is expanding its capabilities with the introduction of “Photoshoot.”
- Problem Solved: Professional product photography is expensive. Photoshoot aims to provide a cost-effective solution.
- Functionality: Businesses upload a product image, select visual themes and templates, and generate polished marketing images. These images integrate directly into Pomelli’s campaign workflow.
- Business DNA Profile: Photoshoot leverages Pomelli’s existing “business DNA profile” to maintain brand consistency.
- Recent Integrations: Pomelli added animated asset generation through VO3.1 integration in January 2026.
- Status: Currently in public beta, but Google is rapidly iterating, suggesting a near-term release of Photoshoot.
IV. Stitch: AI Design Tool – Agents, App Store Assets, & Code Integration
Stitch, Google’s AI design tool (formerly Galileo AI), is evolving with new features and a focus on “agents.”
- Hatter Agent: A new agent, “Hatter,” is designed for high-quality design creation. Google emphasizes that it’s being labeled as an “agent” to signify its potential for handling complex, multi-step design tasks. This relates to “deep design,” a concept similar to “deep think” but applied to UI/UX generation.
- App Store Asset Generation: Stitch can now automatically generate app store screenshots, descriptions, and icons, streamlining the app launch process.
- Native MCP Integration: Stitch now natively integrates with the Model Context Protocol (MCP), allowing direct connection to coding tools like Cursor, Claude, and Gemini CLI. This eliminates the need for third-party connectors like Lovable. “Designers and developers can pull Stitch design straight into their coding environments with minimal friction.”
- Latency: Google is prioritizing low latency, with Lyria Realtime operating under 2 seconds and Stitch integrating directly into coding tools.
V. Competitive Landscape & Future Outlook
- Competition: Lyria 3 competes with companies like Sunno (focused on viral music with multi-stem splitting) and Udio (focused on studio-grade fidelity and long-form tracks). Google’s strengths lie in speed, multimodal input, and integration with Gemini.
- Integration as a Key Strategy: Google’s overarching strategy is integration – blurring the lines between music, visuals, design, and deployment into a single creative stack.
- API Potential: While currently lacking a public API, Google has hinted at the possibility of making Lyria 3 accessible via an API, transforming it from a consumer tool into a foundational infrastructure.
Conclusion:
Google is aggressively pushing the boundaries of AI-powered creativity with significant updates to Lyria 3, Pomelli, and Stitch. The emphasis on multimodal input, real-time capabilities, and seamless integration across its ecosystem positions Google as a major player in the future of content creation. The introduction of features like Synth ID and native MCP integration demonstrates a commitment to addressing practical challenges related to attribution, copyright, and workflow efficiency. The shift from model announcements to system building signals a deliberate strategy to create a cohesive and powerful creative stack.
Chat with this Video
AI-PoweredLoad the transcript when you're ready to chat so the initial page stays lighter.
Related Videos

Seedance 2.0 4K: The New AI Video King?
Zubair Trabzada | AI Workshop

I Used Higgsfield Inside Photoshop and It Changed Everything
Zubair Trabzada | AI Workshop

AI tools for human creativity
Google for Developers

GPT 5.6, Mythos ban lifted, realtime avatars, Seedance 2.5, brain ultrasound: AI NEWS
AI Search

Claude Design 2.0 Major Upgrades Explained
Zubair Trabzada | AI Workshop

Figma Unveils Full-Stack Canvas for the AI Era
Bloomberg Technology

What's new with Gemini from Google DeepMind
Google Cloud Tech