product update

Google's Gemini adds Lyria 3 music generation from text and images

TL;DR

Google has integrated Lyria 3, its music generation model, directly into the Gemini app. Users can now create custom 30-second music tracks from text descriptions and images without additional tools or subscriptions.

2 min read
0

Google Integrates Lyria 3 Music Generation Into Gemini App

Google has rolled out Lyria 3 music generation capabilities within the Gemini app, allowing users to create 30-second audio tracks from text prompts and image inputs. The feature is now available to Gemini app users.

What Lyria 3 Does

Lyria 3 generates original music compositions based on user-provided text descriptions and images. Users input creative prompts—specifying genre, mood, instrumentation, or other characteristics—and the model produces corresponding audio files. The system can also generate music inspired by visual inputs, adding another dimension to music creation workflows.

Integration Details

The feature is embedded directly in the Gemini app interface, eliminating the need for users to navigate to separate music generation tools. This positions music creation as a native capability alongside Gemini's text, image, and document analysis features.

Google has not disclosed specific technical specifications for Lyria 3 in this announcement, including model size, training data composition, or detailed performance benchmarks against competing music generation systems.

Broader Context

This release follows Google's earlier introduction of Lyria and Music AI Essentials framework. Google DeepMind has been developing music generation technology as part of its broader multimodal AI strategy. The integration into Gemini represents a shift toward making generative music capabilities as accessible as text and image generation.

The 30-second generation limit suggests design decisions around computational efficiency and user experience, though Google has not explained the rationale for this constraint.

Industry Position

Music generation remains an emerging category in generative AI. Competitors including OpenAI-backed projects, Meta's research divisions, and specialized startups have developed similar tools. Google's approach of integrating music generation into its flagship AI assistant positions music creation as a mainstream feature rather than a specialized tool.

The lack of disclosed pricing or usage limits suggests the feature may be part of Gemini's existing subscription tier structure, though specific details remain unavailable.

What This Means

Google is treating music generation as a standard multimodal AI capability, not a separate product. By embedding Lyria 3 into Gemini, the company signals confidence in both the technology's stability and its appeal to mainstream users. For music creators, marketers, and content producers, this lowers barriers to generative music experimentation. For competitors, it intensifies pressure to include music generation in their own platforms. The feature's limitations—30-second maximum length, unspecified quality standards—remain important constraints relative to specialized music generation tools.

Related Articles

product update

Google Expands Gemini-Powered Ask Maps Globally With Personal Intelligence, Real-Time Transit, Agentic Ordering

Google Maps' Gemini-powered Ask Maps chat is rolling out globally to English speakers in Australia, Brazil, Canada, Indonesia, Japan, Mexico, and over 150 other countries and territories. The update adds Gmail-based Personal Intelligence, real-time transit data, conversation memory, and agentic capabilities for ordering food and booking hotels.

product update

Google Maps' Ask Maps Adds Agentic Food Ordering, Hotel Booking, and Gmail-Based Personalization

Google is adding agentic capabilities to Google Maps' Ask Maps feature, letting users order food, book hotels, and buy event tickets directly through the app. A new Personal Intelligence feature, off by default, lets Ask Maps pull context from Gmail and Calendar to personalize responses.

product update

Anthropic Adds Cross-Session Messaging to Claude Code v2.1.224

Claude Code v2.1.224 introduces cross-session messaging, letting separate Claude Code instances on macOS and Linux send each other summaries to coordinate work. The feature does not support approving permissions or executing commands remotely.

product update

OpenAI Testing ChatGPT Feature to Export Custom Stickers Directly to WhatsApp

An APK teardown of ChatGPT's Android app reveals a hidden 'ChatGPT Stickers' feature that would let users create custom stickers and export them directly into WhatsApp as sticker packs. The feature is unreleased and its public launch timeline is unknown.

Comments

Loading...