ARTIFICIAL INTELLIGENCE

NOW EVERYONE CAN GENERATE MUSIC WITH GOOGLE'S GEMINI

Google has added its DeepMind-developed Lyria 3 music generation model to the Gemini app. The integration lets users create 30-second tracks complete with instrumentals, vocals, and automatically generated lyrics from text prompts or uploaded images and videos. Outputs include custom cover art produced by the Nano Banana image model.

The process uses a dedicated “Create music” button or direct prompt in the Gemini chat interface. Examples include a text request for an “Afrobeat track for my mother about the great times we had growing up” or a “comical R&B slow jam about a sock finding their match.” Users can also upload a photo or short video for the system to match the mood and generate fitting lyrics and music.

Lyria 3 improves on prior versions by handling lyric creation internally, increasing musical complexity, and providing finer control over style, vocal delivery, and tempo. The model is set to prioritize original expression; prompts naming specific artists are treated as broad style references only.

The feature is live in beta on the Gemini web interface today and will roll out to the mobile app in the coming days. It is available to users aged 18 and older in English, German, Spanish, French, Hindi, Japanese, Korean, and Portuguese, with additional languages planned. Generation limits are standard for free users and higher for Gemini paid subscribers (Plus, Pro, Ultra tiers). Google also extended Lyria 3 to YouTube’s Dream Track tool for Shorts creators.

The Gemini app reported over 750 million monthly active users in Alphabet’s most recent earnings. This addition completes the current set of multimodal tools in Gemini (text, image, video, and now audio) and extends the same capability to Google Workspace users for custom soundtrack needs.

Business and investment context

The move places music generation inside a general-purpose AI platform with massive distribution rather than a standalone app. It directly competes on accessibility with dedicated services such as Suno and Udio, which focus on longer-form music and advanced editing but lack Google’s integrated ecosystem across Search, YouTube, and Android.

For Alphabet, the launch targets three measurable outcomes: higher daily engagement in the Gemini app, increased conversion to paid subscriptions via usage limits, and stronger positioning of YouTube as a creation platform. Audio content remains highly shareable and sticky, which supports ad inventory and creator tools that generated the majority of YouTube’s revenue last year.

Enterprise access remains available through Vertex AI, where Google provides IP indemnification for commercial use cases. The 30-second limit and beta status indicate a controlled initial deployment focused on short-form and ideation use rather than full professional production.

Quality and consistency will be tracked through user feedback, as with all new Gemini features. The company continues to apply safety filters and watermarking to generated content.

This fits Alphabet’s pattern of shipping research from DeepMind into consumer products at scale. Execution on refinement (track length, editing tools, licensing clarity) will determine the contribution to overall AI-driven growth. No further assumptions are required at this stage; the data on user base, rollout, and feature scope are public and direct from the announcement.