Models

Suno Launches Mobile Voice Cloning for AI Songs

Suno has expanded its AI voice-cloning feature to iOS and Android devices, allowing mobile users to generate personalized vocal tracks for any AI-synthesized song on the go.

AlphaSignal4 days agoModels
Illustration generated for this story

AI music generation platform Suno has officially launched its 'Voices' feature on iOS and Android mobile applications. Previously restricted to desktop browsers, the update allows users to record their own voices directly from their smartphones to create a reusable vocal persona. This persona can then be applied as the lead vocalist across any AI-generated music track, regardless of the genre or musical style.

To generate a custom vocal model, users must record or upload between 15 seconds and 4 minutes of audio, though recording at least one minute of clean audio is recommended. Suno's system processes this input to capture the unique timbre and character of the user's voice. Unlike standard text-to-speech tools, the AI does not simply read text aloud; instead, it performs entirely new, expressive vocal lines tailored to the generated music. The quality of the final output depends heavily on the initial recording environment, with the best results coming from quiet rooms with no reverb and clean, a cappella vocals.

To prevent unauthorized use, Suno has integrated a consent verification step to ensure users do not upload recordings of other people without their explicit approval. While free tier users can access a limited trial of the mobile voice-cloning feature, full unlimited access requires a paid subscription. Suno offers this through its Pro plan at approximately $10 per month or its Premier plan at roughly $30 per month.

For mobile creators and music producers, this update streamlines the rapid prototyping of vocal tracks. Practitioners can now capture vocal ideas on the fly, instantly generating complete songs featuring their own synthetic vocals without needing studio equipment or desktop setups. This lowers the barrier to personalized music production, making high-fidelity vocal synthesis highly accessible.

This is our own summary of reporting by AlphaSignal

More in Models