Suno debuts Speech, which generates spoken voices with optional background music using scripts or prompts, in public beta, as its music product faces lawsuits
First reported by The Verge ·
AI-generated spoken voice with integrated music now has a readily accessible, eight-minute-long option for content creators.
Suno has launched Speech, a public beta feature that generates spoken voices from scripts or prompts, optionally accompanied by AI-composed background music. This expansion into spoken audio technology comes as Suno's AI music generation platform faces multiple lawsuits. Speech offers two modes: Simple, for descriptive prompts, and Advanced, for custom scripts with adjustable voice gender, style, and variety. The generated audio tracks can be up to eight minutes long. Suno acknowledges the feature is in beta and subject to ongoing improvement based on user feedback, with potential for unexpected outputs like accent drift or exaggerated pauses. This move aims to diversify Suno's offerings beyond music generation, leveraging its existing AI audio capabilities.
Suno's introduction of Speech, a text-to-speech and voice generation tool with optional AI music, signals a strategic pivot to diversify its generative audio capabilities amidst legal challenges. By integrating voice and music generation into a single cohesive track, Suno aims to capture use cases that benefit from synchronized audio elements, distinguishing itself from pure text-to-speech services.
The public beta of Speech, offering both prompt-based and script-based generation, along with customizable voice parameters and a generous eight-minute maximum duration, positions Suno to attract a broad user base seeking integrated audio content. This diversification allows the company to explore new revenue streams and applications, potentially mitigating risks associated with its music generation lawsuits by expanding its product portfolio.
AI-written summary. May contain errors.