
What happened
Creatune's Alex wrote that the current AI cover screen offers two paths — a Voice Cover, which keeps the song's core structure and swaps the singing voice, and a Style Cover, which changes genre, arrangement and mood.
Why it matters
Changing voice, arrangement and tempo at once makes it hard to tell why a result worked, so fixing one variable and holding the other as a comparison condition appears to be the point of the split.
What to watch
The test is whether users actually commit to a keep/change/leave-alone table before generating, since the upload rules — MP3, WAV, M4A, up to 50MB and 8 minutes — say nothing about permission to use a track or voice.
WHO IT HITSIndependent musicians and cover creators using AI voice tools are the ones who would work through this keep/change/leave-alone table, and labels or rights holders deciding whether a track may be used at all sit on the other side of the same check.
Summaries like this, in your inbox every morning.
The piece is framed as a way to organize judgment rather than a showcase of generated results, and Alex is explicit that it is not a measured report on any particular cover output. The stated premise is that you work only with audio you hold rights to and voices you have permission to use.
The method rests on separating decisions. A three-column table — keep, change, and leave untouched for now — is meant to make covers comparable, and Alex notes that the 'leave untouched' column is easily dismissed even though it is what keeps a comparison meaningful; making everything variable from the first attempt leaves no information to guide the next revision. Instructions for the voice are broken into musical roles rather than a person's name, covering range, attack, dynamics, texture and harmony, so the request can be reused as a production note. Style is likewise split into rhythm, instrumentation, density, space and the role of each section, with a worked example of dark rock described in concrete terms. Comparison is done on the same three segments of the original and the cover, and fixes move one column at a time, depending on whether the original character vanished, the change felt weak, or the results became incomparable.
What this hinges on is whether the separation actually survives contact with users' habits, since the tool itself allows either direction and neither forces the table or the permission checks. For creators the payoff would be revisions that teach them something, and for rights holders the risk sits in the step Alex places last: confirming the original song, the master, the voice and commercial-use terms for wherever the cover is published.
Pick your industry and the AI tools you use, and get news related to your work every day.
Free · 30 seconds with Google · unsubscribe anytimeWhat is AIToday? →
Ask AI anything about this article. The AI reads this article, earlier AIToday articles, and Wikipedia, and cites its sources. Q&As are published on this page for other readers too.
Microsoft launched MAI-Transcribe-2-Streaming, its first streaming transcription model, priced at 54 cents per…
Microsoft AI said it released MAI-Transcribe-2-Streaming, which returns provisional results in just over 100 m…

Microsoft released MAI-Transcribe-2-Streaming on October 1, 2026

Starkey announced Omega AI+, succeeding last year's Omega AI

Airbnb CEO Brian Chesky said he has used Airbnb on the agents Muse and Instinct and "it doesn't work very well…

A Reddit analysis of audio.cpp mapped shared building blocks across 100+ audio models
