Microsoft's Azure AI Speech needs just seconds of audio to spit out a convincing deepfake
ID: f8870ee6-9a9a-5619-b768-247bc17f3dfa
STIX ID: report--f8870ee6-9a9a-5619-b768-247bc17f3dfa
Feed Name: The Register (Security)
Microsoft upgraded Azure AI Speech’s personal voice feature to the zero-shot “DragonV2.1Neural” model, delivering more natural, expressive voice cloning across 100+ languages from only seconds of audio, and added measures like audio watermarks and strict usage policies (consent, disclosure, no impersonation). While the capability enables applications such as customized chatbot voices and multilingual dubbing, the report underscores growing risks of audio deepfakes, citing Consumer Reports’ concerns over insufficient safeguards and an FBI warning about scammers using synthetic voices in fraud campaigns.
Your team is not currently subscribed to this feed. You must subscribe to it in order to see this post.
