Skip to main content
  1. Home
  2. Computing
  3. News

Digital Trends may earn a commission when you buy through links on our site. Why trust us?

The new AI tool that was deemed ‘too dangerous’ to release

Add as a preferred source on Google

Back in 2019, OpenAI refused to release its full research into the development of GPT2 over fears that it was “too dangerous” to release publicly. On Thursday, OpenAI’s biggest financial backer, Microsoft, made a similar pronouncement about its new VALL-E 2 voice synthesizer AI.

The VALL-E 2 system is a zero-shot text-to-speech synthesis (TTS) AI, meaning that it can recreate hyper-realistic speech based on just a few seconds of sample audio. Per the research team, VALL-E 2 “surpasses previous systems in speech robustness, naturalness, and speaker similarity. It is the first of its kind to reach human parity on these benchmarks.”

Recommended Videos

The system reportedly can even handle sentences that are difficult to pronounce because of their structural complexity or repetitive phrasing, such as tongue twisters.

There are a host of potential beneficial uses for such a system, like enabling people suffering from aphasia or Amyotrophic lateral sclerosis (commonly known as ALS or Lou Gehrig’s disease) to speak again, albeit through a computer, as well as use in education, entertainment, journalism, chatbots and translation, or as accessibility features and “interactive voice response systems,” like Siri. However, the team also recognizes numerous opportunities for the public to misuse its technology, “such as spoofing voice identification or impersonating a specific speaker.”

As such the AI will only be available for research purposes. “Currently, we have no plans to incorporate VALL-E 2 into a product or expand access to the public,” the team wrote. ” If you suspect that VALL-E 2 is being used in a manner that is abusive or illegal or infringes on your rights or the rights of other people, you can report it at the Report Abuse Portal.”

Microsoft is hardly alone in its efforts to train computers to speak as humans do. Google’s Chirp, ElevenLabs’ Iconic Voices, and Voicebox from Meta all aim to perform similar functions.

However, such systems have come under ethical scrutiny as they have repeatedly been used to scam unsuspecting victims by emulating the voice of a loved one or a well-known celebrity. And unlike generated images, there’s currently no way to effectively “watermark” AI generated audio.

Andrew Tarantola
Former Computing Writer
Andrew Tarantola is a journalist with more than a decade reporting on emerging technologies ranging from robotics and machine…
Tokyo lets AI play matchmaker, and hundreds of couples have already tied the knot
Tokyo's government-built AI matchmaking app has led to 265 marriages so far.
ai-dating-apps-chatting

Government-run dating apps sound like something out of a policy white paper. Tokyo built one anyway, and it's already helping turn AI matches into real marriages (via Tech Xplore).

So how does this app actually work?

Read more
Apple’s new Upgrade program lets you lease an iPhone, Mac, Watch, and iPad
Launched in partnership with Klarna, the new program gets you an iPhone for as little as $17.99 a month.
Apple Upgrade site on a MacBook Neo

Apple has officially launched Apple Upgrade, a leasing program that lets customers pay monthly for an iPhone, iPad, Mac, or Apple Watch instead of buying the device outright. The program went live today on Apple's website in partnership with Klarna, replacing the iPhone Upgrade Program that Apple has offered since 2015.

What you can lease and what it costs

Read more
Chuwi CoreBook Air review: A rewarding budget laptop that wants you to cross the brand boundary
Chuwi's CoreBook Air is a relatively unknown, but a heavy-hitting budget machine for the AI-first computing era.
Computer, Electronics, Laptop

View at Chuwi

Quick verdict

Read more