OpenAI rolls out voice mode after delaying it for safety reasons

399
SHARES
2.3k
VIEWS


SAN FRANCISCO — ChatGPT maker OpenAI mentioned Tuesday it would start rolling out its new voice mode to clients, a month after delaying the launch to do extra safety testing on the instrument.

OpenAI in May confirmed off the conversational voice mode, which may detect totally different tones of voice and reply to being interrupted, very similar to a human. But some researchers rapidly criticized the corporate for displaying off a synthetic intelligence product that hewed to sexist stereotypes about feminine assistants being flirty and compliant. Actor Scarlett Johansson alleged the corporate had copied her voice from the film “Her,” by which an AI bot develops a romantic relationship with a person.

OpenAI’s information present it labored with a very totally different actor, and it pulled the voice, known as Sky, from its product. In June, it mentioned it would delay the launch of voice mode to conduct extra safety testing. The new voice mode launching Tuesday doesn’t embrace the Sky voice, an OpenAI spokesperson confirmed.

Tech firms have labored to make conversational AI chatbots for years. Amazon’s Alexa and Apple’s Siri are ubiquitous and utilized by thousands and thousands of individuals to set timers and lookup the climate however aren’t succesful sufficient for complicated duties. Now, OpenAI, Google, Microsoft, Apple and a number of different tech firms try to make use of breakthroughs in generative AI to lastly construct the type of assistant that has been a fixture of science fiction for a long time.

OpenAI’s followers and clients have clamored for the voice mode, with some complaining on-line when the corporate delayed the launch in June. The new function shall be out there to a small variety of customers at first, and the corporate will step by step open it as much as all of OpenAI’s paying clients by the autumn.

Previous variations of ChatGPT have had the flexibility to take heed to spoken questions and reply with audio by transcribing the questions into textual content, working them by its AI algorithm, after which studying its textual content response out loud. But the brand new voice options are constructed on OpenAI’s newest AI mannequin, which instantly processes audio with no need to transform it to textual content first. That permits the bot to take heed to a number of voices directly and decide an individual’s tone of voice, responding otherwise primarily based on what it thinks the individual’s feelings are.

That opens up an entire new set of questions, akin to how cultural variations come into play, or whether or not folks would possibly develop relationships with bots which can be skilled to reply to their feelings in particular methods. OpenAI mentioned it labored with folks representing 45 languages and 29 “geographies” to enhance the AI mannequin’s capabilities.

Only 4 distinctive voices shall be out there to make use of, and the instrument will block makes an attempt to get the bot to generate voices of actual folks, the corporate mentioned.



Source hyperlink