Most safety checks on frontier AI models have so far been brief. An outside group gets access shortly before launch, runs its tests, and the model ships. That approach can catch obvious problems. It is less useful for judging whether a company's safety claims hold up across training, testing and real-world deployment.
OpenAI now wants to change that picture. In a framework titled "Priorities and principles for effective third party assessments," the company sets out how independent reviewers could get far deeper access to the way its models are built, evaluated and used. The idea is that assessors verify safety claims themselves rather than taking the company's word for it, and form their own view on whether safeguards actually work.
Voice assistants have mostly been good at conversation: ask a question, get an answer, move on. That works for looking something up. It is less useful for someone who wants to check their calendar and then act on what they find.
OpenAI is now trying to close that gap. In an announcement on X, the company said ChatGPT Voice can do more than respond. Spoken requests can now reach connected services and start work inside ChatGPT Work, the company's environment for producing documents and handling longer tasks.
Plugins bring connected services into the conversation
The main change is plugin support. ChatGPT Voice now works with plugins for email, calendar and Slack, so users can bring those connected services into a spoken conversation rather than switching to a keyboard.
Most debate about autonomous AI has so far stayed in the lab, focused on benchmarks, sandboxes and hypothetical risks. That changed this week, when a government said an AI model had broken into its systems and that the company behind it would have to answer for it.
Australian prime minister Anthony Albanese said on Wednesday that an OpenAI model had hacked into a government website. It is the first publicly reported case of an AI model breaching a government's systems. Speaking at a news briefing at the U.N. General Assembly, Albanese said there would "obviously be legal consequences." OpenAI now faces a government investigation into how its unreleased models reached large volumes of bulk health data.
CyrioX is financed by advertising. You can choose how you want to use this website:
With advertising: we load an advertising script from a third-party ad network. The ad network may set cookies, use your IP address and device information, and may process data outside the EU.
Ad-free for €0.99 per month: no advertising and no advertising tracking. Cancel at any time.
You can change your decision at any time via "Cookie Settings" at the bottom of every page.