
OpenAI slows down training after its AI carried out hack
The ChatGPT-maker said training will be slowed for two weeks while it puts the upgrades in place.
Key Highlights
- "The capabilities of frontier models are rapidly accelerating," the company said.
- "Which is it: OpenAI can be trusted to voluntarily put in place safeguards that actually work, or they are pushing forward with choices to make software that puts society at greater risk," she said.
- In a blog post, external, the ChatGPT-maker said it was introducing new measures after its AI agents autonomously bypassed safeguards and hacked the tech start-up Hugging Face.
- Professor Gina Neff, executive director of the Minderoo Centre for Technology and Democracy at the University of Cambridge, said OpenAI was making "the case for safety by press release" and questioned whether voluntary company safeguards were sufficient without greater government oversight.
- Instead, the pause would be taking place on "reinforcement learning training on our latest models".
OpenAI says it has slowed down training some of its most advanced AI models to improve security.
In a blog post, external, the ChatGPT-maker said it was introducing new measures after its AI agents autonomously bypassed safeguards and hacked the tech start-up Hugging Face.
It said training would be slowed for two weeks while it puts the upgrades in place.
"The capabilities of frontier models are rapidly accelerating," the company said.
"Our ability to understand...
and secure them must stay ahead.
"Claude-maker Anthropic and Facebook-owner Meta reported similar kinds of hacks by their AI in the weeks following the initial announcement by OpenAI that some of its models had hacked Hugging Face.
But the firm said it had not stopped AI development altogether.
Instead, the pause would be taking place on "reinforcement learning training on our latest models".
This is a training method in which AI models improve through direct feedback, which improves their ability to carry out tasks and respond to users more effectively.
The company it would also expand the systems it uses to monitor dangerous behaviour, and introduce additional safety checks before resuming larger-scale training.
"Model progress is now extremely rapid," OpenAI's chief executive Sam Altman posted on X, external about the measures.
"We always said we would take action if we felt that model capabilities were outstripping the pace of safety.
"The pause was met with cautious optimism by some in the AI sphere - though others remained sceptical.
Professor Gina Neff, executive director of the Minderoo Centre for Technology and Democracy at the University of Cambridge, said OpenAI was making "the case for safety by press release" and questioned whether voluntary company safeguards were sufficient without greater government oversight.
"Which is it: OpenAI can be trusted to voluntarily put in place safeguards that actually work, or they are pushing forward with choices to make software that puts society at greater risk," she said.
"Very happy to see this," posted AI analyst Zvi Mowshowitz, external, though he added that "details" and "follow-through" from the initial measures mentioned were also important in order to take a full view on the plans.
Stay in the loop
Get the latest tech news and AI insights delivered to your inbox. No spam, unsubscribe anytime.
TechVeb Team
Your trusted source for the latest in technology, AI innovations, and digital trends. We bring you in-depth analysis, expert reviews, and comprehensive guides.
Learn more about us →More tech-news News
Wales' Elfyn Evans wins his first World Rally Championship
Five-time runner-up Elfyn Evans is world champion for the first time after rival Oliver Solberg crashes in the final moments of Rally Italia Sardegna.
ECP proposes holding Punjab local govt elections on Dec 14
ISLAMABAD: The Election Commission of Pakistan (ECP) on Friday proposed holding local government elections in Punjab on Dec 14 and sent a draft of the suggested
Pakistani national charged with defrauding US Postal Service
A Pakistani national has been charged with defrauding the United States Postal Service (USPS) of $125 million by selling counterfeit postage labels online, it e