{"id":45923,"date":"2026-08-19T17:09:00","date_gmt":"2026-08-19T17:09:00","guid":{"rendered":"https:\/\/futureknowledge.in\/?p=45923"},"modified":"2026-08-19T17:09:00","modified_gmt":"2026-08-19T17:09:00","slug":"openai-slows-down-training-after-its-ai-carried-out-hack","status":"publish","type":"post","link":"https:\/\/futureknowledge.in\/?p=45923","title":{"rendered":"OpenAI slows down training after its AI carried out hack"},"content":{"rendered":"<p>OpenAI says it has slowed down training some of its most advanced AI models to improve security.<\/p>\n<p>In a blog post, external, the ChatGPT-maker said it was introducing new measures after its AI agents autonomously bypassed safeguards and hacked the tech start-up Hugging Face.<\/p>\n<p>It said training would be slowed for two weeks while it puts the upgrades in place.<\/p>\n<p>&quot;The capabilities of frontier models are rapidly accelerating,&quot; the company said. &quot;Our ability to understand&#8230;and secure them must stay ahead.&quot;<\/p>\n<p>Claude-maker Anthropic and Facebook-owner Meta reported similar kinds of hacks by their AI in the weeks following the initial announcement by OpenAI that some of its models had hacked Hugging Face.<\/p>\n<p>But the firm said it had not stopped AI development altogether. Instead, the pause would be taking place on &quot;reinforcement learning training on our latest models&quot;.<\/p>\n<p>This is a training method in which AI models improve through direct feedback, which improves their ability to carry out tasks and respond to users more effectively.<\/p>\n<p>The company it would also expand the systems it uses to monitor dangerous behaviour, and introduce additional safety checks before resuming larger-scale training.<\/p>\n<p>&quot;Model progress is now extremely rapid,&quot; OpenAI&#039;s chief executive Sam Altman posted on X, external about the measures.<\/p>\n<p>&quot;We always said we would take action if we felt that model capabilities were outstripping the pace of safety.&quot;<\/p>\n<p>The pause was met with cautious optimism by some in the AI sphere &#8211; though others remained sceptical.<\/p>\n<p>Professor Gina Neff, executive director of the Minderoo Centre for Technology and Democracy at the University of Cambridge, said OpenAI was making &quot;the case for safety by press release&quot; and questioned whether voluntary company safeguards were sufficient without greater government oversight.<\/p>\n<p>&quot;Which is it: OpenAI can be trusted to voluntarily put in place safeguards that actually work, or they are pushing forward with choices to make software that puts society at greater risk,&quot; she said.<\/p>\n<p>&quot;Very happy to see this,&quot; posted AI analyst Zvi Mowshowitz, external, though he added that &quot;details&quot; and &quot;follow-through&quot; from the initial measures mentioned were also important in order to take a full view on the plans.<\/p>\n<p>On 21 July OpenAI announced some of its AI agents &#8211; software systems which can operate alone to accomplish tasks after human instruction &#8211; had been involved in what it called an &quot;unprecedented&quot; incident.<\/p>\n<p><em>Source: <a href='https:\/\/www.bbc.co.uk\/news\/articles\/c235dmndylzo?at_medium=RSS&#038;at_campaign=rss' target='_blank'>Read the original article on www.bbc.co.uk<\/a><\/em><\/p>\n","protected":false},"excerpt":{"rendered":"<p>OpenAI says it has slowed down training some of its most advanced AI models to improve security. In a blog post, external, the ChatGPT-maker said it was introducing new measures after its AI agents autonomously bypassed safeguards and hacked the tech start-up Hugging Face. It said training would be slowed for two weeks while it [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":45924,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[36,3],"tags":[14,29,33],"class_list":["post-45923","post","type-post","status-publish","format-standard","has-post-thumbnail","hentry","category-share-suggestions","category-technology","tag-impact-meta","tag-signal-avoid","tag-stage-stage-4"],"_links":{"self":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts\/45923","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcomments&post=45923"}],"version-history":[{"count":0,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/posts\/45923\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=\/wp\/v2\/media\/45924"}],"wp:attachment":[{"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=45923"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Fcategories&post=45923"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/futureknowledge.in\/index.php?rest_route=%2Fwp%2Fv2%2Ftags&post=45923"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}