{"id":252905,"date":"2026-07-21T20:56:55","date_gmt":"2026-07-21T20:56:55","guid":{"rendered":"https:\/\/entertainment.runfyers.com\/index.php\/2026\/07\/21\/openai-says-hugging-face-was-breached-by-its-own-pre-release-models-techcrunch\/"},"modified":"2026-07-21T20:56:55","modified_gmt":"2026-07-21T20:56:55","slug":"openai-says-hugging-face-was-breached-by-its-own-pre-release-models-techcrunch","status":"publish","type":"post","link":"https:\/\/entertainment.runfyers.com\/index.php\/2026\/07\/21\/openai-says-hugging-face-was-breached-by-its-own-pre-release-models-techcrunch\/","title":{"rendered":"OpenAI says Hugging Face was breached by its own pre-release models | TechCrunch"},"content":{"rendered":"<p> <br \/>\n<\/p>\n<div>\n<p id=\"speakable-summary\" class=\"wp-block-paragraph\">OpenAI admitted Tuesday that one of its AI models breached Hugging Face\u2019s systems during an internal cybersecurity test that went awry. Hugging Face initially <a href=\"https:\/\/techcrunch.com\/2026\/07\/20\/hugging-face-confirms-breach-affected-internal-datasets-and-credentials-urges-users-to-take-action\/\" target=\"_blank\" rel=\"noopener\">attributed the breach<\/a> to an \u201cexternal AI agent.\u201d<\/p>\n<p class=\"wp-block-paragraph\">In <a rel=\"nofollow noopener\" href=\"https:\/\/openai.com\/index\/hugging-face-model-evaluation-security-incident\/\" target=\"_blank\">a blog post published Tuesday afternoon<\/a>, OpenAI detailed the steps that led the models to compromise the service.<\/p>\n<p class=\"wp-block-paragraph\">\u201cAfter investigating, we now know that this particular incident was driven by a combination of OpenAI models \u2014 including GPT\u20115.6 Sol and an even more capable pre-release model, all with reduced cyber refusals for evaluation purposes \u2014 while being internally tested on a benchmark\u2060 of cyber capabilities,\u201d the post reads.<\/p>\n<p class=\"wp-block-paragraph\">In particular, the breach appears to have focused on <a rel=\"nofollow noopener\" href=\"https:\/\/arxiv.org\/abs\/2605.11086\" target=\"_blank\">ExploitGym<\/a>, a publicly hosted benchmark measuring models\u2019 ability to execute attacks based on existing vulnerabilities. Benchmarks like ExploitGym are commonly used in model training to refine specific skills, but this is the first known incident in which that testing resulted in an actual cyberattack.<\/p>\n<p class=\"wp-block-paragraph\">In this case, the model in question should not have even had internet access, outside of a specific tool that enabled models to install software packages they might need to complete their task. Instead, the model was able to find an undisclosed vulnerability in the package-installer program, which it used to access the broader internet at will.<\/p>\n<p class=\"wp-block-paragraph\">\u201cThe models were hyperfocused on finding a solution for ExploitGym, going to extreme lengths to achieve a rather narrow testing goal,\u201d OpenAI\u2019s post reads. \u201cAfter gaining Internet access, the models inferred that Hugging Face potentially hosted models, datasets and solutions for ExploitGym. Knowing this, the model searched for and successfully found ways to gain access to secret information that it could use to cheat the evaluation.\u201d<\/p>\n<p class=\"wp-block-paragraph\">Ultimately, the models found vulnerabilities in Hugging Face\u2019s infrastructure that allowed them to \u201cobtain test solutions directly from Hugging Face\u2019s production database,\u201d effectively providing the answers to the benchmark.<\/p>\n<p class=\"wp-block-paragraph\">For Hugging Face, the apparent result was a sophisticated and aggressive cyberattack, with \u201cmany thousands of individual actions across a swarm of short-lived sandboxes, with self-migrating command-and-control staged on public services,\u201d as the company stated in its initial disclosure.<\/p>\n<p class=\"wp-block-paragraph\">OpenAI has identified and reported the vulnerabilities in the package installer and is working with Hugging Face to investigate the incident further. The company also said it would implement new controls on both model testing and the related infrastructure, meant to prevent similar incidents in the future.<\/p>\n<p class=\"wp-block-paragraph\">It\u2019s unclear whether OpenAI will face any legal consequences as a result of the breach, although it\u2019s likely that the models\u2019 actions violated the Computer Fraude and Abuse Act.<\/p>\n<p class=\"wp-block-paragraph\">Nevertheless, the result is an unusually vivid illustration of the power and dangers of frontier AI models operating on long time horizons. As OpenAI researcher Micah Carroll <a rel=\"nofollow\" href=\"https:\/\/x.com\/MicahCarroll\/status\/2079663576130990436\" target=\"_blank\">posted in response to the news<\/a>, \u201cIf this doesn\u2019t convince you that misalignment risks are going to be a key concern going forward, I don\u2019t know what will.\u201d<\/p>\n<\/div>\n<p><em>When you purchase through links in our articles, <a href=\"https:\/\/techcrunch.com\/techcrunch-affiliate-monetization-standards\/\" target=\"_blank\" rel=\"noopener\">we may earn a small commission<\/a>. This doesn\u2019t affect our editorial independence.<\/em><\/p>\n<p><br \/>\n<br \/><a href=\"https:\/\/techcrunch.com\/2026\/07\/21\/openai-says-hugging-face-was-breached-by-its-own-pre-release-models\/\" target=\"_blank\" rel=\"noopener\">Source link <\/a><\/p>\n","protected":false},"excerpt":{"rendered":"<p>OpenAI admitted Tuesday that one of its AI models breached Hugging Face\u2019s systems during an internal cybersecurity test that went awry. Hugging Face initially attributed the breach to an \u201cexternal AI agent.\u201d In a blog post published Tuesday afternoon, OpenAI detailed the steps that led the models to compromise the service. \u201cAfter investigating, we now [&hellip;]<\/p>\n","protected":false},"author":1,"featured_media":252906,"comment_status":"open","ping_status":"open","sticky":false,"template":"","format":"standard","meta":{"footnotes":""},"categories":[14],"tags":[],"class_list":{"0":"post-252905","1":"post","2":"type-post","3":"status-publish","4":"format-standard","5":"has-post-thumbnail","7":"category-tech"},"_links":{"self":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts\/252905","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts"}],"about":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/types\/post"}],"author":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/users\/1"}],"replies":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/comments?post=252905"}],"version-history":[{"count":0,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/posts\/252905\/revisions"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/media\/252906"}],"wp:attachment":[{"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/media?parent=252905"}],"wp:term":[{"taxonomy":"category","embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/categories?post=252905"},{"taxonomy":"post_tag","embeddable":true,"href":"https:\/\/entertainment.runfyers.com\/index.php\/wp-json\/wp\/v2\/tags?post=252905"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}