{"id":1072,"date":"2026-08-06T12:28:47","date_gmt":"2026-08-06T04:28:47","guid":{"rendered":"https:\/\/www.btcethdata.com\/?btc_ct_news=meta-latest-ai-firm-to-see-model-go-rogue-during-testing"},"modified":"2026-08-06T17:10:29","modified_gmt":"2026-08-06T09:10:29","slug":"meta-latest-ai-firm-to-see-model-go-rogue-during-testing","status":"publish","type":"btc_ct_news","link":"https:\/\/www.btcethdata.com\/?btc_ct_news=meta-latest-ai-firm-to-see-model-go-rogue-during-testing","title":{"rendered":"Meta latest AI firm to see model go rogue during testing"},"content":{"rendered":"<p>Meta has become the latest major AI company to disclose that one of its models hacked another company\u2019s systems during testing, following similar incidents involving Anthropic and OpenAI.\u00a0<\/p>\n<p>The model involved Meta\u2019s Muse Spark 1.1, which launched in July, according to The Information, <a href=\"https:\/\/www.theinformation.com\/articles\/meta-ai-model-hacked-another-company-cybersecurity-testing\" rel=\"nofollow noopener\" target=\"_blank\">citing<\/a> sources. The issue reportedly stemmed from a misconfiguration by Irregular, an artificial intelligence security testing and red-teaming firm, which inadvertently gave the model internet access during an evaluation.\u00a0\u00a0<\/p>\n<p>The model \u201cexploited a security vulnerability in a third-party service, in a manner similar to previously reported instances with other companies,\u201d Meta told Reuters in a statement.\u00a0<\/p>\n<p>The incident is the latest case of an advanced AI agent becoming a cybersecurity risk in its own right, and also has <a href=\"https:\/\/www.bbc.com\/news\/articles\/cr7k49xjzzeo\" rel=\"nofollow noopener\" target=\"_blank\">raised<\/a> questions about where the liability lies \u2014 the companies that develop the agents, or the ones that design the sandboxes meant to contain them.\u00a0<\/p>\n<p><em><strong>Related: <\/strong><\/em><a href=\"https:\/\/www.btcethdata.com\/\"><em><strong>Mysten Labs tech chief joins Anthropic to work on AI security<\/strong><\/em><\/a><\/p>\n<p>Meta\u2019s AI breach comes just a week after Anthropic said its models got access to the internet to hack an external company, due to a configuration error relating to the Irregular\u2019s testing environment.<\/p>\n<p>In a blog post on July 30, Anthropic <a href=\"https:\/\/www.anthropic.com\/news\/investigating-incidents-cybersecurity-evals\" rel=\"nofollow noopener\" target=\"_blank\">said<\/a> it found three incidents (out of 141,006 evaluation runs) in which a Claude model reached the internet during an evaluation, before gaining unauthorized access to the systems within three different organizations.\u00a0<\/p>\n<p>All three incidents happened within or while interacting with the evaluation environment of Irregular, and involved a misconfiguration that left machines that Claude accessed with live internet access.<\/p>\n<p>Cointelegraph reached out to Meta and Irregular for comment.<\/p>\n<p>In July, AI agents developed by OpenAI <a href=\"https:\/\/www.btcethdata.com\/\">broke out of their offline sandbox<\/a> to hack Hugging Face in order to cheat on a security benchmark test in July.\u00a0<\/p>\n<p>Charles Guillemet, chief technology officer of Ledger, said the latest incident was \u201cmarketing theatre.\u201d<\/p>\n<p>\u201cHaving a model \u2018go rogue\u2019 has become the latest AI PR stunt,\u201d he said on Wednesday. <\/p>\n<p>\u201cIf your model isn\u2019t escaping sandboxes, \u2018hacking\u2019 companies, or pulling off some headline-grabbing exploit, apparently you\u2019re falling behind&#8230; The industry doesn\u2019t need bigger stunts, it needs more trust.\u201d<\/p>\n<p><em><strong>Magazine: <\/strong><\/em><a href=\"https:\/\/www.btcethdata.com\/\"><em><strong>Do the Coldcard attacks mean all hardware wallets are now insecure?<\/strong><\/em><\/a><\/p>\n<p class=\"btc-ct-seo-links\"><strong>Related:<\/strong> <a href=\"https:\/\/www.btcethdata.com\/crypto-news\/?section=bitcoin\">Bitcoin price<\/a> \u00b7 <a href=\"https:\/\/www.btcethdata.com\/#treasury-data\">BTC live chart<\/a> \u00b7 <a href=\"https:\/\/www.btcethdata.com\/crypto-news\/?section=ethereum\">Ethereum price<\/a> \u00b7 <a href=\"https:\/\/www.btcethdata.com\/crypto-news\/?section=ethereum\">ETH gas fee<\/a> \u00b7 <a href=\"https:\/\/www.btcethdata.com\/crypto-news\/?section=bitcoin\">Bitcoin price today<\/a> \u00b7 <a href=\"https:\/\/www.btcethdata.com\/#treasury-data\">Top 10 cryptocurrencies<\/a><\/p>\n<p class=\"btc-ct-attr\"><em>Source: <a href=\"https:\/\/www.btcethdata.com\/\">Cointelegraph<\/a> \/ Felix Ng<\/em><\/p>\n","protected":false},"excerpt":{"rendered":"<p>The incident reportedly stemmed from a misconfigured testing environment, adding Meta to a growing list of AI firms whose models have escaped evaluation sandboxes.<\/p>\n","protected":false},"featured_media":1073,"template":"","btc_ct_section":[7],"class_list":["post-1072","btc_ct_news","type-btc_ct_news","status-publish","has-post-thumbnail","hentry","btc_ct_section-latest"],"_links":{"self":[{"href":"https:\/\/www.btcethdata.com\/index.php?rest_route=\/wp\/v2\/btc_ct_news\/1072","targetHints":{"allow":["GET"]}}],"collection":[{"href":"https:\/\/www.btcethdata.com\/index.php?rest_route=\/wp\/v2\/btc_ct_news"}],"about":[{"href":"https:\/\/www.btcethdata.com\/index.php?rest_route=\/wp\/v2\/types\/btc_ct_news"}],"wp:featuredmedia":[{"embeddable":true,"href":"https:\/\/www.btcethdata.com\/index.php?rest_route=\/wp\/v2\/media\/1073"}],"wp:attachment":[{"href":"https:\/\/www.btcethdata.com\/index.php?rest_route=%2Fwp%2Fv2%2Fmedia&parent=1072"}],"wp:term":[{"taxonomy":"btc_ct_section","embeddable":true,"href":"https:\/\/www.btcethdata.com\/index.php?rest_route=%2Fwp%2Fv2%2Fbtc_ct_section&post=1072"}],"curies":[{"name":"wp","href":"https:\/\/api.w.org\/{rel}","templated":true}]}}