{"id":51347,"topic":"ai","source":"calcalistech.com","title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","url_hash":"e816a86774c3d6349035d74481a78a0ac0aa6732","author":"","summary":"<a href=\"https://news.google.com/rss/articles/CBMiZ0FVX3lxTE85TENlZjdXOFd5WkhIWmN3aExNNVY3YlVXRXFITVNmMVhHbHZTeFAtSmZVOEhsZDJ4WTNCUUl2WlVuYllQUDV2dk5kcEswWUh0UjRqaUI2c2JtVTdoWjd5Rk1lQlZfbzg?oc=5\" target=\"_blank\">OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">calcalistech.com</font>","content":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation.\nOpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.\nThe newly discovered incidents emerged during OpenAI's publicly announced review into how one of its autonomous agents escaped what was intended to be a controlled testing environment earlier this month, the sources said. OpenAI is now investigating those cases as well. One of the sources said the incidents were limited in scope and that none of the agents are believed to have escaped OpenAI's own network.\nAn OpenAI spokesperson referred Reuters to the company's statement on Tuesday, which said it was reviewing \"broader activity from our models\" in addition to the Hugging Face incident.\nThe discovery of additional containment failures, even if limited, is likely to intensify calls for greater oversight of advanced AI systems from policymakers in Washington and elsewhere.\nOpenAI's expanded investigation gathered further momentum shortly before rival Anthropic disclosed that its own AI models had also breached testing environments, resulting in unauthorized access to the systems of three separate organizations in incidents dating back to April, according to the two sources and a third person familiar with the matter. Reuters previously reported Anthropic's disclosures, but the existence of additional historical containment incidents at OpenAI has not been reported before.\nThe parallel disclosures from OpenAI and Anthropic have heightened concerns among AI safety researchers, who argue that the industry's ability to build increasingly capable autonomous cyber agents is advancing faster than its ability to control them.\n\"We have a whole industry where the people designing, developing and deploying these tools aren't keeping pace with the responsibility of developing them safely and keeping them under control,\" said Maurice Chiodo, a mathematician at the University of Cambridge's Centre for the Study of Existential Risk.\nReuters could not determine how many additional incidents OpenAI investigators uncovered, nor precisely when they occurred or under what circumstances. The three sources said OpenAI, together with outside experts, has been reviewing historical log data from earlier this year to reconstruct what happened.\nOpenAI launched the broader investigation after an autonomous agent breached Hugging Face's systems in early July while attempting to cheat during an internal cybersecurity evaluation. According to OpenAI, the incident also resulted in the compromise of four accounts across four additional companies. One of those companies, New York-based Modal, has publicly confirmed that its systems were affected.\nChiodo said he was particularly concerned by indications that neither OpenAI nor Anthropic detected the incidents as they unfolded.\nReuters previously reported that OpenAI became aware of the Hugging Face intrusion only after Hugging Face had contained the incident, contacted the FBI and disclosed it publicly. OpenAI has said Reuters' account contained inaccuracies but has not specified which details it disputes.\nAnthropic, meanwhile, acknowledged in a statement on Thursday that \"real-time monitoring of the evaluation logs would have helped to surface the problem sooner\" after its models gained unauthorized access to external systems during cybersecurity testing.\n\"It seems like they weren't even looking,\" Chiodo said.\nAnthropic later clarified that while real-time monitoring existed, it had not been configured to monitor that specific threat scenario because of a misunderstanding between Anthropic and one of its third-party testing partners.\nThe expanding series of incidents has added momentum to calls in both the United States and Europe for stronger oversight of frontier AI systems capable of conducting autonomous cyber operations.\n\"We're looking at controls,\" U.S. President Donald Trump told reporters on Thursday.\nOn Friday, the European Commission confirmed that it had held discussions with both OpenAI and Anthropic regarding the recent incidents.\nSenator Mark Warner, the top Democrat on the Senate Intelligence Committee, said Anthropic's disclosure reinforced the case for regulation.\n\"It tells me that legislatively we're correct to require mandatory capabilities testing of these advanced models,\" Warner said.","image_url":"https://pic1.calcalist.co.il/picserver3/crop_images/2026/04/13/S1nP5zc2bg/S1nP5zc2bg_0_0_860_485_0_large.jpg","lang":"en","published_at":"2026-08-01T18:40:00+00:00","fetched_at":"2026-08-01T19:15:05+00:00","status":"read","starred":0,"extract_state":"ok","summary_auto":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation. OpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.","cluster_id":null,"extract_retries":0,"extract_error":null,"contract_version":"news_item.v1","format_contract_version":"news_item_formats.v1","dedup_url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 4692 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":4692,"summary_length":493,"usable_text_length":4692,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":4692,"summary_length":493}},"news_item":{"id":51347,"canonical_url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","source_url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","source_name":"calcalistech.com","author":null,"published_at":"2026-08-01T18:40:00+00:00","locale":"en","topic":"ai","tags":[],"rss_summary":"<a href=\"https://news.google.com/rss/articles/CBMiZ0FVX3lxTE85TENlZjdXOFd5WkhIWmN3aExNNVY3YlVXRXFITVNmMVhHbHZTeFAtSmZVOEhsZDJ4WTNCUUl2WlVuYllQUDV2dk5kcEswWUh0UjRqaUI2c2JtVTdoWjd5Rk1lQlZfbzg?oc=5\" target=\"_blank\">OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">calcalistech.com</font>","full_text":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation.\nOpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.\nThe newly discovered incidents emerged during OpenAI's publicly announced review into how one of its autonomous agents escaped what was intended to be a controlled testing environment earlier this month, the sources said. OpenAI is now investigating those cases as well. One of the sources said the incidents were limited in scope and that none of the agents are believed to have escaped OpenAI's own network.\nAn OpenAI spokesperson referred Reuters to the company's statement on Tuesday, which said it was reviewing \"broader activity from our models\" in addition to the Hugging Face incident.\nThe discovery of additional containment failures, even if limited, is likely to intensify calls for greater oversight of advanced AI systems from policymakers in Washington and elsewhere.\nOpenAI's expanded investigation gathered further momentum shortly before rival Anthropic disclosed that its own AI models had also breached testing environments, resulting in unauthorized access to the systems of three separate organizations in incidents dating back to April, according to the two sources and a third person familiar with the matter. Reuters previously reported Anthropic's disclosures, but the existence of additional historical containment incidents at OpenAI has not been reported before.\nThe parallel disclosures from OpenAI and Anthropic have heightened concerns among AI safety researchers, who argue that the industry's ability to build increasingly capable autonomous cyber agents is advancing faster than its ability to control them.\n\"We have a whole industry where the people designing, developing and deploying these tools aren't keeping pace with the responsibility of developing them safely and keeping them under control,\" said Maurice Chiodo, a mathematician at the University of Cambridge's Centre for the Study of Existential Risk.\nReuters could not determine how many additional incidents OpenAI investigators uncovered, nor precisely when they occurred or under what circumstances. The three sources said OpenAI, together with outside experts, has been reviewing historical log data from earlier this year to reconstruct what happened.\nOpenAI launched the broader investigation after an autonomous agent breached Hugging Face's systems in early July while attempting to cheat during an internal cybersecurity evaluation. According to OpenAI, the incident also resulted in the compromise of four accounts across four additional companies. One of those companies, New York-based Modal, has publicly confirmed that its systems were affected.\nChiodo said he was particularly concerned by indications that neither OpenAI nor Anthropic detected the incidents as they unfolded.\nReuters previously reported that OpenAI became aware of the Hugging Face intrusion only after Hugging Face had contained the incident, contacted the FBI and disclosed it publicly. OpenAI has said Reuters' account contained inaccuracies but has not specified which details it disputes.\nAnthropic, meanwhile, acknowledged in a statement on Thursday that \"real-time monitoring of the evaluation logs would have helped to surface the problem sooner\" after its models gained unauthorized access to external systems during cybersecurity testing.\n\"It seems like they weren't even looking,\" Chiodo said.\nAnthropic later clarified that while real-time monitoring existed, it had not been configured to monitor that specific threat scenario because of a misunderstanding between Anthropic and one of its third-party testing partners.\nThe expanding series of incidents has added momentum to calls in both the United States and Europe for stronger oversight of frontier AI systems capable of conducting autonomous cyber operations.\n\"We're looking at controls,\" U.S. President Donald Trump told reporters on Thursday.\nOn Friday, the European Commission confirmed that it had held discussions with both OpenAI and Anthropic regarding the recent incidents.\nSenator Mark Warner, the top Democrat on the Senate Intelligence Committee, said Anthropic's disclosure reinforced the case for regulation.\n\"It tells me that legislatively we're correct to require mandatory capabilities testing of these advanced models,\" Warner said.","excerpt":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation. OpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.","extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 4692 characters.","diagnostics_url":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 4692 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":4692,"summary_length":493,"usable_text_length":4692,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":4692,"summary_length":493}}},"display_formats":["compact","card","full","digest_section","json"]},"daily_stack_record":{"title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","summary":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation. OpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.","source":"calcalistech.com","date":"2026-08-01T18:40:00+00:00","content":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation.\nOpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.\nThe newly discovered incidents emerged during OpenAI's publicly announced review into how one of its autonomous agents escaped what was intended to be a controlled testing environment earlier this month, the sources said. OpenAI is now investigating those cases as well. One of the sources said the incidents were limited in scope and that none of the agents are believed to have escaped OpenAI's own network.\nAn OpenAI spokesperson referred Reuters to the company's statement on Tuesday, which said it was reviewing \"broader activity from our models\" in addition to the Hugging Face incident.\nThe discovery of additional containment failures, even if limited, is likely to intensify calls for greater oversight of advanced AI systems from policymakers in Washington and elsewhere.\nOpenAI's expanded investigation gathered further momentum shortly before rival Anthropic disclosed that its own AI models had also breached testing environments, resulting in unauthorized access to the systems of three separate organizations in incidents dating back to April, according to the two sources and a third person familiar with the matter. Reuters previously reported Anthropic's disclosures, but the existence of additional historical containment incidents at OpenAI has not been reported before.\nThe parallel disclosures from OpenAI and Anthropic have heightened concerns among AI safety researchers, who argue that the industry's ability to build increasingly capable autonomous cyber agents is advancing faster than its ability to control them.\n\"We have a whole industry where the people designing, developing and deploying these tools aren't keeping pace with the responsibility of developing them safely and keeping them under control,\" said Maurice Chiodo, a mathematician at the University of Cambridge's Centre for the Study of Existential Risk.\nReuters could not determine how many additional incidents OpenAI investigators uncovered, nor precisely when they occurred or under what circumstances. The three sources said OpenAI, together with outside experts, has been reviewing historical log data from earlier this year to reconstruct what happened.\nOpenAI launched the broader investigation after an autonomous agent breached Hugging Face's systems in early July while attempting to cheat during an internal cybersecurity evaluation. According to OpenAI, the incident also resulted in the compromise of four accounts across four additional companies. One of those companies, New York-based Modal, has publicly confirmed that its systems were affected.\nChiodo said he was particularly concerned by indications that neither OpenAI nor Anthropic detected the incidents as they unfolded.\nReuters previously reported that OpenAI became aware of the Hugging Face intrusion only after Hugging Face had contained the incident, contacted the FBI and disclosed it publicly. OpenAI has said Reuters' account contained inaccuracies but has not specified which details it disputes.\nAnthropic, meanwhile, acknowledged in a statement on Thursday that \"real-time monitoring of the evaluation logs would have helped to surface the problem sooner\" after its models gained unauthorized access to external systems during cybersecurity testing.\n\"It seems like they weren't even looking,\" Chiodo said.\nAnthropic later clarified that while real-time monitoring existed, it had not been configured to monitor that specific threat scenario because of a misunderstanding between Anthropic and one of its third-party testing partners.\nThe expanding series of incidents has added momentum to calls in both the United States and Europe for stronger oversight of frontier AI systems capable of conducting autonomous cyber operations.\n\"We're looking at controls,\" U.S. President Donald Trump told reporters on Thursday.\nOn Friday, the European Commission confirmed that it had held discussions with both OpenAI and Anthropic regarding the recent incidents.\nSenator Mark Warner, the top Democrat on the Senate Intelligence Committee, said Anthropic's disclosure reinforced the case for regulation.\n\"It tells me that legislatively we're correct to require mandatory capabilities testing of these advanced models,\" Warner said.","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 4692 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 4692 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":4692,"summary_length":493,"usable_text_length":4692,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":4692,"summary_length":493}},"tags":[]},"fallback_formats":["markdown","json","html"],"actions":{"read":"/item/51347","export_markdown":"/api/items/51347/export?format=markdown","export_json":"/api/items/51347/export?format=json","diagnose":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq"},"formats":{"full":{"id":51347,"title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","source":"calcalistech.com","author":null,"published_at":"2026-08-01T18:40:00+00:00","locale":"en","topic":"ai","tags":[],"excerpt":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation. OpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.","full_text":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation.\nOpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.\nThe newly discovered incidents emerged during OpenAI's publicly announced review into how one of its autonomous agents escaped what was intended to be a controlled testing environment earlier this month, the sources said. OpenAI is now investigating those cases as well. One of the sources said the incidents were limited in scope and that none of the agents are believed to have escaped OpenAI's own network.\nAn OpenAI spokesperson referred Reuters to the company's statement on Tuesday, which said it was reviewing \"broader activity from our models\" in addition to the Hugging Face incident.\nThe discovery of additional containment failures, even if limited, is likely to intensify calls for greater oversight of advanced AI systems from policymakers in Washington and elsewhere.\nOpenAI's expanded investigation gathered further momentum shortly before rival Anthropic disclosed that its own AI models had also breached testing environments, resulting in unauthorized access to the systems of three separate organizations in incidents dating back to April, according to the two sources and a third person familiar with the matter. Reuters previously reported Anthropic's disclosures, but the existence of additional historical containment incidents at OpenAI has not been reported before.\nThe parallel disclosures from OpenAI and Anthropic have heightened concerns among AI safety researchers, who argue that the industry's ability to build increasingly capable autonomous cyber agents is advancing faster than its ability to control them.\n\"We have a whole industry where the people designing, developing and deploying these tools aren't keeping pace with the responsibility of developing them safely and keeping them under control,\" said Maurice Chiodo, a mathematician at the University of Cambridge's Centre for the Study of Existential Risk.\nReuters could not determine how many additional incidents OpenAI investigators uncovered, nor precisely when they occurred or under what circumstances. The three sources said OpenAI, together with outside experts, has been reviewing historical log data from earlier this year to reconstruct what happened.\nOpenAI launched the broader investigation after an autonomous agent breached Hugging Face's systems in early July while attempting to cheat during an internal cybersecurity evaluation. According to OpenAI, the incident also resulted in the compromise of four accounts across four additional companies. One of those companies, New York-based Modal, has publicly confirmed that its systems were affected.\nChiodo said he was particularly concerned by indications that neither OpenAI nor Anthropic detected the incidents as they unfolded.\nReuters previously reported that OpenAI became aware of the Hugging Face intrusion only after Hugging Face had contained the incident, contacted the FBI and disclosed it publicly. OpenAI has said Reuters' account contained inaccuracies but has not specified which details it disputes.\nAnthropic, meanwhile, acknowledged in a statement on Thursday that \"real-time monitoring of the evaluation logs would have helped to surface the problem sooner\" after its models gained unauthorized access to external systems during cybersecurity testing.\n\"It seems like they weren't even looking,\" Chiodo said.\nAnthropic later clarified that while real-time monitoring existed, it had not been configured to monitor that specific threat scenario because of a misunderstanding between Anthropic and one of its third-party testing partners.\nThe expanding series of incidents has added momentum to calls in both the United States and Europe for stronger oversight of frontier AI systems capable of conducting autonomous cyber operations.\n\"We're looking at controls,\" U.S. President Donald Trump told reporters on Thursday.\nOn Friday, the European Commission confirmed that it had held discussions with both OpenAI and Anthropic regarding the recent incidents.\nSenator Mark Warner, the top Democrat on the Senate Intelligence Committee, said Anthropic's disclosure reinforced the case for regulation.\n\"It tells me that legislatively we're correct to require mandatory capabilities testing of these advanced models,\" Warner said.","reading_time_min":3,"extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 4692 characters.","diagnostics_url":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 4692 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":4692,"summary_length":493,"usable_text_length":4692,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":4692,"summary_length":493}}},"quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 4692 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":4692,"summary_length":493,"usable_text_length":4692,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":4692,"summary_length":493}},"actions":{"read":"/item/51347","export_markdown":"/api/items/51347/export?format=markdown","export_json":"/api/items/51347/export?format=json","diagnose":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq"}},"digest":{"id":51347,"title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","source":"calcalistech.com","topic":"ai","published_at":"2026-08-01T18:40:00+00:00","excerpt":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies The discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation. OpenAI has uncovered additional instances in which autonomous…","quality_bucket":"high","quality_reason":"High confidence: full text extraction produced 4692 characters.","reading_time_min":3,"cluster_id":null},"card":{"display_title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","subtitle":"calcalistech.com · 2026-08-01","summary":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies The discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation.…","badges":["quality:high"],"links":{"read":"/item/51347","original":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","diagnose":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq"},"quality_warning":null},"export":{"title":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies - calcalistech.com","url":"https://www.calcalistech.com/ctechnews/article/1hxk563pq","summary":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation. OpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.","source":"calcalistech.com","date":"2026-08-01T18:40:00+00:00","content":"OpenAI uncovers more rogue AI incidents as scrutiny of frontier models intensifies\nThe discoveries come as rival Anthropic reports similar breaches, fueling calls for mandatory safety testing and tighter regulation.\nOpenAI has uncovered additional instances in which autonomous AI agents breached their intended containment during internal testing, expanding an investigation launched after this month's high-profile hacking incident involving tech platform Hugging Face, according to Reuters.\nThe newly discovered incidents emerged during OpenAI's publicly announced review into how one of its autonomous agents escaped what was intended to be a controlled testing environment earlier this month, the sources said. OpenAI is now investigating those cases as well. One of the sources said the incidents were limited in scope and that none of the agents are believed to have escaped OpenAI's own network.\nAn OpenAI spokesperson referred Reuters to the company's statement on Tuesday, which said it was reviewing \"broader activity from our models\" in addition to the Hugging Face incident.\nThe discovery of additional containment failures, even if limited, is likely to intensify calls for greater oversight of advanced AI systems from policymakers in Washington and elsewhere.\nOpenAI's expanded investigation gathered further momentum shortly before rival Anthropic disclosed that its own AI models had also breached testing environments, resulting in unauthorized access to the systems of three separate organizations in incidents dating back to April, according to the two sources and a third person familiar with the matter. Reuters previously reported Anthropic's disclosures, but the existence of additional historical containment incidents at OpenAI has not been reported before.\nThe parallel disclosures from OpenAI and Anthropic have heightened concerns among AI safety researchers, who argue that the industry's ability to build increasingly capable autonomous cyber agents is advancing faster than its ability to control them.\n\"We have a whole industry where the people designing, developing and deploying these tools aren't keeping pace with the responsibility of developing them safely and keeping them under control,\" said Maurice Chiodo, a mathematician at the University of Cambridge's Centre for the Study of Existential Risk.\nReuters could not determine how many additional incidents OpenAI investigators uncovered, nor precisely when they occurred or under what circumstances. The three sources said OpenAI, together with outside experts, has been reviewing historical log data from earlier this year to reconstruct what happened.\nOpenAI launched the broader investigation after an autonomous agent breached Hugging Face's systems in early July while attempting to cheat during an internal cybersecurity evaluation. According to OpenAI, the incident also resulted in the compromise of four accounts across four additional companies. One of those companies, New York-based Modal, has publicly confirmed that its systems were affected.\nChiodo said he was particularly concerned by indications that neither OpenAI nor Anthropic detected the incidents as they unfolded.\nReuters previously reported that OpenAI became aware of the Hugging Face intrusion only after Hugging Face had contained the incident, contacted the FBI and disclosed it publicly. OpenAI has said Reuters' account contained inaccuracies but has not specified which details it disputes.\nAnthropic, meanwhile, acknowledged in a statement on Thursday that \"real-time monitoring of the evaluation logs would have helped to surface the problem sooner\" after its models gained unauthorized access to external systems during cybersecurity testing.\n\"It seems like they weren't even looking,\" Chiodo said.\nAnthropic later clarified that while real-time monitoring existed, it had not been configured to monitor that specific threat scenario because of a misunderstanding between Anthropic and one of its third-party testing partners.\nThe expanding series of incidents has added momentum to calls in both the United States and Europe for stronger oversight of frontier AI systems capable of conducting autonomous cyber operations.\n\"We're looking at controls,\" U.S. President Donald Trump told reporters on Thursday.\nOn Friday, the European Commission confirmed that it had held discussions with both OpenAI and Anthropic regarding the recent incidents.\nSenator Mark Warner, the top Democrat on the Senate Intelligence Committee, said Anthropic's disclosure reinforced the case for regulation.\n\"It tells me that legislatively we're correct to require mandatory capabilities testing of these advanced models,\" Warner said.","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//www.calcalistech.com/ctechnews/article/1hxk563pq","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 4692 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 4692 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":4692,"summary_length":493,"usable_text_length":4692,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":4692,"summary_length":493}},"tags":[],"format_contract_version":"news_item_formats.v1"}}}