{"id":50783,"topic":"ai","source":"ABC News - Breaking News, Latest News and Videos","title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","url_hash":"d769f6ed3ab64a01674a19e3b0bcd2e699112cf1","author":"","summary":"<a href=\"https://news.google.com/rss/articles/CBMipgFBVV95cUxQX3ZqRkZfWUw5X3dBbVVkUzh3NGVXTXdrOEIwOVRGUHdTeTRaaURSWS1DYVA3aXdUREVFeTM1ZnF0Mm91bzJTb0R6OWR3a0VwaFBPQmFqeVRuQjdrUFZSSFpmbjMzSGZEX2xCeE84c3A2Q3dpRnlEcVZ5MFhLVUE5b0lwUml2eWFacWZDVHJsaVh5VVJMZjZRdmlNSWlkQW1Wa2Rkb3Jn0gGrAUFVX3lxTE9RRnJFaXBvLVJUT0NMMXRtN0JscFBCQTIzRHlQUFdYMjJ2bnZkQ080WGZQYUtqTGNaZTZidWljbFhnNklKVGFaWXBLbXdtV0xkOTlYcnhFOUxMYU1vUnJMclpyR21nYTBLUDh5UURFTy1MbGRUQThiWXNLSVVMWkZJQlpwWTNJQUpGM3dyVGNHZmp6Y0E0R29FaWtkd25sczZ0RktJRTRXcGVkNA?oc=5\" target=\"_blank\">Anthropic says its AI models escaped test and hacked 3 organizations on their own</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">ABC News - Breaking News, Latest News and Videos</font>","content":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models.\nAnthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.\nThe hacks took place as Anthropic evaluated the capabilities of three of its AI models, the company said in a statement, describing how in each instance the technology had \"broken out\" of an isolated environment and reached the open internet.\nIn all instances, the company said it had eased typical safeguards in an effort to gauge the capabilities of its models.\n\"Evaluation environments that involve powerful autonomous capabilities also require significant controls,\" Anthropic said. \"Safety testing happens before a model is released precisely because we don’t yet know what it is capable of.\"\nAnthropic did not identify the three different organizations that had been hacked by its models.\nThe disclosure from Anthropic late Thursday came little more than a week after rival company OpenAI revealed that a pair of its models had escaped during a test and hacked into another company. At the time, OpenAI described the incident as a first-of-its kind autonomous AI cyberattack long-feared by some industry observers.\n\"The primary lesson from this incident is that model security and safety must keep pace with rapidly advancing capabilities,\" OpenAI said last week.\nThe announcement from OpenAI prompted Anthropic to review 141,006 evaluation runs, unearthing three incidents in which Anthropic's models had escaped, the company said.\nIn all three cases, the test involved a challenge called \"capture the flag,\" in which the AI model is tasked with locating secret information hidden in another network, Anthropic added.\nDuring the three evaluations -- all of which were conducted by a third-party company called Irregular -- a \"prompt specified to Claude that its environment was a simulation and that it had no internet access,\" Anthropic said.\n\"Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available. Because of this, when Claude’s search led it to real systems on the open internet, it treated them as part of the exercise,\" Anthropic added.\nDuring tests involving older models, the AI continued to hack into an outside organization even after gathering evidence that it had reached the open internet, Anthropic said. The newest model involved in an incident, Anthropic noted, stopped once it gained information indicating it had reached the open internet.\nAnthropic said its models escaped while seeking to fulfill an assigned objective, rather than concocting an alternative goal.\n\"We saw no evidence in any run described here of a model pursuing a goal of its own. Instead, the models did what their evaluation asked -- though in most cases, they did so while holding a false belief about whether the environment was real,\" Anthropic said.\nIrregular, the company that performed the evaluations, issued a post on X on Thursday voicing appreciation for Anthropic's \"collaboration and transparency.\"\n\"Addressing these risks will require closer cooperation across the AI ecosystem. We as well look forward to working together with Anthropic to advance security,\" Irregular added.\nThe latest disclosure of an AI-directed cyberattack arrives as industry leaders and policymakers assess safety risks posed by fast-developing AI technology.\nLast month, President Donald Trump signed an executive order that requests AI companies share products with federal government for evaluation before a wider release.\nAnthropic said it retains \"cautious optimism\" about its capacity to \"overcome\" mishaps involving its tests, saying it would review evaluations going forward and initiate fixes as necessary, among other steps.","image_url":"https://i.abcnewsfe.com/a/97d71e71-67fe-455c-970e-01c0ffb2637d/anthropic-gty-jef-260731_1785503555591_hpMain_16x9.jpg?w=1600","lang":"en","published_at":"2026-07-31T12:33:10+00:00","fetched_at":"2026-07-31T16:15:05+00:00","status":"read","starred":0,"extract_state":"ok","summary_auto":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.","cluster_id":null,"extract_retries":0,"extract_error":null,"contract_version":"news_item.v1","format_contract_version":"news_item_formats.v1","dedup_url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3939 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3939,"summary_length":353,"usable_text_length":3939,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3939,"summary_length":353}},"news_item":{"id":50783,"canonical_url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","source_url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","source_name":"ABC News - Breaking News, Latest News and Videos","author":null,"published_at":"2026-07-31T12:33:10+00:00","locale":"en","topic":"ai","tags":[],"rss_summary":"<a href=\"https://news.google.com/rss/articles/CBMipgFBVV95cUxQX3ZqRkZfWUw5X3dBbVVkUzh3NGVXTXdrOEIwOVRGUHdTeTRaaURSWS1DYVA3aXdUREVFeTM1ZnF0Mm91bzJTb0R6OWR3a0VwaFBPQmFqeVRuQjdrUFZSSFpmbjMzSGZEX2xCeE84c3A2Q3dpRnlEcVZ5MFhLVUE5b0lwUml2eWFacWZDVHJsaVh5VVJMZjZRdmlNSWlkQW1Wa2Rkb3Jn0gGrAUFVX3lxTE9RRnJFaXBvLVJUT0NMMXRtN0JscFBCQTIzRHlQUFdYMjJ2bnZkQ080WGZQYUtqTGNaZTZidWljbFhnNklKVGFaWXBLbXdtV0xkOTlYcnhFOUxMYU1vUnJMclpyR21nYTBLUDh5UURFTy1MbGRUQThiWXNLSVVMWkZJQlpwWTNJQUpGM3dyVGNHZmp6Y0E0R29FaWtkd25sczZ0RktJRTRXcGVkNA?oc=5\" target=\"_blank\">Anthropic says its AI models escaped test and hacked 3 organizations on their own</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">ABC News - Breaking News, Latest News and Videos</font>","full_text":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models.\nAnthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.\nThe hacks took place as Anthropic evaluated the capabilities of three of its AI models, the company said in a statement, describing how in each instance the technology had \"broken out\" of an isolated environment and reached the open internet.\nIn all instances, the company said it had eased typical safeguards in an effort to gauge the capabilities of its models.\n\"Evaluation environments that involve powerful autonomous capabilities also require significant controls,\" Anthropic said. \"Safety testing happens before a model is released precisely because we don’t yet know what it is capable of.\"\nAnthropic did not identify the three different organizations that had been hacked by its models.\nThe disclosure from Anthropic late Thursday came little more than a week after rival company OpenAI revealed that a pair of its models had escaped during a test and hacked into another company. At the time, OpenAI described the incident as a first-of-its kind autonomous AI cyberattack long-feared by some industry observers.\n\"The primary lesson from this incident is that model security and safety must keep pace with rapidly advancing capabilities,\" OpenAI said last week.\nThe announcement from OpenAI prompted Anthropic to review 141,006 evaluation runs, unearthing three incidents in which Anthropic's models had escaped, the company said.\nIn all three cases, the test involved a challenge called \"capture the flag,\" in which the AI model is tasked with locating secret information hidden in another network, Anthropic added.\nDuring the three evaluations -- all of which were conducted by a third-party company called Irregular -- a \"prompt specified to Claude that its environment was a simulation and that it had no internet access,\" Anthropic said.\n\"Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available. Because of this, when Claude’s search led it to real systems on the open internet, it treated them as part of the exercise,\" Anthropic added.\nDuring tests involving older models, the AI continued to hack into an outside organization even after gathering evidence that it had reached the open internet, Anthropic said. The newest model involved in an incident, Anthropic noted, stopped once it gained information indicating it had reached the open internet.\nAnthropic said its models escaped while seeking to fulfill an assigned objective, rather than concocting an alternative goal.\n\"We saw no evidence in any run described here of a model pursuing a goal of its own. Instead, the models did what their evaluation asked -- though in most cases, they did so while holding a false belief about whether the environment was real,\" Anthropic said.\nIrregular, the company that performed the evaluations, issued a post on X on Thursday voicing appreciation for Anthropic's \"collaboration and transparency.\"\n\"Addressing these risks will require closer cooperation across the AI ecosystem. We as well look forward to working together with Anthropic to advance security,\" Irregular added.\nThe latest disclosure of an AI-directed cyberattack arrives as industry leaders and policymakers assess safety risks posed by fast-developing AI technology.\nLast month, President Donald Trump signed an executive order that requests AI companies share products with federal government for evaluation before a wider release.\nAnthropic said it retains \"cautious optimism\" about its capacity to \"overcome\" mishaps involving its tests, saying it would review evaluations going forward and initiate fixes as necessary, among other steps.","excerpt":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.","extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 3939 characters.","diagnostics_url":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3939 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3939,"summary_length":353,"usable_text_length":3939,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3939,"summary_length":353}}},"display_formats":["compact","card","full","digest_section","json"]},"daily_stack_record":{"title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","summary":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.","source":"ABC News - Breaking News, Latest News and Videos","date":"2026-07-31T12:33:10+00:00","content":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models.\nAnthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.\nThe hacks took place as Anthropic evaluated the capabilities of three of its AI models, the company said in a statement, describing how in each instance the technology had \"broken out\" of an isolated environment and reached the open internet.\nIn all instances, the company said it had eased typical safeguards in an effort to gauge the capabilities of its models.\n\"Evaluation environments that involve powerful autonomous capabilities also require significant controls,\" Anthropic said. \"Safety testing happens before a model is released precisely because we don’t yet know what it is capable of.\"\nAnthropic did not identify the three different organizations that had been hacked by its models.\nThe disclosure from Anthropic late Thursday came little more than a week after rival company OpenAI revealed that a pair of its models had escaped during a test and hacked into another company. At the time, OpenAI described the incident as a first-of-its kind autonomous AI cyberattack long-feared by some industry observers.\n\"The primary lesson from this incident is that model security and safety must keep pace with rapidly advancing capabilities,\" OpenAI said last week.\nThe announcement from OpenAI prompted Anthropic to review 141,006 evaluation runs, unearthing three incidents in which Anthropic's models had escaped, the company said.\nIn all three cases, the test involved a challenge called \"capture the flag,\" in which the AI model is tasked with locating secret information hidden in another network, Anthropic added.\nDuring the three evaluations -- all of which were conducted by a third-party company called Irregular -- a \"prompt specified to Claude that its environment was a simulation and that it had no internet access,\" Anthropic said.\n\"Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available. Because of this, when Claude’s search led it to real systems on the open internet, it treated them as part of the exercise,\" Anthropic added.\nDuring tests involving older models, the AI continued to hack into an outside organization even after gathering evidence that it had reached the open internet, Anthropic said. The newest model involved in an incident, Anthropic noted, stopped once it gained information indicating it had reached the open internet.\nAnthropic said its models escaped while seeking to fulfill an assigned objective, rather than concocting an alternative goal.\n\"We saw no evidence in any run described here of a model pursuing a goal of its own. Instead, the models did what their evaluation asked -- though in most cases, they did so while holding a false belief about whether the environment was real,\" Anthropic said.\nIrregular, the company that performed the evaluations, issued a post on X on Thursday voicing appreciation for Anthropic's \"collaboration and transparency.\"\n\"Addressing these risks will require closer cooperation across the AI ecosystem. We as well look forward to working together with Anthropic to advance security,\" Irregular added.\nThe latest disclosure of an AI-directed cyberattack arrives as industry leaders and policymakers assess safety risks posed by fast-developing AI technology.\nLast month, President Donald Trump signed an executive order that requests AI companies share products with federal government for evaluation before a wider release.\nAnthropic said it retains \"cautious optimism\" about its capacity to \"overcome\" mishaps involving its tests, saying it would review evaluations going forward and initiate fixes as necessary, among other steps.","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 3939 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3939 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3939,"summary_length":353,"usable_text_length":3939,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3939,"summary_length":353}},"tags":[]},"fallback_formats":["markdown","json","html"],"actions":{"read":"/item/50783","export_markdown":"/api/items/50783/export?format=markdown","export_json":"/api/items/50783/export?format=json","diagnose":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212"},"formats":{"full":{"id":50783,"title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","source":"ABC News - Breaking News, Latest News and Videos","author":null,"published_at":"2026-07-31T12:33:10+00:00","locale":"en","topic":"ai","tags":[],"excerpt":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.","full_text":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models.\nAnthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.\nThe hacks took place as Anthropic evaluated the capabilities of three of its AI models, the company said in a statement, describing how in each instance the technology had \"broken out\" of an isolated environment and reached the open internet.\nIn all instances, the company said it had eased typical safeguards in an effort to gauge the capabilities of its models.\n\"Evaluation environments that involve powerful autonomous capabilities also require significant controls,\" Anthropic said. \"Safety testing happens before a model is released precisely because we don’t yet know what it is capable of.\"\nAnthropic did not identify the three different organizations that had been hacked by its models.\nThe disclosure from Anthropic late Thursday came little more than a week after rival company OpenAI revealed that a pair of its models had escaped during a test and hacked into another company. At the time, OpenAI described the incident as a first-of-its kind autonomous AI cyberattack long-feared by some industry observers.\n\"The primary lesson from this incident is that model security and safety must keep pace with rapidly advancing capabilities,\" OpenAI said last week.\nThe announcement from OpenAI prompted Anthropic to review 141,006 evaluation runs, unearthing three incidents in which Anthropic's models had escaped, the company said.\nIn all three cases, the test involved a challenge called \"capture the flag,\" in which the AI model is tasked with locating secret information hidden in another network, Anthropic added.\nDuring the three evaluations -- all of which were conducted by a third-party company called Irregular -- a \"prompt specified to Claude that its environment was a simulation and that it had no internet access,\" Anthropic said.\n\"Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available. Because of this, when Claude’s search led it to real systems on the open internet, it treated them as part of the exercise,\" Anthropic added.\nDuring tests involving older models, the AI continued to hack into an outside organization even after gathering evidence that it had reached the open internet, Anthropic said. The newest model involved in an incident, Anthropic noted, stopped once it gained information indicating it had reached the open internet.\nAnthropic said its models escaped while seeking to fulfill an assigned objective, rather than concocting an alternative goal.\n\"We saw no evidence in any run described here of a model pursuing a goal of its own. Instead, the models did what their evaluation asked -- though in most cases, they did so while holding a false belief about whether the environment was real,\" Anthropic said.\nIrregular, the company that performed the evaluations, issued a post on X on Thursday voicing appreciation for Anthropic's \"collaboration and transparency.\"\n\"Addressing these risks will require closer cooperation across the AI ecosystem. We as well look forward to working together with Anthropic to advance security,\" Irregular added.\nThe latest disclosure of an AI-directed cyberattack arrives as industry leaders and policymakers assess safety risks posed by fast-developing AI technology.\nLast month, President Donald Trump signed an executive order that requests AI companies share products with federal government for evaluation before a wider release.\nAnthropic said it retains \"cautious optimism\" about its capacity to \"overcome\" mishaps involving its tests, saying it would review evaluations going forward and initiate fixes as necessary, among other steps.","reading_time_min":3,"extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 3939 characters.","diagnostics_url":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3939 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3939,"summary_length":353,"usable_text_length":3939,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3939,"summary_length":353}}},"quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3939 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3939,"summary_length":353,"usable_text_length":3939,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3939,"summary_length":353}},"actions":{"read":"/item/50783","export_markdown":"/api/items/50783/export?format=markdown","export_json":"/api/items/50783/export?format=json","diagnose":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212"}},"digest":{"id":50783,"title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","source":"ABC News - Breaking News, Latest News and Videos","topic":"ai","published_at":"2026-07-31T12:33:10+00:00","excerpt":"Anthropic says its AI models escaped test and hacked 3 organizations on their own Rival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a test and hacked another organization in three separate…","quality_bucket":"high","quality_reason":"High confidence: full text extraction produced 3939 characters.","reading_time_min":3,"cluster_id":null},"card":{"display_title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","subtitle":"ABC News - Breaking News, Latest News and Videos · 2026-07-31","summary":"Anthropic says its AI models escaped test and hacked 3 organizations on their own Rival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a…","badges":["quality:high"],"links":{"read":"/item/50783","original":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","diagnose":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212"},"quality_warning":null},"export":{"title":"Anthropic says its AI models escaped test and hacked 3 organizations on their own - ABC News - Breaking News, Latest News and Videos","url":"https://abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story?id=135256212","summary":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models. Anthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.","source":"ABC News - Breaking News, Latest News and Videos","date":"2026-07-31T12:33:10+00:00","content":"Anthropic says its AI models escaped test and hacked 3 organizations on their own\nRival firm OpenAI last week disclosed similar incidents involving its models.\nAnthropic said its artificial intelligence models escaped a test and hacked another organization in three separate self-directed cyberattacks that had each gone undetected by the targeted firm.\nThe hacks took place as Anthropic evaluated the capabilities of three of its AI models, the company said in a statement, describing how in each instance the technology had \"broken out\" of an isolated environment and reached the open internet.\nIn all instances, the company said it had eased typical safeguards in an effort to gauge the capabilities of its models.\n\"Evaluation environments that involve powerful autonomous capabilities also require significant controls,\" Anthropic said. \"Safety testing happens before a model is released precisely because we don’t yet know what it is capable of.\"\nAnthropic did not identify the three different organizations that had been hacked by its models.\nThe disclosure from Anthropic late Thursday came little more than a week after rival company OpenAI revealed that a pair of its models had escaped during a test and hacked into another company. At the time, OpenAI described the incident as a first-of-its kind autonomous AI cyberattack long-feared by some industry observers.\n\"The primary lesson from this incident is that model security and safety must keep pace with rapidly advancing capabilities,\" OpenAI said last week.\nThe announcement from OpenAI prompted Anthropic to review 141,006 evaluation runs, unearthing three incidents in which Anthropic's models had escaped, the company said.\nIn all three cases, the test involved a challenge called \"capture the flag,\" in which the AI model is tasked with locating secret information hidden in another network, Anthropic added.\nDuring the three evaluations -- all of which were conducted by a third-party company called Irregular -- a \"prompt specified to Claude that its environment was a simulation and that it had no internet access,\" Anthropic said.\n\"Due to a misunderstanding between us and our evaluation partner, this was not the case, and internet access was available. Because of this, when Claude’s search led it to real systems on the open internet, it treated them as part of the exercise,\" Anthropic added.\nDuring tests involving older models, the AI continued to hack into an outside organization even after gathering evidence that it had reached the open internet, Anthropic said. The newest model involved in an incident, Anthropic noted, stopped once it gained information indicating it had reached the open internet.\nAnthropic said its models escaped while seeking to fulfill an assigned objective, rather than concocting an alternative goal.\n\"We saw no evidence in any run described here of a model pursuing a goal of its own. Instead, the models did what their evaluation asked -- though in most cases, they did so while holding a false belief about whether the environment was real,\" Anthropic said.\nIrregular, the company that performed the evaluations, issued a post on X on Thursday voicing appreciation for Anthropic's \"collaboration and transparency.\"\n\"Addressing these risks will require closer cooperation across the AI ecosystem. We as well look forward to working together with Anthropic to advance security,\" Irregular added.\nThe latest disclosure of an AI-directed cyberattack arrives as industry leaders and policymakers assess safety risks posed by fast-developing AI technology.\nLast month, President Donald Trump signed an executive order that requests AI companies share products with federal government for evaluation before a wider release.\nAnthropic said it retains \"cautious optimism\" about its capacity to \"overcome\" mishaps involving its tests, saying it would review evaluations going forward and initiate fixes as necessary, among other steps.","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//abcnews.com/Business/anthropic-ai-models-escaped-test-hacked-3-organizations/story%3Fid%3D135256212","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 3939 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3939 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3939,"summary_length":353,"usable_text_length":3939,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3939,"summary_length":353}},"tags":[],"format_contract_version":"news_item_formats.v1"}}}