{"id":45352,"topic":"ai","source":"Rensselaer Polytechnic Institute (RPI)","title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","url_hash":"0e01056248339582cf4105ecf1fdee0a2b65bba5","author":"","summary":"<a href=\"https://news.google.com/rss/articles/CBMilAFBVV95cUxPTmplZl9BNmMwY2dWUHdJVVN4Mi1qdFBHSzhRUG5BNGpJZkM0cUNaVmg4LTNDaTZITnBMNnBmN3ZtVWVGNjItMVBVRWlzZVFieWg2ZXJGNHpud3lkeXdPa3lBb0ludTRfZXFsMnVEb29XZFFQaEJKMDhsUXBONXdsWFNIbXhva0R2dl9FVFJiOWsxNFZ5?oc=5\" target=\"_blank\">AI Growing Pains Reach Scientific Labs, New RPI Study Finds</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">Rensselaer Polytechnic Institute (RPI)</font>","content":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.\nThe paper, authored by George I. Makhatadze, professor of biological sciences and Constellation Endowed Chair at RPI, evaluated widely used deep learning tools for predicting how flat sequences of amino acids fold into the three-dimensional structures that determine a protein's function. Makhatadze found that these tools frequently overlook the underlying scientific rules of protein folding – and, notably, that every tool tested rated its own accuracy higher than the results warranted.\n“The major conclusion of the paper essentially is: trust but verify,” Makhatadze explained. “You have to verify [AI outputs] using physics-based methods.”\nAI has become indispensable for analyzing the massive data sets used to predict protein folds. For example, Google’s DeepMind AI laboratory – called AlphaFold2 – shared the 2024 Nobel Prize in Chemistry for its contributions to protein structure prediction.\nBut according to Makhatadze's work, AlphaFold2 and RoseTTAFold2 — a similar deep learning-based prediction platform developed at the University of Washington — both produced “implausible structures for variant sequences” by “[prioritizing] statistical patterns over the underlying thermodynamic principles of folding.” Both tools are trained on evolutionary data and structural databases.\n“AlphaFold is considered the gospel of the field,” Makhatadze said. “It is very good, and it does many things well. But occasionally it makes mistakes, because there simply isn’t enough of the right kind of data in the model yet.”\nMakhatadze found fewer scientific impossibilities in a different class of tools – “transformer-based protein language models” that rely on protein sequences rather than structural data. Tools in this class included OmegaFold and the Meta-developed ESMFold. However, neither category of model performed well when proteins contained ionizable residues, meaning amino acid side chains that can gain or lose a proton depending on their surrounding environment.\nAccording to Makhatadze, current predictive AI tools simply haven’t been trained to account for ionizable residues, sometimes leading to scientifically impossible predictions.\nTo address these shortcomings, Makhatadze recommends pairing AI-generated predictions with computer simulations of molecular dynamics to validate protein structures. That combination, the paper notes, “strengthens confidence in AI-generated structures and underscores the need to couple machine learning with physics-based refinement.”\nWith the paper’s publication, Makhatadze expects the developers of AI predictive folding models to address this blind spot in existing platforms by incorporating additional levels of physicochemical validation. “It has to be a combination of pattern recognition and physics,” he said.\nIn the meantime, Makhatadze offers one piece of advice for AI users in scientific labs and beyond: “You cannot blindly believe everything the model predicts.”","image_url":null,"lang":"en","published_at":"2026-07-24T13:29:43+00:00","fetched_at":"2026-07-24T14:15:03+00:00","status":"read","starred":0,"extract_state":"ok","summary_auto":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.","cluster_id":1713350,"extract_retries":0,"extract_error":null,"contract_version":"news_item.v1","format_contract_version":"news_item_formats.v1","dedup_url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3484 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3484,"summary_length":547,"usable_text_length":3484,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3484,"summary_length":547}},"news_item":{"id":45352,"canonical_url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","source_url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","source_name":"Rensselaer Polytechnic Institute (RPI)","author":null,"published_at":"2026-07-24T13:29:43+00:00","locale":"en","topic":"ai","tags":[],"rss_summary":"<a href=\"https://news.google.com/rss/articles/CBMilAFBVV95cUxPTmplZl9BNmMwY2dWUHdJVVN4Mi1qdFBHSzhRUG5BNGpJZkM0cUNaVmg4LTNDaTZITnBMNnBmN3ZtVWVGNjItMVBVRWlzZVFieWg2ZXJGNHpud3lkeXdPa3lBb0ludTRfZXFsMnVEb29XZFFQaEJKMDhsUXBONXdsWFNIbXhva0R2dl9FVFJiOWsxNFZ5?oc=5\" target=\"_blank\">AI Growing Pains Reach Scientific Labs, New RPI Study Finds</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">Rensselaer Polytechnic Institute (RPI)</font>","full_text":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.\nThe paper, authored by George I. Makhatadze, professor of biological sciences and Constellation Endowed Chair at RPI, evaluated widely used deep learning tools for predicting how flat sequences of amino acids fold into the three-dimensional structures that determine a protein's function. Makhatadze found that these tools frequently overlook the underlying scientific rules of protein folding – and, notably, that every tool tested rated its own accuracy higher than the results warranted.\n“The major conclusion of the paper essentially is: trust but verify,” Makhatadze explained. “You have to verify [AI outputs] using physics-based methods.”\nAI has become indispensable for analyzing the massive data sets used to predict protein folds. For example, Google’s DeepMind AI laboratory – called AlphaFold2 – shared the 2024 Nobel Prize in Chemistry for its contributions to protein structure prediction.\nBut according to Makhatadze's work, AlphaFold2 and RoseTTAFold2 — a similar deep learning-based prediction platform developed at the University of Washington — both produced “implausible structures for variant sequences” by “[prioritizing] statistical patterns over the underlying thermodynamic principles of folding.” Both tools are trained on evolutionary data and structural databases.\n“AlphaFold is considered the gospel of the field,” Makhatadze said. “It is very good, and it does many things well. But occasionally it makes mistakes, because there simply isn’t enough of the right kind of data in the model yet.”\nMakhatadze found fewer scientific impossibilities in a different class of tools – “transformer-based protein language models” that rely on protein sequences rather than structural data. Tools in this class included OmegaFold and the Meta-developed ESMFold. However, neither category of model performed well when proteins contained ionizable residues, meaning amino acid side chains that can gain or lose a proton depending on their surrounding environment.\nAccording to Makhatadze, current predictive AI tools simply haven’t been trained to account for ionizable residues, sometimes leading to scientifically impossible predictions.\nTo address these shortcomings, Makhatadze recommends pairing AI-generated predictions with computer simulations of molecular dynamics to validate protein structures. That combination, the paper notes, “strengthens confidence in AI-generated structures and underscores the need to couple machine learning with physics-based refinement.”\nWith the paper’s publication, Makhatadze expects the developers of AI predictive folding models to address this blind spot in existing platforms by incorporating additional levels of physicochemical validation. “It has to be a combination of pattern recognition and physics,” he said.\nIn the meantime, Makhatadze offers one piece of advice for AI users in scientific labs and beyond: “You cannot blindly believe everything the model predicts.”","excerpt":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.","extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 3484 characters.","diagnostics_url":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3484 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3484,"summary_length":547,"usable_text_length":3484,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3484,"summary_length":547}}},"display_formats":["compact","card","full","digest_section","json"]},"daily_stack_record":{"title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","summary":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.","source":"Rensselaer Polytechnic Institute (RPI)","date":"2026-07-24T13:29:43+00:00","content":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.\nThe paper, authored by George I. Makhatadze, professor of biological sciences and Constellation Endowed Chair at RPI, evaluated widely used deep learning tools for predicting how flat sequences of amino acids fold into the three-dimensional structures that determine a protein's function. Makhatadze found that these tools frequently overlook the underlying scientific rules of protein folding – and, notably, that every tool tested rated its own accuracy higher than the results warranted.\n“The major conclusion of the paper essentially is: trust but verify,” Makhatadze explained. “You have to verify [AI outputs] using physics-based methods.”\nAI has become indispensable for analyzing the massive data sets used to predict protein folds. For example, Google’s DeepMind AI laboratory – called AlphaFold2 – shared the 2024 Nobel Prize in Chemistry for its contributions to protein structure prediction.\nBut according to Makhatadze's work, AlphaFold2 and RoseTTAFold2 — a similar deep learning-based prediction platform developed at the University of Washington — both produced “implausible structures for variant sequences” by “[prioritizing] statistical patterns over the underlying thermodynamic principles of folding.” Both tools are trained on evolutionary data and structural databases.\n“AlphaFold is considered the gospel of the field,” Makhatadze said. “It is very good, and it does many things well. But occasionally it makes mistakes, because there simply isn’t enough of the right kind of data in the model yet.”\nMakhatadze found fewer scientific impossibilities in a different class of tools – “transformer-based protein language models” that rely on protein sequences rather than structural data. Tools in this class included OmegaFold and the Meta-developed ESMFold. However, neither category of model performed well when proteins contained ionizable residues, meaning amino acid side chains that can gain or lose a proton depending on their surrounding environment.\nAccording to Makhatadze, current predictive AI tools simply haven’t been trained to account for ionizable residues, sometimes leading to scientifically impossible predictions.\nTo address these shortcomings, Makhatadze recommends pairing AI-generated predictions with computer simulations of molecular dynamics to validate protein structures. That combination, the paper notes, “strengthens confidence in AI-generated structures and underscores the need to couple machine learning with physics-based refinement.”\nWith the paper’s publication, Makhatadze expects the developers of AI predictive folding models to address this blind spot in existing platforms by incorporating additional levels of physicochemical validation. “It has to be a combination of pattern recognition and physics,” he said.\nIn the meantime, Makhatadze offers one piece of advice for AI users in scientific labs and beyond: “You cannot blindly believe everything the model predicts.”","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 3484 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3484 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3484,"summary_length":547,"usable_text_length":3484,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3484,"summary_length":547}},"tags":[]},"fallback_formats":["markdown","json","html"],"actions":{"read":"/item/45352","export_markdown":"/api/items/45352/export?format=markdown","export_json":"/api/items/45352/export?format=json","diagnose":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds"},"formats":{"full":{"id":45352,"title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","source":"Rensselaer Polytechnic Institute (RPI)","author":null,"published_at":"2026-07-24T13:29:43+00:00","locale":"en","topic":"ai","tags":[],"excerpt":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.","full_text":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.\nThe paper, authored by George I. Makhatadze, professor of biological sciences and Constellation Endowed Chair at RPI, evaluated widely used deep learning tools for predicting how flat sequences of amino acids fold into the three-dimensional structures that determine a protein's function. Makhatadze found that these tools frequently overlook the underlying scientific rules of protein folding – and, notably, that every tool tested rated its own accuracy higher than the results warranted.\n“The major conclusion of the paper essentially is: trust but verify,” Makhatadze explained. “You have to verify [AI outputs] using physics-based methods.”\nAI has become indispensable for analyzing the massive data sets used to predict protein folds. For example, Google’s DeepMind AI laboratory – called AlphaFold2 – shared the 2024 Nobel Prize in Chemistry for its contributions to protein structure prediction.\nBut according to Makhatadze's work, AlphaFold2 and RoseTTAFold2 — a similar deep learning-based prediction platform developed at the University of Washington — both produced “implausible structures for variant sequences” by “[prioritizing] statistical patterns over the underlying thermodynamic principles of folding.” Both tools are trained on evolutionary data and structural databases.\n“AlphaFold is considered the gospel of the field,” Makhatadze said. “It is very good, and it does many things well. But occasionally it makes mistakes, because there simply isn’t enough of the right kind of data in the model yet.”\nMakhatadze found fewer scientific impossibilities in a different class of tools – “transformer-based protein language models” that rely on protein sequences rather than structural data. Tools in this class included OmegaFold and the Meta-developed ESMFold. However, neither category of model performed well when proteins contained ionizable residues, meaning amino acid side chains that can gain or lose a proton depending on their surrounding environment.\nAccording to Makhatadze, current predictive AI tools simply haven’t been trained to account for ionizable residues, sometimes leading to scientifically impossible predictions.\nTo address these shortcomings, Makhatadze recommends pairing AI-generated predictions with computer simulations of molecular dynamics to validate protein structures. That combination, the paper notes, “strengthens confidence in AI-generated structures and underscores the need to couple machine learning with physics-based refinement.”\nWith the paper’s publication, Makhatadze expects the developers of AI predictive folding models to address this blind spot in existing platforms by incorporating additional levels of physicochemical validation. “It has to be a combination of pattern recognition and physics,” he said.\nIn the meantime, Makhatadze offers one piece of advice for AI users in scientific labs and beyond: “You cannot blindly believe everything the model predicts.”","reading_time_min":2,"extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 3484 characters.","diagnostics_url":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3484 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3484,"summary_length":547,"usable_text_length":3484,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3484,"summary_length":547}}},"quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3484 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3484,"summary_length":547,"usable_text_length":3484,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3484,"summary_length":547}},"actions":{"read":"/item/45352","export_markdown":"/api/items/45352/export?format=markdown","export_json":"/api/items/45352/export?format=json","diagnose":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds"}},"digest":{"id":45352,"title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","source":"Rensselaer Polytechnic Institute (RPI)","topic":"ai","published_at":"2026-07-24T13:29:43+00:00","excerpt":"July 24, 2026 Researchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI…","quality_bucket":"high","quality_reason":"High confidence: full text extraction produced 3484 characters.","reading_time_min":2,"cluster_id":1713350},"card":{"display_title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","subtitle":"Rensselaer Polytechnic Institute (RPI) · 2026-07-24","summary":"July 24, 2026 Researchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and…","badges":["quality:high"],"links":{"read":"/item/45352","original":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","diagnose":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds"},"quality_warning":null},"export":{"title":"AI Growing Pains Reach Scientific Labs, New RPI Study Finds - Rensselaer Polytechnic Institute (RPI)","url":"https://news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","summary":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.","source":"Rensselaer Polytechnic Institute (RPI)","date":"2026-07-24T13:29:43+00:00","content":"July 24, 2026\nResearchers at Rensselaer Polytechnic Institute (RPI) have found that today's leading artificial intelligence tools for predicting protein structures routinely generate results that are physically and chemically impossible, exposing critical blind spots in how AI is being applied across scientific research. The work, published in the Proceedings of the National Academy of Sciences (PNAS), serves as a cautionary reminder that AI still requires human oversight and physics-based verification to produce reliable results in the lab.\nThe paper, authored by George I. Makhatadze, professor of biological sciences and Constellation Endowed Chair at RPI, evaluated widely used deep learning tools for predicting how flat sequences of amino acids fold into the three-dimensional structures that determine a protein's function. Makhatadze found that these tools frequently overlook the underlying scientific rules of protein folding – and, notably, that every tool tested rated its own accuracy higher than the results warranted.\n“The major conclusion of the paper essentially is: trust but verify,” Makhatadze explained. “You have to verify [AI outputs] using physics-based methods.”\nAI has become indispensable for analyzing the massive data sets used to predict protein folds. For example, Google’s DeepMind AI laboratory – called AlphaFold2 – shared the 2024 Nobel Prize in Chemistry for its contributions to protein structure prediction.\nBut according to Makhatadze's work, AlphaFold2 and RoseTTAFold2 — a similar deep learning-based prediction platform developed at the University of Washington — both produced “implausible structures for variant sequences” by “[prioritizing] statistical patterns over the underlying thermodynamic principles of folding.” Both tools are trained on evolutionary data and structural databases.\n“AlphaFold is considered the gospel of the field,” Makhatadze said. “It is very good, and it does many things well. But occasionally it makes mistakes, because there simply isn’t enough of the right kind of data in the model yet.”\nMakhatadze found fewer scientific impossibilities in a different class of tools – “transformer-based protein language models” that rely on protein sequences rather than structural data. Tools in this class included OmegaFold and the Meta-developed ESMFold. However, neither category of model performed well when proteins contained ionizable residues, meaning amino acid side chains that can gain or lose a proton depending on their surrounding environment.\nAccording to Makhatadze, current predictive AI tools simply haven’t been trained to account for ionizable residues, sometimes leading to scientifically impossible predictions.\nTo address these shortcomings, Makhatadze recommends pairing AI-generated predictions with computer simulations of molecular dynamics to validate protein structures. That combination, the paper notes, “strengthens confidence in AI-generated structures and underscores the need to couple machine learning with physics-based refinement.”\nWith the paper’s publication, Makhatadze expects the developers of AI predictive folding models to address this blind spot in existing platforms by incorporating additional levels of physicochemical validation. “It has to be a combination of pattern recognition and physics,” he said.\nIn the meantime, Makhatadze offers one piece of advice for AI users in scientific labs and beyond: “You cannot blindly believe everything the model predicts.”","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//news.rpi.edu/2026/07/24/ai-growing-pains-reach-scientific-labs-new-rpi-study-finds","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 3484 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3484 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3484,"summary_length":547,"usable_text_length":3484,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3484,"summary_length":547}},"tags":[],"format_contract_version":"news_item_formats.v1"}}}