{"id":44464,"topic":"ai","source":"The Conversation","title":"What is open-source AI? A software engineering researcher explains - The Conversation","url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","url_hash":"63a106a21a56ca509416a8b37dbbf387a32e9944","author":"","summary":"<a href=\"https://news.google.com/rss/articles/CBMiogFBVV95cUxOM3Y4U2t5T2I0cDVaYzM1eHBxZ2RHZVV5SzR3bHR0WkFPLXFzMWRDSW40TW4tdEJ0S3BwVkkxVkZsMko1RTd0NTVsRlJmVjZOaDluS3UtVV9TRG82ZVNUQkhNT3RCTF9vUFhIaEozMXhPNGhjb3QyX0p1bGhobXlONEZXYkdXdkZnN1hXZUprMk9vUE54UzlZaWZwOXJLNEpUV3c?oc=5\" target=\"_blank\">What is open-source AI? A software engineering researcher explains</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">The Conversation</font>","content":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.\nThe labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.\nOpen-source software\nThe concept of open-source software originated in the free software movement of the 1980s and ’90s. The movement’s founders believed that software creators and users had the right to “four freedoms” – to run the program, to study and modify it, to distribute copies of the original, and to distribute copies of subsequently modified versions. The fundamental requirement was that the source code – the basic instructions – for a program should be made available.\nIn the late 1990s, software developers associated with projects such as the Netscape web browser and the Linux operating system coined and promoted the term “open source” to refer to these ideals.\nAs part of the evolving movement, certain organizations developed open-source licenses that specified how a particular piece of source code could be used and distributed, including the Gnu General Public License, Apache License, MIT License and the Berkeley Software Distribution. Each type of license also specified any potential restrictions on how software patents applied to the source code.\nOpen source or open weight?\nThe open-source idea has risen to prominence again in the past several years as artificial intelligence large language models have surged, notably OpenAI’s ChatGPT, released in 2022. Developers first train new models on large datasets, then deploy the models for use by other people.\nMeta was one of the first large companies to release an open-source large language model, called LLaMa. The company released LLaMa on Feb. 24, 2023, and made available the “inference” source code – the instructions that run the model. And it released the so-called weights, the encoded knowledge the model learned during training. However, open-source organizations such as the Open Source Initiative have stated that the LLaMa licensing guidelines prohibit commercial reuse, which the initiative maintains is not truly open source.\nOther companies have released “open weight” models, such as DeepSeek from DeepSeek AI and Qwen from Alibaba. The models have less restrictive terms for reuse, and the AI community has adopted them rapidly. Still, many developers believe that a true open-source AI model must not only include the source code and weights but also the data that is used to train the model.\nA lot to open up\nThe Open Source Initiative’s definition of a fully open-source AI model includes the training data as a key element. Some developers wonder, however, how feasible it is to distribute the enormous datasets required.","image_url":"https://images.theconversation.com/files/749316/original/file-20260721-69-wu09rg.jpg?ixlib=rb-4.1.1&rect=0%2C1085%2C7127%2C3563&q=45&auto=format&w=1356&h=668&fit=crop","lang":"en","published_at":"2026-07-23T12:24:00+00:00","fetched_at":"2026-07-23T13:15:04+00:00","status":"read","starred":0,"extract_state":"ok","summary_auto":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.","cluster_id":1699595,"extract_retries":0,"extract_error":null,"contract_version":"news_item.v1","format_contract_version":"news_item_formats.v1","dedup_url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3019 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3019,"summary_length":258,"usable_text_length":3019,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3019,"summary_length":258}},"news_item":{"id":44464,"canonical_url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","source_url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","title":"What is open-source AI? A software engineering researcher explains - The Conversation","source_name":"The Conversation","author":null,"published_at":"2026-07-23T12:24:00+00:00","locale":"en","topic":"ai","tags":[],"rss_summary":"<a href=\"https://news.google.com/rss/articles/CBMiogFBVV95cUxOM3Y4U2t5T2I0cDVaYzM1eHBxZ2RHZVV5SzR3bHR0WkFPLXFzMWRDSW40TW4tdEJ0S3BwVkkxVkZsMko1RTd0NTVsRlJmVjZOaDluS3UtVV9TRG82ZVNUQkhNT3RCTF9vUFhIaEozMXhPNGhjb3QyX0p1bGhobXlONEZXYkdXdkZnN1hXZUprMk9vUE54UzlZaWZwOXJLNEpUV3c?oc=5\" target=\"_blank\">What is open-source AI? A software engineering researcher explains</a>&nbsp;&nbsp;<font color=\"#6f6f6f\">The Conversation</font>","full_text":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.\nThe labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.\nOpen-source software\nThe concept of open-source software originated in the free software movement of the 1980s and ’90s. The movement’s founders believed that software creators and users had the right to “four freedoms” – to run the program, to study and modify it, to distribute copies of the original, and to distribute copies of subsequently modified versions. The fundamental requirement was that the source code – the basic instructions – for a program should be made available.\nIn the late 1990s, software developers associated with projects such as the Netscape web browser and the Linux operating system coined and promoted the term “open source” to refer to these ideals.\nAs part of the evolving movement, certain organizations developed open-source licenses that specified how a particular piece of source code could be used and distributed, including the Gnu General Public License, Apache License, MIT License and the Berkeley Software Distribution. Each type of license also specified any potential restrictions on how software patents applied to the source code.\nOpen source or open weight?\nThe open-source idea has risen to prominence again in the past several years as artificial intelligence large language models have surged, notably OpenAI’s ChatGPT, released in 2022. Developers first train new models on large datasets, then deploy the models for use by other people.\nMeta was one of the first large companies to release an open-source large language model, called LLaMa. The company released LLaMa on Feb. 24, 2023, and made available the “inference” source code – the instructions that run the model. And it released the so-called weights, the encoded knowledge the model learned during training. However, open-source organizations such as the Open Source Initiative have stated that the LLaMa licensing guidelines prohibit commercial reuse, which the initiative maintains is not truly open source.\nOther companies have released “open weight” models, such as DeepSeek from DeepSeek AI and Qwen from Alibaba. The models have less restrictive terms for reuse, and the AI community has adopted them rapidly. Still, many developers believe that a true open-source AI model must not only include the source code and weights but also the data that is used to train the model.\nA lot to open up\nThe Open Source Initiative’s definition of a fully open-source AI model includes the training data as a key element. Some developers wonder, however, how feasible it is to distribute the enormous datasets required.","excerpt":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.","extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 3019 characters.","diagnostics_url":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3019 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3019,"summary_length":258,"usable_text_length":3019,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3019,"summary_length":258}}},"display_formats":["compact","card","full","digest_section","json"]},"daily_stack_record":{"title":"What is open-source AI? A software engineering researcher explains - The Conversation","url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","summary":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.","source":"The Conversation","date":"2026-07-23T12:24:00+00:00","content":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.\nThe labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.\nOpen-source software\nThe concept of open-source software originated in the free software movement of the 1980s and ’90s. The movement’s founders believed that software creators and users had the right to “four freedoms” – to run the program, to study and modify it, to distribute copies of the original, and to distribute copies of subsequently modified versions. The fundamental requirement was that the source code – the basic instructions – for a program should be made available.\nIn the late 1990s, software developers associated with projects such as the Netscape web browser and the Linux operating system coined and promoted the term “open source” to refer to these ideals.\nAs part of the evolving movement, certain organizations developed open-source licenses that specified how a particular piece of source code could be used and distributed, including the Gnu General Public License, Apache License, MIT License and the Berkeley Software Distribution. Each type of license also specified any potential restrictions on how software patents applied to the source code.\nOpen source or open weight?\nThe open-source idea has risen to prominence again in the past several years as artificial intelligence large language models have surged, notably OpenAI’s ChatGPT, released in 2022. Developers first train new models on large datasets, then deploy the models for use by other people.\nMeta was one of the first large companies to release an open-source large language model, called LLaMa. The company released LLaMa on Feb. 24, 2023, and made available the “inference” source code – the instructions that run the model. And it released the so-called weights, the encoded knowledge the model learned during training. However, open-source organizations such as the Open Source Initiative have stated that the LLaMa licensing guidelines prohibit commercial reuse, which the initiative maintains is not truly open source.\nOther companies have released “open weight” models, such as DeepSeek from DeepSeek AI and Qwen from Alibaba. The models have less restrictive terms for reuse, and the AI community has adopted them rapidly. Still, many developers believe that a true open-source AI model must not only include the source code and weights but also the data that is used to train the model.\nA lot to open up\nThe Open Source Initiative’s definition of a fully open-source AI model includes the training data as a key element. Some developers wonder, however, how feasible it is to distribute the enormous datasets required.","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 3019 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3019 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3019,"summary_length":258,"usable_text_length":3019,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3019,"summary_length":258}},"tags":[]},"fallback_formats":["markdown","json","html"],"actions":{"read":"/item/44464","export_markdown":"/api/items/44464/export?format=markdown","export_json":"/api/items/44464/export?format=json","diagnose":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668"},"formats":{"full":{"id":44464,"title":"What is open-source AI? A software engineering researcher explains - The Conversation","url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","source":"The Conversation","author":null,"published_at":"2026-07-23T12:24:00+00:00","locale":"en","topic":"ai","tags":[],"excerpt":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.","full_text":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.\nThe labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.\nOpen-source software\nThe concept of open-source software originated in the free software movement of the 1980s and ’90s. The movement’s founders believed that software creators and users had the right to “four freedoms” – to run the program, to study and modify it, to distribute copies of the original, and to distribute copies of subsequently modified versions. The fundamental requirement was that the source code – the basic instructions – for a program should be made available.\nIn the late 1990s, software developers associated with projects such as the Netscape web browser and the Linux operating system coined and promoted the term “open source” to refer to these ideals.\nAs part of the evolving movement, certain organizations developed open-source licenses that specified how a particular piece of source code could be used and distributed, including the Gnu General Public License, Apache License, MIT License and the Berkeley Software Distribution. Each type of license also specified any potential restrictions on how software patents applied to the source code.\nOpen source or open weight?\nThe open-source idea has risen to prominence again in the past several years as artificial intelligence large language models have surged, notably OpenAI’s ChatGPT, released in 2022. Developers first train new models on large datasets, then deploy the models for use by other people.\nMeta was one of the first large companies to release an open-source large language model, called LLaMa. The company released LLaMa on Feb. 24, 2023, and made available the “inference” source code – the instructions that run the model. And it released the so-called weights, the encoded knowledge the model learned during training. However, open-source organizations such as the Open Source Initiative have stated that the LLaMa licensing guidelines prohibit commercial reuse, which the initiative maintains is not truly open source.\nOther companies have released “open weight” models, such as DeepSeek from DeepSeek AI and Qwen from Alibaba. The models have less restrictive terms for reuse, and the AI community has adopted them rapidly. Still, many developers believe that a true open-source AI model must not only include the source code and weights but also the data that is used to train the model.\nA lot to open up\nThe Open Source Initiative’s definition of a fully open-source AI model includes the training data as a key element. Some developers wonder, however, how feasible it is to distribute the enormous datasets required.","reading_time_min":2,"extraction":{"state":"ok","confidence":0.9,"error":null,"explanation":"High confidence: full text extraction produced 3019 characters.","diagnostics_url":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3019 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3019,"summary_length":258,"usable_text_length":3019,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3019,"summary_length":258}}},"quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3019 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3019,"summary_length":258,"usable_text_length":3019,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3019,"summary_length":258}},"actions":{"read":"/item/44464","export_markdown":"/api/items/44464/export?format=markdown","export_json":"/api/items/44464/export?format=json","diagnose":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668"}},"digest":{"id":44464,"title":"What is open-source AI? A software engineering researcher explains - The Conversation","url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","source":"The Conversation","topic":"ai","published_at":"2026-07-23T12:24:00+00:00","excerpt":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.","quality_bucket":"high","quality_reason":"High confidence: full text extraction produced 3019 characters.","reading_time_min":2,"cluster_id":1699595},"card":{"display_title":"What is open-source AI? A software engineering researcher explains - The Conversation","subtitle":"The Conversation · 2026-07-23","summary":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have…","badges":["quality:high"],"links":{"read":"/item/44464","original":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","diagnose":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668"},"quality_warning":null},"export":{"title":"What is open-source AI? A software engineering researcher explains - The Conversation","url":"https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","summary":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.","source":"The Conversation","date":"2026-07-23T12:24:00+00:00","content":"You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.\nThe labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.\nOpen-source software\nThe concept of open-source software originated in the free software movement of the 1980s and ’90s. The movement’s founders believed that software creators and users had the right to “four freedoms” – to run the program, to study and modify it, to distribute copies of the original, and to distribute copies of subsequently modified versions. The fundamental requirement was that the source code – the basic instructions – for a program should be made available.\nIn the late 1990s, software developers associated with projects such as the Netscape web browser and the Linux operating system coined and promoted the term “open source” to refer to these ideals.\nAs part of the evolving movement, certain organizations developed open-source licenses that specified how a particular piece of source code could be used and distributed, including the Gnu General Public License, Apache License, MIT License and the Berkeley Software Distribution. Each type of license also specified any potential restrictions on how software patents applied to the source code.\nOpen source or open weight?\nThe open-source idea has risen to prominence again in the past several years as artificial intelligence large language models have surged, notably OpenAI’s ChatGPT, released in 2022. Developers first train new models on large datasets, then deploy the models for use by other people.\nMeta was one of the first large companies to release an open-source large language model, called LLaMa. The company released LLaMa on Feb. 24, 2023, and made available the “inference” source code – the instructions that run the model. And it released the so-called weights, the encoded knowledge the model learned during training. However, open-source organizations such as the Open Source Initiative have stated that the LLaMa licensing guidelines prohibit commercial reuse, which the initiative maintains is not truly open source.\nOther companies have released “open weight” models, such as DeepSeek from DeepSeek AI and Qwen from Alibaba. The models have less restrictive terms for reuse, and the AI community has adopted them rapidly. Still, many developers believe that a true open-source AI model must not only include the source code and weights but also the data that is used to train the model.\nA lot to open up\nThe Open Source Initiative’s definition of a fully open-source AI model includes the training data as a key element. Some developers wonder, however, how feasible it is to distribute the enormous datasets required.","confidence":0.9,"diagnostics_url":"/api/diagnose?url=https%3A//theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668","quality_bucket":"high","failure_kind":"none","retryable":false,"quality_reason":"High confidence: full text extraction produced 3019 characters.","quality_profile":{"profile_version":"extraction_quality.v2","bucket":"high","confidence":0.9,"failure_kind":"none","retryable":false,"retry_after_attempts":0,"reason":"High confidence: full text extraction produced 3019 characters.","operator_guidance":{"severity":"ok","recommended_action":"trust_full_text","next_step":"Use the extracted full text as the primary article source.","operator_label":"Ready","can_retry":false,"can_use_summary":false,"diagnostics_required":false},"content_depth":{"contract_version":"content_depth.v1","category":"full_text","label":"Full text","has_full_text":true,"has_summary":true,"content_length":3019,"summary_length":258,"usable_text_length":3019,"source_field":"content"},"legacy_collapsed":false,"signals":{"extract_state":"ok","extract_error":null,"extract_retries":0,"content_length":3019,"summary_length":258}},"tags":[],"format_contract_version":"news_item_formats.v1"}}}