<node id="692240">
  <nid>692240</nid>
  <type>news</type>
  <uid>
    <user id="27469"><![CDATA[27469]]></user>
  </uid>
  <created>1788440889</created>
  <changed>1788441212</changed>
  <title><![CDATA[What Is Open‑Source AI? A Software Engineering Researcher Explains]]></title>
  <body><![CDATA[<div class="theconversation-article-body"><p>You’ve probably heard artificial intelligence models described as “open” or “closed.” These are not descriptions of the model’s personality. Large language model AIs like the one under the hood of ChatGPT don’t have actual personalities, despite appearances.</p><p>The labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.</p><h2>Open-Source Software</h2><p>The concept of open-source software originated in the <a href="https://os-sci.com/blog/our-blog-posts-1/the-history-of-open-source-132">free software movement</a> of the 1980s and ’90s. The movement’s founders believed that software creators and users had the right to “<a href="https://os-sci.com/blog/our-blog-posts-1/four-freedoms-of-open-source-42">four freedoms</a>” – to run the program, to study and modify it, to distribute copies of the original, and to distribute copies of subsequently modified versions. The fundamental requirement was that the source code – <a href="https://www.techtarget.com/searchapparchitecture/definition/source-code">the basic instructions</a> – for a program should be made available.</p><p>In the late 1990s, software developers associated with projects such as the <a href="https://www.zdnet.com/home-and-office/networking/how-netscape-lives-on-30-years-of-shaping-the-web-open-source-and-business/">Netscape web browser</a> and the <a href="https://www.linux.com/what-is-linux/">Linux operating system</a> coined and promoted the term “open source” to refer to these ideals.</p><p>As part of the evolving movement, certain organizations developed open-source licenses that specified how a particular piece of source code could be used and distributed, including the <a href="https://www.gnu.org/licenses/gpl-3.0.en.html">Gnu General Public License</a>, Apache License, MIT License and the Berkeley Software Distribution. Each type of license also specified any potential restrictions on how software patents applied to the source code.</p><h2>Open Source or Open Weight?</h2><p>The open-source idea has risen to prominence again in the past several years as artificial intelligence large language models have surged, notably OpenAI’s ChatGPT, released in 2022. Developers first train new models on large datasets, then deploy the models for use by other people.</p><figure><p><iframe width="440" height="260" src="https://www.youtube.com/embed/Gmz24X86-5E?wmode=transparent&amp;start=0" frameborder="0" allowfullscreen=""></iframe></p><figcaption><span class="caption">Open-source artificial intelligence is explained in two minutes.</span></figcaption></figure><p>Meta was one of the first large companies to release an open-source large language model, called <a href="https://techcrunch.com/2025/10/06/meta-llama-everything-you-need-to-know-about-the-open-generative-ai-model/">LLaMa</a>. The company released LLaMa on Feb. 24, 2023, and made available the “inference” source code – the instructions that run the model. And it released the so-called weights, the encoded knowledge the model learned during training. However, open-source organizations such as the <a href="https://opensource.org/">Open Source Initiative</a> have stated that the LLaMa licensing guidelines prohibit commercial reuse, which the initiative maintains is not truly open source.</p><p>Other companies have released “<a href="https://opensource.org/ai/open-weights">open weight” models</a>, such as DeepSeek from DeepSeek AI and Qwen from Alibaba. The models have less restrictive terms for reuse, and the AI community has adopted them rapidly. Still, many developers believe that a true open-source AI model must not only include the source code and weights but also the data that is used to train the model.</p><h2>A Lot to Open Up</h2><p>The Open Source Initiative’s definition of a <a href="https://opensource.org/ai/open-source-ai-definition">fully open-source AI model</a> includes the training data as a key element. Some developers wonder, however, how feasible it is to distribute the enormous datasets required.<!-- Below is The Conversation's page counter tag. Please DO NOT REMOVE. --><img style="border-color:!important;border-style:none;box-shadow:none !important;margin:0 !important;max-height:1px !important;max-width:1px !important;min-height:1px !important;min-width:1px !important;opacity:0 !important;outline:none !important;padding:0 !important;" src="https://counter.theconversation.com/content/236668/count.gif?distributor=republish-lightbox-basic" alt="The Conversation" width="1" height="1" referrerpolicy="no-referrer-when-downgrade"><!-- End of code. If you don't see any code above, please get new code from the Advanced tab after you click the republish button. The page counter does not collect any personal data. More info: https://theconversation.com/republishing-guidelines --></p><p>&nbsp;</p><p><em>This article is republished from </em><a href="https://theconversation.com"><em>The Conversation</em></a><em> under a Creative Commons license. Read the </em><a href="https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668"><em>original article</em></a><em>.</em></p></div>]]></body>
  <field_subtitle>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_subtitle>
  <field_dateline>
    <item>
      <value>2026-07-23T00:00:00-04:00</value>
      <timezone><![CDATA[America/New_York]]></timezone>
    </item>
  </field_dateline>
  <field_summary_sentence>
    <item>
      <value><![CDATA[The labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.]]></value>
    </item>
  </field_summary_sentence>
  <field_summary>
    <item>
      <value><![CDATA[<p>The labels refer to whether all of the information about how an AI model works is publicly available and the model can be modified, or whether the model’s developer keeps its inner workings secret and the model itself private property.</p>]]></value>
    </item>
  </field_summary>
  <field_media>
          <item>
        <nid>
          <node id="681056">
            <nid>681056</nid>
            <type>image</type>
            <title><![CDATA[Open-source AI models allow anyone to read their inner workings like a book. Utamaru Kido/Moment via Getty Images]]></title>
            <body><![CDATA[<p>Open-source AI models allow anyone to read their inner workings like a book. <a href="https://www.gettyimages.com/detail/photo/books-royalty-free-image/957583536">Utamaru Kido/Moment via Getty Images</a></p>]]></body>
                          <field_image>
                <item>
                  <fid>265407</fid>
                  <filename><![CDATA[file-20260721-69-wu09rg.jpg]]></filename>
                  <filepath><![CDATA[/sites/default/files/2026/09/03/file-20260721-69-wu09rg.jpg]]></filepath>
                  <file_full_path><![CDATA[http://hg.gatech.edu//sites/default/files/2026/09/03/file-20260721-69-wu09rg.jpg]]></file_full_path>
                  <filemime>image/jpeg</filemime>
                  <image_740><![CDATA[]]></image_740>
                  <image_alt><![CDATA[Open-source AI models allow anyone to read their inner workings like a book. Utamaru Kido/Moment via Getty Images]]></image_alt>
                </item>
              </field_image>
            
                      </node>
        </nid>
      </item>
      </field_media>
  <field_contact_email>
    <item>
      <email><![CDATA[]]></email>
    </item>
  </field_contact_email>
  <field_location>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_location>
  <field_contact>
    <item>
      <value><![CDATA[<h5>Author:</h5><p><a href="https://theconversation.com/profiles/jeffrey-young-1631845">Jeffrey Young</a>, Principal Research Scientist, Partnership for an Advanced Computing Environment, <a href="https://theconversation.com/institutions/georgia-institute-of-technology-1310">Georgia Institute of Technology</a></p><h5>Media Contact:</h5><p>Shelley Wunder-Smith<br><a href="mailto:shelley.wunder-smith@research.gatech.edu">shelley.wunder-smith@research.gatech.edu</a></p>]]></value>
    </item>
  </field_contact>
  <field_sidebar>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_sidebar>
  <field_boilerplate>
    <item>
      <nid><![CDATA[]]></nid>
    </item>
  </field_boilerplate>
  <!--  TO DO: correct to not conflate categories and news room topics  -->
  <!--  Disquisition: it's funny how I write these TODOs and then never
         revisit them. It's as though the act of writing the thing down frees me
         from the responsibility to actually solve the problem. But what can I
         say? There are more problems than there's time to solve.  -->
  <links_related> </links_related>
  <files> </files>
  <og_groups>
          <item>47223</item>
          <item>1214</item>
      </og_groups>
  <og_groups_both>
      </og_groups_both>
  <field_categories>
      </field_categories>
  <core_research_areas>
      </core_research_areas>
  <field_news_room_topics>
          <item>
        <tid>71881</tid>
        <value><![CDATA[Science and Technology]]></value>
      </item>
      </field_news_room_topics>
  <links_related>
          <link>
      <url>https://theconversation.com/what-is-open-source-ai-a-software-engineering-researcher-explains-236668`</url>
      <title></title>
      </link>
      </links_related>
  <files>
      </files>
  <og_groups>
          <item>47223</item>
          <item>1214</item>
      </og_groups>
  <og_groups_both>
          <item><![CDATA[College of Computing]]></item>
          <item><![CDATA[News Room]]></item>
      </og_groups_both>
  <field_keywords>
          <item>
        <tid>194974</tid>
        <value><![CDATA[go-theconversation]]></value>
      </item>
      </field_keywords>
  <field_userdata><![CDATA[]]></field_userdata>
</node>
