<node id="675843">
  <nid>675843</nid>
  <type>event</type>
  <uid>
    <user id="27707"><![CDATA[27707]]></user>
  </uid>
  <created>1723064112</created>
  <changed>1723064112</changed>
  <title><![CDATA[PhD Defense by Cheng Wan]]></title>
  <body><![CDATA[<p><strong>Title: Optimizing Sparsity in Distributed Machine Learning Training</strong></p><p><strong>&nbsp;</strong></p><p><strong>Date:&nbsp;</strong>Wednesday, August 14th, 2024</p><p><strong>Time:&nbsp;</strong>10:00 AM - 11:30 AM EST</p><p><strong>Location:&nbsp;</strong>KACB 3126</p><p><strong>Virtual:</strong> <a href="https://gatech.zoom.us/j/95651978956?pwd=cNa98ds2wGysW5SlajJDBAkouSiAvH.1">https://gatech.zoom.us/j/95651978956?pwd=cNa98ds2wGysW5SlajJDBAkouSiAvH.1</a></p><p>&nbsp;</p><p><strong>Cheng Wan</strong></p><p>School of Computer Science</p><p>College of Computing</p><p>Georgia Institute of Technology</p><p><strong>&nbsp;</strong></p><p><strong>Committee Members:</strong></p><p>Dr. Yingyan (Celine) Lin (Advisor) – School of Computer Science, Georgia Institute of Technology</p><p>Dr. Pan Li&nbsp;– School of Computational Science and Engineering, Georgia Institute of Technology</p><p>Dr.&nbsp;Alexey Tumanov – School of Computer Science, Georgia Institute of Technology</p><p>Dr. Anand Iyer&nbsp;– School of Computer Science, Georgia Institute of Technology</p><p><strong>&nbsp;</strong></p><p><strong>Abstract:</strong></p><p>As machine learning models and datasets continue to grow, distributed machine learning has become essential for meeting the computational demands of large-scale training. While recent frameworks have improved scalability and throughput, the optimization of sparsity within distributed deep neural networks remains under-explored. Sparsity, however, poses a significant challenge in advanced models such as graph neural networks and mixture-of-experts models.</p><p>&nbsp;</p><p>This thesis proposal systematically investigates three types of sparsity in distributed machine learning training: sparse data access, sparse operations, and sparse workflows. We introduce both system-level and algorithm-level innovations to address inefficiencies caused by these types of sparsity. Through theoretical insights and practical implementations, we demonstrate how these optimizations can significantly reduce communication overhead and enhance scalability, thereby improving the efficiency and performance of distributed machine learning systems. Our work advances the understanding of sparsity optimization and lays the groundwork for developing more efficient distributed training architectures in the future.</p>]]></body>
  <field_summary_sentence>
    <item>
      <value><![CDATA[Optimizing Sparsity in Distributed Machine Learning Training]]></value>
    </item>
  </field_summary_sentence>
  <field_summary>
    <item>
      <value><![CDATA[<p><strong>&nbsp;Optimizing Sparsity in Distributed Machine Learning Training</strong></p>]]></value>
    </item>
  </field_summary>
  <field_time>
    <item>
      <value><![CDATA[2024-08-14T10:00:00-04:00]]></value>
      <value2><![CDATA[2024-08-14T11:30:00-04:00]]></value2>
      <rrule><![CDATA[]]></rrule>
      <timezone><![CDATA[America/New_York]]></timezone>
    </item>
  </field_time>
  <field_fee>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_fee>
  <field_extras>
      </field_extras>
  <field_audience>
          <item>
        <value><![CDATA[Public]]></value>
      </item>
      </field_audience>
  <field_media>
      </field_media>
  <field_contact>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_contact>
  <field_location>
    <item>
      <value><![CDATA[KACB 3126]]></value>
    </item>
  </field_location>
  <field_sidebar>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_sidebar>
  <field_phone>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_phone>
  <field_url>
    <item>
      <url><![CDATA[]]></url>
      <title><![CDATA[]]></title>
            <attributes><![CDATA[]]></attributes>
    </item>
  </field_url>
  <field_email>
    <item>
      <email><![CDATA[]]></email>
    </item>
  </field_email>
  <field_boilerplate>
    <item>
      <nid><![CDATA[]]></nid>
    </item>
  </field_boilerplate>
  <links_related>
      </links_related>
  <files>
      </files>
  <og_groups>
          <item>221981</item>
      </og_groups>
  <og_groups_both>
          <item><![CDATA[Graduate Studies]]></item>
      </og_groups_both>
  <field_categories>
          <item>
        <tid>1788</tid>
        <value><![CDATA[Other/Miscellaneous]]></value>
      </item>
      </field_categories>
  <field_keywords>
          <item>
        <tid>100811</tid>
        <value><![CDATA[Phd Defense]]></value>
      </item>
      </field_keywords>
  <field_userdata><![CDATA[]]></field_userdata>
</node>
