<node id="676016">
  <nid>676016</nid>
  <type>event</type>
  <uid>
    <user id="27707"><![CDATA[27707]]></user>
  </uid>
  <created>1723724785</created>
  <changed>1723724823</changed>
  <title><![CDATA[PhD Defense by Andy Ko]]></title>
  <body><![CDATA[<p><strong>Title:</strong>&nbsp;Statistical Theory for Neural Network-Based Learning</p><p>&nbsp;</p><p><strong>Committee:</strong></p><p>Dr. Xiaoming Huo (advisor), School of Industrial and Systems Engineering, Georgia Institute of Technology</p><p>Dr. Yao Xie, School of&nbsp;Industrial and Systems Engineering, Georgia Institute of Technology</p><p>Dr. Shihao Yang, School of&nbsp;Industrial and Systems Engineering, Georgia Institute of Technology</p><p>Dr. Mayya Zhilova, School of Mathematics, Georgia Institute of Technology</p><p>Dr. Yajun Mei, School of Global Public Health, New York University</p><p>&nbsp;</p><p><strong>Date and Time:&nbsp;</strong>Wednesday, August 28th, 09:00 AM - 10:30 AM EDT</p><p><strong>Location:&nbsp;</strong>Groseclose 303</p><p><strong>Virtual Link:</strong>&nbsp;Microsoft Teams&nbsp;</p><p>Meeting ID: 253 390 754 819</p><p>Passcode: YtgR27</p><p><strong>Abstract:</strong></p><p>Data have now become an indispensable part of our lives. This of course was made possible with the ease at which we can now collect, curate, display and manipulate data. However, the power and usefulness of data are only realized when they guide us to make better decisions. Statistical decision theory is a theoretical framework created for the purpose of answering this question, and in this thesis, we adopt this viewpoint to study the problem of pattern recognition and classification. In particular, we focus on understanding the performance of neural network-based learning for classification problems.</p><p>&nbsp;</p><p>In Chapter 2, we provide the necessary technical preparations for the statement and proofs of the results in the rest of the thesis along with a discussion of related works. We conclude with a discussion of the relationship between regression and classification.</p><p>&nbsp;</p><p>In Chapter 3, we show that random classifiers based on finitely wide and deep neural networks are &nbsp;consistent for a very general class of distributions. Consistency is a highly desirable property for a sequence of classifiers that guarantees that the classification risk converges to the smallest possible risk. This result improves the current literature by extending the known consistency property of shallow, underparametrized neural networks with sigmoid activations to wide and deep ReLU neural networks without complexity constraints.</p><p>&nbsp;</p><p>In Chapter 4, we give several convergence rate guarantees of the excess classification risk for a semiparametric model of distributions indexed by Borel probability measures and regression functions belonging to L2 class of functions with finite Kolmogorov-Donoho optimal exponents. Furthermore, we give explicit characterizations of distributional regimes in which neural network classifiers are minimax optimal.</p><p>&nbsp;</p><p>In Chapter 5, we show that for a semiparametric model of distributions characterized by regression functions that locally belong to the Barron approximation space, neural network classifiers achieve a minimax optimal rate of convergence up to a logarithmic factor.</p><p>&nbsp;</p>]]></body>
  <field_summary_sentence>
    <item>
      <value><![CDATA[Statistical Theory for Neural Network-Based Learning]]></value>
    </item>
  </field_summary_sentence>
  <field_summary>
    <item>
      <value><![CDATA[<p>Statistical Theory for Neural Network-Based Learning</p>]]></value>
    </item>
  </field_summary>
  <field_time>
    <item>
      <value><![CDATA[2024-08-28T09:00:00-04:00]]></value>
      <value2><![CDATA[2024-08-28T10:30:00-04:00]]></value2>
      <rrule><![CDATA[]]></rrule>
      <timezone><![CDATA[America/New_York]]></timezone>
    </item>
  </field_time>
  <field_fee>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_fee>
  <field_extras>
      </field_extras>
  <field_audience>
          <item>
        <value><![CDATA[Public]]></value>
      </item>
      </field_audience>
  <field_media>
      </field_media>
  <field_contact>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_contact>
  <field_location>
    <item>
      <value><![CDATA[Groseclose 303]]></value>
    </item>
  </field_location>
  <field_sidebar>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_sidebar>
  <field_phone>
    <item>
      <value><![CDATA[]]></value>
    </item>
  </field_phone>
  <field_url>
    <item>
      <url><![CDATA[]]></url>
      <title><![CDATA[]]></title>
            <attributes><![CDATA[]]></attributes>
    </item>
  </field_url>
  <field_email>
    <item>
      <email><![CDATA[]]></email>
    </item>
  </field_email>
  <field_boilerplate>
    <item>
      <nid><![CDATA[]]></nid>
    </item>
  </field_boilerplate>
  <links_related>
      </links_related>
  <files>
      </files>
  <og_groups>
          <item>221981</item>
      </og_groups>
  <og_groups_both>
          <item><![CDATA[Graduate Studies]]></item>
      </og_groups_both>
  <field_categories>
          <item>
        <tid>1788</tid>
        <value><![CDATA[Other/Miscellaneous]]></value>
      </item>
      </field_categories>
  <field_keywords>
          <item>
        <tid>100811</tid>
        <value><![CDATA[Phd Defense]]></value>
      </item>
      </field_keywords>
  <field_userdata><![CDATA[]]></field_userdata>
</node>
