MATH - Seminar on Statistics and Data Science - Generalization error of linearized neural networks: staircase and double-descent

7月24日

11:00am - 12:00pm

研討會, 演講, 講座

Deep learning methods operate in regimes that defy the traditional statistical mindset. Despite the non-convexity of empirical risks and the huge complexity of neural network architectures, stochastic gradient algorithms can often find the global minimizer of the training loss and achieve small generalization error on test data. As one possible explanation to the training efficiency of neural networks, tangent kernel theory shows that a multi-layers neural network — in a proper large width limit — can be well approximated by its linearization. As a consequence, the gradient flow of the empirical risk turns into a linear dynamics and converges to a global minimizer. Since last year, linearization has become a popular approach in analyzing training dynamics of neural networks. However, this naturally raises the question of whether the linearization perspective can also explain the observed generalization efficacy. In this talk, I will discuss the generalization error of linearized neural networks, which reveals two interesting phenomena: the staircase decay and the double-descent curve. Through the lens of these phenomena, I will also address the benefits and limitations of the linearization approach for neural networks.

7月24日

11:00am - 12:00pm

立即登記

地點

https://hkust.zoom.us/j/5616960008

講者/表演者

Prof. Song MEI
UC Berkeley

主辦單位

Department of Mathematics

聯絡方法

mathseminar@ust.hk

付款詳情

對象

Alumni, Faculty and Staff, PG Students, UG Students

語言

英語

其他活動

6月16日

研討會, 演講, 講座

IAS / School of Science Joint Lecture - Shaping Tumor Cell Plasticity and Therapy Resistance in Glioblastoma

Abstract Tumor heterogeneity fueled by plasticity and genetic diversification of cancer cells is key to therapy failure of malignant glioma. The speaker's team implemented spatial and genetic p...

5月11日

研討會, 演講, 講座

IAS / School of Science Joint Lecture - Regioselective Pyridine C-H-Functionalization and Skeletal Editing

Abstract Pyridines belong to the most abundant heteroarenes in medicinal chemistry and in agrochemical industry. In the lecture, highly regioselective pyridine C-H functionalization through a d...