Understanding modern machine learning models through the lens of high-dimensional statistics
MBZUAI · Notable
Summary
This talk explores modern machine learning through high-dimensional statistics, using random matrix theory to analyze learning models. The speaker, Denny Wu from University of Toronto and the Vector Institute, presents two examples: hyperparameter selection in overparameterized models and gradient-based representation learning in neural networks. The analysis reveals insights such as the possibility of negative optimal ridge penalty and the advantages of feature learning over random features. Why it matters: This research provides a deeper theoretical understanding of deep learning phenomena, with potential implications for optimizing training and improving model performance in the region.
Keywords
high-dimensional statistics · machine learning · neural networks · ridge regression · generalization error
Get the weekly digest
Top AI stories from the GCC region, every week.