CRIS Current Research Information System

Generalised linear models for multi-class classification problems are one of the fundamental building blocks of modern machine learning tasks. In this manuscript, we characterise the learning of a mixture of K Gaussians with generic means and covariances via empirical risk minimisation (ERM) with any convex loss and regularisation. In particular, we prove exact asymptotics characterising the ERM estimator in high-dimensions, extending several previous results about Gaussian mixture classification in the literature. We exemplify our result in two tasks of interest in statistical learning: a) classification for a mixture with sparse means, where we study the efficiency of ℓ1 penalty with respect to ℓ2; b) max-margin multiclass classification, where we characterise the phase transition on the existence of the multi-class logistic maximum likelihood estimator for K > 2. Finally, we discuss how our theory can be applied beyond the scope of synthetic data, showing that in different cases Gaussian mixtures capture closely the learning curve of classification tasks in real data sets.

Loureiro, B., Sicuro, G., Gerbelot, C., Pacco, A., Krzakala, F., Zdeborova, L. (2021). Learning Gaussian Mixtures with Generalised Linear Models: Precise Asymptotics in High-dimensions. Neural information processing systems foundation.

Learning Gaussian Mixtures with Generalised Linear Models: Precise Asymptotics in High-dimensions

Loureiro B.;Sicuro G.;Gerbelot C.;Pacco A.;Krzakala F.;Zdeborova L.

2021

Abstract

Generalised linear models for multi-class classification problems are one of the fundamental building blocks of modern machine learning tasks. In this manuscript, we characterise the learning of a mixture of K Gaussians with generic means and covariances via empirical risk minimisation (ERM) with any convex loss and regularisation. In particular, we prove exact asymptotics characterising the ERM estimator in high-dimensions, extending several previous results about Gaussian mixture classification in the literature. We exemplify our result in two tasks of interest in statistical learning: a) classification for a mixture with sparse means, where we study the efficiency of ℓ1 penalty with respect to ℓ2; b) max-margin multiclass classification, where we characterise the phase transition on the existence of the multi-class logistic maximum likelihood estimator for K > 2. Finally, we discuss how our theory can be applied beyond the scope of synthetic data, showing that in different cases Gaussian mixtures capture closely the learning curve of classification tasks in real data sets.

Scheda breve

Scheda completa

Scheda completa (DC)

	Anno
	
				2021
			
	Titolo del volume
	
				Advances in Neural Information Processing Systems
			
	Pagina iniziale
	
				10144
			
	Pagina finale
	
				10157
			
	Collana/Serie
	
				ADVANCES IN NEURAL INFORMATION PROCESSING SYSTEMS
			
	Citazione
	
				Loureiro, B., Sicuro, G., Gerbelot, C., Pacco, A., Krzakala, F., Zdeborova, L. (2021). Learning Gaussian Mixtures with Generalised Linear Models: Precise Asymptotics in High-dimensions. Neural information processing systems foundation.
			
	Tutti gli autori
	
						Loureiro, B.; Sicuro, G.; Gerbelot, C.; Pacco, A.; Krzakala, F.; Zdeborova, L.

File in questo prodotto:

Eventuali allegati, non sono esposti

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11585/958761

Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo

Citazioni

ND

32

0

social impact