In recent years, bilevel approaches have become very popular to efficiently estimate highdimensional hyperparameters of machine learning models. However, to date, binary parameters are handled by continuous relaxation and rounding strategies, which could lead to inconsistent solutions. In this context, we tackle the challenging optimization of mixed-binary hyperparameters by resorting to an equivalent continuous bilevel reformulation based on an appropriate penalty term. We propose an algorithmic framework that, under suitable assumptions, is guaranteed to provide mixed-binary solutions. Moreover, the generality of the method allows to safely use existing continuous bilevel solvers within the proposed framework. We evaluate the performance of our approach for two specific machine learning problems, i.e., the estimation of the group-sparsity structure in regression problems and the data distillation problem. The reported results show that our method is competitive with state-of-the-art approaches based on relaxation and rounding.

Relax and penalize: a new bilevel approach to mixed-binary hyperparameter optimization / Venturini, Sara; De Santis, Marianna; Patracone, Jordan; Schmidt, Martin; Rinaldi, Francesco; Salzo, Saverio. - In: TRANSACTIONS ON MACHINE LEARNING RESEARCH. - ISSN 2835-8856. - (2025), pp. 1-27.

Relax and penalize: a new bilevel approach to mixed-binary hyperparameter optimization

Saverio Salzo
2025

Abstract

In recent years, bilevel approaches have become very popular to efficiently estimate highdimensional hyperparameters of machine learning models. However, to date, binary parameters are handled by continuous relaxation and rounding strategies, which could lead to inconsistent solutions. In this context, we tackle the challenging optimization of mixed-binary hyperparameters by resorting to an equivalent continuous bilevel reformulation based on an appropriate penalty term. We propose an algorithmic framework that, under suitable assumptions, is guaranteed to provide mixed-binary solutions. Moreover, the generality of the method allows to safely use existing continuous bilevel solvers within the proposed framework. We evaluate the performance of our approach for two specific machine learning problems, i.e., the estimation of the group-sparsity structure in regression problems and the data distillation problem. The reported results show that our method is competitive with state-of-the-art approaches based on relaxation and rounding.
2025
Bilevel optimization; hyperparameter optimization; mixed-integer programming
01 Pubblicazione su rivista::01a Articolo in rivista
Relax and penalize: a new bilevel approach to mixed-binary hyperparameter optimization / Venturini, Sara; De Santis, Marianna; Patracone, Jordan; Schmidt, Martin; Rinaldi, Francesco; Salzo, Saverio. - In: TRANSACTIONS ON MACHINE LEARNING RESEARCH. - ISSN 2835-8856. - (2025), pp. 1-27.
File allegati a questo prodotto
File Dimensione Formato  
Venturini_Relax_2025.pdf

accesso aperto

Note: https://openreview.net/forum?id=A1R1cQ93Cb
Tipologia: Versione editoriale (versione pubblicata con il layout dell'editore)
Licenza: Creative commons
Dimensione 787.6 kB
Formato Adobe PDF
787.6 kB Adobe PDF

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11573/1741127
Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 0
  • ???jsp.display-item.citation.isi??? ND
social impact