Equivalence of the Empirical Risk Minimization to Regularization on the Family of f-Divergences
Daunas, Francisco, Esnaola, Iñaki, Perlaza, Samir M., Poor, H. Vincent
–arXiv.org Artificial Intelligence
The solution to empirical risk minimization with $f$-divergence regularization (ERM-$f$DR) is presented under mild conditions on $f$. Under such conditions, the optimal measure is shown to be unique. Examples of the solution for particular choices of the function $f$ are presented. Previously known solutions to common regularization choices are obtained by leveraging the flexibility of the family of $f$-divergences. These include the unique solutions to empirical risk minimization with relative entropy regularization (Type-I and Type-II). The analysis of the solution unveils the following properties of $f$-divergences when used in the ERM-$f$DR problem: $i\bigl)$ $f$-divergence regularization forces the support of the solution to coincide with the support of the reference measure, which introduces a strong inductive bias that dominates the evidence provided by the training data; and $ii\bigl)$ any $f$-divergence regularization is equivalent to a different $f$-divergence regularization with an appropriate transformation of the empirical risk function.
arXiv.org Artificial Intelligence
Feb-1-2024
- Country:
- Oceania > French Polynesia (0.04)
- North America
- United States
- Wisconsin > Dane County
- Madison (0.04)
- Tennessee > Davidson County
- Nashville (0.04)
- New York > New York County
- New York City (0.04)
- New Jersey > Mercer County
- Princeton (0.04)
- Massachusetts > Middlesex County
- Burlington (0.04)
- California > Alameda County
- Berkeley (0.04)
- Wisconsin > Dane County
- Canada > British Columbia
- United States
- Europe
- France > Provence-Alpes-Côte d'Azur (0.04)
- Finland (0.04)
- United Kingdom > England
- South Yorkshire > Sheffield (0.04)
- Switzerland > Vaud
- Lausanne (0.04)
- Asia > Taiwan
- Taiwan Province > Taipei (0.04)
- Genre:
- Research Report (0.50)
- Technology: