Generalization Analysis of Machine Learning Algorithms via the Worst-Case Data-Generating Probability Measure
Zou, Xinying, Perlaza, Samir M., Esnaola, Iñaki, Altman, Eitan
–arXiv.org Artificial Intelligence
In this paper, the worst-case probability measure over the data is introduced as a tool for characterizing the generalization capabilities of machine learning algorithms. More specifically, the worst-case probability measure is a Gibbs probability measure and the unique solution to the maximization of the expected loss under a relative entropy constraint with respect to a reference probability measure. Fundamental generalization metrics, such as the sensitivity of the expected loss, the sensitivity of the empirical risk, and the generalization gap are shown to have closed-form expressions involving the worst-case data-generating probability measure. Existing results for the Gibbs algorithm, such as characterizing the generalization gap as a sum of mutual information and lautum information, up to a constant factor, are recovered. A novel parallel is established between the worst-case data-generating probability measure and the Gibbs algorithm. Specifically, the Gibbs probability measure is identified as a fundamental commonality of the model space and the data space for machine learning algorithms.
arXiv.org Artificial Intelligence
Dec-19-2023
- Country:
- Oceania > French Polynesia (0.04)
- North America > United States
- New York > New York County
- New York City (0.04)
- New Jersey > Mercer County
- Princeton (0.04)
- New York > New York County
- Europe
- Finland (0.04)
- United Kingdom > England
- Cambridgeshire > Cambridge (0.04)
- Spain > Andalusia
- Cádiz Province > Cadiz (0.04)
- France
- Provence-Alpes-Côte d'Azur (0.05)
- Île-de-France > Paris
- Paris (0.04)
- Asia
- Taiwan > Taiwan Province
- Taipei (0.04)
- Japan > Honshū
- Chūbu > Ishikawa Prefecture > Kanazawa (0.04)
- China > Guangdong Province
- Guangzhou (0.04)
- Taiwan > Taiwan Province
- Genre:
- Research Report (0.40)
- Technology: