UQ-ARMED: Uncertainty quantification of adversarially-regularized mixed effects deep learning for clustered non-iid data
Treacher, Alex, Nguyen, Kevin, Owens, Dylan, Heitjan, Daniel, Montillo, Albert
–arXiv.org Artificial Intelligence
This work demonstrates the ability to produce readily interpretable statistical metrics for model fit, fixed effects covariance coefficients, and prediction confidence. Importantly, this work compares 4 suitable and commonly applied epistemic UQ approaches, BNN, SWAG, MC dropout, and ensemble approaches in their ability to calculate these statistical metrics for the ARMED MEDL models. In our experiment, not only do the UQ methods provide these benefits, but several UQ methods maintain the high performance of the original ARMED method, some even provide a modest (but not statistically significant) performance improvement. The ensemble models, especially the ensemble method with a 90% subsampling, performed well across all metrics we tested with (1) high performance that was comparable to the non-UQ ARMED model, (2) properly deweights the confounds probes and assigns them statistically insignificant p-values, (3) attains relatively high calibration of the output prediction confidence. The MC dropout models showed the lowest performance, and failed to provide non-statistically significant fixed effects covariate coefficients. The SWAG model's performance was dependent on the learning rate.
arXiv.org Artificial Intelligence
Nov-28-2022
- Country:
- North America > United States > New York (0.04)
- Genre:
- Research Report
- Experimental Study (1.00)
- New Finding (0.88)
- Research Report
- Industry:
- Health & Medicine > Therapeutic Area > Neurology (1.00)
- Technology: