TABCF: Counterfactual Explanations for Tabular Data Using a Transformer-Based VAE
Panagiotou, Emmanouil, Heurich, Manuel, Landgraf, Tim, Ntoutsi, Eirini
–arXiv.org Artificial Intelligence
In the field of Explainable AI (XAI), counterfactual (CF) explanations are one prominent method to interpret a black-box model by suggesting changes to the input that would alter a prediction. In real-world applications, the input is predominantly in tabular form and comprised of mixed data types and complex feature interdependencies. These unique data characteristics are difficult to model, and we empirically show that they lead to bias towards specific feature types when generating CFs. To overcome this issue, we introduce TABCF, a CF explanation method that leverages a transformer-based Variational Autoencoder (VAE) tailored for modeling tabular data. Our approach uses transformers to learn a continuous latent space and a novel Gumbel-Softmax detokenizer that enables precise categorical reconstruction while preserving end-to-end differentiability. Extensive quantitative evaluation on five financial datasets demonstrates that TABCF does not exhibit bias toward specific feature types, and outperforms existing methods in producing effective CFs that align with common CF desiderata.
arXiv.org Artificial Intelligence
Oct-14-2024
- Country:
- North America
- United States > New York
- Kings County > New York City (0.05)
- New York County > New York City (0.04)
- Canada > Ontario
- Toronto (0.04)
- United States > New York
- Europe
- Spain > Basque Country
- Biscay Province > Bilbao (0.04)
- Germany
- Berlin (0.04)
- North Rhine-Westphalia > Upper Bavaria
- Munich (0.04)
- Bavaria > Upper Bavaria
- Munich (0.04)
- Spain > Basque Country
- North America
- Genre:
- Research Report (0.82)
- Industry:
- Banking & Finance > Credit (0.68)
- Technology: