Initialization of Feature Selection Search for Classification
Luque-Rodriguez, Maria (Universidad de Cordoba) | Molina-Baena, Jose (Universidad de Cordoba) | Jimenez-Vilchez, Alfonso (Universidad de Cordoba) | Arauzo-Azofra, Antonio (Universidad de Cordoba)
–Journal of Artificial Intelligence Research
Selecting the best features in a dataset improves accuracy and efficiency of classifiers in a learning process. Datasets generally have more features than necessary, some of them being irrelevant or redundant to others. For this reason, numerous feature selection methods have been developed, in which different evaluation functions and measures are applied. This paper proposes the systematic application of individual feature evaluation methods to initialize search-based feature subset selection methods. An exhaustive review of the starting methods used by genetic algorithms from 2014 to 2020 has been carried out. Subsequently, an in-depth empirical study has been carried out evaluating the proposal for different search-based feature selection methods (Sequential forward and backward selection, Las Vegas filter and wrapper, Simulated Annealing and Genetic Algorithms). Since the computation time is reduced and the classification accuracy with the selected features is improved, the initialization of feature selection proposed in this work is proved to be worth considering while designing any feature selection algorithms.
Journal of Artificial Intelligence Research
Nov-27-2022
- Country:
- South America (0.04)
- North America
- Central America (0.04)
- United States
- New Mexico > Bernalillo County
- Albuquerque (0.04)
- Nevada > Clark County
- Las Vegas (0.24)
- New Mexico > Bernalillo County
- Asia
- Indonesia (0.04)
- Malaysia (0.04)
- India > Uttar Pradesh (0.04)
- Middle East
- Republic of Türkiye (0.04)
- Iran (0.04)
- China > Shanghai
- Shanghai (0.04)
- Industry:
- Information Technology > Security & Privacy (1.00)
- Health & Medicine
- Pharmaceuticals & Biotechnology (1.00)
- Diagnostic Medicine > Imaging (1.00)
- Therapeutic Area
- Oncology (1.00)
- Neurology (0.68)
- Endocrinology > Diabetes (0.46)
- Technology: