Efficient Sentiment Analysis: A Resource-Aware Evaluation of Feature Extraction Techniques, Ensembling, and Deep Learning Models
Kamruzzaman, Mahammed, Kim, Gene Louis
–arXiv.org Artificial Intelligence
While reaching for NLP systems that maximize accuracy, other important metrics of system performance are often overlooked. Prior models are easily forgotten despite their possible suitability in settings where large computing resources are unavailable or relatively more costly. In this paper, we perform a broad comparative evaluation of document-level sentiment analysis models with a focus on resource costs that are important for the feasibility of model deployment and general climate consciousness. Our experiments consider different feature extraction techniques, the effect of ensembling, task-specific deep learning modeling, and domain-independent large language models (LLMs). We find that while a fine-tuned LLM achieves the best accuracy, some alternate configurations provide huge (up to 24, 283 *) resource savings for a marginal (<1%) loss in accuracy. Furthermore, we find that for smaller datasets, the differences in accuracy shrink while the difference in resource consumption grows further.
arXiv.org Artificial Intelligence
Aug-3-2023
- Country:
- Oceania > Australia
- North America > United States
- Oregon (0.04)
- New York > New York County
- New York City (0.04)
- Minnesota > Hennepin County
- Minneapolis (0.14)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- Florida > Hillsborough County
- Tampa (0.14)
- University (0.04)
- California > San Francisco County
- San Francisco (0.14)
- Arizona > Maricopa County
- Scottsdale (0.04)
- Europe
- Spain > Valencian Community
- Valencia Province > Valencia (0.04)
- Middle East > Malta
- Port Region > Southern Harbour District > Valletta (0.04)
- Spain > Valencian Community
- Asia > Bangladesh
- Dhaka Division > Dhaka District > Dhaka (0.05)
- Genre:
- Research Report (0.50)
- Industry:
- Energy (0.48)
- Information Technology > Services (0.47)
- Technology: