cantnlp@DravidianLangTech2025: A Bag-of-Sounds Approach to Multimodal Hate Speech Detection
–arXiv.org Artificial Intelligence
This paper presents the systems and results for the Multimodal Social Media Data Analysis in Dravidian Languages (MSMDA-DL) shared task at the Fifth Workshop on Speech, Vision, and Language Technologies for Dravidian Languages (DravidianLangTech-2025). We took a `bag-of-sounds' approach by training our hate speech detection system on the speech (audio) data using transformed Mel spectrogram measures. While our candidate model performed poorly on the test set, our approach offered promising results during training and development for Malayalam and Tamil. With sufficient and well-balanced training data, our results show that it is feasible to use both text and speech (audio) data in the development of multimodal hate speech detection systems.
arXiv.org Artificial Intelligence
Mar-16-2025
- Country:
- Oceania > New Zealand (0.04)
- South America > Chile
- North America > United States
- New York > New York County
- New York City (0.04)
- New Mexico > Bernalillo County
- Albuquerque (0.04)
- Louisiana > Orleans Parish
- New Orleans (0.04)
- New York > New York County
- Europe
- Bulgaria (0.04)
- Germany > Berlin (0.04)
- Ukraine > Kyiv Oblast
- Kyiv (0.04)
- Spain > Valencian Community
- Valencia Province > Valencia (0.04)
- Middle East > Malta
- Eastern Region > Northern Harbour District > St. Julian's (0.05)
- Italy > Tuscany
- Florence (0.04)
- Ireland > Leinster
- County Dublin > Dublin (0.04)
- Croatia > Dubrovnik-Neretva County
- Dubrovnik (0.04)
- Asia > Thailand
- Genre:
- Research Report > New Finding (0.86)
- Industry:
- Education (0.46)
- Technology: