MAP-Music2Vec: A Simple and Effective Baseline for Self-Supervised Music Audio Representation Learning

Li, Yizhi, Yuan, Ruibin, Zhang, Ge, Ma, Yinghao, Lin, Chenghua, Chen, Xingran, Ragni, Anton, Yin, Hanzhi, Hu, Zhijie, He, Haoyu, Benetos, Emmanouil, Gyenge, Norbert, Liu, Ruibo, Fu, Jie

Dec-5-2022–arXiv.org Artificial Intelligence

The deep learning community has witnessed an exponentially growing interest in self-supervised learning (SSL). However, it still remains unexplored how to build a framework for learning useful representations of raw music waveforms in a self-supervised manner. In this work, we design Music2Vec, a framework exploring different SSL algorithmic components and tricks for music audio recordings. Our model achieves comparable results to the state-of-the-art (SOTA) music SSL model Jukebox, despite being significantly smaller with less than 2% of parameters of the latter. The model will be released on Huggingface(Please refer to: https://huggingface.co/m-a-p/music2vec-v1)

artificial intelligence, machine learning, representation, (10 more...)

arXiv.org Artificial Intelligence

Dec-5-2022

arXiv.org PDF

Add feedback

Country:
- North America > United States
  - Michigan > Washtenaw County > Ann Arbor (0.04)
- Europe
  - United Kingdom > England
    - South Yorkshire > Sheffield (0.05)
    - Greater London > London (0.04)
  - Spain > Andalusia
    - Málaga Province > Málaga (0.04)
  - Germany > Baden-Württemberg
    - Tübingen Region > Tübingen (0.04)
- Asia
  - India > Karnataka
    - Bengaluru (0.04)
  - China > Beijing
    - Beijing (0.04)

Genre:
- Research Report (0.50)

Industry:
- Media > Music (0.71)
- Leisure & Entertainment (0.71)
- Education (0.49)

Technology:
- Information Technology > Artificial Intelligence > Machine Learning > Neural Networks (0.34)

Duplicate Docs Excel Report

Title
None found

Similar Docs Excel Report more

Title	Similarity	Source
None found