Author:
Pandeya Yagya Raj,Bhattarai Bhuwan,Lee Joonwhoan
Abstract
AbstractAffective computing has suffered by the precise annotation because the emotions are highly subjective and vague. The music video emotion is complex due to the diverse textual, acoustic, and visual information which can take the form of lyrics, singer voice, sounds from the different instruments, and visual representations. This can be one reason why there has been a limited study in this domain and no standard dataset has been produced before now. In this study, we proposed an unsupervised method for music video emotion analysis using music video contents on the Internet. We also produced a labelled dataset and compared the supervised and unsupervised methods for emotion classification. The music and video information are processed through a multimodal architecture with audio–video information exchange and boosting method. The general 2D and 3D convolution networks compared with the slow–fast network with filter and channel separable convolution in multimodal architecture. Several supervised and unsupervised networks were trained in an end-to-end manner and results were evaluated using various evaluation metrics. The proposed method used a large dataset for unsupervised emotion classification and interpreted the results quantitatively and qualitatively in the music video that had never been applied in the past. The result shows a large increment in classification score using unsupervised features and information sharing techniques on audio and video network. Our best classifier attained 77% accuracy, an f1-score of 0.77, and an area under the curve score of 0.94 with minimum computational cost.
Funder
National Research Foundation of Korea
Publisher
Springer Science and Business Media LLC
Reference77 articles.
1. Montagu, J. How music and instruments began: A brief overview of the origin and entire development of music, from its earliest stages. Front. Sociol. 2, 8. https://doi.org/10.3389/fsoc.2017.00008 (2017).
2. Hallam, S., Cross, I. & Thaut, M. The Oxford Handbook of Music Psychology. Part 1 the Origins and Functions of Music 3–62 (Oxford University Press, 2016).
3. Welch, G. F., Biasutti, M., MacRitchie, J., McPherson, G. E. & Himonides, E. Editorial: The impact of music on human development and well-being. Front. Psychol. 11, 1246. https://doi.org/10.3389/fpsyg.2020.01246 (2020).
4. Juslin, P. N. & Laukka, P. Expression, perception, and induction of musical emotions: A review and a questionnaire study of everyday listening. J. New Music Res. 33, 217–238 (2004).
5. North, A. C. Individual differences in musical taste. Am. J. Psychol. 123(2), 199–208. https://doi.org/10.5406/amerjpsyc.123.2.0199 (2021).
Cited by
10 articles.
订阅此论文施引文献
订阅此论文施引文献,注册后可以免费订阅5篇论文的施引文献,订阅后可以查看论文全部施引文献