Deep-Learning-Based Multimodal Emotion Classification for Music Videos-Reference-Cited by-同舟云学术

Deep-Learning-Based Multimodal Emotion Classification for Music Videos

Published:2021-07-20 Issue:14 Volume:21 Page:4927
ISSN:1424-8220
Container-title:Sensors
language:en
Short-container-title:Sensors

Author:

Pandeya Yagya Raj^ORCID,Bhattarai Bhuwan^ORCID,Lee Joonwhoan

Abstract

Music videos contain a great deal of visual and acoustic information. Each information source within a music video influences the emotions conveyed through the audio and video, suggesting that only a multimodal approach is capable of achieving efficient affective computing. This paper presents an affective computing system that relies on music, video, and facial expression cues, making it useful for emotional analysis. We applied the audio–video information exchange and boosting methods to regularize the training process and reduced the computational costs by using a separable convolution strategy. In sum, our empirical findings are as follows: (1) Multimodal representations efficiently capture all acoustic and visual emotional clues included in each music video, (2) the computational cost of each neural network is significantly reduced by factorizing the standard 2D/3D convolution into separate channels and spatiotemporal interactions, and (3) information-sharing methods incorporated into multimodal representations are helpful in guiding individual information flow and boosting overall performance. We tested our findings across several unimodal and multimodal networks against various evaluation metrics and visual analyzers. Our best classifier attained 74% accuracy, an f1-score of 0.73, and an area under the curve score of 0.926.

Funder

National Research Foundation of Korea

Publisher

MDPI AG

Subject

Electrical and Electronic Engineering,Biochemistry,Instrumentation,Atomic and Molecular Physics, and Optics,Analytical Chemistry

Link

https://www.mdpi.com/1424-8220/21/14/4927/pdf

Reference88 articles.

1. Machine Recognition of Music Emotion

2. Expression, Perception, and Induction of Musical Emotions: A Review and a Questionnaire Study of Everyday Listening

3. Music listening as self-enhancement: Effects of empowering music on momentary explicit and implicit self-esteem

4. Effects of music and music therapy on mood in neurological patients

5. Music as a Mood Modulator. Retrospective Theses and Dissertations, 1992, 17311https://lib.dr.iastate.edu/rtd/17311

Cited by 46 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. Video2Music: Suitable music generation from videos using an Affective Multimodal Transformer model;Expert Systems with Applications;2024-09

2. Using artificial intelligence to analyze and classify music emotion;Journal of Computational Methods in Sciences and Engineering;2024-08-14

3. Cascaded cross-modal transformer for audio–textual classification;Artificial Intelligence Review;2024-08-02

4. A dataset for multimodal music information retrieval of Sotho-Tswana musical videos;Data in Brief;2024-08

5. Detecting and Explaining Emotions in Video Advertisements;Proceedings of the 47th International ACM SIGIR Conference on Research and Development in Information Retrieval;2024-07-10