Understanding image-text relations and news values for multimodal news analysis-Reference-Cited by-同舟云学术

Understanding image-text relations and news values for multimodal news analysis

Published:2023-05-02 Issue: Volume:6 Page:
ISSN:2624-8212
Container-title:Frontiers in Artificial Intelligence
language:
Short-container-title:Front. Artif. Intell.

Author:

Cheema Gullal S.,Hakimov Sherzod,Müller-Budack Eric,Otto Christian,Bateman John A.,Ewerth Ralph

Abstract

The analysis of news dissemination is of utmost importance since the credibility of information and the identification of disinformation and misinformation affect society as a whole. Given the large amounts of news data published daily on the Web, the empirical analysis of news with regard to research questions and the detection of problematic news content on the Web require computational methods that work at scale. Today's online news are typically disseminated in a multimodal form, including various presentation modalities such as text, image, audio, and video. Recent developments in multimodal machine learning now make it possible to capture basic “descriptive” relations between modalities–such as correspondences between words and phrases, on the one hand, and corresponding visual depictions of the verbally expressed information on the other. Although such advances have enabled tremendous progress in tasks like image captioning, text-to-image generation and visual question answering, in domains such as news dissemination, there is a need to go further. In this paper, we introduce a novel framework for the computational analysis of multimodal news. We motivate a set of more complex image-text relations as well as multimodal news values based on real examples of news reports and consider their realization by computational approaches. To this end, we provide (a) an overview of existing literature from semiotics where detailed proposals have been made for taxonomies covering diverse image-text relations generalisable to any domain; (b) an overview of computational work that derives models of image-text relations from data; and (c) an overview of a particular class of news-centric attributes developed in journalism studies called news values. The result is a novel framework for multimodal news analysis that closes existing gaps in previous work while maintaining and combining the strengths of those accounts. We assess and discuss the elements of the framework with real-world examples and use cases, setting out research directions at the intersection of multimodal learning, multimodal analytics and computational social sciences that can benefit from our approach.

Publisher

Frontiers Media SA

Subject

Artificial Intelligence

Reference151 articles.

1. “Analyzing user modeling on twitter for personalized news recommendations,”;Abel,2011

2. “Twitter-based user modeling for news recommendations,”;Abel,2013

3. “Fact vs. opinion: the role of argumentation features in news classification,”;Alhindi,2020

4. “Cross-modal coherence modeling for caption generation,”;Alikhani;Proceedings of the 58th Annual Meeting of the Association for Computational Linguistics,2020

Cited by 3 articles. 订阅此论文施引文献订阅此论文施引文献，注册后可以免费订阅5篇论文的施引文献，订阅后可以查看论文全部施引文献

1. How verbal text guides the interpretation of advertisement images: a predictive typology of verbal anchoring;Communication Theory;2024-07-22

2. Multimodal prosody: gestures and speech in the perception of prominence in Spanish;Frontiers in Communication;2024-03-27

3. Multimodal news discourse and COVID-19: on the interplay between stylistic features and images in three UK newsbrands’ early framing of (the pandemic in) Italy;Language and Intercultural Communication;2024-03-27