Recent years have witnessed the proliferation of offensive content online such as fake news, propaganda, misinformation, and disinformation. While initially this was mostly about textual content, over time images and videos gained popularity, as they are much easier to consume, attract more attention, and spread further than text. As a result, researchers started leveraging different modalities and combinations thereof to tackle online multimodal offensive content. In this study, we offer a survey on the state-of-the-art on multimodal disinformation detection covering various combinations of modalities: text, images, speech, video, social media network structure, and temporal information. Moreover, while some studies focused on factuality, others investigated how harmful the content is. While these two components in the definition of disinformation – (i) factuality, and (ii) harmfulness –, are equally important, they are typically studied in isolation. Thus, we argue for the need to tackle disinformation detection by taking into account multiple modalities as well as both factuality and harmfulness, in the same framework. Finally, we discuss current challenges and future research directions.

A Survey on Multimodal Disinformation Detection / Alam, Firoj; Cresci, Stefano; Chakraborty, Tanmoy; Silvestri, Fabrizio; Dimitrov, Dimiter; Da San Martino, Giovanni; Shaar, Shaden; Firooz, Hamed; Nakov, Preslav. - (2022). (Intervento presentato al convegno COLING 2022 tenutosi a Gyeongju, Republic of Korea) [10.48550/arxiv.2103.12541].

A Survey on Multimodal Disinformation Detection

Fabrizio Silvestri;
2022

Abstract

Recent years have witnessed the proliferation of offensive content online such as fake news, propaganda, misinformation, and disinformation. While initially this was mostly about textual content, over time images and videos gained popularity, as they are much easier to consume, attract more attention, and spread further than text. As a result, researchers started leveraging different modalities and combinations thereof to tackle online multimodal offensive content. In this study, we offer a survey on the state-of-the-art on multimodal disinformation detection covering various combinations of modalities: text, images, speech, video, social media network structure, and temporal information. Moreover, while some studies focused on factuality, others investigated how harmful the content is. While these two components in the definition of disinformation – (i) factuality, and (ii) harmfulness –, are equally important, they are typically studied in isolation. Thus, we argue for the need to tackle disinformation detection by taking into account multiple modalities as well as both factuality and harmfulness, in the same framework. Finally, we discuss current challenges and future research directions.
2022
COLING 2022
multimodal disinformation; transformer models
04 Pubblicazione in atti di convegno::04b Atto di convegno in volume
A Survey on Multimodal Disinformation Detection / Alam, Firoj; Cresci, Stefano; Chakraborty, Tanmoy; Silvestri, Fabrizio; Dimitrov, Dimiter; Da San Martino, Giovanni; Shaar, Shaden; Firooz, Hamed; Nakov, Preslav. - (2022). (Intervento presentato al convegno COLING 2022 tenutosi a Gyeongju, Republic of Korea) [10.48550/arxiv.2103.12541].
File allegati a questo prodotto
Non ci sono file associati a questo prodotto.

I documenti in IRIS sono protetti da copyright e tutti i diritti sono riservati, salvo diversa indicazione.

Utilizza questo identificativo per citare o creare un link a questo documento: https://hdl.handle.net/11573/1678450
 Attenzione

Attenzione! I dati visualizzati non sono stati sottoposti a validazione da parte dell'ateneo

Citazioni
  • ???jsp.display-item.citation.pmc??? ND
  • Scopus 45
  • ???jsp.display-item.citation.isi??? ND
social impact