Next Article in Journal
A One-Step Methodology for Identifying Concrete Pathologies Using Neural Networks—Using YOLO v8 and Dataset Review
Previous Article in Journal
The Total Phenolic Content and Antioxidant Activity of Nine Monofloral Honey Types
 
 
Font Type:
Arial Georgia Verdana
Font Size:
Aa Aa Aa
Line Spacing:
Column Width:
Background:
Article

Limitations of Large Language Models in Propaganda Detection Task

1
Graduate School of Information Science and Technology, Hokkaido University, Sapporo 060-0808, Japan
2
Mateusz Staszków Software Development, 01-234 Warsaw, Poland
3
Faculty of Information Science and Technology, Hokkaido University, Sapporo 060-0808, Japan
*
Author to whom correspondence should be addressed.
Appl. Sci. 2024, 14(10), 4330; https://doi.org/10.3390/app14104330
Submission received: 3 April 2024 / Revised: 26 April 2024 / Accepted: 9 May 2024 / Published: 20 May 2024
(This article belongs to the Section Computing and Artificial Intelligence)

Abstract

Propaganda in the digital era is often associated with online news. In this study, we focused on the use of large language models and their detection of propaganda techniques in the electronic press to investigate whether it is a noteworthy replacement for human annotators. We prepared prompts for generative pre-trained transformer models to find spans in news articles where propaganda techniques appear and name them. Our study was divided into three experiments on different datasets—two based on an annotated SemEval2020 Task 11 corpora and one on an unannotated subset of the Polish Online News Corpus, which we claim to be an even bigger challenge as an example of an under-resourced language. Reproduction of the results of the first experiment resulted in a higher recall of 64.53% than the original run, and the highest precision of 81.82% was achieved for gpt-4-1106-preview CoT. None of our attempts outperformed the baseline F1 score. One of the attempts with gpt-4-0125-preview on original SemEval2020 Task 11 achieved an almost 20% F1 score, but it was below the baseline, which oscillated around 50%. Part of our work that was dedicated to Polish articles showed that gpt-4-0125-preview had a 74% accuracy in the binary detection of propaganda techniques and 69% in propaganda technique classification. The results for SemEval2020 show that the outputs of generative models tend to be unpredictable and are hardly reproducible for propaganda detection. For the time being, these are unreliable methods for this task, but we believe they can help to generate more training data.
Keywords: propaganda detection; media bias; online news analysis; propaganda in online news; propaganda techniques propaganda detection; media bias; online news analysis; propaganda in online news; propaganda techniques

Share and Cite

MDPI and ACS Style

Szwoch, J.; Staszkow, M.; Rzepka, R.; Araki, K. Limitations of Large Language Models in Propaganda Detection Task. Appl. Sci. 2024, 14, 4330. https://doi.org/10.3390/app14104330

AMA Style

Szwoch J, Staszkow M, Rzepka R, Araki K. Limitations of Large Language Models in Propaganda Detection Task. Applied Sciences. 2024; 14(10):4330. https://doi.org/10.3390/app14104330

Chicago/Turabian Style

Szwoch, Joanna, Mateusz Staszkow, Rafal Rzepka, and Kenji Araki. 2024. "Limitations of Large Language Models in Propaganda Detection Task" Applied Sciences 14, no. 10: 4330. https://doi.org/10.3390/app14104330

APA Style

Szwoch, J., Staszkow, M., Rzepka, R., & Araki, K. (2024). Limitations of Large Language Models in Propaganda Detection Task. Applied Sciences, 14(10), 4330. https://doi.org/10.3390/app14104330

Note that from the first issue of 2016, this journal uses article numbers instead of page numbers. See further details here.

Article Metrics

Back to TopTop