Media / Publication
November 18, 2024

Panning for gold: Comparative analysis of cross-platform approaches for automated detection of political content in textual data

© Image: Unsplash/Sebastian Kanczok

This paper compares dictionary-based, conventional machine-learning, and deep-learning methods for detecting political content in German-language online material across platforms.

Abstract

To understand and measure political information consumption in the high-choice media environment, we need new methods to trace individual interactions with online content and novel techniques to analyse and detect politics-related information. In this paper, we report the results of a comparative analysis of the performance of automated content analysis techniques for detecting political content in the German language across different platforms. Using three validation datasets, we compare the performance of three groups of detection techniques relying on dictionaries, classic supervised machine learning, and deep learning. We also examine the impact of different modes of data preprocessing on the low-cost implementations of these techniques using a large set (n = 66) of models. Our results show the limited impact of preprocessing on model performance, with the best results for less noisy data being achieved by deep learning- and classic machine learning-based models, in contrast to the more robust performance of dictionary-based models on noisy data.

To continue reading please visit: https://doi.org/10.1371/journal.pone.0312865

This open access article was published in PLoS ONE on 18 November 2024.

Makhortykh, M., León, E. de, Urman, A., Gil-Lopez, T., Christner, C., Sydorova, M., Adam, S., & Maier, M. (2024). Panning for gold: Comparative analysis of cross-platform approaches for automated detection of political content in textual data. PLOS ONE, 19(11), e0312865. https://doi.org/10.1371/journal.pone.0312865

Keywords: Preprocessing, Language, Social media, Machine learning, Parsers, Supervised machine learning

More results /

/ algosoc
AI insiders warn it could end humanity. Does the public agree?

By Ernesto de León • Claes de Vreese • September 17, 2026

Hou DigiD uit Amerikaanse handen nu het nog kan

By José van Dijck • April 30, 2026

/ health
The expectation game

By Martijn Logtenberg • March 13, 2026

Labour politics and why AI hasn't 'fixed' healthcare

By Martijn Logtenberg • November 20, 2025

/ media
The Label Paradox: Can AI transparency create more distrust?

By Natali Helberger • September 21, 2026

How does political efficacy condition clicks on politics? Understanding information-selection behavior in algorithmic feeds over time

By Jin Wan • Theo Araujo • Natali Helberger • Claes de Vreese • September 17, 2026

A Jug of Settled-Down Juice: AI Guidelines on Transparency Obligations

By João Pedro Quintais • September 03, 2026

Subscribe to our newsletter and receive the latest research results, blogs and news directly in your mailbox.

Subscribe