Identification and explanation of disinformation in wiki data streams

Research Projects

Organizational Units

Journal Issue

Alternative Title

Abstract

Social media platforms, increasingly used as news sources for varied data analytics, have transformed how information is generated and disseminated. However, the unverified nature of this content raises concerns about trustworthiness and accuracy, potentially negatively impacting readers’ critical judgment due to disinformation. This work aims to contribute to the automatic data quality validation field, addressing the rapid growth of online content on wiki pages. Our scalable solution includes stream-based data processing with feature engineering, feature analysis and selection, stream-based classification, and real-time explanation of prediction outcomes. The explainability dashboard is designed for the general public, who may need more specialized knowledge to interpret the model’s prediction. Experimental results on two datasets attain approximately 90% values across all evaluation metrics, demonstrating robust and competitive performance compared to works in the literature. In summary, the system assists editors by reducing their effort and time in detecting disinformation.

Keywords

Wiki data streams

Document Type

Journal article

Citation

Arriba-Pérez, F., García-Méndez, S., Leal, F., Malheiro, B., & Burguillo, J. C. (2025). Identification and explanation of disinformation in wiki data streams. Integrated Computer-Aided Engineering, (published online: 02 February 2025), 1-17. https://doi.org/10.1177/1069250924130658. Repositório Institucional UPT. https://hdl.handle.net/11328/6149

Identifiers

TID

Designation

Access Type

Open Access

Sponsorship

Description