Article
Cross-Dataset Generalization in Urdu Fake News Detection: An Empirical Study with XLM-RoBERTa and a Length Confound Analysis
2026-06-25
Abstract excerpt
<title>Abstract</title> <p> <bold>Urdu fake news detection (FND)</bold> remains an under-resourced problem despite Urdu being spoken by over 231 million people worldwide. While prior work has demonstrated strong in-domain performance on individual Urdu datasets, whether models trained on one corpus generalise to another has received little systematic attention. This paper presents the first <bold>cross-dataset...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- 9631e415-ba1d-5599-a6c7-ccec05d22dd2
- DOI
- 10.21203/rs.3.rs-10135494/v1
