Back to search

Article

Cross-Dataset Generalization in Urdu Fake News Detection: An Empirical Study with XLM-RoBERTa and a Length Confound Analysis

2026-06-25

Abstract excerpt

<title>Abstract</title> <p> <bold>Urdu fake news detection (FND)</bold> remains an under-resourced problem despite Urdu being spoken by over 231 million people worldwide. While prior work has demonstrated strong in-domain performance on individual Urdu datasets, whether models trained on one corpus generalise to another has received little systematic attention. This paper presents the first <bold>cross-dataset...

Topics

Open a Topic to create a Post that cites this publication.

Identifiers and source

Literature Corpus work
9631e415-ba1d-5599-a6c7-ccec05d22dd2
DOI
10.21203/rs.3.rs-10135494/v1
Open publication

Related research

Semantic proximity does not establish scientific evidence.

Click a neighbor to travelStep 1 · 12 closest
Interactive article relationship graphSelect a related publication card to move it into the centre and load its closest explainable connections. Solid lines are source-backed structured connections. Dashed lines are semantic discovery signals and are not scientific evidence.
Cross-Dataset Generalization in Urdu Fake News Detection: An Empirical Study with XLM-RoBERTa and a Length Confound AnalysisDOI 10.21203/rs.3.rs-10135494/v1
Select a neighboring publication to make it the new centre.