Article
Enhancing Transformer-Based Language Models for Hungarian Handwritten Text Recognition
2026-03-19
Abstract excerpt
Optical Character Recognition (OCR) is still working on making a multilingual model that incorporates the Hungarian language. We introduce a hybrid Hungarian and English model, one of the biggest challenges is to recognize handwritten text. We are going to investigate a set of models in this research, such as TrOCR large-handwritten, leveraging PULI-BERT, and Roberta-base with Diet models. The digitization of docu...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- e29563aa-a72c-58af-89b9-5d063ef58323
- DOI
- 10.12688/f1000research.176408.2
