Article
Enhancing Transformer-Based Language Models for Hungarian Handwritten Text Recognition
2026-02-03
Abstract excerpt
Optical Character Recognition (OCR) is still working on making a multilingual model that incorporates the Hungarian language. We introduce a hybrid Hungarian and English model, one of the biggest challenges is to recognize handwritten text. We are going to investigate a set of models in this research, such as TrOCR large-handwritten, leveraging PULI-BERT, and Roberta-base with Diet models. The digitization of docu...
Topics
Open a Topic to create a Post that cites this publication.
Identifiers and source
- Literature Corpus work
- b8d6e714-d3df-50c8-a646-308e3f43fa20
- DOI
- 10.12688/f1000research.176408.1
