Systematic Literature Review of Machine Learning Models and Applications for Text Recognition

📅 2026-08-26
📈 Citations: 0
Influential: 0
📄 PDF
🤖 AI Summary
该文通过系统性文献回顾,评估了近十年OCR技术的发展,分析了97篇相关研究,探讨了解决多语言处理和复杂数据格式的方法及挑战。
📝 Abstract
Optical Character Recognition (OCR) for text recognition using machine vision has significantly improved, particularly when handling heterogeneous textual data. Traditional OCR models struggle with script variations, writing styles, and degraded documents. Advancements in technology are leading to new AI models with improved architecture for handling multiple languages and complex data formats. Despite this progress, a comprehensive evaluation of OCR advancements remains limited. Based on the established preferred reporting items for systematic reviews and meta-analysis (PRISMA) guidelines, this literature review presents an extensive assessment of OCR research to trace the evolution of AI models over the past decade. It explores the transition in AI models, application domains, data types, linguistic coverage, and challenges. Through a detailed analysis of 97 selected studies published during January 2015 - January 2025, key OCR models are identified, and their performance, strengths, and limitations are analyzed. The findings highlight how OCR technologies have evolved to address structured and unstructured text, scene text recognition, and multilingual processing. Unresolved challenges include limited resources for underrepresented languages, high variability in handwritten text, visual similarity among characters, and constraints in real-time OCR applications. To address these issues, several promising approaches are proposed. Key suggestions include self-supervised learning, multimodal AI, automated machine learning (AutoML), AI-assisted postprocessing, tiny machine learning (TinyML), and the creation of joint corpora for script matching. The future recommendations aim to enhance OCR accuracy and tackle the challenges identified for real-time industrial applications. This study will guide future research and establish a foundation for OCR field.
Problem

Research questions and friction points this paper is trying to address.

Optical Character Recognition
handwritten text
multilingual processing
real-time OCR
Innovation

Methods, ideas, or system contributions that make the work stand out.

self-supervised learning
multimodal AI
AutoML
TinyML
🔎 Similar Papers
2024-07-29International Journal of Computer VisionCitations: 2
💼 Related Jobs
No related jobs found.
N
Nuzhat Khan
Faculty of Electrical Engineering, Universiti Teknologi Malaysia, Johor Bahru 81310, Malaysia
A
Ab Al-Hadi Ab Rahman
Faculty of Electrical Engineering, Universiti Teknologi Malaysia, Johor Bahru 81310, Malaysia
S
Shahriyar Masud Rizvi
Department of Electrical and Electronic Engineering, American International University-Bangladesh, 408/1, Kuratoli 1229, Bangladesh
I
Ibrahim Yousef Alshareef
Faculty of Electrical Engineering, Universiti Teknologi Malaysia, Johor Bahru 81310, Malaysia
Muhammad Nadzir Marsono
Muhammad Nadzir Marsono
Faculty of Electrical Engineering, Universiti Teknologi Malaysia, Johor Bahru 81310, Malaysia
M
Muhammad Paend Bakht
Department of Electrical Engineering, Balochistan University of Information Technology, Engineering and Management Sciences, Quetta 87300, Pakistan
M
Mohd Shahrizal Rusli
Faculty of Artificial Intelligence, Universiti Teknologi Malaysia, Kuala Lumpur 54100, Malaysia
S
Shahidatul Sadiah
Faculty of Electrical Engineering, Universiti Teknologi Malaysia, Johor Bahru 81310, Malaysia