Efficient smart OCR solution for banking document digitization
| dc.contributor.advisor | Alam, Md. Golam Rabiul | |
| dc.contributor.author | Islam, Maria | |
| dc.date.accessioned | 2026-01-18T05:15:53Z | |
| dc.date.available | 2026-01-18T05:15:53Z | |
| dc.date.issued | 2025-10 | |
| dc.description | Cataloged from PDF version of internship report. | |
| dc.description | Includes bibliographical references (page 48). | |
| dc.description | This internship report is submitted in partial fulfillment of the requirements for the degree of Bachelor of Science in Computer Science, 2025. | |
| dc.description.abstract | The digitization of multilingual banking documents, particularly those containing handwritten Bengali and English scripts, poses significant challenges due to variable handwriting styles, document noise, and domain-specific terminology. This study presents a hybrid Optical Character Recognition (OCR) and language model–based pipeline designed to achieve high-fidelity text extraction and correction for banking document digitization. The proposed system integrates two stateof- the-art OCR architectures—Tesseract, EasyOCR OCR for robust unstructured Raw text extraction and GPT-3.5,LLaMA-2 for end-toend handwritten text recognition—with advanced language models for post-processing. Bengali text correction is performed using Gemma- 7B and BLOOM-7B, while English text is refined through GPT-3.5 and LLaMA-2 (7B-chat). The dataset comprising paired images and annotations for both languages, undergoes preprocessing, binarization ,noise reduction, skew correction and redundancy filtering before model training and evaluation. Experimental results show substantial improvements in linguistic accuracy and semantic preservation compared to baseline OCR outputs, demonstrating the system’s applicability for real-world multilingual banking document digitization. | |
| dc.identifier.other | ID 20301304 | |
| dc.identifier.other | https://dspace.bracu.ac.bd/server/api/core/items/4de1301a-7f24-44ff-93c0-bb9c5f0f79df | |
| dc.identifier.uri | http://hdl.handle.net/10361/27444 | |
| dc.language.iso | en | |
| dc.publisher | BRAC University | |
| dc.source | BRAC University Institutional Repository | |
| dc.subject | Multilingual documents | |
| dc.subject | Documents digitization | |
| dc.subject | Hybrid OCR | |
| dc.subject | Natural language processing | |
| dc.subject | Text correction | |
| dc.subject | Transformer models | |
| dc.subject | Banking documents | |
| dc.subject | Bengali language | |
| dc.subject | Handwriting recognition | |
| dc.title | Efficient smart OCR solution for banking document digitization | |
| dc.type | Internship Report |
Files
Original bundle
1 - 1 of 1
