To ensure high accuracy across varying document qualities and lighting conditions, I trained and evaluated multiple state-of-the-art models. The core engine utilized Donut and Tesseract for complex document parsing, alongside PaddleOCR, which I specifically converted into a Paddle Lite format to support lightweight, high-performance edge applications. By fine-tuning these models, the system successfully identified, isolated, and extracted critical text regions—translating transaction details on receipts and key product codes on labels into highly accurate, machine-readable formats.