Contra - A professional network for the jobs and skills of the futureA realtor or a business in need of python and LLM AI document processing? Here are some recaps of...
The network for creativity
Join 1.25M professional creatives like you
Connect with clients, get discovered, and run your business 100% commission-free
Creatives on Contra have earned over $150M and we are just getting started
A realtor or a business in need of python and LLM AI document processing? Here are some recaps of my previous internship I did.
Just completed some advanced level python and LLM from extern for a mortgage company.
Check out some of the project I've worked on
Engineered AI-powered document processing pipelines for 200+ page mortgage files using OCR technologies including Tesseract OCR, PaddleOCR, and PyMuPDF, enabling scalable document extraction and parsing.
Developed a Retrieval-Augmented Generation (RAG) system using LlamaIndex for intelligent document search, classification, and contextual retrieval workflows.
Improved retrieval accuracy and system performance through embedding optimization, chunking strategies, vector search tuning, and intelligent routing logic.
Delivered technical documentation, demo UI, and evaluation reports validating OCR accuracy, retrieval precision, and system readiness for production deployment.

github.com

GitHub - Kyl67899/Outamation_Externship_internship

Contribute to Kyl67899/Outamation_Externship_internship development by creating an account on GitHub.

Maty's avatar
The OCR plus RAG pipeline for 200-page mortgage files sounds like a serious test of retrieval quality. Which document type caused the toughest extraction or chunking issue?
Back to feed
The network for creativity
Join 1.25M professional creatives like you
Connect with clients, get discovered, and run your business 100% commission-free
Creatives on Contra have earned over $150M and we are just getting started