Complete stack RAG driven document intelligence assistant developed using FastAPI and React. Features multi-format ingestion support (PDF, DOCX, and scanned images through EasyOCR), with FAISS vector indexing, and 6 LLM driven modes operating on Groq’s Llama 3.3 70B model. Dockerized application via GitHub actions CI pipeline for automatic linting and build validation with sub-2 second response time in local tests. FAISS is chosen over Pinecone for quick in-memory computation without relying on cloud infrastructure, while Groq is selected over other options for quicker and cheaper inference than GPT-4 alternatives.