Production RAG pipeline with FAISS vector search, semantic chunking, and LLM observability. Sub-500ms retrieval on 100K+ documents. FastAPI backend deployed on Render
Built evalflow — an open-source pytest-style quality gate for LLMs. Catches prompt regressions before production by diffing live outputs against a saved baseline. Live on PyPI. github.com/emartai/evalflow