End-to-end engineering of production AI pipelines combined with deep-level backend performance optimization and code audits. I build deterministic LLM orchestration systems with strict schema validation while resolving CPU, memory, and database bottlenecks. For compute-heavy workloads, I build native Rust extensions (PyO3) embedded directly into Python runtimes to bypass the GIL and achieve maximum throughput.