What You Will Get:
A custom real-time AI Voice Assistant backend utilizing FastAPI and highly optimized bidirectional WebSockets. I handle the heavy lifting—from FFmpeg audio normalization to offloading STT/LLM inference to Groq Cloud (Whisper & Llama-3)—ensuring sub-3-second response times and zero server-side OOM crashes.