Listing description:
"If you've got an LLM feature or agent that works in testing but you're not confident shipping it, I'll audit it against what actually breaks in production: missing evals, no guardrails, unhandled edge cases, runaway cost or latency. You'll get a concrete, prioritized list of what to fix before you scale it."