I audit and pressure-test AI agents, workflows, automations, memory systems, and multi-agent architectures to identify brittle assumptions, hidden failure modes, confusing boundaries, and places where a system can behave correctly according to one component while still failing as a whole.