The state machine covers movement (walking, running, jumping, dancing), emotional reactions, and contextual actions such as thinking, writing code, and searching the web. A separate talking layer provides visemes for lip sync only while the character speaks. Drag, fall, and land interactions let the development team control timing and drop height.