AI Coding Agent and Voice Model EvaluationAI Coding Agent and Voice Model Evaluation
The network for creativity
Join 1.25M professional creatives like you
Connect with clients, get discovered, and run your business 100% commission-free
Creatives on Contra have earned over $150M and we are just getting started
AI Coding Agent & Voice Model Evaluation Ongoing paid work evaluating frontier AI models for an AI training platform.
For coding agents, I work through the same real-world bug in a codebase with two different agents, then score each on six quality dimensions plus a head-to-head comparison: correctness, reasoning, tool use, and where each one went wrong.
For voice models, I rate AI narration quality against detailed rubrics and review other evaluators' work for accuracy and consistency.
What this shows: I know how leading AI agents actually fail on real tasks, and I can evaluate them rigorously and consistently, which is exactly what teams need before shipping an AI feature.
Post image
Back to feed
The network for creativity
Join 1.25M professional creatives like you
Connect with clients, get discovered, and run your business 100% commission-free
Creatives on Contra have earned over $150M and we are just getting started