AI Interaction Evaluator (Codex / Claude Code, up to $200/hr)
Miami, Calgary, Mississauga, Vancouver, Boston, Montreal, Denver, Oklahoma City, Nashville, Indianapolis, Chicago, Seattle, Atlanta, Ottawa, Toronto, Austin, Dallas, San Antonio, Philadelphia, Richmond, Columbus, Las Vegas, Winnipeg, Boise, Edmonton, Washington
Workplace: RemoteContractUSD 100 - 200 hourlyFunction: People, HR & TalentSkills: ["Subjective but rigorous judgments","Direct feedback","High engineering standards","Engineering judgment"]Evaluate AI coding agent interactions end-to-end by judging whether responses make sense, reasoning and explanations are useful, and outputs reflect strong engineering judgment. Distinguish response quality levels and provide clear, opinionated feedback on what worked, what didn’t, and what felt misleading. Help define what “great” looks like when interacting with tools like Cursor, focusing on engineering taste rather than syntax correctness.
Loading
Loading job details...
Preparing the role view and application actions.

