AI Interaction Evaluator (Codex / Claude Code, up to $200/hr)
Miami
Workplace: RemoteContractUSD 50 - 200 hourlyFunction: People, HR & TalentSkills: ["Opinionated feedback","Subjective but rigorous judgment","Quality evaluation"]Evaluate how AI coding agents (OpenAI Codex and Claude Code) behave in real-world scenarios. You’ll review end-to-end interactions, judge whether responses are useful and correct at a high level, and assess the quality of explanations and reasoning. The focus is on engineering taste—subjective but rigorous judgments—providing clear, opinionated feedback and helping define what “great” looks like when using tools like Cursor.

