AI Evaluation Engineer
Sofia
Workplace: HybridFull timeFunction: Data Science & Machine LearningSkills: ["Data-driven mindset","Experiment design","Root-cause analysis","Safety mindset","Healthy scepticism"]Own how SiteGround’s LLM-powered AI products perform in production. You’ll build and maintain evaluation datasets with ground truth, define task-specific metrics, run prompt/agent experiments, and ensure changes are measurable—not guessed. Partner with backend engineers and product teams to root-cause quality and cost issues via trace analysis and observability, iterating on prompts, agent behavior, and safety guardrails across OpenAI, Gemini, and Anthropic ecosystems.
Loading
Loading job details...
Preparing the role view and application actions.

