Principal AI Ops Architect, GPS

Scale AI
Doha, London
Workplace: OnsiteFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 6+ yearsSkills: ["Communication","Leadership","Problem-solving"]

Lead the production lifecycle and reliability of full-stack AI applications for international government partners, overseeing end-to-end integration, real-time observability, sovereign data orchestration, and secure cloud infrastructure. Requires deep experience with public sector deployments, modern AI stacks, and translating complex performance metrics for senior officials while driving architectural evolution.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Scale AI
Scale AI
4 months ago

Principal AI Ops Architect, GPS

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Live

Job Summary

Lead the production lifecycle and reliability of full-stack AI applications for international government partners, overseeing end-to-end integration, real-time observability, sovereign data orchestration, and secure cloud infrastructure. Requires deep experience with public sector deployments, modern AI stacks, and translating complex performance metrics for senior officials while driving architectural evolution.
Location: Doha, London
Workplace: Onsite
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Own the production outcome: Take full accountability for the long-term performance and reliability of AI use cases deployed across international government agencies.
  • •Ensure Full-Stack integrity: Oversee the end-to-end health of the platform, ensuring seamless integration between the AI core and all full-stack components, from APIs to UI, to maintain a responsive and production-ready environment.
  • •Scale the feedback loop: Build automated systems to monitor model performance and data drift across geographically dispersed environments, ensuring the right levels of reliability.
  • •Navigate global compliance: Manage the technical lifecycle within diverse regulatory frameworks.
  • •Incident command: Lead the response for production issues in mission-critical environments, ensuring rapid resolution and building guardrails to prevent recurrence.

Key Requirements

  • •6+ years in a high-impact technical role (SRE, FDE or MLOps) with public sector experience.
Experience:6+ yearsPublic sectorGovernment
Skills:CommunicationLeadershipProblem-solving
Languages:English
Tech Stack:KubernetesVector databasesAgentic developmentLLM observability toolsCloudAPIsFrontendBackend

Company Brief

Scale AI
Provides data labeling, annotation, and infrastructure services to accelerate machine learning and AI development. Supplies high-quality training data, tooling, and APIs for customers in autonomous vehicles, mapping, robotics, and enterprise AI applications.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Scaleup
Valuation: Unicorn (USD 1B+)
Funding: Series E+
Headquarters: San Francisco, United States
Founded: 2016
WebsiteLinkedIn