Associate ML Ops Engineer

AppliedAI / Opus
Abu Dhabi
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 1-2 yearsSkills: ["Communication","Problem-solving","Teamwork","Proactive","Continuous learning"]

Support the reliability, performance, and security of production, development, and staging environments by monitoring systems, collaborating with DevOps/MLOps teams, and helping troubleshoot incidents. Improve observability, SLO/SLI/SLA practices, incident response, and CI/CD pipelines while assisting with capacity planning, performance optimization, and cloud cost optimization. Work with AWS infrastructure patterns, infrastructure as code, and event-driven architectures in an early-career, mentorship-driven on-site role.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
AppliedAI / Opus
AppliedAI / Opus
4 hours ago

Associate ML Ops Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Support the reliability, performance, and security of production, development, and staging environments by monitoring systems, collaborating with DevOps/MLOps teams, and helping troubleshoot incidents. Improve observability, SLO/SLI/SLA practices, incident response, and CI/CD pipelines while assisting with capacity planning, performance optimization, and cloud cost optimization. Work with AWS infrastructure patterns, infrastructure as code, and event-driven architectures in an early-career, mentorship-driven on-site role.
Location: Abu Dhabi
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Entry level

Key Responsibilities

  • •Monitor and maintain production, development, and staging environments to support high availability and performance.
  • •Collaborate with DevOps, MLOps, and development teams to troubleshoot and resolve issues.
  • •Support observability by helping maintain the observability stack with guidance from senior engineers.
  • •Help maintain and improve incident response processes, including participating in on-call rotation.
  • •Assist with capacity planning, performance optimization, SLO/SLI/SLAs, and cloud cost optimization, and support CI/CD deployment processes.

Pay and Benefits

Perks:Health InsurancePaid LeaveVisa Sponsorship

Key Requirements

  • •1–2 years of experience in SRE, DevOps, or a similar role (internships and hands-on projects considered).
  • •Hands-on familiarity with AWS services such as Lambda, ECS/Fargate, ALB/ELB, API Gateway, Route 53, CloudFront, AppSync, DynamoDB, RDS (PostgreSQL), Aurora, EventBridge, SNS/SQS, IAM, and Secrets Manager.
  • •Basic understanding of infrastructure as code and tooling such as CDK and/or Terraform, including modularity and versioning concepts.
  • •Knowledge of observability patterns (logging, metrics, tracing) and exposure to monitoring/observability tools.
  • •Basic scripting and automation skills with a systematic approach to debugging and problem-solving.
Experience:1-2 yearsSREDevOpsCloud infrastructureFinOpsML/LLM operations
Skills:CommunicationProblem-solvingTeamworkProactiveContinuous learning
Certifications:AWS certificationAzure certificationGCP certification
Tech Stack:AWSAzureGCPLambdaECSFargateALBELBAPI GatewayRoute53CloudFrontAppSyncDynamoDBRDSPostgreSQLAuroraEventBridgeSNSSQSSecurity Groups

Eligibility

Work Authorization:Sponsorship available.

Company Brief

AppliedAI / Opus
AppliedAI builds Opus, an agentic AI workflow platform for regulated enterprises. It helps organizations discover, build, run, and optimize governed business processes with human oversight, auditability, and compliance controls, especially in banking, healthcare, insurance, and other regulated sectors.
Industry: Enterprise Software
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Funding: Series B
Headquarters: Abu Dhabi, United Arab Emirates
Founded: 2021
WebsiteLinkedIn