Associate - AI Tooling Ops - Platform Reliability Engineer
Pune
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningExperience: 3+ yearsEducation: bachelorsSkills: ["Analytical","Problem-solving","Communication","Stakeholder engagement","Self-motivated"]Join the global Platform Reliability Engineering team as an Associate Platform Reliability Engineer (AI Tooling Ops). You’ll design, build, and maintain AI tooling infrastructure on AWS Kubernetes, monitor health and availability, triage incidents, and run post-incident reviews. Partner across engineering and business teams to improve resilience and operational visibility, reduce toil through automation, and strengthen deployment, monitoring, alerting, and observability using Grafana, Datadog, Prometheus, and OpenTelemetry.
Loading
Loading job details...
Preparing the role view and application actions.

