Site Reliability Engineer

Procter & Gamble
Philippines
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 3-6 yearsSkills: ["Problem-solving","Troubleshooting","Communication","Collaboration","Attention to detail"]

Lead incident response and root-cause analysis for critical Warehousing IT systems, minimizing downtime and user impact. Own reliability, monitoring, and alerting by implementing robust observability and automated incident response. Optimize system architecture and configurations for performance, scalability, and fault tolerance, collaborating with software engineers, DevOps teams, and customers. Continuously improve incident management practices through post-incident reviews, preventive measures, and ongoing upskilling.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Procter & Gamble
Procter & Gamble
1 day ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 16 hours agoStatus: Live
Reposted: similar role first listed 2 weeks ago

Job Summary

Lead incident response and root-cause analysis for critical Warehousing IT systems, minimizing downtime and user impact. Own reliability, monitoring, and alerting by implementing robust observability and automated incident response. Optimize system architecture and configurations for performance, scalability, and fault tolerance, collaborating with software engineers, DevOps teams, and customers. Continuously improve incident management practices through post-incident reviews, preventive measures, and ongoing upskilling.
Location: Philippines
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Lead incident response and resolve critical incidents to minimize downtime and user impact.
  • •Implement incident management processes with clear communication, coordination, and documentation.
  • •Perform root cause analysis and implement preventive measures to drive continuous improvement.
  • •Build and maintain monitoring solutions and monitoring/alerting tools for proactive incident response and visibility.
  • •Optimize system architecture and configurations for improved performance, scalability, and fault tolerance while collaborating cross-functionally.

Pay and Benefits

Perks:Health InsuranceGym MembershipEmployee AssistanceRemote Work

Key Requirements

  • •3-6 years of system administration experience with Linux/Unix, cloud platforms (AWS, Azure, or GCP), and SAP.
  • •Experience with configuration management and infrastructure-as-code (e.g., Terraform).
  • •Proficiency in at least one programming/scripting language for automation (e.g., Python or C#).
  • •Networking fundamentals including protocols, infrastructure, load balancing, and DNS management.
  • •Familiarity with containers/orchestration (Docker, Kubernetes), databases with SQL, and monitoring/observability tools (Prometheus, Grafana).
Experience:3-6 years
Skills:Problem-solvingTroubleshootingCommunicationCollaborationAttention to detail
Tech Stack:LinuxUnixAWSAzureGCPSAPTerraformPythonC#DockerKubernetesSQLPrometheusGrafana

Company Brief

Procter & Gamble
Procter & Gamble is a global consumer goods company that develops, manufactures, and markets branded household, personal care, and health products across categories like beauty, grooming, health care, fabric & home care, and baby & family care.
Industry: Consumer Goods
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Cincinnati, United States
Founded: 1837
WebsiteLinkedIn