Site Reliability Engineer - Warehousing IT Operations

Procter & Gamble
Philippines
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Problem-solving","Troubleshooting","Communication","Collaboration","Continuous learning"]

Lead incident response for critical Warehousing IT systems, ensuring rapid resolution with minimal downtime. Build reliability through robust monitoring, alerting, automated incident response, and resilient system design. Perform root-cause analysis, implement preventive measures, and drive post-incident improvements. Collaborate with software engineers and DevOps teams to optimize architecture and configurations for performance, scalability, and fault tolerance, while mentoring and upskilling within the SRE incident response team.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Procter & Gamble
Procter & Gamble
4 hours ago

Site Reliability Engineer - Warehousing IT Operations

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Lead incident response for critical Warehousing IT systems, ensuring rapid resolution with minimal downtime. Build reliability through robust monitoring, alerting, automated incident response, and resilient system design. Perform root-cause analysis, implement preventive measures, and drive post-incident improvements. Collaborate with software engineers and DevOps teams to optimize architecture and configurations for performance, scalability, and fault tolerance, while mentoring and upskilling within the SRE incident response team.
Location: Philippines
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Entry level

Key Responsibilities

  • •Lead incident response efforts to swiftly resolve critical incidents and minimize downtime.
  • •Implement and improve incident management processes, including clear communication, coordination, and documentation.
  • •Conduct root cause analysis and implement preventive measures to drive continuous improvement.
  • •Ensure reliability via robust monitoring, alerting, and automated incident response systems; configure monitoring tools and improve alert accuracy.
  • •Optimize system architecture and configurations for performance, scalability, and fault tolerance; collaborate cross-functionally and support rotating on-call coverage.

Key Requirements

  • •Familiarity with system administration (Linux/Unix), cloud platforms (AWS, Azure, GCP), and SAP.
  • •Experience with configuration management and infrastructure-as-code (e.g., Terraform).
  • •Proficiency in at least one programming/scripting language (e.g., Python or C#) for automation.
  • •Knowledge of networking fundamentals (protocols, load balancing, DNS) and container/orchestration technologies (Docker, Kubernetes).
  • •Experience with monitoring/observability (e.g., Prometheus, Grafana) and incident response/root-cause analysis; security best practices preferred.
Experience:Site reliability engineeringIncident responseCloudDevOpsInfrastructure as codeWarehousing operations
Skills:Problem-solvingTroubleshootingCommunicationCollaborationContinuous learning
Tech Stack:Linux/UnixAWSAzureGCPSAPTerraformPythonC#Networking protocolsLoad balancingDNS managementDockerKubernetesSQLPrometheusGrafana

Company Brief

Procter & Gamble
Procter & Gamble is a global consumer goods company that develops, manufactures, and markets branded household, personal care, and health products across categories like beauty, grooming, health care, fabric & home care, and baby & family care.
Industry: Consumer Goods
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Cincinnati, United States
Founded: 1837
WebsiteLinkedIn