Site Reliability Engineer - SAP Cloud Ops (São Leopoldo, BR, 93022-718)

SAP
Brazil
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Collaboration","Problem-solving","Code-quality focus","Continuous improvement"]

Build and run highly reliable SAP cloud systems as part of a high-performance SRE team. You’ll monitor and troubleshoot critical services, develop tooling and automation, and improve code quality through reviews and pair programming. Partner with development and operations teams on deployments, CI/CD pipelines, observability, SLIs/SLOs, and post-incident root-cause analysis. Automate provisioning and infrastructure on AWS and private data centers using modern IaC and configuration management.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
SAP
SAP
2 hours ago

Site Reliability Engineer - SAP Cloud Ops (São Leopoldo, BR, 93022-718)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 2 hours agoStatus: Live

Job Summary

Build and run highly reliable SAP cloud systems as part of a high-performance SRE team. You’ll monitor and troubleshoot critical services, develop tooling and automation, and improve code quality through reviews and pair programming. Partner with development and operations teams on deployments, CI/CD pipelines, observability, SLIs/SLOs, and post-incident root-cause analysis. Automate provisioning and infrastructure on AWS and private data centers using modern IaC and configuration management.
Location: Brazil
Workplace: Hybrid
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Continuously improve reliability by monitoring, troubleshooting, and identifying engineering defects in the code base.
  • •Collaborate with development teams to implement and deploy new features that meet reliability and performance standards.
  • •Build and automate end-to-end provisioning, including CI/CD pipeline configurations to orchestrate deployments.
  • •Automate monitoring and observability to maintain system health and support high uptime, using SLIs/SLOs and engineering metrics.
  • •Perform post-incident analyses, capacity planning, and maintain documentation for system architecture and troubleshooting procedures.
Travel: Low travel

Key Requirements

  • •Full understanding of DevOps, SRE, and agile software development concepts.
  • •Ability to use one or more languages: Python, Typescript/Javascript, Golang, Java, or C#.
  • •Strong knowledge of Git and software development best practices (GitOps).
  • •Knowledge of IaC and configuration management using Cloud Formation, Terraform, Puppet, and Ansible.
  • •Strong observability and monitoring knowledge (e.g., Dynatrace, New Relic, Prometheus, Grafana, ELK, Splunk) plus experience with CI/CD pipelines and containerized deployments.
Skills:CollaborationProblem-solvingCode-quality focusContinuous improvement
Tech Stack:AWSPythonTypescriptJavascriptGolangJavaC#GitGitOpsLinuxUnixCloudFormationTerraformPuppetAnsibleVPCEC2IAMAPI GatewayAutoscaling

Company Brief

SAP
Global enterprise software company best known for ERP systems and business applications covering finance, supply chain, procurement, HR, analytics, and customer management for large organizations.
Industry: Enterprise Software
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Walldorf, Germany
Founded: 1972
Glassdoor
Glassdoor: 4.1
WebsiteLinkedInGlassdoor