Site Reliability Engineer- REMOTE - Onsite Training

NTT
Memphis
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 8+ yearsEducation: bachelorsSkills: ["Continuous improvement","Operational excellence","Troubleshooting","Technical leadership","Automation mindset"]

Design and support CI/CD pipelines and containerized deployments across the organization, with a focus on Kubernetes-based reliability. Monitor production and non-production systems, troubleshoot using APM and observability tools (e.g., Datadog, Splunk), and define SLIs to measure service health. Lead automation and testing strategies, improve monitorability, and participate in an on-call rotation while building SRE competency within the Digital organization.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
NTT
NTT
1 week ago

Site Reliability Engineer- REMOTE - Onsite Training

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 15 hours agoStatus: Live

Job Summary

Design and support CI/CD pipelines and containerized deployments across the organization, with a focus on Kubernetes-based reliability. Monitor production and non-production systems, troubleshoot using APM and observability tools (e.g., Datadog, Splunk), and define SLIs to measure service health. Lead automation and testing strategies, improve monitorability, and participate in an on-call rotation while building SRE competency within the Digital organization.
Location: Memphis
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Implement and support CI/CD tools and pipelines across the organization, with GitLab preferred.
  • •Monitor production/non-production systems and troubleshoot issues using observability and logging tools such as Datadog and Splunk.
  • •Maintain and improve containerized applications using Docker and Kubernetes-based deployments.
  • •Collaborate with product and development teams to scale applications/infrastructure and define SLIs for service health.
  • •Create automation solutions and testing/quality gates for deployment, and participate in an on-call rotation.

Key Requirements

  • •8+ years in DevOps/SRE environments, including CI/CD platforms, containerization technologies, and Kubernetes-based deployments.
  • •5+ years troubleshooting using APM tools such as Datadog and Dynatrace.
  • •4+ years Unix/Linux shell scripting and experience with NoSQL databases such as Couchbase.
  • •3+ years with static code analysis tools such as Checkmarx and SonarQube.
  • •Experience with CI/CD tooling, strong log analysis (e.g., Splunk), and OWASP knowledge; Bachelor’s in Computer Science or equivalent experience.
Experience:8+ yearsDevOpsSRECI/CDKubernetesContainerizationObservability
Education:Bachelor's in Computer Science
Skills:Continuous improvementOperational excellenceTroubleshootingTechnical leadershipAutomation mindset
Tech Stack:GitLabCI/CDEKSBambooJenkinsDockerSplunkDatadogDynatraceKubernetesCouchbaseCheckmarxSonarQubeNexus RepositoryServiceNowJiraGroovyShellPythonTerraform

Company Brief

NTT
Dimension Data, operating under NTT Ltd, provides managed IT services, cloud and data center solutions, networking, cybersecurity, and digital transformation services to enterprise customers worldwide.
Industry: Professional Services
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Valuation: Public Company (Market Cap in USD)
Headquarters: London, United Kingdom
Founded: 1983
WebsiteLinkedIn