Site Reliability Engineer

Truist Financial
Atlanta, Charlotte
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 7+ yearsEducation: bachelorsSkills: ["Leadership","Mentoring","Coaching","Cross-team collaboration","Communication"]

Enhance the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. Lead improvements in automation, observability, and incident/problem management, including driving major incident responses and problem management to closure. Standardize enterprise reliability frameworks and observability practices, implement scalable recovery and runbook automation, mentor SRE engineers, and collaborate across delivery, architecture, security, and risk teams to embed resilience into product execution.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Truist Financial
Truist Financial
1 day ago

Site Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 hours agoStatus: Live

Job Summary

Enhance the reliability and operational excellence of enterprise platforms across hybrid cloud and on-premises environments. Lead improvements in automation, observability, and incident/problem management, including driving major incident responses and problem management to closure. Standardize enterprise reliability frameworks and observability practices, implement scalable recovery and runbook automation, mentor SRE engineers, and collaborate across delivery, architecture, security, and risk teams to embed resilience into product execution.
Location: Atlanta, Charlotte
Workplace: Onsite
Employment Type: Full time · Permanent
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Lead major and high-severity incident response efforts, diagnosing root causes and coordinating multi-team technical resolution.
  • •Drive problem management to closure so systemic fixes replace recurring operational risks.
  • •Architect and deliver automation to reduce toil, minimize MTTR, and improve service resilience.
  • •Enhance observability by standardizing telemetry practices across logs, metrics, traces, and events using tools like Dynatrace and Splunk.
  • •Mentor and coach SRE team members, conduct design reviews/knowledge sharing, and support enterprise reliability frameworks and runbooks.

Pay and Benefits

Perks:Health InsuranceDentalVision401kPaid HolidaysPaid LeaveLife InsuranceDisability

Key Requirements

  • •Bachelor’s degree in Computer Science, Software Engineering, or related field.
  • •Minimum 7 years of professional experience in software development.
  • •Deep knowledge of software architecture, design principles, and the software development lifecycle.
  • •Deep understanding of testing, deployment, and security practices.
  • •Strong leadership in major incident management and cross-team technical coordination.
Experience:7+ yearsSite reliability engineeringDevopsPlatform engineeringInfrastructure operationsDistributed systemsCloud-nativeMicroservices
Education:Bachelor's in Computer Science, Software Engineering, or related field
Skills:LeadershipMentoringCoachingCross-team collaborationCommunication
Certifications:Certified Software Development Professional (CSDP)
Languages:English
Tech Stack:KubernetesPythonGoPowerShellAnsibleSplunkDynatraceCI/CDAgileMicroservicesDevOpsAIOpsAI

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.

Company Brief

Truist Financial
Provides consumer and commercial banking, wealth management, insurance, lending, and payments services through a large U.S. financial services platform formed by the merger of BB&T and SunTrust.
Industry: Banking
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Charlotte, United States
Founded: 2019
WebsiteLinkedIn