Senior/Lead Site Reliability Engineer – Federal

C3 AI
Virginia, Redwood City
Full timeUSD 159,000 - 230,000Function: DevOps, Cloud & InfrastructureEducation: bachelorsSkills: ["Problem-solving","Critical thinking","Communication"]

Design and implement customized C3 AI Platform installations for Federal environments, ensuring required access and security. Own reliability outcomes by maximizing uptime and availability, building end-to-end monitoring/alerting, and automating incident prevention. Set up infrastructure, tools, and frameworks to streamline deployment cycles, collaborating with Services and Engineering. Deploy and operate scalable, fault-tolerant Kubernetes infrastructure across AWS/Azure (with on-prem experience preferred), traveling to customer sites up to 50%.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
C3 AI
C3 AI
11 months ago

Senior/Lead Site Reliability Engineer – Federal

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 27 minutes agoStatus: Live

Job Summary

Design and implement customized C3 AI Platform installations for Federal environments, ensuring required access and security. Own reliability outcomes by maximizing uptime and availability, building end-to-end monitoring/alerting, and automating incident prevention. Set up infrastructure, tools, and frameworks to streamline deployment cycles, collaborating with Services and Engineering. Deploy and operate scalable, fault-tolerant Kubernetes infrastructure across AWS/Azure (with on-prem experience preferred), traveling to customer sites up to 50%.
Location: Virginia, Redwood City
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Work with Federal customers to design and implement customized C3 AI Platform installations that meet access and security requirements.
  • •Maximize system uptime and availability to meet functional and performance SLAs.
  • •Establish end-to-end monitoring and alerting for critical systems and services.
  • •Build automation and scripting to streamline system updates, upgrades, and prevent problem recurrence.
  • •Set up critical infrastructure, tools, and frameworks to streamline the deployment cycle, collaborating with Services and Engineering.
Travel: High travel

Pay and Benefits

Salary: USD 159,000 - 230,000
Equity and Bonus:Equity

Key Requirements

  • •Bachelor’s degree in a STEM field or comparable area of study.
  • •Active U.S. Government security clearance (Top Secret preferred).
  • •Experience deploying, managing, and operating scalable, fault-tolerant Kubernetes-based infrastructure in AWS and Azure (on-premise preferred).
  • •Expertise in Linux, networking, and database concepts.
  • •Infrastructure-as-Code experience (Terraform, Ansible, or Puppet) and scripting in Ruby, Bash, or Python to automate and monitor systems.
Experience:SaaSKubernetesAWSAzureFederalDevOps
Education:Bachelor's in STEM or comparable area of study
Skills:Problem-solvingCritical thinkingCommunication
Tech Stack:KubernetesAWSAzureGCPLinuxTerraformAnsiblePuppetRubyBashPythonInfrastructure-as-Code

Eligibility

Work Authorization:Authorization required. Sponsorship not provided.
Security Clearance:SecretTop Secret

Company Brief

C3 AI
Provides an enterprise AI platform and applications that help organizations build, deploy, and operate machine learning and generative AI solutions across industries such as manufacturing, energy, defense, and finance.
Industry: AI & Machine Learning
Company Size: Enterprise (1,001+ employees)
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Redwood City, United States
Founded: 2009
WebsiteLinkedIn