Senior Site Reliability Engineer, CCIP

Chainlink Labs
Charlotte, Brazil, Canada, Phoenix, Argentina, Las Vegas, Colombia, Mexico
Workplace: RemoteFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Technical leadership","Cross-team collaboration"]

Own reliability for the CCIP Platform powering Chainlink’s Cross-Chain Interoperability Protocol. Improve deployment safety and delivery velocity, implement distributed tracing for faster incident investigation, and automate to reduce operational toil. Drive adoption of SLOs/SLIs and error budgets to guide engineering decisions, while scaling Kubernetes-based production systems as CCIP grows.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Chainlink Labs
Chainlink Labs
2 months ago

Senior Site Reliability Engineer, CCIP

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 20 hours agoStatus: Live

Job Summary

Own reliability for the CCIP Platform powering Chainlink’s Cross-Chain Interoperability Protocol. Improve deployment safety and delivery velocity, implement distributed tracing for faster incident investigation, and automate to reduce operational toil. Drive adoption of SLOs/SLIs and error budgets to guide engineering decisions, while scaling Kubernetes-based production systems as CCIP grows.
Location: Charlotte, Brazil, Canada, Phoenix, Argentina, Las Vegas, Colombia, Mexico
Workplace: Remote
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure
Seniority: Mid level

Key Responsibilities

  • •Improve deployment safety and increase delivery velocity through production engineering practices.
  • •Establish distributed tracing across the platform to improve observability and accelerate incident investigation.
  • •Eliminate operational toil with automation to increase engineering efficiency and platform reliability.
  • •Drive adoption of SLOs, SLIs, and error budgets to guide engineering decisions and improve service health.
  • •Increase CCIP platform scalability and operational readiness as the protocol grows.

Key Requirements

  • •Demonstrated experience in Site Reliability Engineering, Production Engineering, or operating large-scale distributed systems.
  • •Expertise defining, implementing, and driving adoption of SLOs, SLIs, and error budgets across engineering organizations.
  • •Built and operated production Kubernetes environments supporting critical services.
  • •Applied OpenTelemetry to improve observability across distributed systems.
  • •Experience improving reliability, scalability, and operability of production infrastructure.
Experience:DeFiWeb3Distributed systemsProduction engineeringCrypto-native
Skills:Technical leadershipCross-team collaboration
Tech Stack:KubernetesOpenTelemetry

Company Brief

Chainlink Labs
Chainlink Labs develops decentralized oracle networks and Web3 services that connect smart contracts to real-world data, enabling secure off-chain inputs, outputs, and computation for blockchains and enterprise applications.
Industry: Blockchain & Web3
Company Size: Large (251 to 1,000 employees)
Growth: Scaleup
Headquarters: San Francisco, United States
Founded: 2014
Glassdoor
Glassdoor: 3.6
WebsiteLinkedInGlassdoor