Site Reliability Engineer II
Costa Rica
Workplace: RemoteFull timeCRC 16,932,850 - 30,479,150 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 2+ yearsEducation: bachelorsSkills: ["Collaboration","Accountability","Continuous improvement","Operational excellence","Incident response"]Design and improve automation to enhance reliability, scalability, and efficiency across systems and teams. Partner with engineering to optimize workflows, infrastructure, and applications, and to handle deployment, monitoring, and incident resolution. Build and maintain automated tools that reduce operational toil and improve incident response, while improving observability with SLOs, Prometheus/Grafana, and distributed tracing. Support on-call rotations and contribute to capacity planning for AI compute infrastructure.

