Senior System Reliability Engineer

Beyond ONE
Lahore
Workplace: HybridFull timeFunction: Solutions Engineering & Sales EngineeringExperience: 6+ yearsEducation: bachelorsSkills: ["Communication","Collaboration","Ownership"]

As a System Reliability Engineer-II, you will improve the reliability, scalability, and performance of critical platform services, lead automation and observability initiatives, and reduce manual tasks through tooling and infrastructure-as-code. You’ll collaborate with platform, product, and security teams, manage on-call incident response, and drive reliability best practices in a high-scale environment from day one.

This position is no longer accepting applications.

  • See live roles at Beyond ONE
  • Search all live jobs
  • Browse companies, collections, and locations hiring now
Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa

This position is no longer accepting applications.

See live roles at Beyond ONESearch all live jobsBrowse companies, collections, and locations hiring now

Beyond ONE
Beyond ONE
6 months ago

Senior System Reliability Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 5 hours agoStatus: Closed

Job Summary

As a System Reliability Engineer-II, you will improve the reliability, scalability, and performance of critical platform services, lead automation and observability initiatives, and reduce manual tasks through tooling and infrastructure-as-code. You’ll collaborate with platform, product, and security teams, manage on-call incident response, and drive reliability best practices in a high-scale environment from day one.
Location: Lahore
Workplace: Hybrid
Employment Type: Full time
Job Function: Solutions Engineering & Sales Engineering
Seniority: Sr. Manager level

Key Responsibilities

  • •Lead the development of automation and tooling to improve deployment, monitoring, and incident resolution, ensuring greater system efficiency and scalability.
  • •Collaborate with platform, product, and security teams, driving the adoption of reliability best practices across services.
  • •Manage on-call responsibilities and incident management processes, ensuring rapid detection, resolution, and follow-up improvements.
  • •Drive the implementation of observability strategies including metrics, logging, and alerting.
  • •Contribute to infrastructure-as-code, self-healing systems, and continuous improvement in system performance and availability.

Pay and Benefits

Perks:Health InsuranceHybrid Work

Key Requirements

  • •Proficiency in programming/scripting (e.g., Python, Go, Bash).
  • •Strong Linux and networking fundamentals.
  • •Experience with cloud platforms (AWS, GCP, or Azure) and Kubernetes.
  • •Experience with observability tools like Prometheus, Grafana, or ELK; Terraform.
  • •Education: Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent practical experience.
Experience:6+ yearsHigh-scale environmentsDevOpsSite Reliability Engineering
Education:Bachelor's in Computer Science
Skills:CommunicationCollaborationOwnership
Languages:English
Tech Stack:PythonGoBashLinuxNetworkingAWSGCPAzureKubernetesPrometheusGrafanaELKTerraform

Company Brief

Beyond ONE
Provides Web3 services and solutions focused on decentralized identity, digital wallets, and blockchain-based asset management, enabling users and developers to interact with distributed ledger networks and manage credentials and tokens.
Industry: Blockchain & Web3
Website