Lead Engineer

Brillio
Bengaluru
Workplace: OnsiteContractFunction: Administration & Executive AssistanceExperience: 5+ yearsSkills: ["Root cause analysis","Troubleshooting","Incident management","Collaboration","Security compliance"]

Design, provision, and support public-cloud infrastructure for high-performance computing (HPC) grid workloads. Migrate complex multi-tier applications to AWS/Azure/GSP at scale, lead monitoring/alerting design, and perform incident/problem troubleshooting with deep root-cause analysis. Collaborate with Tech Risk on security compliance and work across a distributed Agile DevOps fleet with stakeholders, vendors, and internal customers. Requires strong Linux, Kubernetes, and Python/shell SRE skills plus Infrastructure as Code.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Brillio
Brillio
15 hours ago

Lead Engineer

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 13 hours agoStatus: Live
Reposted: similar role first listed 8 months ago

Job Summary

Design, provision, and support public-cloud infrastructure for high-performance computing (HPC) grid workloads. Migrate complex multi-tier applications to AWS/Azure/GSP at scale, lead monitoring/alerting design, and perform incident/problem troubleshooting with deep root-cause analysis. Collaborate with Tech Risk on security compliance and work across a distributed Agile DevOps fleet with stakeholders, vendors, and internal customers. Requires strong Linux, Kubernetes, and Python/shell SRE skills plus Infrastructure as Code.
Location: Bengaluru
Workplace: Onsite
Employment Type: Contract
Job Function: Administration & Executive Assistance
Seniority: Mid level

Key Responsibilities

  • •Design, provision, and support cloud infrastructure for large-scale distributed computing (HPC grids).
  • •Migrate complex multi-tier applications from on-premises to AWS/Azure/GSP at scale.
  • •Design and operate monitoring and alerting solutions.
  • •Troubleshoot production issues and perform in-depth root-cause analysis and mitigation.
  • •Collaborate with Tech Risk on security compliance and provide operational support for a globally distributed team.

Key Requirements

  • •5+ years of experience working with AWS and/or Azure and/or GSP.
  • •Advanced Kubernetes skills.
  • •Experience with Ops support, including incident and problem management.
  • •Expertise in Linux OS internals and administration.
  • •Scripting skills in Python and/or Linux Shell, plus Infrastructure as Code tooling (Terraform, Ansible, others).
Experience:5+ yearsCloudSREHPCInfrastructure
Skills:Root cause analysisTroubleshootingIncident managementCollaborationSecurity compliance
Tech Stack:AWSAzureGSPKubernetesLinuxPythonLinux ShellTerraformAnsibleMonitoringAlertingMetricsLogsInfrastructure as CodeAgileDevOps

Company Brief

Brillio
Provides digital transformation and IT consulting services, including cloud migration, analytics, digital engineering, and managed services to accelerate business outcomes for enterprises across industries.
Industry: Consulting
Company Size: Enterprise (1,001+ employees)
Growth: Established Company
Funding: Private Equity Backed
Headquarters: Santa Clara, United States
Founded: 2014
WebsiteLinkedIn