Digital - Principal SRE (AI Engineer)

Huntington Bancshares
Columbus
Workplace: HybridFull timeFunction: Data Science & Machine LearningExperience: 5+ yearsEducation: bachelorsSkills: ["Problem-solving","Attention to detail","Communication","Documentation","Collaboration"]

Design, deploy, and maintain AI-driven systems with an SRE focus on reliability, scalability, and performance for mission-critical digital platforms. Own monitoring and incident response using SRE best practices, establish SLOs and error budgets, and integrate machine learning models into production. Partner with cross-functional teams to build AI platform integration abstractions, automate operations, and continuously improve through post-incident analysis, observability, and evaluation metrics.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Huntington Bancshares
Huntington Bancshares
4 months ago

Digital - Principal SRE (AI Engineer)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 4 hours agoStatus: Live

Job Summary

Design, deploy, and maintain AI-driven systems with an SRE focus on reliability, scalability, and performance for mission-critical digital platforms. Own monitoring and incident response using SRE best practices, establish SLOs and error budgets, and integrate machine learning models into production. Partner with cross-functional teams to build AI platform integration abstractions, automate operations, and continuously improve through post-incident analysis, observability, and evaluation metrics.
Location: Columbus
Workplace: Hybrid
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Sr. Manager level

Key Responsibilities

  • •Design, develop, and implement AI-driven systems and automation tools to improve reliability and efficiency of digital platforms.
  • •Monitor health, availability, and performance of AI-enabled applications and infrastructure using SRE best practices.
  • •Integrate machine learning models into production environments to ensure seamless deployment and operation.
  • •Establish and enforce SLOs, error budgets, and incident response procedures for AI-driven services.
  • •Troubleshoot complex AI-related incidents using observability and monitoring tools, then drive continuous improvement via post-incident analysis and automation.

Key Requirements

  • •Bachelor’s degree in computer science, engineering, data science, or a related field.
  • •5+ years hands-on experience in AI/ML engineering, SRE, DevOps, or related roles.
  • •Hands-on programming skills in Python or Java, developing and deploying machine learning models.
  • •Hands-on experience with cloud platforms (AWS, GCP) and containerization (Docker, Kubernetes).
  • •Familiarity with observability tools (Prometheus, Grafana, ELK) and ServiceNow incident management, plus SRE principles (monitoring, alerting, SLOs, error budgets, automation).
Experience:5+ yearsAI/MLMLOpsDevOpsSRE
Education:Bachelor's
Skills:Problem-solvingAttention to detailCommunicationDocumentationCollaboration
Tech Stack:PythonJavaAWSGCPDockerKubernetesPrometheusGrafanaELK stackServiceNowSLOsError budgetsTerraformAnsibleCI/CDMachine learningLLMsOpenAIGoogle

Company Brief

Huntington Bancshares
Regional bank holding company providing commercial and consumer banking, payments, wealth management, and lending services across the Midwestern and select other U.S. markets. It serves individuals, small businesses, and corporate clients through branches and digital channels.
Industry: Banking
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Columbus, United States
Founded: 1866
WebsiteLinkedIn