Principal Forward Deployed Engineer - AI PlatformPrincipal Forward Deployed Engineer - AI Platform/Kubernetes/Pytorch

Red Hat
Singapore
Workplace: OnsiteFull timeFunction: DevOps, Cloud & InfrastructureExperience: 8+ yearsSkills: ["Technical mentorship","Communication","Consultative approach","Ambiguity tolerance","Open-source stewardship"]

Build and deploy enterprise AI platform extensions in APAC, solving high-stakes integration blockers with strategic customers and governments. Architect future-proof, secure, and scalable systems for distributed inference, RAG pipelines, and agentic workflows—then upstream 70% of field-built primitives into global open-source communities. Collaborate with global engineering and R&D teams, improve SDLC with new methodologies, and mentor engineers while tuning performance using real-world telemetry.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Red Hat
Red Hat
1 month ago

Principal Forward Deployed Engineer - AI PlatformPrincipal Forward Deployed Engineer - AI Platform/Kubernetes/Pytorch

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 10 days agoStatus: Live

Job Summary

Build and deploy enterprise AI platform extensions in APAC, solving high-stakes integration blockers with strategic customers and governments. Architect future-proof, secure, and scalable systems for distributed inference, RAG pipelines, and agentic workflows—then upstream 70% of field-built primitives into global open-source communities. Collaborate with global engineering and R&D teams, improve SDLC with new methodologies, and mentor engineers while tuning performance using real-world telemetry.
Location: Singapore
Workplace: Onsite
Employment Type: Full time
Job Function: DevOps, Cloud & Infrastructure

Key Responsibilities

  • •Embed with strategic customers and partners to resolve deep engineering blockers such as custom identity/SSO, legacy data pipelines, and air-gapped networking challenges.
  • •Architect and design future-proof software solutions across multiple subsystems that shape the overall architecture of customer AI platforms.
  • •Establish and monitor testing practices for large-scale AI systems to ensure reliability and long-term operational stability for field-developed integrations.
  • •Rapidly prototype and build secure integrations using real-world data, validating distributed inference, advanced RAG pipelines, and autonomous AI agent architectures; harden POCs/MVPs into scalable production architectures.
  • •Act as a global technical bridge and open-source upstream steward by collaborating across global hubs, committing upstream code, evolving SDLC practices, and mentoring engineers across APAC.

Key Requirements

  • •8+ years of experience in system engineering, distributed computing, platform engineering, or AI/ML software development with a history of technical strategy leadership.
  • •Exceptional proficiency in C/C++, Go, and Python, with a track record of shipping production-ready, optimized, and robustly tested code.
  • •Hands-on experience with deep learning frameworks, model fine-tuning (LoRA, QLoRA, SFT), large-scale model serving, and LLM orchestration (LangChain, LlamaIndex).
  • •Strong Kubernetes/container architecture experience, including Custom Resource Definitions (CRDs), custom operators, and multi-subsystem integrations.
  • •Familiarity with hardware-level optimization such as CUDA, ROCm, driver compilation, and GPU/NPU operator configuration.
Experience:8+ yearsAI/MLDistributed computingPlatform engineeringOpen source
Skills:Technical mentorshipCommunicationConsultative approachAmbiguity toleranceOpen-source stewardship
Tech Stack:C/C++GoPythonKubernetesCRDsCustom operatorsDeep learning frameworksLoRAQLoRASFTModel fine-tuningModel servingLangChainLlamaIndexCUDAROCmVLLMDistributed inferenceRAGAgentic AI

Company Brief

Red Hat
Provides enterprise open-source software solutions, including Red Hat Enterprise Linux, middleware, cloud, container, and Kubernetes technologies, delivering support, services, and solutions for hybrid cloud environments.
Industry: Enterprise Software
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Established Company
Valuation: Unicorn (USD 1B+)
Funding: IPO / Publicly Listed
Headquarters: Raleigh, United States
Founded: 1993
WebsiteLinkedIn