AI Infrastructure Engineer
Santa Clara, Austin
Workplace: HybridFull timeUSD 170,500 - 315,490 annuallyFunction: DevOps, Cloud & InfrastructureExperience: 4+ yearsEducation: bachelorsSkills: ["Performance optimization","Profiling","Problem-solving","Open-source collaboration","Cross-stack debugging"]Own end-to-end performance optimization for LLM inference on Intel next-generation GPU architectures. Profile and resolve cross-stack bottlenecks, design and integrate custom GPU kernels for attention, MoE, quantization, and operator fusions, and upstream improvements into open-source inference frameworks like vLLM, SGLang, and PyTorch. Use systematic profiling and roofline analysis to guide hardware roadmap decisions based on real GenAI workload data.
Loading
Loading job details...
Preparing the role view and application actions.

