AI Frameworks Engineer - Model and Kernel Optimization

Intel
Shanghai
Workplace: OnsiteFull timeFunction: Data Science & Machine LearningEducation: mastersSkills: ["Problem-solving","Proactive thinking","Communication"]

Develop high-performance deep learning solutions for customer frameworks by optimizing model and use-case performance, debugging accuracy and memory issues, and designing model deployment architectures. Implement and extend features in vLLM/SGLang (including inference acceleration approaches) and develop/debug high-performance kernels for Intel accelerators. Collaborate closely with teammates, collaborators, and architects to propose solutions, share status, and improve products.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
Intel
Intel
5 days ago

AI Frameworks Engineer - Model and Kernel Optimization

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 8 hours agoStatus: Live

Job Summary

Develop high-performance deep learning solutions for customer frameworks by optimizing model and use-case performance, debugging accuracy and memory issues, and designing model deployment architectures. Implement and extend features in vLLM/SGLang (including inference acceleration approaches) and develop/debug high-performance kernels for Intel accelerators. Collaborate closely with teammates, collaborators, and architects to propose solutions, share status, and improve products.
Location: Shanghai
Workplace: Onsite
Employment Type: Full time
Job Function: Data Science & Machine Learning
Seniority: Graduate level

Key Responsibilities

  • •Optimize performance for key use cases/models and debug/resolution issues related to accuracy and memory management.
  • •Design and develop model deployment architectures, including implementing new features on vLLM/SGLang to accelerate inference.
  • •Develop and debug high-performance kernels for Intel accelerators.
  • •Communicate with teammates, collaborators, and architects to discuss issues, propose solutions, provide status updates, and gather feedback.
  • •Apply innovative ideas to enhance products.

Key Requirements

  • •Master's or Ph.D. in Computer Science, Artificial Intelligence, Software Engineering, or a related field.
  • •Strong programming skills in C++ and Python.
  • •Solid understanding of deep learning fundamentals with hands-on experience.
  • •Proficient in both written and spoken English.
  • •Passion for problem-solving and proactive thinking.
Education:Master's in Computer Science, Artificial Intelligence, Software Engineering, or related field
Skills:Problem-solvingProactive thinkingCommunication
Languages:English
Tech Stack:C++PythonDeep learningLLMsAIGCPyTorchVLLMSGLangGPU kernel development

Company Brief

Intel
Designs and manufactures semiconductor chips, processors, and related hardware for PCs, data centers, networking, and embedded applications, while providing software and services to accelerate computing across industries globally.
Industry: Electronics Manufacturing
Company Size: Enterprise (1,001+ employees)
Revenue: USD 1B+
Growth: Public Company
Valuation: Public Company (Market Cap in USD)
Funding: IPO / Publicly Listed
Headquarters: Santa Clara, United States
Founded: 1968
Glassdoor
Glassdoor: 3.8
WebsiteLinkedInGlassdoor