Software Engineer, AI Model Enablement & Inference
Seoul
Workplace: HybridFull timeFunction: Software EngineeringSkills: ["Technical communication","Cross-team collaboration","Analysis","Debugging"]Build production-ready inference support for frontier LLMs on FuriosaAI’s Tensor Contraction Processor (TCP) architecture. You’ll develop and optimize TCL (Tensor Contraction Language) model kernels (including attention and mixture-of-experts), integrate new models into Furiosa-LLM, and validate correctness on NPUs with reference comparisons and regression tests. You’ll also improve onboarding via reusable analysis/integration tools and evaluate approaches from vLLM and SGLang to enhance kernel integration and performance.
Loading
Loading job details...
Preparing the role view and application actions.

