Software - Software Engineer, Compiler (Kernel Optimization)

FuriosaAI
Seoul
Full timeFunction: Software EngineeringEducation: bachelorsSkills: ["Performance analysis","Benchmarking","Collaboration"]

Own performance of critical AI kernels integrated into Furiosa-LLM by analyzing end-to-end LLM serving workloads, profiling and identifying bottlenecks, and implementing highly optimized kernels in TCL. Develop techniques to support dynamic serving workloads via scheduling, specialization, and algorithmic approaches. Integrate, benchmark, and validate kernels across models and serving scenarios while collaborating with compiler and serving teams to improve production capabilities.

Loading

Loading job details...

Preparing the role view and application actions.

FursaFursa
FuriosaAI
FuriosaAI
1 day ago

Software - Software Engineer, Compiler (Kernel Optimization)

✓ Verified Job

Canonical indexed version, validated from employer's careers page.

Source: Company careers pageValidated by: Fursa AI
Last checked: 14 hours agoStatus: Live

Job Summary

Own performance of critical AI kernels integrated into Furiosa-LLM by analyzing end-to-end LLM serving workloads, profiling and identifying bottlenecks, and implementing highly optimized kernels in TCL. Develop techniques to support dynamic serving workloads via scheduling, specialization, and algorithmic approaches. Integrate, benchmark, and validate kernels across models and serving scenarios while collaborating with compiler and serving teams to improve production capabilities.
Location: Seoul
Employment Type: Full time
Job Function: Software Engineering

Key Responsibilities

  • •Analyze end-to-end LLM serving workloads and identify performance bottlenecks that can be addressed via kernel-level optimization.
  • •Design, implement, and optimize high-performance kernels in TCL for critical operations in Furiosa-LLM.
  • •Develop algorithmic techniques to efficiently support dynamic serving workloads.
  • •Integrate, benchmark, and validate optimized kernels across representative models, input shapes, and serving scenarios.
  • •Collaborate with compiler and serving teams to improve compiler capabilities and ensure optimized kernels work effectively in the production software stack.

Key Requirements

  • •BS in Computer Science, Artificial Intelligence, Electrical Engineering, or a related field.
  • •Experience developing low-level or performance-critical software.
  • •Experience analyzing performance bottlenecks using profiling, benchmarking, and hardware performance characteristics.
  • •Understanding of parallel computation, memory hierarchies, and data movement on modern architectures.
Experience:LLM servingAI accelerators
Education:Bachelor's
Skills:Performance analysisBenchmarkingCollaboration
Tech Stack:TCLFuriosa-LLMRNGDTritonCuTilePallasGPUTPUTSMCBroadcomVirtual ISA

Company Brief

FuriosaAI
Designs and develops data‑center AI inference accelerators (RNGD) and a full stack hardware‑software platform to deliver energy‑efficient, high‑performance AI compute for enterprise and cloud customers.
Industry: Hardware Devices
Company Size: Medium (51 to 250 employees)
Growth: Scaleup
Valuation: USD 500M to 1B
Funding: Series C
Headquarters: Seoul, South Korea
Founded: 2017
Glassdoor
Glassdoor: 4.6
WebsiteLinkedInGlassdoor