Member of Technical Staff - Edge Inference Engineer
San Francisco, Boston, Cambridge
Workplace: HybridFull timeFunction: Software EngineeringExperience: 5+ yearsSkills: ["Autonomy","Problem-solving","Collaboration","Communication","Focus on performance"]Seeking an autonomous edge inference engineer to develop and optimize inference kernels for CPU, NPU, and GPU on resource-constrained devices. You will bridge ML and systems, work on quantization strategies, contribute to llama.cpp, and ship production code that impacts model performance on real devices. Collaborative, high-ownership role with exposure to open-source frameworks and edge hardware considerations.

