Software Engineer, ML Inference Platform
Dialpad
Buenos Aires
Workplace: OnsiteFull timeFunction: IT Operations (Systems/Network Admin)Experience: 6+ yearsSkills: ["Scrappy","Curious","Optimistic","Persistent","Empathetic","Operational judgment","Systems thinking","Collaboration","Performance awareness"]Build the production inference platform that serves Dialpad’s in-house AI models at scale. You’ll design and improve low-latency, high-throughput inference serving, GPU utilization, and deployment safety using Kubernetes/GCP and NVIDIA GPUs. Collaborate with model developers and engineers to integrate model serving runtimes (vLLM, Triton, TGI), strengthen observability and debugging, and deliver benchmarking, evaluation, traffic management, and rollback-safe releases.

