Senior Site Reliability Engineer, AI Inference
Dublin
Workplace: HybridFull timeFunction: DevOps, Cloud & InfrastructureSkills: ["Python","C++","Rust","Golang","VLLM","TensorRT","Llama.cpp","Ollama","Docker","Kubernetes","AWS","GCP","Azure","CUDA","TPU","CoreML","NVIDIA GPUs","TTFT","Latency"]Senior AI inference-focused SRE responsible for building and optimizing high-performance inference engines, deploying scalable, low-latency AI services, and ensuring reliability across GPU-rich data centers and edge environments. You’ll work with vLLM, TensorRT, and Kubernetes to maximize throughput, minimize latency, and maintain model accuracy in production.
Loading
Loading job details...
Preparing the role view and application actions.

