30 sep
|
Dialpad
|
Buenos Aires
30 sep
Dialpad
Buenos Aires
Dialpad is seeking ML Inference Platform Engineers to build the production systems that serve our in-house AI models at scale. This role lies at the intersection of model development, high-performance runtime systems, and cloud infrastructure.
You will help turn trained models and AI capabilities into reliable, observability-rich, low-latency services running on NVIDIA GPUs in GCP. This is an implementation-heavy role focused on inference machinery, deployment safety, and scalable production
📌 ML Inference Platform Engineer: Low-Latency AI on GPUs (Buenos Aires)
🏢 Dialpad
📍 Buenos Aires