ML Inference Platform Engineer: Low-Latency AI on GPUs (Buenos Aires)

ML Inference Platform Engineer: Low-Latency AI on GPUs (Buenos Aires)

30 sep
|
Dialpad
|
Buenos Aires

30 sep

Dialpad

Buenos Aires

Dialpad is seeking ML Inference Platform Engineers to build the production systems that serve our in-house AI models at scale. This role lies at the intersection of model development, high-performance runtime systems, and cloud infrastructure.
You will help turn trained models and AI capabilities into reliable, observability-rich, low-latency services running on NVIDIA GPUs in GCP. This is an implementation-heavy role focused on inference machinery, deployment safety, and scalable production

📌 ML Inference Platform Engineer: Low-Latency AI on GPUs (Buenos Aires)
🏢 Dialpad
📍 Buenos Aires

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: ml inference platform engineer: low-latency ai on gpus (buenos aires) / buenos aires

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: ml inference platform engineer: low-latency ai on gpus (buenos aires) / buenos aires