Ml Inference Platform Engineer — Low-Latency Gpu Systems (Buenos Aires)

Ml Inference Platform Engineer — Low-Latency Gpu Systems (Buenos Aires)

09 sep
|
Dialpad
|
Buenos Aires

09 sep

Dialpad

Buenos Aires

Dialpad is hiring a Software Engineer for the ML Inference Platform to build production systems that serve in-house AI models at scale.
You will work at the intersection of model development, runtime systems, and cloud infrastructure on NVIDIA GPUs in GCP.
This is an implementation-heavy role focused on model serving, deployment safety, telemetry, and cost-aware optimization.
You'll help define packaging, deployment, benchmarking, and operations across environments.
#J-*****-Ljbffr

📌 Ml Inference Platform Engineer — Low-Latency Gpu Systems (Buenos Aires)
🏢 Dialpad
📍 Buenos Aires

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: ml inference platform engineer — low-latency gpu systems (buenos aires) / buenos aires

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: ml inference platform engineer — low-latency gpu systems (buenos aires) / buenos aires