We Are Hirning Remote: Site Reliability Engineer (SRE) || Argentina || Contract
Role: PRODUCTION ENGINEER / SRE
Location: Córdoba, Argentina – Remote-friendly with US time-zone overlap
Employment Type: Contract
We are looking for an experienced Production Engineer / Site Reliability Engineer (SRE) with strong hands-on expertise in Grafana, Prometheus, PromQL, Kubernetes, Docker, production on-call, incident troubleshooting, observability, and AWS/Azure, who can support complex production environments, drive operational reliability, and maintain clear runbooks and documentation.
Key Responsibilities
Participate in a week-long rotational on-call and resolve production incidents.
Investigate recurring production issues and reduce alert fatigue/noise.
Develop and maintain runbooks, operational playbooks, and post-incident documentation.
Build and enhance Grafana dashboards and Prometheus alerting rules.
Partner with engineering teams to improve observability and deployment reliability.
Support services across multiple regions, datacenters, and deployment environments.
Required Skills
3+ years of experience in Backend Engineering, SRE, Infrastructure, or Production Engineering.
Hands-on experience with production on-call, incident troubleshooting, and resolution.
Strong expertise in Grafana, Prometheus, and PromQL.
Hands-on experience with Kubernetes and Docker.
Strong documentation and written communication skills.
Nice to Have
Java, Go, Python, or Shell Scripting.
AWS and/or Azure experience.
Multi-region or multi-cloud deployment experience.
Exposure to compliance-driven environments such as FedRAMP / SOC 2.
Ready for your next SRE opportunity?
Send your updated resume today at
[email protected] with your contact details, expected rate, and availability/notice period to be considered for this urgent contract opportunity
📌 Site Reliability Engineer (Córdoba)
🏢 Arkhya Tech.
📍 Córdoba