We Are Hirning Remote: Site Reliability Engineer (SRE) || Argentina || Contract
Role: PRODUCTION ENGINEER / SRE
Location: Córdoba, Argentina – Remote-friendly with US time-zone overlap
Employment Type: Contract
We are looking for an experienced Production Engineer / Site Reliability Engineer (SRE) with strong hands-on expertise in Grafana, Prometheus, PromQL, Kubernetes, Docker, production on-call, incident troubleshooting, observability, and AWS/Azure, who can support complex production environments, drive operational reliability, and maintain clear runbooks and documentation.
Key Responsibilities
- Participate in a week-long rotational on-call and resolve production incidents.
- Investigate recurring production issues and reduce alert fatigue/noise.
- Develop and maintain runbooks, operational playbooks, and post-incident documentation.
- Build and enhance Grafana dashboards and Prometheus alerting rules.
- Partner with engineering teams to improve observability and deployment reliability.
- Support services across multiple regions, datacenters, and deployment environments.
Required Skills
- 3+ years of experience in Backend Engineering, SRE, Infrastructure, or Production Engineering.
- Hands-on experience with production on-call, incident troubleshooting, and resolution.
- Strong expertise in Grafana, Prometheus, and PromQL.
- Hands-on experience with Kubernetes and Docker.
- Strong documentation and written communication skills.
Nice to Have
- Java, Go, Python, or Shell Scripting.
- AWS and/or Azure experience.
- Multi-region or multi-cloud deployment experience.
- Exposure to compliance-driven environments such as FedRAMP / SOC 2.
Ready for your next SRE opportunity?
Send your updated resume today at
[email protected] with your contact details, expected rate, and availability/notice period to be considered for this urgent contract opportunity!
📌 Site Reliability Engineer (Córdoba)
🏢 Arkhya Tech.
📍 Córdoba