Site Reliability Engineer (Buenos Aires)

Site Reliability Engineer (Buenos Aires)

31 ago
|
Peterson Technology Partners
|
Buenos Aires

31 ago

Peterson Technology Partners

Buenos Aires

Our client seeking a Senior Site Reliability Engineer (Sr. SRE) with strong Microsoft Azure experience to help operate, scale, and improve the reliability of our critical cloud-based services. This role will anchor an SRE Team that will focus on production reliability, observability, automation, incident response, cloud operations, and continuous improvement.

The Sr. SRE partners with software engineering, product, security, and operations leadership to embed reliability principles into the software delivery lifecycle, mentors the SRE team, and serves as the technical authority for observability and reliability across mission‐critical healthcare workloads.

Role and Responsibilities

Reliability Engineering

- Define, implement, and improve Service Level Indicators, Service Level Objectives, and error budgets for critical applications and platform services.

- Partner with application teams to improve service reliability, fault tolerance, scalability, and operational readiness.

- Identify and eliminate recurring reliability issues through root cause analysis, automation, and architectural improvements.

- Help design systems that are resilient to Azure region, zone, network, dependency, and deployment failures.

- Participate in production readiness reviews for new services, major releases, and infrastructure changes.

Observability and Monitoring

- Build and improve observability across applications, infrastructure, networks, and cloud services. Implement monitoring for the four golden signals of Latency, Traffic, Errors, and Saturation

- Develop dashboards, alerts, logs, traces, and metrics using tools such as Azure Monitor, Log Analytics, Elastic/ELK, Grafana, OpenTelemetry, Datadog, Dynatrace, New Relic, or similar APM platforms

- Create service health dashboards for engineering, operations, and leadership audiences.

Performance, Capacity, and Resilience:

- Analyze system performance, bottlenecks, saturation trends, and capacity risks.

- Improve backup, disaster recovery, failover, and business continuity practices.

- Partner with engineering teams to implement resiliency patterns such as retries, circuit breakers, bulkheads, graceful degradation, and queue-based decoupling.

Azure Cloud Operations

- Support and improve production workloads running on Microsoft Azure.

- Collaborate with cloud and network teams on secure, scalable Azure architecture.

- Help enforce Azure operational standards, including tagging, monitoring, backup, recovery, identity, security, and cost awareness.

Incident Management and Response

- Conduct blameless post-incident reviews and document root causes, contributing factors, corrective actions, and prevention plans.

- Work with Incident Management, NOC, Help Desk, and application teams to improve response processes and runbooks.

Security and Compliance Support





- Work with Security Engineering to ensure production systems follow cloud security and compliance standards.

- Support operational controls for identity, access, encryption, secrets management, vulnerability remediation, logging, and auditability.

Required Qualifications

- 7 years in site reliability, including hands‐on ownership of mission‐critical services, through a combination of applicable work experience, training, military experience, or education.

- Deep Azure experience including Azure Monitor, Application Insights, AKS, and cloud‐native operations across hybrid infrastructure.

- Proven track record designing and rolling out an SLO program with SLIs, SLOs, and error budget policy in production environments.

Preferred Qualifications

- Prior experience standing up or anchoring an SRE practice, including operating model, rituals, and adoption across multiple engineering teams.

- Strong background in observability, including APM (Datadog, Dynatrace, New Relic), Prometheus, Grafana, and OpenTelemetry.

- Strong incident command experience and a track record of running blameless postmortems that drive measurable improvement.

- Healthcare IT experience and familiarity with HIPAA, HITRUST, or equivalent compliance frameworks.

- Multi‐cloud reliability experience (AWS) in addition to Azure.

- Chaos engineering and resiliency testing experience (e.g., Gremlin, Chaos Mesh, Azure Chaos Studio).

- Infrastructure as Code expertise (Terraform, Bicep, Ansible) for observability and remediation automation.

- Scripting and programming proficiency (Python, PowerShell, Go) for automation, tooling, and integration work.

- Experience with security controls, vulnerability management, compliance audits, and cloud governance.

- Recognized industry certifications (e.g., Azure Solutions Architect, Google SRE certificate, CKA).

Salary/Rate: $20-$25/HR (depends on experience level). This is a contract position with candidates expected to work 40 hours/ week.

About Us

Peterson Technology Partners (PTP) is an Equal Opportunity Employer committed to creating a transparent, inclusive, and human-centered hiring experience.

For more than 28 years, PTP has operated as one of the top IT staffing and recruiting firms in the USA—built on trust, long-term partnerships, and technical excellence.

Based in the Chicago suburb of Park Ridge, IL, our team of more than 500 employees and consultants is dedicated to:

- Helping every client make the best hiring decisions possible





- Matching professionals with the right IT jobs and career opportunities

As part of that commitment, we believe in providing clear information about how our hiring technologies work and how your data is used. The following section outlines our AI-assisted interview process and your rights as a candidate.

AI-Assisted Interview Experience (Pete & Gabi – Rebecca):

To provide a consistent, fair, and versátil experience for all candidates, we use AI-assisted tools to support parts of the interview process. This includes our proprietary AI platform Pete & Gabi, which includes AI recruiter Rebecca.

These AI hiring tools help us:

- Conduct recorded video interviews

- Transcribe interviews

- Summarize candidate responses

- Generate job-related insights

- Streamline communication and scheduling

Please note that

- The AI does NOT make hiring decisions; all decisions are made by our human recruiters, hiring managers, or client partners.

- The AI does not evaluate facial expressions, emotions, or physical traits; it is used only to support fairness, consistency, and efficiency.

If you prefer a non-AI interview format, we will gladly provide an alternative.

Technical or Case Interviews (Role-Dependent): −

When applying for certain tech jobs, you may participate in:

- A technical interview
- A coding challenge
- A case study
- A client-specific assessment

We will always explain what to expect in advance so you can prepare with confidence.

Human Review & Selection:

Every candidate’s profile—including interviews, conversations, and assessments—is reviewed by experienced recruiters and hiring leaders.

AI insights may assist with organization and evaluation, but final decisions are always human-driven.

Your Rights as a Candidate:

At PTP, every candidate has the right to:

- Request a non-AI interview path

- Ask how your data is being used

- Request access to transcripts or interview recordings

- Request deletion of your AI-recorded interview

- Receive clear, timely communication

Our goal is to ensure you feel respected, informed, and supported throughout your experience.

Our Commitment

For more than 28 years, PTP has focused on putting people first—candidates, consultants, employees, and clients.

We’re committed to a hiring process that is:

- Transparent

- Compliant

- Equitable

- Powered by innovative technology that enhances—not replaces—human judgment

Welcome to the future of hiring at Peterson Technology Partners.

We’re excited to learn more about you.

Equal Employment Opportunity

Peterson Technology Partners is an Equal Opportunity Employer. All qualified applicants will receive consideration without regard to race, color, religion, national origin, gender identity, sexual orientation, disability, veteran status, or any other protected characteristic.

📌 Site Reliability Engineer (Buenos Aires)
🏢 Peterson Technology Partners
📍 Buenos Aires

Postulate a este anuncio

Muestra tus habilidades a la empresa, rellenar el formulario y deja un toque personal en la carta, ayudará el reclutador en la elección del candidato.

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (buenos aires) / buenos aires

Suscribete a esta alerta:

Recibe por email las nuevas ofertas de trabajo para: site reliability engineer (buenos aires) / buenos aires