Senior Site Reliability Engineer

🕒 July 1

🌐 Brazil, Honduras, +1 more countries – Remote

infoinfo

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 45%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Zipdev

Zipdev

51 - 200 employees

Founded 2017

💼 Consulting

📦 Logistics

📣 Marketing

Consulting • Logistics • Marketing

Zipdev is a company that specializes in providing remote talent solutions, particularly focused on hiring top Latin American professionals. Zipdev enables companies to build high-performing teams in their own time zones while reducing costs compared to local hiring. The company offers a streamlined hiring process that includes defining hiring needs, selecting top candidates, and onboarding new team members, handling payroll and HR overhead. Zipdev fills a variety of technical and professional roles, such as software engineers, project managers, and designers. The company is recognized for its ability to provide cultural alignment and scalable, flexible hiring solutions.

📋 Description

• Support observability tooling implementation (Datadog and/or Azure Monitor/App Insights) and help build SLO definitions, alert rules, and synthetic checks • Participate in a PagerDuty on-call rotation, including escalation handling and incident documentation • Build and maintain operational runbooks for incident response, rollback, and recovery scenarios • Contribute to deployment automation work (blue/green or canary patterns) and Infrastructure as Code • Work across Azure SQL and Cosmos DB environments, supporting performance and cost optimization initiatives • Collaborate closely with US-based engineers during overlapping working hours

🎯 Requirements

• 5+ years in SRE, DevOps, or cloud infrastructure roles • Strong hands-on experience with Microsoft Azure (Azure SQL, Cosmos DB, Container Apps, App Service) • Experience with observability tooling (Datadog, Azure Monitor, or similar) and on- call/incident response • Familiarity with Infrastructure as Code (Terraform preferred) • Strong written and spoken English; you'll be in daily communication with US-based team members and, at times, client stakeholders • **Availability with meaningful overlap with US Eastern or Mountain time zones** • Experience working in HIPAA-regulated environments, including handling PHI under a Business Associate Agreement (BAA) and working within least-privilege, audited access controls • Willingness to complete a healthcare-industry-standard background check prior to production access • **On-Call Expectations** • This role includes participation in a pager-based on-call rotation via PagerDuty, covering SEV- 1/SEV-2 incidents on a shared schedule with the SRE team. This is a core, required part of the role, not an occasional ask.

🏖️ Benefits

• Work remotely • Vacation: 10 business days a year • Holidays: 5 National Holidays a year • Company Holidays: 5 Company Holidays a year (Christmas Eve, Christmas Day, New Year's Eve, New Year's Day, Zipdev Day) • Parental Leave • Health Care Reimbursement • Active Lifestyle Reimbursement • Quarterly Home Office Reimbursement • Payroll Deduction Purchase Plans • Longevity Bonus • Continuous Learning Bonus • Access to Training and Professional Development Platforms • Did we mention it's REMOTE?!!

Apply Now

Similar Jobs

🕒 June 23

Segware

51 - 200

💼 Consulting

📦 Logistics

🏭 Manufacturing

SRE / SecOps Senior role to enhance security and performance at Segware. Collaborating with teams to implement innovative solutions in monitoring software for client growth.

🗣️🇧🇷🇵🇹 Portuguese Required

Apache

AWS

Docker

Jenkins

Kafka

Kubernetes

Linux

MongoDB

MySQL

Redis

SQL

🕒 June 12

In All Media

1001 - 5000

💼 Consulting

📣 Marketing

☁️ SaaS

Senior DevOps Engineer focusing on migrating workloads from AWS to Azure for a clean energy solutions provider. Leading optimization of cloud environments and deployment workflows.

AWS

Azure

Cloud

Docker

EC2

Jenkins

Kubernetes

Terraform

🕒 June 12

Swile

201 - 500

💳 Fintech

👥 HR Tech

🤝 B2B

Senior Site Reliability Engineer at Swile providing innovative solutions in Fintech, Travel, HR, and Employee Benefits. Focused on problem-solving while enhancing the developer experience in a remote role.

🗣️🇧🇷🇵🇹 Portuguese Required

🕒 May 11

Jusbrasil

201 - 500

💼 Consulting

🏥 Healthcare

📣 Marketing

Senior Site Reliability Engineer at Jusbrasil improving the integrity and performance of product systems. Focusing on data-driven SRE practices and collaborating closely with product teams.

🗣️🇧🇷🇵🇹 Portuguese Required

ElasticSearch

Google Cloud Platform

Grafana

Kubernetes

Prometheus

Terraform

🕒 May 8

Keyrus

1001 - 5000

🤝 B2B

💼 Consulting

DevOps Analyst responsible for designing and maintaining CI/CD pipelines within Keyrus. Collaborating with teams to enhance cloud environments and implement best practices in security and reliability.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

AWS

Azure

Cloud

Docker

Google Cloud Platform

Jenkins

Kubernetes

Linux

Terraform