SRE Engineer

🔥 3 minutes ago

🇧🇷 Brazil – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 12%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Verity Group

Verity Group

51 - 200 employees

Founded 2010

💼 Consulting

🤖 Artificial Intelligence

🔒 Cybersecurity

Consulting • Artificial Intelligence • Cybersecurity

Verity Group is a digital transformation and innovation consulting firm that focuses on delivering real results through technology. With expertise in areas such as AI migration, hyperautomation, cybersecurity, and digital engineering, Verity Group partners with ambitious companies to develop strategic business solutions. The firm prides itself on prioritizing depth over volume, ensuring tailored services that drive efficiency and growth throughout the technology journey, from strategy to execution.

📋 Description

• Work as an SRE Engineer at Verity, a digital transformation and engineering consultancy • Define and track reliability metrics • Implement observability, monitoring, alerting, and APM • Monitor latency, traffic, errors, saturation, availability, and performance • Prevent, identify, and resolve incidents • Conduct root cause analyses and define preventive actions • Identify risks, bottlenecks, and single points of failure • Support the design of resilient, scalable, and highly available solutions • Automate operational activities and reduce manual tasks • Operate and evolve Kubernetes and Docker environments • Support capacity planning, business continuity, and disaster recovery strategies • Participate in deployments and support application stabilization • Collaborate with teams to improve reliability from the design stage onward • Create and maintain dashboards, alerts, procedures, and operational documentation

🎯 Requirements

• Define and track SLIs, SLOs, SLAs, MTTR, and MTTD • Implement observability, monitoring, alerting, and APM • Monitor latency, traffic, errors, saturation, availability, and performance • Prevent, identify, and resolve incidents • Conduct root cause analyses and define actions to prevent recurrence • Identify risks, bottlenecks, and single points of failure • Support the design of resilient, scalable, and highly available solutions • Automate operational activities and reduce manual tasks • Operate and evolve Kubernetes and Docker environments • Support capacity planning, business continuity, and disaster recovery strategies • Participate in deployments and support application stabilization • Work with teams to improve reliability from the design stage onward • Create and maintain dashboards, alerts, procedures, and operational documentation • Hands-on experience with SRE and reliability metrics such as SLI, SLO, SLA, MTTR, and MTTD (inferred from a required application question)

🏖️ Benefits

• Meal allowance • Food allowance • Work-from-home allowance • Medical insurance • Dental insurance • Life insurance • Birthday day off • Total Pass / Wellhub • Boon Saúde app • Discount partnerships • Discounts at participating establishments and educational institutions • Welcome kit • Onboarding • Verity Learning • Verity Break • #VerityComVocê • Access to professional development courses

Apply Now

Similar Jobs

🔥 8 minutes ago

Méliuz

201 - 500

💼 Consulting

📣 Marketing

📦 Logistics

Analista SRE sênior garantindo escalabilidade, resiliência e automação da infraestrutura AWS e Kubernetes do Méliuz. Trabalho remoto em qualquer parte do Brasil.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

DNS

Google Cloud Platform

Grafana

JavaScript

Kubernetes

Linux

NGINX

Node.js

Python

TCP/IP

Terraform

Go

🔥 4 hours ago

Runtalent

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Engenheiro SRE sênior operando e evoluindo plataformas de produção de alta disponibilidade. Monitoramento, observabilidade, incidentes e resiliência em ambientes AWS.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Go

🔥 4 hours ago

Runtalent

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Engenheiro DevSecOps sênior projetando pipelines CI/CD e infraestrutura Terraform para operações AWS/ECS. Integrando segurança, automação e confiabilidade ao ciclo de entrega.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Terraform

🕒 Yesterday

Truelogic Software

501 - 1000

☁️ SaaS

🤝 B2B

🏢 Enterprise

Senior SRE coordinating incident response, observability, and infrastructure operations for a leading financial services client. Supporting mission-critical systems in a regulated environment.

AWS

Jenkins

Linux

Python

🕒 Yesterday

Truelogic Software

501 - 1000

☁️ SaaS

🤝 B2B

🏢 Enterprise

DevOps and Security Engineer building Google Cloud infrastructure, delivery pipelines, and security for a global news organization. Remote role supporting new front-end, backend, and authentication services.

Cloud

DNS

Kubernetes

Terraform