Senior SRE

🕒 August 6

🇧🇷 Brazil – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 17%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Raízen

Raízen

10,000+ employees

Founded 2011

🍽️ Food & Beverage

📦 Logistics

🏭 Manufacturing

Food & Beverage • Logistics • Manufacturing

Raízen is a leading company in the production of ethanol and sugar, as well as the distribution of fuels and related services under the Shell brand in Brazil, Argentina, and Paraguay. The company focuses on renewable energy solutions, producing ethanol from sugarcane for various applications, including fuel and pharmaceutical uses. Raízen also serves the B2B market and is engaged in logistics and sustainability initiatives, operating convenience stores like OXXO and Shell Select to adapt to changing consumer habits and lead trends in mobility and energy efficiency.

📋 Description

• Support distributed environments in Azure and AWS, advancing corporate observability with Grafana, Prometheus, OpenTelemetry, Loki, Tempo and Zabbix • Support initiatives for monitoring applications, infrastructure, logs, metrics and distributed tracing • Perform troubleshooting of critical environments and conduct incident analysis • Define dashboards, alerts and operational indicators • Collaborate with development, architecture and infrastructure teams to ensure application quality, availability and performance • Improve automation, GitOps and monitoring practices for Kubernetes environments (AKS/EKS) • Promote continuous platform improvement with focus on governance, security, operational efficiency and internal team experience

🎯 Requirements

• Bachelor's degree (completed) • Strong experience with observability and monitoring of critical environments using Grafana, Prometheus, Zabbix, OpenTelemetry or similar tools • Experience in troubleshooting and root cause analysis (RCA) in distributed, high-availability environments • Advanced knowledge of Kubernetes, preferably AKS and/or EKS • Experience with public cloud providers, especially Azure and/or AWS • Experience with application monitoring, including collection and analysis of logs, metrics and traces • Solid knowledge of Linux systems, containers and Docker • Experience with network infrastructure, DNS, load balancers, connectivity and network troubleshooting • Ability to develop automations and scripts in Bash, PowerShell and/or Python • Experience with GitOps, CI/CD and operating modern observability-oriented platforms • Experience building, maintaining and optimizing pipelines with GitHub Actions • Nice to have: Grafana stack with Loki, Tempo, Mimir and Prometheus • Nice to have: SRE practices, including SLI, SLO and Error Budgets • Nice to have: ArgoCD, FinOps, Terraform, IaC, hybrid and multi-cloud environments, Dynatrace, Datadog or New Relic • Advantage: Argo Workflows, Argo Events, Service Mesh/Istio, AIOps, Terragrunt, Crossplane and reliability, resilience and scalability patterns

🏖️ Benefits

• Open to applicants of any sexual orientation, gender identity, race, ethnicity and age, including people with disabilities • Hands-on learning • Selection process with an intelligent interview supported by artificial intelligence

Apply Now

Similar Jobs

🕒 August 6

Stefanini Brasil

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Arquiteto DevSecOps evoluindo plataformas críticas OpenShift e Kubernetes na Stefanini. Administrando clusters, automação, segurança, confiabilidade e troubleshooting em escala corporativa.

🗣️🇧🇷🇵🇹 Portuguese Required

Kubernetes

Linux

OpenShift

🕒 August 5

Grupo Adriano Cobuccio

1001 - 5000

📦 Logistics

💼 Consulting

🏭 Manufacturing

DevOps Engineer Sênior construindo pipelines CI/CD e orquestrando Kubernetes. Automatização de infraestrutura segura e escalável para instituição brasileira de pagamentos.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

Cloud

Docker

Google Cloud Platform

Grafana

Kubernetes

Linux

Prometheus

Terraform

🕒 July 31

Experian

10,000+ employees

💼 Consulting

📣 Marketing

📦 Logistics

Seeking a Principal Site Reliability Engineer to lead SRE for a critical product at Experian. Drive technical strategy and operational excellence in a global data technology company.

🗣️🇧🇷🇵🇹 Portuguese Required

🕒 July 31

Quality Digital

1001 - 5000

💼 Consulting

📣 Marketing

📦 Logistics

DevOps Specialist at Quality Digital focusing on Infrastructure as Code and cloud environments management. Hands-on role in automating solutions with Terraform and Ansible.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

AWS

Azure

Cloud

Google Cloud Platform

Kubernetes

Linux

Python

Terraform

🕒 July 30

Experian

10,000+ employees

💼 Consulting

📣 Marketing

📦 Logistics

Principal Site Reliability Engineer leading SRE discipline for critical product at Experian. Collaborating across teams to ensure reliability, efficiency, and operational excellence.

🗣️🇧🇷🇵🇹 Portuguese Required