Mid-Level SRE

Job not on LinkedIn

🔥 13 hours ago

🇧🇷 Brazil – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 10%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of OZmap

OZmap

11 - 50 employees

☁️ SaaS

📡 Telecommunications

🤝 B2B

SaaS • Telecommunications • B2B

OZmap is a SaaS platform that provides GIS-based mapping and management for fiber-optic (FTTH) and hybrid network infrastructures, tailored for internet service providers (ISPs). It centralizes georeferenced network documentation, planning, monitoring (including OTDR and switch integration), mobile field apps, and open APIs to streamline provisioning, reduce operational costs, and speed repairs. OZmap also offers data migration, integrations with CRMs/ERPs, dashboards, and tools for commercial viability checks to support network growth, M&A and day-to-day operations.

📋 Description

• Prevent production incidents by identifying operational risks, points of failure, and noisy alerts before they become problems. • Participate in building the observability platform (logs, metrics, and tracing), contributing to the definition of processes, SLIs, SLOs, and error budgets. • Conduct root cause analyses and lead postmortems, documenting lessons learned and tracking action plans through to completion. • Resolve critical incidents through troubleshooting in AWS and on-premises production environments. • Identify FinOps opportunities and contribute to cloud cost predictability. • Collaborate with the development team on the continuous improvement of application reliability and performance. • Support scalability initiatives and the creation of new infrastructure, with a focus on automation. • Contribute to the team’s SRE maturity through practices, documentation, and incident management culture.

🎯 Requirements

• Solid experience troubleshooting production environments and distributed systems. • Production experience with AWS (EC2, networking, load balancing, IAM); experience with on-premises environments is a plus. • Knowledge of Docker/Docker Compose, including running containers in production. • Experience with observability and monitoring tools (e.g., Grafana, Prometheus, Datadog, SigNoz, or similar) and OpenTelemetry (logs, metrics, and tracing). • Understanding of SRE practices: SLI/SLO, error budgets, incident management and resolution, and postmortems. • Strong knowledge of Linux, networking, and protocols (HTTP, TCP/IP, DNS). • Strong communication skills, autonomy, and resilience when working during critical incidents, including occasional direct interaction with customers. • Experience with automation (Python, Bash, Terraform, Ansible, or similar) is a plus. • Familiarity with DevOps practices, including CI/CD and infrastructure as code (IaC), is a plus.

Apply Now

Similar Jobs

🔥 15 hours ago

FCamara Consulting & Training

1001 - 5000

🏥 Healthcare

🛡️ Insurance

📦 Logistics

Engenheiro SRE Sênior administrando RDS, Aurora e DynamoDB para a FCamara, ecossistema brasileiro de tecnologia e inovação. Garantindo disponibilidade, performance, segurança e evolução dos ambientes de dados.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Cloud

DynamoDB

EC2

JavaScript

Node.js

NoSQL

Python

SQL

Terraform

.NET

🕒 Yesterday

Franq

51 - 200

🛡️ Insurance

💼 Consulting

💳 Fintech

DevOps/SRE administrando Kubernetes, AWS/GCP e pipelines CI/CD na Franq. Garantindo alta disponibilidade, observabilidade e confiabilidade para inovação financeira.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

Apache

AWS

Cloud

Docker

Google Cloud Platform

Grafana

Java

Kubernetes

Linux

NoSQL

Prometheus

Python

SQL

Terraform

🕒 Yesterday

Compass

10,000+ employees

🏠 Real Estate

📱 Media

DevOps Engineer building Azure DevOps pipelines and Databricks automation for Compass UOL. Managing cloud infrastructure, IaC, security, monitoring, and resilient deployments.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

🕒 2 days ago

Tether.to

11 - 50

₿ Crypto

💳 Fintech

💸 Finance

DevOps Engineer architecting multi-language CI/CD, IaC, and release pipelines. Supporting Tether’s blockchain-powered digital finance products across web, desktop, and mobile platforms.

Android

Ansible

AWS

Distributed Systems

Docker

Firewalls

Grafana

iOS

JavaScript

Linux

Microservices

Prometheus

PyTorch

Shell Scripting

Tensorflow

Terraform

TypeScript

C++

🕒 2 days ago

Verity Group

51 - 200

💼 Consulting

🏥 Healthcare

🛡️ Insurance

SRE Engineer fortalecendo a confiabilidade, observabilidade e disponibilidade de sistemas da consultoria de transformação e engenharia digital Verity. Automatizando operações e evoluindo ambientes Cloud, Kubernetes e Docker.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

AWS

Azure

Cloud

Docker

ElasticSearch

Google Cloud Platform

Grafana

Kubernetes

Linux

Prometheus

Terraform