Senior Site Reliability Engineer

Job not on LinkedIn

🔥 1 minute ago

🇧🇷 Brazil – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 21%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Runtalent

Runtalent

501 - 1000 employees

Founded 2003

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Runtalent is a technology services company that specializes in the allocation of qualified professionals across various tech platforms for project needs. They focus on agile squads and managed services, ensuring operational efficiency and project success. With over 20 years in the market, Runtalent has established a strong talent pool and collaborates with multiple industries including finance, healthcare, and entertainment to provide tailored solutions for each client.

📋 Description

• Operate and evolve post-Go-Live environments • Monitor the availability, performance, and health of production services • Implement and enhance observability practices • Ensure that the platform is understandable, available, and recoverable

🎯 Requirements

• Solid experience as a Site Reliability Engineer (SRE), DevOps Engineer, or in a similar role • Proven experience in highly available production environments • Strong knowledge of AWS • Hands-on experience with monitoring, observability, and alerting • Knowledge of metrics, logs, and distributed tracing • Experience with incident management and response • Knowledge of capacity, performance, availability, and resilience • Experience automating operational processes • Knowledge of SRE practices, SLI, SLO, and SLA • Ability to perform troubleshooting and root cause analysis • Experience with applications and services handling real-time traffic • Hands-on, analytical, and problem-solving mindset • Advanced English proficiency

Apply Now

Similar Jobs

🕒 Yesterday

Truelogic Software

501 - 1000

☁️ SaaS

🤝 B2B

🏢 Enterprise

Senior SRE coordinating incident response, observability, and infrastructure operations for a leading financial services client. Supporting mission-critical systems in a regulated environment.

AWS

Jenkins

Linux

Python

🕒 Yesterday

Truelogic Software

501 - 1000

☁️ SaaS

🤝 B2B

🏢 Enterprise

DevOps and Security Engineer building Google Cloud infrastructure, delivery pipelines, and security for a global news organization. Remote role supporting new front-end, backend, and authentication services.

Cloud

DNS

Kubernetes

Terraform

🕒 2 days ago

NDD Tech | Brasil

501 - 1000

📦 Logistics

🏭 Manufacturing

💼 Consulting

Analista SRE Pleno garantindo disponibilidade e confiabilidade da NDDPay, instituição de pagamento da NDD. Monitorando incidentes, infraestrutura Cloud, banco de dados e mudanças operacionais.

🗣️🇧🇷🇵🇹 Portuguese Required

Cloud

SQL

🕒 3 days ago

Verity Group

51 - 200

💼 Consulting

🤖 Artificial Intelligence

🔒 Cybersecurity

SRE Engineer improving cloud reliability, observability, and incident response for Verity, a digital transformation and engineering consultancy. Automating resilient Kubernetes and Docker environments.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

AWS

Azure

Cloud

Docker

ElasticSearch

Google Cloud Platform

Grafana

Kubernetes

Linux

Prometheus

Terraform

🕒 3 days ago

Franq

51 - 200

🛡️ Insurance

💼 Consulting

💳 Fintech

DevOps/SRE administrando Kubernetes, AWS/GCP e automação de infraestrutura na Franq. Fortalecendo observabilidade, CI/CD, disponibilidade e confiabilidade de sistemas financeiros.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

Apache

AWS

Cloud

Docker

Google Cloud Platform

Grafana

Java

Kubernetes

Linux

NoSQL

Prometheus

Python

SQL

Terraform