Senior Site Reliability Engineer

Job not on LinkedIn

🕒 July 21

🇧🇷 Brazil – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 30%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Sólides

Sólides

501 - 1000 employees

💼 Consulting

🏥 Healthcare

📣 Marketing

Consulting • Healthcare • Marketing

Sólides is a Brazilian HR technology company that runs Escola de Pessoas, a large online learning platform focused on human resources, leadership and people management. It offers free and paid courses, certifications, and expert-led training in areas like talent attraction, development, retention, payroll/departmental personnel, and behavioral management. The platform combines technology and educational content to support HR professionals and organizations in upskilling teams and improving people management practices.

📋 Description

• Build, maintain and evolve clear infrastructure dashboards and intelligent alerts for business rules • Analyze, parse and centralize logs • Configure and monitor APM metrics to optimize performance • Develop and sustain automation solutions for provisioning, configuration and infrastructure deployment using Infrastructure as Code (IaC) • Participate in troubleshooting and production incident resolution • Use observability data for rapid diagnostics and root cause analysis • Collaborate with development teams to promote best practices for resilience, instrumentation and metrics collection

🎯 Requirements

• Proven experience working as an SRE, DevOps Engineer or in infrastructure roles with a strong observability focus • Strong knowledge of monitoring and observability tools • Hands-on experience with Grafana, Prometheus, Alertmanager, Thanos, Fluent Bit, Promtail and the Elastic stack (Elasticsearch, Kibana) • Proficiency in log analysis and parsing, ensuring standardization and quality of structured data • Practical experience with APM tools for detailed application diagnostics • Ability to create effective dashboards and predictive/reactive alerts while avoiding alert fatigue • Good knowledge of Unix/Linux operating systems • Experience with Cloud Computing (AWS, GCP or Azure) • Experience with Docker and Kubernetes • Knowledge of CI/CD practices • Experience with Infrastructure as Code (IaC) tools such as Terraform, Ansible and Puppet • Nice to have: familiarity with Python, Go or Bash; cloud or observability certifications; experience with AI agents and automated workflows; development of AI-assisted solutions

🏖️ Benefits

• Meal allowance (food/meal card) of R$45.00 per working day (Sólides Benefits Card) • Transportation voucher or fuel allowance • Unimed health plan with copayment, no monthly fee • OdontoPrev dental plan, fixed monthly fee of R$21.91 • Therapy: partnership with Psicologia Viva - 3 free sessions per month • Online courses ranging from culinary to postgraduate programs (Qualifica) • Access to all courses from the Escola de Pessoas • Home office allowance of R$60.00 (Sólides Benefits Card) • On-site perks (company manicure, healthy snacks, and more) • Day off during your birthday month • TotalPass • Childcare assistance for parents with children up to 3 years old • Assistance for dependents with special needs (also extended to parents) • Payssego (salary advance program) • Ânima ecosystem agreement (discounts on undergraduate and graduate courses within the group's institutions) • Partnerships with OnHappy and SESC • Annual awards/bonus • Super flexible dress code • Guapeco - pet health insurance

Apply Now

Similar Jobs

🕒 July 20

Verity Group

51 - 200

💼 Consulting

🤖 Artificial Intelligence

🔒 Cybersecurity

SRE/DevOps Engineer managing cloud-native platforms and CI/CD pipelines for Verity's digital transformation projects. Focused on reliability, performance, and automation in high-availability environments.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

AWS

Azure

Cloud

Docker

ElasticSearch

Google Cloud Platform

Grafana

Jenkins

Kubernetes

Linux

Prometheus

Terraform

🕒 July 20

Truelogic Software

501 - 1000

☁️ SaaS

🤝 B2B

🏢 Enterprise

Semi-Senior DevOps Engineer responsible for AWS infrastructure design and maintenance for software company. Collaborating with teams to enhance platform efficiency and reliability.

AWS

Cloud

Grafana

Jenkins

Prometheus

Python

Ruby

Terraform

🕒 July 18

Experian

10,000+ employees

💼 Consulting

📣 Marketing

📦 Logistics

SRE specialist ensuring high availability and scalability in cloud infrastructure at Experian. Collaborating with IT teams and business areas to deliver optimal results.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Cloud

ElasticSearch

Kubernetes

MongoDB

NoSQL

SQL

🕒 July 16

Alloha Fibra

5001 - 10000

📡 Telecommunications

👥 B2C

🤝 B2B

ANALISTA DEVOPS Sênior role at Alloha Fibra combines automation, cloud, and AI practices. Focused on building scalable and resilient technology environments.

🗣️🇧🇷🇵🇹 Portuguese Required

Ansible

AWS

Azure

Cloud

Docker

ETL

Google Cloud Platform

Grafana

Jenkins

Kafka

Kubernetes

Linux

Prometheus

Python

PyTorch

RabbitMQ

Scikit-Learn

Tensorflow

Terraform

🕒 July 14

SecurityScorecard

501 - 1000

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Senior Site Reliability Engineer driving design and optimization of Kubernetes infrastructure. Build AI tooling and optimize CI/CD systems for SecurityScorecard's global cybersecurity platform.

Grafana

Jenkins

Kafka

Kubernetes

Prometheus

Python

Terraform

Go