SRE – Incidents & Monitoring

Job not on LinkedIn

🕒 July 16

🇧🇷 Brazil – Remote

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 14%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of CIAL Dun & Bradstreet

CIAL Dun & Bradstreet

201 - 500 employees

💼 Consulting

📦 Logistics

📣 Marketing

Consulting • Logistics • Marketing

CIAL Dun & Bradstreet is a leading provider of business data and analysis in Latin America and the Caribbean, offering comprehensive solutions for improving B2B business decisions globally. They provide a wide range of services including supply management, credit decisioning, lead generation, compliance, and ESG risk management. Their platforms allow for efficient supplier management, automated credit decision processes, and mitigation of financial, legal, operational, and compliance risks. With a global database including insights from over 587 million companies, CIAL Dun & Bradstreet empowers businesses to streamline operations, enhance credibility, and optimize their decision-making workflows using tools like the D-U-N-S Number and Dunsguide. Their tools and insights offer significant improvements in efficiency, risk management, and data-driven decision-making for their customers.

📋 Description

• Respond to critical production incidents, leading diagnosis through to resolution and post-mortem. • Design and maintain the observability stack (metrics, logs, tracing, alerts) for applications and services. • Reduce MTTR and increase monitoring coverage and service availability. • Participate in on-call rotations and continuously improve runbooks.

🎯 Requirements

• Senior experience in SRE / DevOps / Production Engineering. • Strong knowledge of observability tools (e.g., Datadog, Grafana, Prometheus, or equivalents). • Solid experience with cloud environments and production infrastructure. • Proven experience in incident management and on-call response. • Scripting for automation (Python, Bash, or similar).

🏖️ Benefits

• CIAL provides equal employment opportunities to all employees and applicants for employment and prohibits discrimination and harassment of any type without regard to race, color, religion, age, sex, national origin, disability status, genetics, protected veteran status, sexual orientation, gender identity or expression, or any other characteristic protected by federal, state, or local laws.

Apply Now

Similar Jobs

🕒 April 1

Reproduzindo Talentos

1 - 10

📣 Marketing

🤝 B2B

📱 Media

Analista de automação e agentes de IA desenvolvendo workflows, integrações e chatbots para agência de marketing. Gerenciando n8n, APIs, LLMs, RAG e infraestrutura DevOps.

🗣️🇧🇷🇵🇹 Portuguese Required

Cloud

Docker

GraphQL

JavaScript

Node.js

Python

🕒 March 31

Intuition Machines

51 - 200

🤖 Artificial Intelligence

🔒 Cybersecurity

☁️ SaaS

Senior Site Reliability Engineer at Intuition Machines, enhancing performance and security solutions for large-scale AI-driven systems with global impact.

Cloud

Distributed Systems

JavaScript

Kubernetes

Python

Rust

Go

🕒 March 24

4 Smart Cloud

11 - 50

💼 Consulting

📦 Logistics

🏥 Healthcare

Analista DevOps na 4Smart Cloud desenvolvendo pipelines CI/CD e automação cloud. Garantindo monitoramento, escalabilidade, segurança e estabilidade operacional.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Azure

Cloud

Kubernetes