Site Reliability Engineer – Core Streaming

🕒 il y a 9 jours

🇨🇦 Canada – Télétravail

💵 $135 000 - $185 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

👻 Score fantôme 0%

infoinfo

🗣️🇺🇸🇬🇧 Anglais requis

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of Yelp

Yelp

1001 - 5000 employés

Fondée en 2004

🍽️ Alimentation et boissons

🏨 Hôtellerie

📣 Marketing

Food & Beverage • Hospitality • Marketing

Yelp est une plateforme qui connecte les utilisateurs à d'excellentes entreprises locales en fournissant des avis et des informations générés par les utilisateurs. Elle favorise une communauté mondiale en ligne où les individus peuvent partager leurs expériences et recommandations concernant les restaurants, les services et bien plus encore. Yelp fonctionne comme un lieu de travail à distance inclusif, mettant l'accent sur le développement personnel, la communauté et le bien-être de ses employés, tout en s'efforçant constamment de maintenir la confiance et l'authenticité des consommateurs dans ses opérations.

Description

• Own the reliability, scalability, and operational health of Kafka clusters across multi-cloud and hybrid environments • Build and maintain automation for cluster operations, upgrades, capacity scaling, and incident recovery • Partner with engineering teams to enable new streaming use cases, advise on best practices, and ensure data pipeline reliability • Troubleshoot complex issues affecting data flow, performance, or stability • Lead root cause analyses • Execute Kafka version upgrades and platform migrations with minimal disruption to critical services • Participate in on-call rotations using a geographically distributed follow-the-sun model • Drive automation and self-service for deploying, upgrading, and scaling streaming infrastructure

🎯 Exigences

• Solid SRE or infrastructure engineering foundation • Experience with infrastructure-as-code, especially Terraform • Experience with configuration management tools such as Puppet, Ansible, or equivalent • Experience with cloud platforms; AWS preferred • Linux operations experience • Production-level experience with Kafka or similar technologies at scale • Experience with cluster upgrades, migrations, and capacity planning • Programming proficiency in Python, Java, or similar • Strong debugging and systems-thinking skills across distributed systems • Experience with Apache Flink or other stream processing frameworks (nice to have) • Familiarity with Kafka Client APIs, including Producer, Consumer, and Streams (nice to have) • Experience building internal self-service tooling or developer platforms (nice to have) • Experience with incident response and management (nice to have)

🏖️ Avantages

• Fully remote work across Canada • Follow-the-sun on-call model; no one needs to be on-call 24 hours a day • Support from managers, mentors, and teams • Five star benefits (linked in the posting) • Reasonable accommodations for individuals with disabilities in the job application process

Postuler Maintenant

Emplois Similaires

🕒 il y a 10 jours

CloudFactory

1001 - 5000

💼 Conseil

📦 Logistique

📣 Marketing

Senior SRE building scalable infrastructure, CI/CD pipelines, and observability systems at CloudFactory. Improving reliability for production environments supporting AI data operations.

🇨🇦 Canada – Télétravail

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 10 jours

InnoData

2 - 10

🤝 B2B

💼 Conseil

🌍 Impact social

Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.

🇨🇦 Canada – Télétravail

💵 $80 000 - $150 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 13 jours

WinAir

51 - 200

📦 Logistique

💼 Conseil

🏭 Fabrication

DevOps Specialist automating CI/CD and infrastructure for WinAir’s aviation maintenance software. Improving Jenkins, Ansible, Linux environments, deployments, and monitoring across development and production systems.

🇨🇦 Canada – Télétravail

💵 $54 000 - $76 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 14 jours

Software Mind

1001 - 5000

🤖 Intelligence artificielle

☁️ SaaS

📡 Télécommunications

Senior SRE maintaining Kubernetes-based UI and AI service reliability for an enterprise cloud software company. Managing incidents, observability, deployments, and runtime troubleshooting in production.

🇨🇦 Canada – Télétravail

💰 Private Equity Round en 2020-12

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 14 jours

Software Mind

1001 - 5000

🤖 Intelligence artificielle

☁️ SaaS

📡 Télécommunications

Senior SRE supporting Kubernetes production reliability for an enterprise cloud software company. Troubleshooting distributed services, observability, incidents, CI/CD, Node.js, and JVM/Java runtimes.

🇨🇦 Canada – Télétravail

💰 Private Equity Round en 2020-12

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis