Senior Site Reliability Engineer

🕒 il y a 1 mois

🇨🇦 Canada – Télétravail

💵 $197 500 - $225 000 / an

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

👻 Score fantôme 1%

infoinfo

🗣️🇺🇸🇬🇧 Anglais requis

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of SecurityScorecard

SecurityScorecard

501 - 1000 employés

Fondée en 2013

💼 Conseil

🏥 Santé

🛡️ Assurance

💰 €180 000 000 Series E en 2021-03

Consulting • Healthcare • Insurance

SecurityScorecard est une entreprise spécialisée dans la cybersécurité et la gestion des risques. Elle propose des solutions de détection et de réponse au sein de la supply chain, de gestion des risques cyber des tiers (TPRM) et de gestion de la surface d’attaque externe (EASM). En s’appuyant sur l’IA et des Security Ratings, SecurityScorecard aide les organisations à améliorer leur posture de cybersécurité et à gérer efficacement leurs risques. Sa plateforme permet de surveiller et de remédier aux vulnérabilités, de collaborer avec les fournisseurs et d’assurer la conformité aux exigences réglementaires. SecurityScorecard accompagne un large éventail de secteurs, notamment le secteur public, la technologie, la santé, les services financiers, et bien d’autres. Sa suite complète d’outils et de services, tels que Security Ratings et MAX, permet aux organisations de gérer de manière proactive les risques cyber et de renforcer leur dispositif de sécurité global.

Description

• Design, build, and scale Kubernetes infrastructure for secure, multi-tenant, high-availability applications. • Build and operate AI tooling infrastructure — stand up MCP servers and establish secure, governed AI access and guardrails for production systems. • Optimize and maintain CI/CD pipelines, improving reliability, speed, and rollback safety. • Implement progressive delivery strategies such as blue/green and canary deployments. • Advance Infrastructure as Code with Terraform, Helm, and Argo CD, defining reusable patterns for the org. • Operate and optimize streaming and analytics infrastructure: Kafka, Flink, and ClickHouse. • Build automated testing into the CI/CD lifecycle. • Improve system observability — define SLOs, alerts, and dashboards. • Lead incident response and postmortems, focusing on root cause and durable fixes. • Mentor engineers across teams on Kubernetes, CI/CD, and cloud infrastructure.

🎯 Exigences

• 6+ years in SRE, DevOps, or Infrastructure roles, with significant production Kubernetes experience. • Hands-on experience integrating AI/LLM tooling into engineering or operational workflows (e.g., MCP servers, AI agents acting on infrastructure), and a clear grasp of the security and governance considerations of giving AI access to production. • Proven success building CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI, or similar). • Strong with Kubernetes internals and managed services like EKS, GKE, or AKS. • Expertise with Infrastructure as Code (Terraform, Helm, Pulumi) and GitOps. • Proficient in Python, Bash, or Go. • Knowledge of observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry). • Production experience with Kafka, Flink, and ClickHouse. • Strong communication and cross-team collaboration skills.

🏖️ Avantages

• competitive salary • stock options • Health benefits • unlimited PTO • parental leave • tuition reimbursements

Postuler Maintenant

Emplois Similaires

🕒 il y a 1 mois

Jonas Software

1001 - 5000

🏗️ Construction

🏥 Santé

🏭 Fabrication

AI-First DevOps Engineer leading AWS infrastructure deployment automation for Computrition. Driving cloud practices and improving DevOps workflows with AI adoption in engineering delivery.

🇨🇦 Canada – Télétravail

💵 $155 000 - $165 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 1 mois

Carbon60

51 - 200

💼 Conseil

🏥 Santé

📦 Logistique

Managed Services Reliability Engineer supporting Canadian customers’ AWS cloud infrastructure at OpsGuru. Leading incident response, troubleshooting, security, backup, and reliability operations.

🇨🇦 Canada – Télétravail

💵 $140 000 / an

💰 Private Equity Round en 2019-01

⏰ Temps Plein

🟠 Senior

🔴 Expert

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 1 mois

Smile Digital Health

201 - 500

💼 Conseil

📦 Logistique

📣 Marketing

Site Reliability Engineer responsible for performance and reliability of cloud services at Smile Digital Health. Collaborating with teams to develop and improve performance testing frameworks and systems.

🇨🇦 Canada – Télétravail

💵 $110 000 - $125 000 / an

💰 €30 000 000 Series B en 2023-01

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 1 mois

Ping Identity

1001 - 5000

💼 Conseil

🏥 Santé

📦 Logistique

Site Reliability Engineer managing AWS accounts and cloud infrastructure deployment. Collaborating with teams to ensure security and efficiency of cloud operations at Ping Identity.

🇨🇦 Canada – Télétravail

💵 $87 000 - $105 000 / an

💰 €35 000 000 Series F - Ping Identity en 2014-09

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

Ansible

AWS

Chef

Cloud

Linux

Puppet

Python

Ruby

SaltStack

Terraform

Unix

Go

🕒 il y a 1 mois

Aequilibrium

51 - 200

💳 Fintech

🏦 Banque

🥽 RA/RV

Azure DevOps Lead architecting secure Azure infrastructure for a new digital banking platform. Building networks, IaC, CI/CD pipelines, and disaster recovery environments.

🇨🇦 Canada – Télétravail

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis