Senior Site Reliability Engineer

🕒 il y a 8 jours

🇨🇦 Canada – Télétravail

💵 CA$131 000 - CA$164 250 / an

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

👻 Score fantôme 2%

infoinfo

🗣️🇺🇸🇬🇧 Anglais requis

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of Blackpoint Cyber

Blackpoint Cyber

51 - 200 employés

💼 Conseil

🎖️ Défense

🔒 Cybersecurity

💰 €190 000 000 Series C en 2023-06

Consulting • Defense • Cybersecurity

Blackpoint Cyber est une entreprise de cybersécurité axée sur la technologie et basée dans le Maryland, aux États-Unis. La société a été fondée par d'anciens experts en sécurité du Département de la Défense et des services de renseignement américains et utilise son expérience réelle en cybersécurité et sa connaissance des pratiques malveillantes pour aider les MSP à protéger leur infrastructure et leurs opérations.

Description

• Design, develop, and maintain highly scalable infrastructure using Infrastructure as Code (Terraform and Terragrunt) for automated cloud resource provisioning and orchestration • Own and optimize the AWS cloud environment for cost efficiency, security best practices, and high availability • Manage and optimize Kubernetes cluster environments using Helm, ArgoCD, Istio, and Kustomize • Administer and scale data streaming infrastructure using Confluent Cloud and Apache Kafka • Deploy, configure, and maintain Redis for caching and real-time data processing • Implement and maintain monitoring, alerting, and incident response frameworks using Prometheus, Grafana, Alert Manager, and OpsGenie/PagerDuty • Facilitate controlled feature deployments and progressive rollouts through LaunchDarkly/PostHog • Partner with software development teams to integrate new services, applications, and features into existing infrastructure • Diagnose and resolve complex system-level issues while maintaining performance and maximizing uptime • Drive continuous improvement of automation tooling, operational processes, and engineering methodologies • Stay current on emerging SRE trends and tools and help adopt relevant industry advancements and best practices

🎯 Exigences

• 5+ years of experience in a Senior Site Reliability Engineer role or equivalent, with substantial emphasis on cloud infrastructure management and automation • Expertise in Infrastructure as Code using Terraform and Terragrunt for enterprise-scale deployments • Comprehensive knowledge of AWS, including designing, implementing, and maintaining secure, scalable, resilient cloud architectures • Extensive hands-on experience with distributed data streaming using Confluent Cloud and Apache Kafka • Proven experience with Redis for caching and Amazon RDS for relational database management • Experience with enterprise search and analytics platforms including OpenSearch, Elasticsearch, and ChaosSearch • Proficiency designing and implementing monitoring and alerting infrastructure using Prometheus, Grafana, Alert Manager, and OpsGenie/PagerDuty • Practical experience with feature flag systems including LaunchDarkly/PostHog for controlled release management • Extensive experience administering production-grade Kubernetes with Helm, ArgoCD, and Istio; working knowledge of Kustomize • Strong problem-solving skills, with the ability to troubleshoot complex systems in production • Strong communication and collaboration skills, with experience working in Agile environments • Are you authorized to work in Canada without restriction? • Will you now or in the future require sponsorship to work within Canada?

🏖️ Avantages

• Equity participation available to employees globally, with program details varying by location and employment structure • Competitive Health, Vision, Dental, and Life Insurance plans for eligible employees in the US • Robust 401k plan for eligible employees in the US • Discretionary Time Off for eligible employees in the US • Other minor perks • International employees receive competitive benefits in accordance with local market standards and applicable country requirements

Postuler Maintenant

Emplois Similaires

🕒 il y a 9 jours

Yelp

1001 - 5000

🍽️ Alimentation et boissons

🏨 Hôtellerie

📣 Marketing

Site Reliability Engineer operating Yelp’s Kafka-based streaming infrastructure. Automating upgrades, scaling, incident recovery, and reliable data pipelines across Canada.

🇨🇦 Canada – Télétravail

💵 $135 000 - $185 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 10 jours

CloudFactory

1001 - 5000

💼 Conseil

📦 Logistique

📣 Marketing

Senior SRE building scalable infrastructure, CI/CD pipelines, and observability systems at CloudFactory. Improving reliability for production environments supporting AI data operations.

🇨🇦 Canada – Télétravail

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 10 jours

InnoData

2 - 10

🤝 B2B

💼 Conseil

🌍 Impact social

Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.

🇨🇦 Canada – Télétravail

💵 $80 000 - $150 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 13 jours

WinAir

51 - 200

📦 Logistique

💼 Conseil

🏭 Fabrication

DevOps Specialist automating CI/CD and infrastructure for WinAir’s aviation maintenance software. Improving Jenkins, Ansible, Linux environments, deployments, and monitoring across development and production systems.

🇨🇦 Canada – Télétravail

💵 $54 000 - $76 000 / an

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 14 jours

Software Mind

1001 - 5000

🤖 Intelligence artificielle

☁️ SaaS

📡 Télécommunications

Senior SRE maintaining Kubernetes-based UI and AI service reliability for an enterprise cloud software company. Managing incidents, observability, deployments, and runtime troubleshooting in production.

🇨🇦 Canada – Télétravail

💰 Private Equity Round en 2020-12

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis