Senior Site Reliability Engineer

🕒 il y a 2 mois

🇨🇦 Canada – Télétravail

💵 $120 400 - $216 600 / an

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

👻 Score fantôme 9%

infoinfo

🗣️🇺🇸🇬🇧 Anglais requis

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of Akamai Technologies

Akamai Technologies

5001 - 10000 employés

🔒 Cybersecurity

💰 Post-IPO Equity en 2001-07

Cloud Computing • Cybersecurity • Content Delivery

Akamai Technologies est une plateforme mondiale en périphérie et une société de services cloud qui alimente, sécurise et accélère les applications en ligne, les médias et les API. Elle propose des services de diffusion de contenu (CDN), de calcul en périphérie, d'infrastructure cloud et une offre complète de sécurité incluant la protection contre les attaques DDoS, la sécurité des applications web et des API, l'atténuation des bots, les services DNS et des solutions zero-trust pour les entreprises. Akamai sert de grandes entreprises dans les secteurs des médias, de la finance, du commerce de détail, du jeu vidéo et du secteur public pour améliorer la performance, la fiabilité et la sécurité à grande échelle.

Description

• Owning the SRE infrastructure lifecycle from design reviews and pre-rollout readiness assessments through production sign-off and ongoing reliability management • Designing and implementing frameworks that reflect customer experience for load balancing services and driving action when error budgets are at risk • Building and maintaining observability pipelines from load-balancing components and system-level sources to dashboards that enable rapid incident triage • Leading technical incident response for complex NB/NLB failures, acting as the technical commander and driving root cause analysis and preventive follow-through • Developing and automating safe deployment workflows for phased releases, including bake-period monitoring, feature flag management, and validation across global datacenter rollouts • Reviewing design documents, product-requirement documents and producing actionable SRE input on operational risks, capacity implications, Day-2 concerns, and product strategy gaps • Building automation and tooling using Python or Go that reduces operational toil and improves team-wide operational capability

🎯 Exigences

• 8+ years of experience in SRE, infrastructure engineering, or platform engineering, working with large-scale distributed systems • Demonstrate deep expertise with Linux networking fundamentals and diagnosing at the packet level using tcpdump, netstat, and similar tools • Have hands-on experience with L4/L7 load balancing technologies covering configuration, health checking, high availability, and failure modes at scale • Show a track record of defining SLO/SLI frameworks, building observability platforms from scratch, and running incident management processes at scale • Demonstrate expertise in Kubernetes and containerization at scale including workload scheduling, networking, resource management, and operating stateful or network-intensive workloads in a cluster environment • Build automation and tooling using Python or Go, with infrastructure-as-code experience (SaltStack, Ansible, or Terraform) and deployment safety instincts.

🏖️ Avantages

• healthcare • RRSP • company holidays • vacation (in the form of PTO) • sick time • family friendly benefits including employee assistance program including a focus on mental and financial wellness

Postuler Maintenant

Emplois Similaires

🕒 il y a 2 mois

Capgemini

10 000+ employés

💼 Conseil

🏥 Santé

📦 Logistique

Software Change Management Consultant supporting application migration projects using IBM’s DBB/Git/IDD Solutions. Guiding clients through the conversion process and providing migration expertise and training.

🇨🇦 Canada – Télétravail

💵 $62 874 - $147 504 / an

⏰ Temps Plein

🟠 Senior

🔴 Expert

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

Groovy

🕒 il y a 4 mois

Workiy Inc.

11 - 50

💼 Conseil

📣 Marketing

🛍️ eCommerce

Senior Salesforce DevOps & Release Manager at Workiy overseeing enterprise Salesforce deployments and CI/CD processes across multiple environments. Driving best practices in release management while collaborating with cross-functional teams.

🇨🇦 Canada – Télétravail

⏰ Temps Plein

🟠 Senior

🔴 Expert

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

Workiy Inc.

11 - 50

💼 Conseil

📣 Marketing

🛍️ eCommerce

Senior Salesforce DevOps Consultant driving DevOps best practices and managing deployment strategies for an IT solutions company. Supporting seamless Salesforce releases across multiple environments and business teams.

🇨🇦 Canada – Télétravail

⏰ Temps Plein

🟠 Senior

🔴 Expert

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

Movable Ink

501 - 1000

📣 Marketing

✈️ Tourisme

🏨 Hôtellerie

Lead Site Reliability Engineer ensuring scalable, resilient services for Movable Ink at high volume content platform. Design and drive automation strategies while mentoring engineering teams.

🇨🇦 Canada – Télétravail

💵 $154 000 - $200 000 / an

💰 €55 000 000 Series D en 2022-04

⏰ Temps Plein

🟠 Senior

⛑ Ingénieur DevOps & SRE

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 4 mois

Tecsys Inc.

501 - 1000

🏥 Santé

☁️ SaaS

📦 Logistique

Ingénieur fiabilité des infrastructures pour soutenir les services SaaS critiques. Collaborer, innover et optimiser la fiabilité et la performance des systèmes cloud sur AWS et Kubernetes.

🇨🇦 Canada – Télétravail

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

⛑ Ingénieur DevOps & SRE