Senior Site Reliability Engineer

🕒 July 14

🇨🇦 Canada – Remote

💵 $197.5k - $225k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of SecurityScorecard

SecurityScorecard

501 - 1000 employees

Founded 2013

💼 Consulting

🏥 Healthcare

🛡️ Insurance

💰 $180M Series E on 2021-03

Consulting • Healthcare • Insurance

SecurityScorecard is a company that focuses on cybersecurity and risk management. It provides solutions for supply chain detection and response, third-party cyber risk management, and external attack surface management. By leveraging AI and security ratings, SecurityScorecard helps organizations improve their cybersecurity postures and manage risks effectively. The company's platform allows for monitoring and remediation of vulnerabilities, collaboration with vendors, and compliance with regulatory mandates. SecurityScorecard serves a diverse range of industries, including public sector, technology, healthcare, financial services, and more. Its comprehensive suite of tools and services, such as Security Ratings and MAX, empower organizations to proactively manage cyber risks and enhance their overall security structures.

📋 Description

• Design, build, and scale Kubernetes infrastructure for secure, multi-tenant, high-availability applications. • Build and operate AI tooling infrastructure — stand up MCP servers and establish secure, governed AI access and guardrails for production systems. • Optimize and maintain CI/CD pipelines, improving reliability, speed, and rollback safety. • Implement progressive delivery strategies such as blue/green and canary deployments. • Advance Infrastructure as Code with Terraform, Helm, and Argo CD, defining reusable patterns for the org. • Operate and optimize streaming and analytics infrastructure: Kafka, Flink, and ClickHouse. • Build automated testing into the CI/CD lifecycle. • Improve system observability — define SLOs, alerts, and dashboards. • Lead incident response and postmortems, focusing on root cause and durable fixes. • Mentor engineers across teams on Kubernetes, CI/CD, and cloud infrastructure.

🎯 Requirements

• 6+ years in SRE, DevOps, or Infrastructure roles, with significant production Kubernetes experience. • Hands-on experience integrating AI/LLM tooling into engineering or operational workflows (e.g., MCP servers, AI agents acting on infrastructure), and a clear grasp of the security and governance considerations of giving AI access to production. • Proven success building CI/CD pipelines (GitHub Actions, Jenkins, GitLab CI, or similar). • Strong with Kubernetes internals and managed services like EKS, GKE, or AKS. • Expertise with Infrastructure as Code (Terraform, Helm, Pulumi) and GitOps. • Proficient in Python, Bash, or Go. • Knowledge of observability tooling (Prometheus, Grafana, Datadog, OpenTelemetry). • Production experience with Kafka, Flink, and ClickHouse. • Strong communication and cross-team collaboration skills.

🏖️ Benefits

• competitive salary • stock options • Health benefits • unlimited PTO • parental leave • tuition reimbursements

Apply Now

Similar Jobs

🕒 July 13

Jonas Software

1001 - 5000

🏗️ Construction

🏥 Healthcare

🏭 Manufacturing

AI-First DevOps Engineer leading AWS infrastructure deployment automation for Computrition. Driving cloud practices and improving DevOps workflows with AI adoption in engineering delivery.

🇨🇦 Canada – Remote

💵 $155k - $165k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 10

Carbon60

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Managed Services Reliability Engineer supporting Canadian customers’ AWS cloud infrastructure at OpsGuru. Leading incident response, troubleshooting, security, backup, and reliability operations.

🇨🇦 Canada – Remote

💵 $140k / year

💰 Private Equity Round on 2019-01

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 10

Smile Digital Health

201 - 500

💼 Consulting

📦 Logistics

📣 Marketing

Site Reliability Engineer responsible for performance and reliability of cloud services at Smile Digital Health. Collaborating with teams to develop and improve performance testing frameworks and systems.

🇨🇦 Canada – Remote

💵 $110k - $125k / year

💰 $30M Series B on 2023-01

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 8

Ping Identity

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

Site Reliability Engineer managing AWS accounts and cloud infrastructure deployment. Collaborating with teams to ensure security and efficiency of cloud operations at Ping Identity.

🇨🇦 Canada – Remote

💵 $87k - $105k / year

💰 $35M Series F - Ping Identity on 2014-09

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 8

Intermedia Cloud Communications

1001 - 5000

💼 Consulting

🏥 Healthcare

⚖️ Legal

DevOps Engineer at Intermedia working on CI/CD pipelines and cloud technology for HostPilot. Collaborate with teams to enhance control panel services and infrastructure reliability.

🇨🇦 Canada – Remote

💰 Venture Round on 2017-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)