Director of SRE

🕒 July 16

🇨🇦 Canada – Remote

💵 CA$167k - CA$213k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 2%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Blackpoint Cyber

Blackpoint Cyber

51 - 200 employees

💼 Consulting

🎖️ Defense

🔒 Cybersecurity

💰 $190M Series C on 2023-06

Consulting • Defense • Cybersecurity

Blackpoint Cyber is a technology-focused cybersecurity company headquartered in Maryland, USA. Established by former US Department of Defense and Intelligence security experts, Blackpoint leverages its real-world cyber experience to help Managed Service Providers (MSPs) safeguard their infrastructure and operations. The company offers a proprietary cybersecurity ecosystem, including its SNAP-Defense platform for Managed Detection and Response (MDR) services. Blackpoint's dedicated security analysts work 24/7 to combine various security measures, including network visualization and endpoint security, to monitor and respond to threats. Additionally, Blackpoint is launching LogIC, a logging and integrated compliance service designed to assist MSPs with cyber compliance requirements. The company's mission is to deliver comprehensive detection and response services to help MSPs combat the evolving threat landscape.

📋 Description

• Lead the design, implementation, and management of scalable, reliable, and highly available cloud-based infrastructure (AWS/Azure) • Establish SRE best practices, including monitoring, incident response, capacity planning, and performance tuning • Improve observability, monitoring, and alerting, ensuring quick detection and resolution of reliability issues • Drive automation-first approaches, reducing manual intervention through Infrastructure-as-Code (IaC) and CI/CD pipelines • Lead a team of SREs, applying Blackpoint Cyber's management values of Coach, Model, Care, in defining business-critical outcomes, creating action plans, and supporting the team in achieving them • Continue hands-on contributions in an SRE role • Design, implement, and support key infrastructure, including automated attack infrastructure deployment, isolated identity and productivity environments, and secure data storage • Establish and apply security hygiene and monitoring policies to meet Blackpoint Cyber security requirements • Monitor and optimize cloud spending, ensuring cost-effective resource utilization without compromising reliability • Manage and mentor a global team of SREs, DevOps engineers, and cloud infrastructure specialists • Collaborate with security teams to ensure compliance, security hardening, and disaster recovery readiness

🎯 Requirements

• 10+ years of experience in SRE, DevOps, or Cloud Infrastructure roles • 5+ years of experience in people management, leading SRE team • Strong experience with AWS, Azure, or GCP, with expertise in cost management and scaling strategies • Proficiency in Infrastructure-as-Code (IaC) (e.g., Terraform, CloudFormation, Pulumi) • Hands-on experience with CI/CD pipelines, Kubernetes, and container orchestration • Expertise in monitoring, logging, and observability tools (e.g., Prometheus, Grafana, Datadog, Splunk) • Proven ability to optimize cloud costs (COGS) while maintaining reliability and performance • Strong leadership, collaboration, and problem-solving skills • Experience with SLA/SLO/SLIs will be valuable • A general understanding of the modern AI tooling landscape and how SRE can use that to increase velocity and improve stability

🏖️ Benefits

• Health insurance • Vision insurance • Dental insurance • Life insurance • 401k plan • Discretionary Time Off

Apply Now

Similar Jobs

🕒 July 10

Carbon60

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Managed Services Reliability Engineer supporting Canadian customers’ AWS cloud infrastructure at OpsGuru. Leading incident response, troubleshooting, security, backup, and reliability operations.

🇨🇦 Canada – Remote

💵 $140k / year

💰 Private Equity Round on 2019-01

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 9

Conga

1001 - 5000

☁️ SaaS

💸 Finance

🏢 Enterprise

Senior technical leader responsible for designing cloud infrastructure and advancing DevOps practices at Conga. Collaborating with Engineering, Product, Security, and Operations teams to enhance software delivery.

🕒 July 1

Branch

501 - 1000

💼 Consulting

📣 Marketing

🔌 API

AI DevOps & Reliability Engineer at Branch, focusing on software delivery and operational standards, enhancing DevOps with AI tools for reliability and efficiency.

🇨🇦 Canada – Remote

💵 $123k - $160k / year

💰 $282M Series F on 2022-02

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 June 6

Capgemini

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Software Change Management Consultant supporting application migration projects using IBM’s DBB/Git/IDD Solutions. Guiding clients through the conversion process and providing migration expertise and training.

🇨🇦 Canada – Remote

💵 $62.9k - $147.5k / year

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 April 16

Workiy Inc.

11 - 50

💼 Consulting

📣 Marketing

🛍️ eCommerce

Senior Salesforce DevOps Consultant driving DevOps best practices and managing deployment strategies for an IT solutions company. Supporting seamless Salesforce releases across multiple environments and business teams.