
51 - 200 employees
đź’Ľ Consulting
🎖️ Defense
đź”’ Cybersecurity
đź’° $190M Series C on 2023-06
Consulting • Defense • Cybersecurity
Blackpoint Cyber is a technology-focused cybersecurity company headquartered in Maryland, USA. Established by former US Department of Defense and Intelligence security experts, Blackpoint leverages its real-world cyber experience to help Managed Service Providers (MSPs) safeguard their infrastructure and operations. The company offers a proprietary cybersecurity ecosystem, including its SNAP-Defense platform for Managed Detection and Response (MDR) services. Blackpoint's dedicated security analysts work 24/7 to combine various security measures, including network visualization and endpoint security, to monitor and respond to threats. Additionally, Blackpoint is launching LogIC, a logging and integrated compliance service designed to assist MSPs with cyber compliance requirements. The company's mission is to deliver comprehensive detection and response services to help MSPs combat the evolving threat landscape.
🔥 0 minutes ago
🇨🇦 Canada – Remote
đź’µ CA$131k - CA$164.3k / year
⏰ Full Time
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
đź‘» Ghost score 1%
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
đź’Ľ Consulting
🎖️ Defense
đź”’ Cybersecurity
đź’° $190M Series C on 2023-06
Consulting • Defense • Cybersecurity
Blackpoint Cyber is a technology-focused cybersecurity company headquartered in Maryland, USA. Established by former US Department of Defense and Intelligence security experts, Blackpoint leverages its real-world cyber experience to help Managed Service Providers (MSPs) safeguard their infrastructure and operations. The company offers a proprietary cybersecurity ecosystem, including its SNAP-Defense platform for Managed Detection and Response (MDR) services. Blackpoint's dedicated security analysts work 24/7 to combine various security measures, including network visualization and endpoint security, to monitor and respond to threats. Additionally, Blackpoint is launching LogIC, a logging and integrated compliance service designed to assist MSPs with cyber compliance requirements. The company's mission is to deliver comprehensive detection and response services to help MSPs combat the evolving threat landscape.
• Design, develop, and maintain highly scalable infrastructure using Infrastructure as Code (Terraform and Terragrunt) for automated cloud resource provisioning and orchestration • Own and optimize the AWS cloud environment for cost efficiency, security best practices, and high availability • Manage and optimize Kubernetes cluster environments using Helm, ArgoCD, Istio, and Kustomize • Administer and scale data streaming infrastructure using Confluent Cloud and Apache Kafka • Deploy, configure, and maintain Redis for caching and real-time data processing • Implement and maintain monitoring, alerting, and incident response frameworks using Prometheus, Grafana, Alert Manager, and OpsGenie/PagerDuty • Facilitate controlled feature deployments and progressive rollouts through LaunchDarkly/PostHog • Partner with software development teams to integrate new services, applications, and features into existing infrastructure • Diagnose and resolve complex system-level issues while maintaining performance and maximizing uptime • Drive continuous improvement of automation tooling, operational processes, and engineering methodologies • Stay current on emerging SRE trends and tools and help adopt relevant industry advancements and best practices
• 5+ years of experience in a Senior Site Reliability Engineer role or equivalent, with substantial emphasis on cloud infrastructure management and automation • Expertise in Infrastructure as Code using Terraform and Terragrunt for enterprise-scale deployments • Comprehensive knowledge of AWS, including designing, implementing, and maintaining secure, scalable, resilient cloud architectures • Extensive hands-on experience with distributed data streaming using Confluent Cloud and Apache Kafka • Proven experience with Redis for caching and Amazon RDS for relational database management • Experience with enterprise search and analytics platforms including OpenSearch, Elasticsearch, and ChaosSearch • Proficiency designing and implementing monitoring and alerting infrastructure using Prometheus, Grafana, Alert Manager, and OpsGenie/PagerDuty • Practical experience with feature flag systems including LaunchDarkly/PostHog for controlled release management • Extensive experience administering production-grade Kubernetes with Helm, ArgoCD, and Istio; working knowledge of Kustomize • Strong problem-solving skills, with the ability to troubleshoot complex systems in production • Strong communication and collaboration skills, with experience working in Agile environments • Are you authorized to work in Canada without restriction? • Will you now or in the future require sponsorship to work within Canada?
• Equity participation available to employees globally, with program details varying by location and employment structure • Competitive Health, Vision, Dental, and Life Insurance plans for eligible employees in the US • Robust 401k plan for eligible employees in the US • Discretionary Time Off for eligible employees in the US • Other minor perks • International employees receive competitive benefits in accordance with local market standards and applicable country requirements
Apply Now🔥 22 hours ago
Site Reliability Engineer operating Yelp’s Kafka-based streaming infrastructure. Automating upgrades, scaling, incident recovery, and reliable data pipelines across Canada.
🇨🇦 Canada – Remote
đź’µ $135k - $185k / year
⏰ Full Time
🟡 Mid-level
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
đź•’ 2 days ago
Senior SRE building scalable infrastructure, CI/CD pipelines, and observability systems at CloudFactory. Improving reliability for production environments supporting AI data operations.
đź•’ 2 days ago
Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and enhancing microservices for a global AI data engineering company.
🇨🇦 Canada – Remote
đź’µ $80k - $150k / year
⏰ Full Time
🟡 Mid-level
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
đź•’ 4 days ago
Lead Application Reliability Engineer supporting Innodata’s Google Cloud enterprise applications. Restoring production services, managing deployments, and delivering enhancements across microservices environments.
🇨🇦 Canada – Remote
đź’µ $80k - $150k / year
⏰ Full Time
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
đź•’ 5 days ago
DevOps Specialist automating CI/CD and infrastructure for WinAir’s aviation maintenance software. Improving Jenkins, Ansible, Linux environments, deployments, and monitoring across development and production systems.
🇨🇦 Canada – Remote
đź’µ $54k - $76k / year
⏰ Full Time
🟡 Mid-level
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)