
5001 - 10000 employees
🛡️ Insurance
💼 Consulting
📦 Logistics
Insurance • Consulting • Logistics
BMO U. S. is a diversified financial services company operating in the United States. It offers a broad range of financial products and services including personal and business banking, mortgage services, investments, financial planning, insurance, and wealth management. Additionally, it provides commercial loans, commercial mortgages, and other financial solutions tailored for small businesses and large enterprises. The company places a strong emphasis on customer service and offers digital and cross-border banking solutions to meet the needs of diverse clients. BMO U. S. is also involved in asset management and capital markets operations, making it a full-service financial institution.
🔥 13 hours ago
🇨🇦 Canada – Remote
💵 $70k - $150k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
👻 Ghost score 0%
Improve your chances of getting an interview by checking your resume score before you apply.

5001 - 10000 employees
🛡️ Insurance
💼 Consulting
📦 Logistics
Insurance • Consulting • Logistics
BMO U. S. is a diversified financial services company operating in the United States. It offers a broad range of financial products and services including personal and business banking, mortgage services, investments, financial planning, insurance, and wealth management. Additionally, it provides commercial loans, commercial mortgages, and other financial solutions tailored for small businesses and large enterprises. The company places a strong emphasis on customer service and offers digital and cross-border banking solutions to meet the needs of diverse clients. BMO U. S. is also involved in asset management and capital markets operations, making it a full-service financial institution.
• Ensure 24x7 availability, reliability, performance, security and operational excellence of the CCaaS platform and supporting services • Define and manage SLIs, SLOs and error budgets • Monitor platform health, performance, capacity and stability • Build and maintain monitoring dashboards, alerts and health checks • Configure and optimize CloudWatch, Splunk, Dynatrace, Grafana and other observability tools • Develop operational runbooks and troubleshooting guides • Act as an escalation point for critical incidents and service disruptions • Lead incident triage, impact assessment, communication and recovery activities • Perform root cause analysis and identify preventive actions • Drive reduction of MTTR, MTTD and recurring incidents • Eliminate operational toil through automation and develop self-healing and automated remediation solutions • Build scripts and tools using Python, PowerShell, AWS Lambda and other automation frameworks • Automate operational checks, deployment validation, monitoring and reporting • Support Amazon Connect, AWS services, APIs, middleware, routing, IVR and integration components • Manage platform configurations, certificates, service accounts and connectivity requirements • Support disaster recovery, backup, restoration and failover testing • Conduct capacity planning and performance optimization • Participate in release planning, deployment validation and production implementation • Review changes for operational risk and reliability impact • Support production deployments and rollback planning • Ensure operational readiness before launch • Support vulnerability remediation, audits, security events and compliance requirements • Analyze operational metrics and identify improvement opportunities • Drive SRE best practices across CCaaS teams • Participate in architecture reviews focused on scalability and resiliency • Mentor developers and operations teams on SRE practices • Partner with Development, DevOps, Infrastructure, Security, Business and Vendor teams • Debug production issues across services and technology-stack levels • Improve service health visibility through metrics, logs and traces • Compute the cost and business impact of SLA breaches and downtime
• Typically 5–6 years of relevant experience • Post-secondary degree in a related field of study or an equivalent combination of education and experience • Deep knowledge and technical proficiency gained through extensive education and business experience • Foundational proficiency in DevOps • Foundational proficiency in cybersecurity and privacy concepts, principles and solutions • Intermediate proficiency in IT infrastructure library • Intermediate proficiency in Robot Process Automation • Intermediate proficiency in cloud computing • Intermediate proficiency in configuration management • Intermediate proficiency in container orchestration • Intermediate proficiency in system design and implementation • Intermediate proficiency in incident management • Advanced proficiency in alerting and log configuration with OpenSearch and CloudWatch • Advanced proficiency in Dynatrace or another APM tool configuration and dashboarding • Advanced proficiency in automation and automation pipelines • Advanced proficiency in automated testing • Proficiency in verbal and written communication, collaboration, analytical and problem-solving skills, and data-driven decision making
• Performance-based incentives • Discretionary bonuses • Health insurance • Tuition reimbursement • Accident and life insurance • Retirement savings plans • In-depth training and coaching • Manager support • Network-building opportunities • Accommodations available on request during the selection process
Apply Now🕒 4 days ago
1001 - 5000
DevOps Specialist automating deployments, testing, and operations for Alberta government digital transformation projects. Supporting CI/CD, observability, production operations, and quality delivery.
🇨🇦 Canada – Remote
💰 Grant - PATH on 2025-07
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 5 days ago
Senior DevOps Engineer scaling AWS infrastructure for Sureify’s insurance SaaS platform. Automating deployments, reliability, monitoring, and security across the Americas.
🕒 September 3
Senior IAM Reliability Engineer securing enterprise identity operations at Genesys, an AI-powered customer experience platform. Improving authentication reliability, incident response, observability, and automated remediation.
🇨🇦 Canada – Remote
💵 $92k - $120.8k / year
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🕒 September 3
Senior Azure DevOps Engineer building Azure infrastructure, Kubernetes environments, and CI/CD automation for High Tech Genesis. Maintaining secure, scalable cloud deployment platforms.
🕒 September 2
Senior DevOps Engineer building Azure and OCI infrastructure for Network Solutions’ AI, web, and backend platforms. Operating Kubernetes, CI/CD, networking, and observability.
🇨🇦 Canada – Remote
💰 Venture Round on 2021-01
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)