Site Reliability Engineer

🔥 3 hours ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MaintainX

MaintainX

501 - 1000 employees

Founded 2018

☁️ SaaS

🏭 Manufacturing

🏢 Enterprise

🔥 Funding within the last year

💰 $150M Series D on 2025-08

SaaS • Manufacturing • Enterprise

MaintainX is a cloud-based, AI-powered maintenance and asset management platform that provides a modern CMMS (Computerized Maintenance Management System) and enterprise asset management (EAM) capabilities. The product helps frontline teams create and track work orders, schedule preventive and predictive maintenance, manage parts inventory, run checklists and inspections, and collect IoT and asset data. MaintainX emphasizes AI features — natural-language reporting, AI-assisted summaries and recommendations, anomaly detection, and operational insights — and offers integrations, an API, and an enterprise marketplace. It serves manufacturing and facilities-heavy industries and is positioned for enterprise customers; an acquisition by Autodesk has been announced.

📋 Description

• Assess service maturity and provide insights to development teams • Partner with development teams to implement observability best practices • Enable development teams to become autonomous with their service deployment, support, and infrastructure • Mentor developers on reliability practices, focusing on making them self-sufficient • Act as the bridge, ear and eyes of the Platform Division teams to drive tooling and practice adoption across development teams

🎯 Requirements

• Deep understanding of observability practices in a distributed system environment and how it influences system design and team behaviour • Practical experience with SRE concepts (SLOs, error budgets, incident management) • 3–5+ years in software development, SRE, DevOps, or production development roles with experience operating production systems • Proficient in cloud-native platforms and infrastructure-as-code concepts and tools • Working knowledge of at least one programming language (TypeScript/Node.js is a plus) • Excellent communication and collaboration abilities across technical and non-technical teams • Ability to translate complex reliability concepts into actionable guidance • You enjoy enabling teams to succeed independently and measuring success by reduced dependency on you.

🏖️ Benefits

• Competitive salary and meaningful equity opportunities. • Healthcare, dental, and vision coverage. • 401(k) / RRSP enrollment program. • Take what you need PTO. • A Work Culture where: • You’ll work alongside folks across the globe that reflect the MaintainX values, Smart Humble Optimist. • We believe in meritocracy, where ideas and effort are publicly celebrated.

Apply Now

Similar Jobs

🔥 6 hours ago

Thumbtack

1001 - 5000

🏪 Marketplace

☁️ SaaS

Senior Software Engineer designing resilient systems for availability and scalability at Thumbtack. Collaborating with teams and managing infrastructure effectively.

🇨🇦 Canada – Remote

💵 $180.2k - $233.2k / year

💰 $75M Debt Financing - Thumbtack on 2024-07

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

AWS

Cloud

Distributed Systems

DNS

JavaScript

Linux

Microservices

PHP

Python

SDLC

TCP/IP

Go

🔥 18 hours ago

Really Great Reading

11 - 50

📚 Education

🤝 B2B

Senior DevOps Engineer ensuring stability of AWS and transitioning to a new GCP platform. Working remotely in Quebec, Canada to support educational tools.

AWS

Cloud

Docker

EC2

Google Cloud Platform

Kubernetes

Python

SQL

Terraform

Go

🔥 22 hours ago

Athennian

51 - 200

☁️ SaaS

⚖️ Legal

📋 Compliance

DevOps Engineer optimizing cloud infrastructure for governance software solution at Athennian. Collaborating with engineering teams and driving efficiency to solve operational challenges.

🇨🇦 Canada – Remote

💵 $95k - $130k / year

💰 $33.6M Series B - Athennian on 2022-03

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

AWS

Cloud

Docker

Kubernetes

Linux

MongoDB

Packer

Terraform

🕒 4 days ago

Vena Solutions

501 - 1000

💼 Consulting

📣 Marketing

📦 Logistics

Site Reliability Developer 1 for Vena’s SaaS Technology and Operations team. Enhancing observability, scalability, and security of Vena’s cloud platform.

🇨🇦 Canada – Remote

💵 $85k - $115k / year

💰 Debt Financing - Vena Solutions on 2025-04

⏰ Full Time

🟢 Junior

🟡 Mid-level

⛑ DevOps & Site Reliability Engineer (SRE)

Ansible

AWS

Azure

Cloud

Distributed Systems

Docker

Prometheus

Terraform

🕒 4 days ago

Kong Inc.

201 - 500

💼 Consulting

📦 Logistics

🔌 API

Senior Site Reliability Engineer focused on orchestrating cloud-native systems for Kong's Managed Gateways. Leading a team to ensure robust and scalable infrastructure for enterprise customers.

Ansible

AWS

Azure

Cloud

Distributed Systems

Google Cloud Platform

Grafana

Kubernetes

Prometheus

Terraform

Go