Site Reliability Engineer

🔥 12 hours ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Origami Risk

Origami Risk

501 - 1000 employees

Founded 2009

🏥 Healthcare

🏗️ Construction

📦 Logistics

💰 Private Equity Round on 2018-03

Healthcare • Construction • Logistics

Origami Risk is a comprehensive software platform that simplifies and enhances risk management and insurance processes. It provides tools for data integration, automating workflows, and facilitating collaboration across various sectors such as healthcare, manufacturing, and construction. Origami Risk supports organizations in improving their safety programs, claims administration, policy management, and compliance with governance requirements, ensuring a tailored approach to risk management.

📋 Description

• Leads post-incident investigations for the Site Reliability team. • Conducts in-depth post-incident analyses to identify root causes and develops preventive strategies. • Drafts clear and insightful RCAs for customer delivery. • Cross trains colleagues on how to best leverage observability tools during incident and performance investigations. • Provides visibility to all stakeholders throughout the entire Site Reliability process. • Collaborates with cross-functional teams to implement system enhancements that enhance scalability and stability. • Develops client-focused dashboards/alerts to proactively identify performance challenges. • Monitors and continuously improves our time to resolution metrics. • Maintains and configures core observability tools to ensure optimum performance and key metrics/data are available for incident response and performance investigations. • Provides an actionable feedback loop to Observability and Engineering teams toward improving MELT and development patterns. • Contributes to the development of automation tools to streamline incident response. • Works proactively to prevent incidents and reduce their impact on our platform. • Partners with the larger Cloud Operations, SRE, Engineering teams, and the business-at-large to advance our SaaS platforms. • Other duties as assigned.

🎯 Requirements

• Bachelor's degree in Computer Science or related field (or equivalent experience) • 5+ years of proven experience in a Site Reliability Engineering role. • Strong knowledge of SRE best practices and incident management protocols • Deep experience using and/or configuring New Relic, Data Dog, SumoLogic or similar observability tools • Proficiency in reading and writing code (e.g., JavaScript, .NET, SQL) • Familiarity with cloud platforms (e.g., AWS, Azure) and architectural patterns • Excellent problem-solving skills and a data-driven approach to incident analysis • Prior experience operating within a Public Cloud environment (AWS strongly preferred) • Experience troubleshooting C#/.Net based web applications to identify bugs/performance challenges. • Solid knowledge of SaaS operations • Ability to succeed when facing ambiguity and differing levels of operational maturation • Advanced written and verbal communication skills • Windows and SQL-server troubleshooting skills preferred • Knowledge of Continuous Integration and Continuous Delivery (CI/CD) pipelines preferred • Experience working in an Infrastructure as a Code (IaC) environment preferred

🏖️ Benefits

• Medical and Dental coverage available for employees, dependents, domestic partners, and spouses • Paid Time Off – Flexible options plus 10 paid company holidays where available** • Fully Paid by Origami Risk – Vision insurance, Short & Long-Term Disability Insurance, and Basic Life Insurance • Generous family leave options—including adoption and foster care placements • Pre-Tax Savings Accounts – Flexible Spending Account, Health Savings Account, Commuter Benefits, Dependent Care Savings Account • Retirement Savings – 401(k) with company match up to 4% • Employee Assistance Program (EAP) – Confidential & Free support offered to colleagues facing personal or work-related complications • Education Assistance Program – to help colleagues pursue industry/role-specific certifications • Wellness Benefits – reimbursement program to invest in healthy habits as well as support better colleague productivity and stress management • Additional coverages available – Pet Insurance, Critical Illness Insurance, and Voluntary Life & AD&D coverage

Apply Now

Similar Jobs

🔥 12 hours ago

LMI

1001 - 5000

📦 Logistics

🏥 Healthcare

🎖️ Defense

Senior DevSecOps/Platform Engineer designing and maintaining the Navy logistics platform. Building robust CI/CD pipelines and managing cloud infrastructure in AWS GovCloud.

Ansible

AWS

Cloud

Grafana

Kubernetes

Prometheus

Python

Terraform

🔥 16 hours ago

Oddball

51 - 200

💼 Consulting

📦 Logistics

🎖️ Defense

DevOps Engineer working on a pivotal Federal program at Oddball to improve daily lives through quality software. Building and maintaining CI/CD pipelines and managing AWS environments.

AWS

Docker

Linux

Terraform

🔥 17 hours ago

Pluribus Digital

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead DevOps Engineer designing and governing enterprise cloud architecture within a federal environment. Collaborating with engineering teams to ensure compliance and alignment with mission objectives.

Azure

Cloud

Jenkins

Python

🔥 20 hours ago

Kings Global LLC

-

🏭 Manufacturing

🛒 Retail

🤝 B2B

Senior DevOps/Security Lead for LiftNet managing AWS cloud infrastructure and CI/CD pipelines. Leading information security programs and ensuring SOC 2 compliance for vertical transportation management solutions.

AWS

Cloud

EC2

Linux

Rust

Terraform

🔥 21 hours ago

Direct Care Innovations

51 - 200

🏥 Healthcare

☁️ SaaS

🏢 Enterprise

Senior DevOps & Cloud Engineer managing Azure infrastructure for Direct Care Innovations' SaaS platform. Collaborating on scalable, reliable, and secure applications and systems while mentoring engineers.

Azure

Cloud

Java

PHP

SQL

.NET