Senior Site Reliability Engineer – SRE

🔥 14 hours ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Branch

Branch

501 - 1000 employees

Founded 2014

💼 Consulting

📣 Marketing

🔌 API

💰 $282M Series F on 2022-02

Consulting • Marketing • API

Branch is a mobile growth company that provides a comprehensive growth platform designed to maximize the value of digital strategies. Their services focus on improving customer engagement, optimizing advertising performance through sophisticated attribution, and ensuring compliance with data protection regulations. Serving over 100,000 companies from startups to Fortune 500 brands, Branch helps businesses create seamless user experiences across various channels, drive conversions, and achieve significant growth in mobile apps and engagement metrics.

📋 Description

• Partner with Developers to produce high-performing and robust services through rigorous testing and release procedures • Design infrastructure, monitoring, processes, and standards for systems and applications • Support services through design, development, load testing, and launch phases • Develop, measure, and monitor key performance and service level indicators including availability, latency, and overall system health • Define and establish SLIs, SLOs, and error budgets with service owners, and drive adoption across platform teams • Profile and optimize platform performance, resilience, and efficiency, including latency, throughput, and capacity planning under load • Participate in incident response and root cause analysis • Remediate tasks and develop preventative and automated measures to meet SLAs/SLOs/SLIs • Manage monitoring services utilized by applications

🎯 Requirements

• Bachelor's degree in an appropriate engineering discipline or equivalent experience required • 3+ years experience in site reliability engineering • Strong hands-on experience building and operating Java / Spring Boot services in production • Experience with Terraform, Go, Java, Gradle, Docker, OpenTelemetry and Kubernetes

🏖️ Benefits

• Market-leading medical, dental, and vision insurance • Stock options • Free Premium-Tier Origin Financial Wellness subscription • Monthly home-office stipend • 401k (TransAmerica) • 12-weeks paid parental leave for birthing and non-birthing parents • Flexible time off + sick and safe time • 11 paid company holidays • Branch@Branch Same Day Pay Option

Apply Now

Similar Jobs

🔥 15 hours ago

Neural Earth

11 - 50

🤖 Artificial Intelligence

🛡️ Insurance

🏠 Real Estate

DevOps Lead at Neural Earth responsible for maintaining CI/CD pipelines and monitoring AWS cloud infrastructure. Support incident response and operational work for smooth engineering processes.

AWS

Azure

Cloud

Cyber Security

Docker

Google Cloud Platform

Kubernetes

Python

Terraform

🔥 16 hours ago

PerfectServe

201 - 500

🏥 Healthcare

⚕️ Healthcare Insurance

☁️ SaaS

Forward Deployment Engineer, AI optimizing and deploying AI voice agent solutions for PerfectServe's customers. Collaborating with Product and Customer Success to ensure successful go-live and ongoing support.

🇺🇸 United States – Remote

💵 $120k - $140k / year

💰 Private Equity Round on 2018-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

AWS

Cloud

Python

🔥 17 hours ago

Outlook Group

201 - 500

🏭 Manufacturing

📦 Logistics

🍽️ Food & Beverage

Senior DevOps Engineer overseeing AWS infrastructure and DevOps practices at Outlook Amusements. Collaborating with teams to enhance automation and service reliability in production environments.

Apache

AWS

Azure

Cloud

EC2

Linux

MS SQL Server

MySQL

SQL

🕒 Yesterday

IPolarity

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead SRE Engineer for the Observability team designing and operating cloud observability platforms. Focus on logging, metrics, tracing, and alerting across large-scale cloud infrastructure.

Ansible

AWS

Azure

Cloud

Consul

ElasticSearch

Google Cloud Platform

Grafana

Kafka

Kubernetes

Prometheus

Python

Ruby

Splunk

Terraform

Go

🕒 Yesterday

Aya Healthcare

5001 - 10000

🏥 Healthcare

💼 Consulting

📦 Logistics

Manager of Site Reliability Engineering leading a team for Aya Healthcare's workforce platform. Ensuring product reliability and outstanding user experience through innovative solutions.

AWS

Azure

Google Cloud Platform