Staff Site Reliability Engineer

🔥 14 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Filevine

Filevine

201 - 500 employees

☁️ SaaS

⚖️ Legal

🤖 Artificial Intelligence

💰 $108M Series D on 2022-04

SaaS • Legal • Artificial Intelligence

Filevine is a comprehensive legal technology platform that offers a wide range of services tailored to law firms and legal practitioners. The platform provides solutions for case management, document management, and contract management, as well as lead management and business analytics. Utilizing advanced AI technologies, Filevine enhances the legal workflow with tools such as DemandsAI, ImmigrationAI, and FilevineAI to automate tasks and improve productivity. It integrates seamlessly with popular tools like QuickBooks and Gmail, ensuring a complete legal tech stack. Filevine also offers eSignature capabilities, time and billing features, and a client portal to facilitate communication. The platform is utilized by various types of law practices, including personal injury, family law, mass torts, and more, and is recognized for its robust security standards and compliance certifications such as SOC 2 Type II and HIPAA.

📋 Description

• As a Staff Site Reliability Engineer at Filevine, you are the senior technical authority on the SRE team and a strategic partner to engineering leadership. • You don’t just maintain systems — you shape engineering culture, define the technical standard for how Filevine runs in production, and bridge the gap between high-level business goals and robust, internet-scale technical execution. • You own the roadmap across two critical SRE domains — Observability & Alerting and Platform Infrastructure — and are accountable for ensuring the team solves reliability problems permanently rather than absorbing them as toil. • You operate as the senior IC counterpart to the Engineering Manager: technical correctness lives with you. • You partner with the Reliability Architect and engineering leadership on significant technical decisions, mentor engineers across experience levels, and influence reliability strategy across the broader organization. • Reliability at Filevine protects revenue. • You are the senior technical voice responsible for ensuring that uptime, incident response, and every production change meet the operational standard the business demands.

🎯 Requirements

• 12+ years of experience in software engineering, infrastructure, platform engineering, or SRE, including 6+ years in SRE and 3+ years leading complex, cross-functional technical initiatives for distributed production systems. • Expert-level depth in observability and platform infrastructure, with broad expertise in incident response, capacity planning, automation, and reliability engineering. • Advanced experience with a major container-orchestration platform, preferably Kubernetes, and an observability platform such as New Relic, Datadog, or equivalent. • Strong software-engineering ability in Python, Go, Bash, or another general-purpose language, with experience building production tooling, automation, or platform capabilities. • Proven ability to mentor engineers and communicate technical risk clearly to engineering, product, and executive audiences. • Experience in a regulated environment such as FedRAMP, CJIS, HIPAA, SOC 2, or PCI is strongly preferred.

🏖️ Benefits

• Medical, Dental, & Vision Insurance (for full-time employees) • Competitive & Fair Pay • Maternity & paternity leave (for full-time employees) • Short & long-term disability • Opportunity to learn from a dedicated leadership team • Top-of-the-line company swag

Apply Now

Similar Jobs

🔥 22 hours ago

Hubstaff

51 - 200

⚡ Productivity

☁️ SaaS

🏢 Enterprise

Principal DevOps Engineer overseeing infrastructure and automation for Hubstaff's cloud platform. Collaborating with CTO and engineering teams to improve scalability and performance.

AWS

Cloud

Distributed Systems

Docker

Google Cloud Platform

Grafana

JavaScript

Kubernetes

Linux

Node.js

Postgres

Prometheus

Redis

Ruby

Ruby on Rails

Rust

Terraform

Go

🔥 23 hours ago

Millennium

201 - 500

💼 Consulting

🎖️ Defense

🔒 Cybersecurity

AWS DevOps Engineer responsible for designing and maintaining AWS cloud environments for Marine Corps IT systems. Collaborating across teams to ensure performance, security, and compliance needs are met.

AWS

Cloud

Docker

Kubernetes

Python

Terraform

🕒 Yesterday

Aya Healthcare

5001 - 10000

🏥 Healthcare

💼 Consulting

📦 Logistics

Manager of Site Reliability Engineering leading a team for Aya Healthcare's workforce platform. Ensuring product reliability and outstanding user experience through innovative solutions.

AWS

Azure

Google Cloud Platform

🕒 2 days ago

TalentWerx

11 - 50

🎯 Recruiter

👥 HR Tech

🤝 B2B

DevOps Engineer IV designing and optimizing deployment solutions for Aether Aerospace. Collaborating with developers to enhance software development processes and ensure system security.

Ansible

AWS

Azure

Cloud

Cyber Security

Docker

Kubernetes

Python

Terraform

🕒 2 days ago

Palo Alto Networks

10,000+ employees

🔒 Cybersecurity

🏢 Enterprise

DevOps Platform Developer enhancing development flow and product infrastructure for cybersecurity company. Focused on mission-critical systems and increasing stability and quality while reducing time to market.

AWS

Azure

Cloud

Cyber Security

Docker

Groovy

Jenkins

Kubernetes

Python