Senior Site Reliability Engineer

Job not on LinkedIn

🕒 Yesterday

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Pinterest

Pinterest

1001 - 5000 employees

Founded 2010

📱 Media

👥 B2C

💰 Post IPO equity on 2022-08

Media • B2C

Pinterest is a visual discovery and bookmarking platform that helps users discover, save, and organize images and ideas (Pins) for interests, projects, and shopping. It provides personalized recommendations and visual search so people can plan and explore topics like home design, fashion, recipes, and events, and it supports commerce features and ads for creators and retailers.

📋 Description

• Ensuring the reliability, availability, and performance of production infrastructure and platform services • Operating and scaling Kubernetes platforms, including governance and support for multi-tenant workloads • Managing GitOps-based deployment workflows using ArgoCD and Helm • Driving infrastructure provisioning and change management through Terraform/Terragrunt • Building and supporting CI/CD automation and deployment workflows using GitHub Actions • Leading incident response efforts, root cause analysis, and post-incident improvement initiatives • Reducing operational toil through scripting, tooling, and process automation • Advancing observability practices across logs, metrics, traces, dashboards, and alerting • Supporting secure secrets integration, IAM-aware operations, and platform guardrails • Partnering closely with application, security, and platform teams to improve reliability and delivery outcomes

🎯 Requirements

• 4+ years of experience in Site Reliability Engineering, DevOps, Platform Engineering, or Cloud Infrastructure • Strong hands-on experience operating AWS in production environments • Deep expertise in Kubernetes, including cluster operations, troubleshooting, workload reliability, and platform administration • Proven experience with Kubernetes multi-tenancy, including namespaces, RBAC, quotas, policies, and tenant isolation patterns • Experience implementing and operating ArgoCD within a GitOps delivery model • Strong hands-on experience with Helm • Strong experience with Terraform/Terragrunt for infrastructure provisioning and environment management • Solid scripting and automation skills using Bash and/or Python • Experience building, maintaining, or supporting CI/CD pipelines, ideally using GitHub Actions • Strong troubleshooting skills across Linux, containers, IAM, networking, and distributed systems • Experience with monitoring, alerting, and observability in production environments • Demonstrated ownership mindset with experience handling incidents, resolving production issues, and driving follow-through after outages • Strong collaboration and communication skills, with the ability to work effectively across engineering, security, and platform teams • Bachelor’s degree in computer science, engineering, a related field or equivalent experience • Demonstrated ability to use AI to improve speed and quality in your day-to-day workflow for relevant outputs • Strong track record of critical evaluation and verification of AI-assisted work (e.g., testing, source-checking, data validation, peer review) • High integrity and ownership: you protect sensitive data, avoid over-reliance on AI, and remain accountable for final decisions and deliverables.

🏖️ Benefits

• Equity • Flexibility to do your best work • Professional development opportunities

Apply Now

Similar Jobs

🕒 Yesterday

TalentWerx

11 - 50

🎯 Recruiter

👥 HR Tech

🤝 B2B

DevOps Engineer IV designing and optimizing deployment solutions for Aether Aerospace. Collaborating with developers to enhance software development processes and ensure system security.

Ansible

AWS

Azure

Cloud

Cyber Security

Docker

Kubernetes

Python

Terraform

🕒 Yesterday

CoRelation

2 - 10

DevOps Developer II at Corelation Inc. focusing on automation and lifecycle management of patching and upgrades across environments. Collaborating with teams to ensure reliable and efficient operational processes.

Java

Linux

🕒 Yesterday

Mirantis

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior AI Deployment Engineer deploying AI infrastructure solutions at Mirantis. Collaborating on cloud technologies and mentoring team members while optimizing performance and reliability.

Cloud

Distributed Systems

JavaScript

Kubernetes

Linux

Microservices

Open Source

OpenStack

Python

Go

🕒 Yesterday

T-Rex Solutions, LLC

201 - 500

🔒 Cybersecurity

🏛️ Government

Senior DevSecOps Engineer at T-Rex Solutions supporting IRS Tax modernization by designing and securing CI/CD pipelines. Collaborating across teams to embed security controls throughout the software lifecycle.

AWS

Cloud

Docker

Kubernetes

🕒 Yesterday

T-Rex Solutions, LLC

201 - 500

🔒 Cybersecurity

🏛️ Government

DevSecOps Engineer designing and maintaining secure CI/CD pipelines for IRS program at T-Rex Solutions. Collaborating with development and security teams to enhance software lifecycle security and efficiency.

AWS

Cloud

Docker

Kubernetes