Senior Site Reliability Engineer

Job not on LinkedIn

🕒 July 13

🇺🇸 United States – Remote

💵 $140k - $170k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 7%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Coterie

Coterie

11 - 50 employees

👥 B2C

🛍️ eCommerce

🛒 Retail

B2C • eCommerce • Retail

Coterie is a company dedicated to providing premium diapering solutions for modern parents. Their products are designed to offer supreme softness, high absorbency, and reduced risk of leaks, blowouts, and diaper rash. Coterie's offerings include various diapering options such as The Diaper and The Pant, along with wipes, to ensure both comfort for babies and peace of mind for parents. The company emphasizes safety, offering dermatologist-tested, hypoallergenic products made from apparel-grade materials. Coterie also provides the convenience of an Auto-Renew subscription service, allowing customers to receive regular deliveries and enjoy savings. The brand is recognized with several awards for its diapering solutions, underscoring its commitment to quality and innovation in baby care.

📋 Description

• Manage and maintain cloud infrastructure on Azure, including Azure Kubernetes Service (AKS) clusters and supporting resources • Build, improve, and maintain CI/CD pipelines using GitHub Actions to support reliable and repeatable deployments • Own and enhance our Grafana implementation; designing dashboards, configuring alerts, and supporting incident management workflows • Monitor system health, triage incidents, and drive root cause analysis to prevent recurrence • Collaborate with development teams to define and track SLIs, SLOs, and error budgets that align with business goals • Contribute to infrastructure-as-code practices using Pulumi • Identify and resolve reliability risks through capacity planning, performance tuning, and proactive system improvements • Participate in an on-call rotation to support production systems and respond to incidents • Document runbooks, operational procedures, and architectural decisions to support team knowledge sharing

🎯 Requirements

• 5+ years of experience in a Site Reliability Engineering, DevOps, or Infrastructure role • 3+ years experience working with infrastructure as code • 2+ years of experience architecting CI/CD pipelines and cloud-based infrastructure • Strong hands-on experience with: Azure Cloud services and resource management, Kubernetes and AKS administration, GitHub Actions for CI/CD pipeline development and maintenance, 3+ experience with Grafana • Hands-on experience with Prometheus, Loki, or other observability tools in the Grafana ecosystem • Proficiency in at least one scripting or programming language such as Python or Bash • Understanding of networking fundamentals, DNS, load balancing, and container orchestration concepts • Strong analytical and communication skills; able to diagnose complex system issues and clearly communicate findings • Experience working in an agile environment with modern DevOps practices

🏖️ Benefits

• 100% remote • Health insurance through Aetna (we pay 100% of premiums) • Dental and vision insurance through Guardian (we pay 100% of premiums) • Basic life insurance (we pay 100% of premiums) • Access to flexible spending account (FSA) or health savings account (HSA) (for those using HSA eligible plans) • 401K plan (up 4% match with immediate vest) • Flexible PTO policy offering employees up to 4 weeks of PTO in their first 12 months • 12 company-paid holidays each year • Continuing education annual stipend

Apply Now

Similar Jobs

🕒 July 11

Arista Networks

1001 - 5000

🏢 Enterprise

📡 Telecommunications

Site Reliability Engineer managing Arista's CloudVision service fleet, ensuring scalability, reliability, and stability while operating production systems at scale with a global team.

🕒 July 10

Evio

11 - 50

💼 Consulting

📦 Logistics

🏥 Healthcare

DevOps Engineer supporting software development and operational execution at Evio. Focused on data ingestion cycles and enhancing product delivery in healthcare solutions.

🕒 July 10

Vertical Relevance

51 - 200

💼 Consulting

📣 Marketing

🏥 Healthcare

AWS DevSecOps/Security & Compliance Consultant responsible for implementing security solutions and guiding customers on their cloud journey at Vertical Relevance.

🕒 July 10

Cisco

10,000+ employees

🔧 Hardware

🔐 Security

🏢 Enterprise

Cisco SRE leading developer infrastructure, automation, and reliability for cloud engineering teams. Supporting CI platforms and tools used by over 2,000 engineers.

🕒 July 10

Intel Corporation

10,000+ employees

🏭 Manufacturing

💼 Consulting

📦 Logistics

DevSecOps Engineer securing federal-agency cloud and cybersecurity systems at Easy Dynamics. Automating CI/CD security, compliance, vulnerability management, and infrastructure protection.