Search Remote Jobs

Senior Site Reliability Engineer, Platform

🕒 August 25

🇺🇸 United States – Remote

💵 $148k - $261k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Blue River Technology

Blue River Technology

201 - 500 employees

Founded 2011

🌾 Agriculture

🤖 Artificial Intelligence

🔧 Hardware

Agriculture • Artificial Intelligence • Hardware

Blue River Technology is at the forefront of agricultural innovation, developing intelligent machinery designed to improve farming yields and reduce environmental impacts. By integrating robotics, computer vision, and machine learning, the company is revolutionizing agriculture through a unique blend of Silicon Valley tech and hands-on fieldwork. Their purpose is to tackle significant challenges in agriculture, optimize chemical usage, and drive global sustainability, all while fostering an engaged and collaborative company culture.

📋 Description

• Architect, scale, and own essential infrastructure • Build and maintain a Kubernetes-based platform supporting multiple teams and services • Build backend services and internal tooling in Golang for autonomous systems • Collaborate with product teams to launch new products on the platform • Grow high-availability infrastructure while maintaining uptime and other key metrics • Build tooling for platform and development teams • Perform end-to-end performance analysis and implement robust improvements • Work with cloud vendors and external technical support on upgrades and issue resolution • Participate in on-call rotation, triage and resolve production incidents, and document root causes and postmortems • Design and maintain observability infrastructure, including dashboards, alerts, and log aggregation • Conduct regular risk assessments with the security team • Maintain the risk register and implement mitigation plans • Assess intrusion detection alerts and improve systems that digest threat feeds • Ensure SaaS payment processes are established and maintained with IT and purchasing teams • Drive architectural decisions, mentor engineers across teams, and shape platform direction

🎯 Requirements

• Minimum of six years of experience building and maintaining infrastructure for data-intensive, high-availability applications, including public cloud solutions • Production experience with Kubernetes and Terraform • Understanding of software design methodologies, information systems architecture, object-oriented design, and software design patterns • Experience securing cloud infrastructure, preferably AWS and Kubernetes, in production • Experience with one or more of Golang, Python, JavaScript, or Rust • Production experience with CI/CD tooling, including GitHub Actions, ArgoCD, ArgoCD Image Updater, and Artifactory • Interest in robotic applications and software that assists robots

🏖️ Benefits

• Annual performance bonus • Competitive benefit package • Career development • Learning and development programs • Mentorship programs • Diversity, equity, and inclusion programs • Reasonable accommodation for disabilities

Apply Now

Similar Jobs

🕒 August 24

The Home Depot

10,000+ employees

🏗️ Construction

📦 Logistics

🛒 Retail

Senior Principal Reliability Engineer designing resilient infrastructure for The Home Depot’s store systems, payments, and COM platform. Guiding multiple teams in cloud optimization, chaos engineering, and technology strategy.

🕒 August 24

Symbotic

501 - 1000

🔧 Hardware

📦 Logistics

🤖 Artificial Intelligence

Senior reliability manager scaling maintenance and asset performance across Exol’s automated warehouses. Driving uptime, safety, launches, vendor governance, and enterprise reliability standards.

🕒 August 24

PingWind Inc. (SDVOSB)

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

DevSecOps Engineer building and deploying secure cloud-based IAM systems for federal government clients. Maintaining highly available architectures, automated delivery, compliance, and infrastructure upgrades.

🕒 August 24

Cisco

10,000+ employees

🔧 Hardware

🔐 Security

🏢 Enterprise

Technical SRE leader keeping Splunk Cloud reliable for demanding enterprise customers. Owning critical incidents, customer stacks, automation strategy, and cloud infrastructure architecture.

🕒 August 23

Claritas Rx

51 - 200

🏥 Healthcare

💼 Consulting

📦 Logistics

DevSecOps Engineer securing Claritas Rx’s AWS-hosted digital health SaaS platform for rare-disease treatment access. Managing cloud controls, incident response, vulnerability programs, and healthcare compliance.