Senior Site Reliability Engineer, SRE

🔥 14 hours ago

🏄 California – Remote

infoinfo

💵 $171k - $192k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of LeoLabs

LeoLabs

51 - 200 employees

Founded 2016

🏭 Manufacturing

📦 Logistics

🎖️ Defense

Manufacturing • Logistics • Defense

LeoLabs is a commercial provider of persistent orbital intelligence and space domain awareness services. The company operates a network of ground-based, rapidly deployable radars and a cloud-based, AI-enabled platform to track and catalog objects in low Earth orbit, deliver real-time conjunction alerts, threat assessments, launch support, and space traffic management data for commercial and government customers. LeoLabs combines radar hardware, authoritative datasets, and SaaS analytics to help operators protect assets, avoid collisions, and manage launch and on-orbit operations.

📋 Description

• Design, implement, and maintain scalable and reliable systems • Set up monitoring tools and create incident response plans to identify and resolve issues and implement preventative measures • Develop and maintain scripts and automation tools for deployment, monitoring, and system health checks • Analyze system capacity and performance metrics to forecast future needs and implement scaling solutions • Work closely with development teams to enhance product reliability and streamline deployment • Create and maintain documentation for system architecture, processes, and incident reports • Participate in on-call rotations to provide 24/7 support for critical systems • Implement and enforce security best practices across systems and ensure compliance with industry standards • Independently deploy infrastructure changes using Infrastructure as Code • Identify reliability risks and recommend improvements • Improve dashboards, alerts, and operational runbooks • Improve deployment pipelines, infrastructure provisioning, and self-service capabilities • Optimize infrastructure utilization and cloud costs without compromising reliability • Drive automation to reduce operational toil and improve deployment reliability • Lead cross-functional initiatives to improve availability, scalability, and operational efficiency • Advise on site reliability for new product developments and platform evolution • Mentor junior engineers and foster the development culture

🎯 Requirements

• Bachelor’s degree in Computer Science, Engineering, or a related field, or equivalent work experience • 5+ years of experience in a Site Reliability Engineering, DevOps, or related role • Proficiency in a scripting or programming language, such as Python or Go • Experience with cloud services such as AWS or Azure • Proficiency with containerization using Docker, Kubernetes, or ECS • Proficiency in configuration management tools such as Terraform, Atlantis, or Terragrunt • Familiarity with CI/CD tools such as GitHub Actions, AWS CodeBuild, or CircleCI • Experience with monitoring tools such as Grafana or Datadog • Familiarity with database technologies such as RDS, Aurora, or PostgreSQL • Experience with large-scale distributed systems and microservices architecture • Strong analytical and problem-solving skills with the ability to troubleshoot complex systems • Excellent verbal and written communication skills, with the ability to collaborate effectively across teams • Ability to obtain a security clearance • Active TS/SCI clearance preferred • Must be eligible to obtain required U.S. Department of State ITAR authorizations

🏖️ Benefits

• Bonus and equity • 100 USD monthly wellness benefit to support your health and well-being • $300 USD home office setup stipend, available for use within your first six months • $75 USD monthly home internet allowance, paid directly through your regular paycheck • Individual Development Fund to support training, learning, and professional development • Employee Recognition Program to celebrate contributions and achievements across the team • Employee Referral Program • And much more!

Apply Now

Similar Jobs

🔥 14 hours ago

OpenRouter

1 - 10

💼 Consulting

📣 Marketing

🤖 Artificial Intelligence

Site Reliability Engineer maintaining observability, failover, and incident response for OpenRouter’s AI provider infrastructure. Automating endpoint quality, capacity, and launch-readiness operations.

🔥 16 hours ago

Mirantis

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Data Platform Engineer operating PostgreSQL and Kafka on Kubernetes for Mirantis, a Kubernetes-native AI infrastructure company. Ensuring reliable, secure, multi-region data services for enterprise GPU infrastructure.

🔥 16 hours ago

LMI

1001 - 5000

📦 Logistics

🏥 Healthcare

🎖️ Defense

DevOps Engineer automating cloud infrastructure and CI/CD for LMI’s federal SHEPRD application. Supporting secure, reliable deployments across development, test, and production environments.

🔥 17 hours ago

Encoura

51 - 200

💼 Consulting

📣 Marketing

📚 Education

Senior DevOps Engineer designing and scaling Azure infrastructure for Encoura, a higher-education technology company. Automating deployments, improving reliability, and supporting real-time student and institutional systems.

🔥 20 hours ago

Slate Auto

201 - 500

🚘 Automotive

🏭 Manufacturing

🚗 Transport

Senior DevOps Engineer building AWS and Kubernetes infrastructure for Slate’s affordable, customizable vehicles. Operating platforms, CI/CD pipelines and production reliability systems.