Site Reliability Engineer

Job not on LinkedIn

🕒 4 days ago

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of ArangoDB

ArangoDB

51 - 200 employees

Founded 2015

🏥 Healthcare

📦 Logistics

💼 Consulting

💰 $27.8M Series B on 2021-10

Healthcare • Logistics • Consulting

ArangoDB is a flexible and scalable graph database platform designed for a variety of applications, particularly in a data-intensive environment like Generative AI. It supports multiple data models including graph, vector, document, full-text search, and geospatial databases, all within a single unified system. This allows developers to efficiently navigate complex data interactions and perform advanced analytics, making it ideal for sectors such as healthcare, finance, telecommunications, and more.

📋 Description

• Design, implement, and maintain cloud infrastructure on AWS and Google Cloud platforms • Ensure the scalability, performance, and reliability of Kubernetes-based distributed database systems • Collaborate with developers to write production-grade Golang code for infrastructure automation and system operations • Optimize and automate CI/CD pipelines, deployment processes, and monitoring systems • Develop disaster recovery, high availability, and fault-tolerance strategies • Identify bottlenecks, troubleshoot, and resolve issues across networking, operating systems, and cloud infrastructure • Implement monitoring, logging, and alerting systems • Participate in on-call rotations and respond to production incidents • Collaborate with cross-functional teams to improve reliability and scalability • Collaborate with Customer Success to resolve customer issues

🎯 Requirements

• Proven experience as an SRE or DevOps Engineer in a cloud-native environment • At least 3 years with Kubernetes in a production environment • Minimum 3 years of experience deploying and managing production-level cloud resources • Proficiency with Kubernetes for large-scale distributed systems • Experience with AWS and Google Cloud (GCP) • Understanding of networking, security practices, and troubleshooting • Understanding of Linux internals, including processes and environment variables • Familiarity with Docker and containerization technologies • Knowledge of CI/CD practices and tools such as Jenkins and CircleCI • Familiarity with Prometheus, Grafana, ELK stack, and observability tools • Knowledge of Git and version control systems • Familiarity with Golang or Python, or willingness and capability to learn and apply Golang • Ability to participate in on-call rotations • Strong troubleshooting, problem-solving, communication, and collaboration skills • Ability to self-organize and work independently as part of a remote team • Nice-to-have: distributed database or large-scale data storage experience • Nice-to-have: cloud security best practices • Nice-to-have: Python or Bash scripting • Nice-to-have: Terraform and Infrastructure-as-Code • Nice-to-have: GitOps experience • Nice-to-have: strong Golang programming experience

Apply Now

Similar Jobs

🕒 4 days ago

CampMinder

51 - 200

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Senior DevSecOps Engineer securing Campminder’s software for summer camps. Hardening cloud infrastructure, embedding security in CI/CD, and leading compliance and threat-remediation work.

🇺🇸 United States – Remote

💵 $180k - $200k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 4 days ago

MeridianLink

501 - 1000

💳 Fintech

🏦 Banking

☁️ SaaS

Senior SRE owning reliability, observability, and scalability for MeridianLink’s financial SaaS applications. Designing resilient AWS/Azure infrastructure, automation, incident response, and security practices.

🇺🇸 United States – Remote

💵 $104.1k - $177.6k / year

💰 $485M Post-IPO Debt on 2021-11

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info

🕒 4 days ago

eFinancial

201 - 500

🏥 Healthcare

💼 Consulting

🛡️ Insurance

DevOps Team Lead building AWS infrastructure, CI/CD pipelines, and shared engineering platforms for a life insurance provider. Coaching engineers and improving deployment reliability and automation.

🕒 5 days ago

NEC Software Solutions

5001 - 10000

🏥 Healthcare

💼 Consulting

📦 Logistics

Senior DevOps Engineer managing AWS cloud infrastructure, Kubernetes, Terraform, and CI/CD for NEC SWS public-sector systems. Hybrid role requiring 50% office attendance and SC eligibility.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 5 days ago

Blitzy

51 - 200

🤖 Artificial Intelligence

☁️ SaaS

🤝 B2B

Site Reliability Engineer operating Blitzy's AI software development platform in secure U.S. public-sector cloud environments. Managing Kubernetes, observability, reliability, and customer deployment operations.

🇺🇸 United States – Remote

💵 $140k - $170k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)