Site Reliability Engineer

Job not on LinkedIn

🔥 7 minutes ago

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of ArangoDB

ArangoDB

51 - 200 employees

Founded 2015

🏥 Healthcare

📦 Logistics

💼 Consulting

💰 $27.8M Series B on 2021-10

Healthcare • Logistics • Consulting

ArangoDB is a flexible and scalable graph database platform designed for a variety of applications, particularly in a data-intensive environment like Generative AI. It supports multiple data models including graph, vector, document, full-text search, and geospatial databases, all within a single unified system. This allows developers to efficiently navigate complex data interactions and perform advanced analytics, making it ideal for sectors such as healthcare, finance, telecommunications, and more.

📋 Description

• Design, implement, and maintain cloud infrastructure on AWS and Google Cloud platforms • Ensure the scalability, performance, and reliability of Kubernetes-based distributed database systems • Collaborate with developers to write production-grade Golang code for infrastructure automation and system operations • Optimize and automate CI/CD pipelines, deployment processes, and monitoring systems • Develop disaster recovery, high availability, and fault-tolerance strategies • Identify bottlenecks, troubleshoot, and resolve issues across networking, operating systems, and cloud infrastructure • Implement monitoring, logging, and alerting systems • Participate in on-call rotations and respond to production incidents • Collaborate with cross-functional teams to improve reliability and scalability • Collaborate with Customer Success to resolve customer issues

🎯 Requirements

• Proven experience as an SRE or DevOps Engineer in a cloud-native environment • At least 3 years with Kubernetes in a production environment • Minimum 3 years of experience deploying and managing production-level cloud resources • Proficiency with Kubernetes for large-scale distributed systems • Experience with AWS and Google Cloud (GCP) • Understanding of networking, security practices, and troubleshooting • Understanding of Linux internals, including processes and environment variables • Familiarity with Docker and containerization technologies • Knowledge of CI/CD practices and tools such as Jenkins and CircleCI • Familiarity with Prometheus, Grafana, ELK stack, and observability tools • Knowledge of Git and version control systems • Familiarity with Golang or Python, or willingness and capability to learn and apply Golang • Ability to participate in on-call rotations • Strong troubleshooting, problem-solving, communication, and collaboration skills • Ability to self-organize and work independently as part of a remote team • Nice-to-have: distributed database or large-scale data storage experience • Nice-to-have: cloud security best practices • Nice-to-have: Python or Bash scripting • Nice-to-have: Terraform and Infrastructure-as-Code • Nice-to-have: GitOps experience • Nice-to-have: strong Golang programming experience

Apply Now

Similar Jobs

🔥 1 hour ago

Lucayan Technology Solutions LLC

51 - 200

💼 Consulting

📦 Logistics

🎖️ Defense

DevSecOps / Cloud Engineer automating CI/CD security and migrating legacy applications to USACE CWBI Cloud. Supporting STIG compliance, vulnerability remediation, and AI/ML governance for federal defense missions.

🔥 4 hours ago

CampMinder

51 - 200

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Senior DevSecOps Engineer securing Campminder’s software for summer camps. Hardening cloud infrastructure, embedding security in CI/CD, and leading compliance and threat-remediation work.

🔥 5 hours ago

CACI International Inc

10,000+ employees

💼 Consulting

🎖️ Defense

Cloud DevOps Engineer securing AWS CI/CD pipelines, containers, and infrastructure for CACI’s national-security customers. Building DevSecOps automation, Kubernetes security, monitoring, and incident-response workflows.

🔥 5 hours ago

ZOLL Medical Corporation

1001 - 5000

🏥 Healthcare

🎖️ Defense

📦 Logistics

Deployment Engineer installing and configuring ZOLL medical software and hardware for hospital customers. Training clients and supporting North American hospital sales teams across resuscitation applications.

🔥 6 hours ago

MeridianLink

501 - 1000

💳 Fintech

🏦 Banking

☁️ SaaS

Senior SRE owning reliability, observability, and scalability for MeridianLink’s financial SaaS applications. Designing resilient AWS/Azure infrastructure, automation, incident response, and security practices.