Site Reliability Engineer

🔥 15 hours ago

☕ Washington – Remote

infoinfo

💵 $123k - $150k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 1%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Kong Inc.

Kong Inc.

201 - 500 employees

Founded 2017

💼 Consulting

📦 Logistics

🔌 API

💰 $100M Series D on 2021-02

Consulting • Logistics • API

Kong Inc. is a company that provides a comprehensive API platform designed to facilitate API management, AI integration, and developer productivity. It offers solutions like Kong Gateway, Kong Konnect, and a variety of other tools targeted at managing and optimizing the API lifecycle. Kong's platform supports multi-cloud environments and is built to deliver high performance and security. It is notably recognized by Gartner as a leader in API management and supports innovations across industries like financial services, healthcare, and technology. The company emphasizes flexibility, security, and speed, making it a favored choice for enterprises looking to enhance their digital services through APIs. Kong also supports a robust community of developers and provides extensive integrations and plugins to streamline API management and operations.

📋 Description

• Operate and scale Kong’s global SaaS platform, Konnect, across regions and clouds • Build, automate, and maintain Kubernetes-based infrastructure and deployment workflows using Terraform/Terragrunt, Helm, and ArgoCD • Design, maintain, and optimize multi-region data and caching layers including PostgreSQL, Redis, ClickHouse, and Druid • Operate and improve Kong Gateway and Kong Mesh environments supporting hybrid and distributed architectures • Develop and maintain CI/CD pipelines and GitOps workflows • Enhance observability and incident response readiness using Datadog, Prometheus, Grafana, and Thanos • Define and track service-level objectives • Collaborate with development and security teams to operate SaaS services in compliance with reliability, security, and regulatory standards • Participate in a global 24/7 on-call rotation • Improve operational playbooks and postmortem practices • Lead and contribute to scaling initiatives that improve elasticity, reliability, and cost-efficiency

🎯 Requirements

• BS in Computer Science or equivalent practical experience • Proven experience managing SaaS or PaaS systems at enterprise scale in multi-region, multi-tenant, secure environments • Deep expertise in Kubernetes, including debugging cluster/networking issues and designing for fault tolerance and scalability • Strong proficiency with Terraform or Terragrunt • Experience with CI/CD pipelines and GitOps workflows, including ArgoCD, Atlantis, and Helm • Proficiency in Go, Python, or Bash • Solid understanding of Linux/Unix systems, DNS, TLS/SSL, HTTP, load balancers, and distributed systems • Experience working with API gateway and service mesh technologies • Familiarity with Kafka and observability platforms such as Datadog, Prometheus, and Grafana • Experience working in a 24/7/365 production support environment • Must be legally authorized to work in the country where the position will be worked • Must disclose current or future sponsorship requirements

🏖️ Benefits

• Healthcare benefits • 401(k) plan • Short-term disability benefits • Long-term disability benefits • Basic life insurance • AD&D insurance • Additional rewards may be available depending on the applicable plan and role

Apply Now

Similar Jobs

🔥 15 hours ago

CACI International Inc

10,000+ employees

💼 Consulting

🎖️ Defense

AWS DevSecOps Engineer modernizing CACI’s federal Grant Solutions cloud platform. Automating secure infrastructure, CI/CD, observability, and compliant software delivery.

🔥 17 hours ago

Bitdeer Group

201 - 500

💼 Consulting

📦 Logistics

🏗️ Construction

Senior DevOps Engineer building CI/CD, Kubernetes, and MLOps infrastructure for Bitdeer's AI and Bitcoin cloud platform. Creating an internal developer platform for scalable, governed deployments.

🇺🇸 United States – Remote

💵 $180k - $260k / year

💰 Post-IPO Equity on 2023-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 18 hours ago

Bitwarden

51 - 200

🔒 Cybersecurity

☁️ SaaS

🏢 Enterprise

Senior SRE operating Bitwarden Gov’s FedRAMP-compliant cloud infrastructure. Managing reliability, monitoring, incident response, Kubernetes, and security across multi-cloud environments.

🔥 18 hours ago

URUS Group

1001 - 5000

🌾 Agriculture

🤝 B2B

🧬 Biotechnology

DevOps Team Lead operating VAS’s AWS platform for farm management software. Leading IaC, reliability, security, cost optimization, and globally distributed workloads.

🔥 19 hours ago

VetsEZ

201 - 500

🏥 Healthcare

💼 Consulting

📦 Logistics

Release Train Engineer leading SAFe delivery and DevOps systems for VetsEZ’s federal healthcare IT project. Overseeing Agile Release Train execution, CI/CD, platform operations, and cross-team delivery.