Senior DevOps Engineer – PostgreSQL, Kafka, Kubernetes

🔥 14 hours ago

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 22%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mirantis

Mirantis

501 - 1000 employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Mirantis is a company that specializes in container management and cloud infrastructure solutions. It offers a range of products, including Mirantis Kubernetes Engine (MKE), Mirantis OpenStack for Kubernetes (MOSK), and Mirantis Container Cloud (MCC), which provide enterprise-level Kubernetes and container management platforms. Mirantis also develops tools for secure software supply chains, such as the Mirantis Container Runtime (MCR) and Mirantis Secure Registry (MSR). As an advocate for open source technologies, Mirantis supports various projects and provides resources like Lens Desktop, a popular Kubernetes IDE, and technical support for enterprises adopting cloud-native technologies. Their solutions cater to sectors such as public services, financial services, and broader SaaS and technology services industries.

📋 Description

• Deploy, upgrade, scale, and operate CloudNativePG PostgreSQL and Strimzi Kafka clusters on cloud and bare-metal Kubernetes • Manage infrastructure declaratively with Argo CD or Flux, Helm, and Terraform • Build CI/CD pipelines for database and Kafka changes, schema migrations, and operator upgrades • Design and operate regional high availability and cross-region disaster recovery • Operate PostgreSQL replica clusters, backups, point-in-time recovery, and Kafka MirrorMaker 2 • Conduct regular failover drills with defined RPO/RTO targets • Provide product teams self-service provisioning of databases, users, topics, and ACLs through code • Build monitoring and alerting with Prometheus, Grafana, and OpenTelemetry • Monitor replication lag, backups, consumer lag, and capacity • Participate in on-call and lead incident reviews • Implement TLS/mTLS, secrets management, network policies, Kafka ACLs, and per-tenant isolation • Integrate database and Kafka access with Keycloak/OIDC identity platforms • Operate Debezium and Kafka Connect outbox/change-data-capture pipelines • Write automation and small services in Go or Python • Partner with application teams on schema standards and performance troubleshooting

🎯 Requirements

• 8+ years in DevOps, SRE, platform, or infrastructure engineering • 3+ years running stateful systems (databases or message streaming) in production • Strong hands-on Kubernetes operations, including StatefulSets, persistent volumes and storage classes, pod disruption budgets, scheduling and affinity, network policies, and cluster upgrades • Production experience running PostgreSQL and/or Kafka through Kubernetes operators • Experience authoring and maintaining Helm charts • Day-to-day use of Argo CD or Flux, Terraform, and CI/CD pipelines • Practical PostgreSQL administration, including replication and failover, backup and point-in-time recovery, connection pooling with PgBouncer, version upgrades, and performance troubleshooting • Production Kafka operations, including brokers, topics and partitions, replication, consumer groups, monitoring, and capacity planning • Experience designing and testing high availability and disaster recovery for stateful systems, with defined RPO/RTO • Strong scripting and programming in Go or Python, plus Bash • Bachelor's degree in Computer Science or a related field, or equivalent practical experience • Nice-to-have experience with cross-region replication, Debezium, Kafka Connect, transactional outbox/CDC patterns, Keycloak, OIDC, LDAP/Active Directory, observability stacks, Vault/OpenBao, cert-manager, mTLS, bare-metal infrastructure, Ceph or object storage, GPU/AI infrastructure, open-source contributions, and SOC 2 or ISO 27001 environments

🏖️ Benefits

• Professional development and training • Conference and working group attendance • Company outings, happy hours, hackathons, and tech talks • Competitive compensation package with a strong benefits plan • Remote work arrangement

Apply Now

Similar Jobs

🔥 15 hours ago

LMI

1001 - 5000

📦 Logistics

🏥 Healthcare

🎖️ Defense

DevOps Engineer automating cloud infrastructure and CI/CD for LMI’s federal SHEPRD application. Supporting secure, reliable deployments across development, test, and production environments.

🔥 16 hours ago

Encoura

51 - 200

💼 Consulting

📣 Marketing

📚 Education

Senior DevOps Engineer designing and scaling Azure infrastructure for Encoura, a higher-education technology company. Automating deployments, improving reliability, and supporting real-time student and institutional systems.

🔥 19 hours ago

Slate Auto

201 - 500

🚘 Automotive

🏭 Manufacturing

🚗 Transport

Senior DevOps Engineer building AWS and Kubernetes infrastructure for Slate’s affordable, customizable vehicles. Operating platforms, CI/CD pipelines and production reliability systems.

🔥 20 hours ago

Akamai Technologies

5001 - 10000

🔒 Cybersecurity

Senior Site Reliability Engineer maintaining Akamai's Linux kernels, KVM/QEMU virtualization, and cloud infrastructure. Supporting distributed Compute products that make digital experiences faster and more secure.

🔥 22 hours ago

Centex Technologies

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

Senior DevOps Engineer building secure AWS infrastructure and CI/CD for Centex Technologies’ Air Force Data Mesh Environment. Automating cloud delivery, reliability, and DevSecOps controls.