Senior DevOps Engineer – Storage

🔥 0 minutes ago

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Mirantis

Mirantis

501 - 1000 employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Consulting • Healthcare • Logistics

Mirantis is a company that specializes in container management and cloud infrastructure solutions. It offers a range of products, including Mirantis Kubernetes Engine (MKE), Mirantis OpenStack for Kubernetes (MOSK), and Mirantis Container Cloud (MCC), which provide enterprise-level Kubernetes and container management platforms. Mirantis also develops tools for secure software supply chains, such as the Mirantis Container Runtime (MCR) and Mirantis Secure Registry (MSR). As an advocate for open source technologies, Mirantis supports various projects and provides resources like Lens Desktop, a popular Kubernetes IDE, and technical support for enterprises adopting cloud-native technologies. Their solutions cater to sectors such as public services, financial services, and broader SaaS and technology services industries.

📋 Description

• Integrate NFS-based high-performance storage into Kubernetes clusters via CSI, storage classes, and persistent volumes • Tune NFS mount options, nconnect/RDMA, Linux client, and network settings for high-throughput, low-latency GPU/AI workloads • Deploy and operate storage services and operators; manage capacity, quotas, snapshots, and lifecycle • Configure and optimize Linux systems for storage workloads, including drivers, file systems, networking, and kernel parameters • Deliver storage integration for k0s-based Kubernetes through Cluster API and K0rdent management/child cluster topologies • Operate storage in fully disconnected air-gapped environments, including Harbor artifact/mirror connectivity and PKI/TLS considerations • Automate storage provisioning and configuration with Terraform/OpenTofu and ArgoCD or Flux GitOps pipelines • Build monitoring, alerting, and observability for storage performance, capacity, and health • Diagnose and resolve performance, reliability, and scaling issues across the storage stack • Set operational standards and communicate across teams

🎯 Requirements

• 7+ years of experience in SRE or infrastructure operations • 5+ years of building and operating distributed production storage systems at scale • Hands-on experience with high-performance storage solutions including VAST, Weka, DDN, and PowerScale • Linux and Kubernetes storage fundamentals, including NFS and CSI • Deep Linux storage and networking knowledge, including kernel and NFS-client layers • Infrastructure-as-code and GitOps experience • Preferred: bare-metal host provisioning, raw disk/hardware layout, and physical server storage configurations • Preferred: hands-on experience with VAST and/or Dell PowerScale • Preferred: GPUDirect Storage and RDMA/RoCE data paths • Preferred: Mirantis K0rdent stack, K0rdent Enterprise, K0rdent AI, k0s, MKE, and Cluster API • Preferred: Ceph, object/S3 storage backends, and CSI driver operations • Preferred: sovereign or high-security air-gapped environments experience

🏖️ Benefits

• Professional development and training • Attend conferences and working groups • Company outings, happy hours, hackathons, and tech talks • Competitive compensation package with a strong benefits plan • Remote work arrangement

Apply Now

Similar Jobs

🔥 49 minutes ago

Ping Identity

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Staff SRE building and operating Ping Identity’s cloud identity platform. Automating secure, resilient deployments with Go, GCP, AWS, Kubernetes, and CI/CD.

🇺🇸 United States – Remote

💵 $170k - $227k / year

💰 $35M Series F - Ping Identity on 2014-09

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 2 hours ago

CWILL

51 - 200

☁️ SaaS

🛍️ eCommerce

📣 Marketing

DevOps/SRE Engineer operating cloud infrastructure and compliance for CWILL, a Shopify post-purchase and retention SaaS suite. Responding to North American incidents and collaborating with China-based teams.

🗣️🇨🇳 Chinese Required

🔥 5 hours ago

CareSource

1001 - 5000

🏥 Healthcare

🛡️ Insurance

⚕️ Healthcare Insurance

Cloud, DevOps, and SRE architect designing infrastructure, deployment automation, and reliability strategies for CareSource’s digital healthcare platform. Scaling enterprise technology for national delivery.

🔥 10 hours ago

Cognitive Medical Systems, Inc.

51 - 200

💼 Consulting

🎖️ Defense

📦 Logistics

DevSecOps Lead owning CI/CD, cloud infrastructure, and security compliance for Cognitive’s federal healthcare systems. Driving secure delivery across CMS drug data and payment reconciliation platforms.

🔥 10 hours ago

Cognitive Medical Systems, Inc.

51 - 200

💼 Consulting

🎖️ Defense

📦 Logistics

Release Engineer managing reliable releases, observability, and disaster recovery for federal healthcare IT systems. Supporting CMS DDPS and PRS production environments remotely from designated U.S. states.