Search Remote Jobs

DevOps Engineer III

🔥 0 minutes ago

🤠 Texas – Remote

infoinfo

đź’µ $135k - $165k / year

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź‘» Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Enable Dental

Enable Dental

51 - 200 employees

Founded 2017

🏥 Healthcare

⚕️ Healthcare Insurance

🧬 Biotechnology

Healthcare • Healthcare Insurance • Biotechnology

Enable Dental is a portable, at-home dental care provider that brings dental services directly to patients in their homes, assisted living facilities, and community centers. Specializing in providing comprehensive dental care to seniors, individuals with memory issues, and special needs populations, Enable Dental focuses on eliminating the stress of traveling for dental appointments. Their services include exams, cleanings, fillings, extractions, and more, all performed by licensed professionals in a comfortable and familiar environment, ensuring high-quality care accessible to those who may have limited options.

đź“‹ Description

• Provision, configure and maintain scalable AWS infrastructure across EC2, RDS, S3, VPC and IAM • Own Terraform and Terragrunt code across AWS, Cloudflare and GitHub • Maintain reproducible, secure, version-controlled infrastructure changed through reviewed pull requests • Plan and execute AWS account-structure changes and move workloads and data between accounts with minimal downtime and no data loss • Enforce cloud security practices covering access control, network security and encryption • Help maintain HIPAA compliance and support SOC 2 work • Deploy, manage, upgrade and scale EKS clusters • Ensure Kubernetes high availability, resource allocation and performance • Collaborate with software engineers to containerize applications and optimize Docker images • Run, upgrade and secure self-hosted open-source applications • Advise on self-hosted versus managed services • Design, build and maintain GitHub Actions CI/CD pipelines, including self-hosted runners • Automate testing, staging and production deployments for zero-downtime releases • Identify development bottlenecks and automate repetitive operational tasks with scripts • Operate monitoring, alerting, logging and tracing while keeping PHI out of telemetry • Own database backups and disaster recovery; regularly test restores • Participate in on-call rotation, resolve production issues and conduct blameless root cause analyses • Partner with engineers on architecture, reviews and priorities • Maintain runbooks and documentation, train backups for critical systems and establish break-glass access procedures • Enable engineers to safely deploy and operate their own services

🎯 Requirements

• 5+ years of hands-on experience in a DevOps, Site Reliability (SRE) or Cloud Engineering role, including owning production infrastructure • Production experience managing infrastructure with Terraform; Terragrunt, Pulumi or CloudFormation experience also counts • Experience with state management, imports, refactoring without downtime, and reviewing plans as part of pull requests • Strong production experience running infrastructure in AWS, including IAM, networking and managed databases • Deep understanding of Docker • Production experience running Kubernetes, preferably EKS, including Helm, networking, upgrades and scaling • Track record of building complex, automated pipelines in GitHub Actions, including secure cloud authentication (OIDC) • Solid understanding of cloud networking, including DNS, load balancing, VPCs, subnets, security groups and private connectivity • Strong Python or Bash skills for automating operational work • History of being the primary owner of infrastructure while working closely with application engineers • Ability to document work, share knowledge and design systems without single-person dependencies • Experience running infrastructure that handles PHI in a HIPAA-regulated environment, including encryption, access logging, retention and working with vendors under BAAs • Experience running third-party open-source applications in production, including version pinning, database migrations and rollbacks • Experience with AWS Organizations or moving workloads and data between AWS accounts • Experience preparing for or supporting a SOC 2 audit • Experience with Grafana stack (Loki, Tempo, Mimir), Prometheus, OpenTelemetry or similar

Apply Now

Similar Jobs

🔥 2 hours ago

MCCi

51 - 200

đź’Ľ Consulting

📦 Logistics

🏛️ Government

Cloud Operations Engineer improving reliability, observability, and automation for MCCi’s JustFOIA public-records SaaS platform. Operating Azure infrastructure and leading incident response for government customers.

🔥 5 hours ago

Comet

51 - 200

🤖 Artificial Intelligence

🤝 B2B

Deployment Engineer helping Comet customers deploy and troubleshoot Comet’s AI development platform. Automating Kubernetes, Helm, Terraform, and cloud infrastructure solutions for customer environments.

🔥 6 hours ago

Element 84

51 - 200

đź’Ľ Consulting

📦 Logistics

🏥 Healthcare

Senior DevOps Engineer designing AWS/Azure infrastructure, IaC and CI/CD systems. Supporting Element 84’s cloud-based geospatial and Generative AI software projects.

🔥 7 hours ago

WorkWave

1001 - 5000

đź’Ľ Consulting

📦 Logistics

📣 Marketing

DevOps Engineer building secure, scalable AWS/Azure infrastructure and CI/CD pipelines for WorkWave’s web, mobile, and API applications. Automating deployments, monitoring, testing, and operational processes.

🔥 8 hours ago

Weave

1 - 10

🧬 Biotechnology

🤖 Artificial Intelligence

🏥 Healthcare

Site Reliability Engineer building reliable cloud infrastructure for Weave's business communications platform. Automating, scaling, and monitoring GCP and Kubernetes environments.