Staff DevOps Engineer

🕒 August 15

🌵 Arizona – Remote

infoinfo

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of StrongMind

StrongMind

51 - 200 employees

Founded 2001

📚 Education

☁️ SaaS

🤝 B2B

Education • SaaS • B2B

StrongMind is a K-12 education company that provides digital learning solutions, curriculum, and services for schools and families. It offers an adaptive, engaging digital curriculum and technology platform aimed at strengthening student outcomes and simplifying K-12 digital learning, along with enrollment marketing and support services. StrongMind also provides a homeschool product for families with built-in curriculum and daily planning guidance.

📋 Description

• Own and evolve AWS infrastructure, including ECS, VPC networking, Route 53, secrets management, and cloud security posture • Improve cluster reliability, load balancing, auto-scaling, health checks, and disaster recovery • Maintain and expand Infrastructure-as-Code using Pulumi or similar technologies • Build and standardize GitHub Actions CI/CD pipelines across Ruby, C#, Python, JavaScript, and TypeScript codebases • Establish blue/green and canary deployments, automated rollback, and secure pipeline gates • Build observability with centralized logging, distributed tracing, metrics, dashboards, alerting, and SLOs/SLIs • Strengthen cloud security through least-privilege IAM, encryption, network security, secrets management, vulnerability scanning, and cloud security posture management • Partner with Security and Compliance on SOC 2 evidence, audit trails, security controls, and remediation • Migrate remaining Azure workloads to AWS, including containerization, traffic cutovers, and database migrations • Create architecture decision records, runbooks, technical documentation, and operational standards • Mentor engineers on cloud-native architecture, Infrastructure-as-Code, secure-by-default design, and operational excellence • Collaborate with Release and engineering teams on deployment quality gates, incident response, and post-incident learning

🎯 Requirements

• 8+ years of experience in DevOps, SRE, Platform Engineering, or a related discipline • Deep hands-on AWS experience, including ECS Fargate, RDS, DynamoDB, SQS/SNS, Lambda, API Gateway, EventBridge, CloudWatch, IAM, VPC networking, and S3 • Proven experience improving observability, uptime, reliability, and operational maturity in production cloud environments • Strong Infrastructure-as-Code experience managing multi-environment infrastructure, including Pulumi or similar tools and policy-as-code/IaC scanning • Experience securing production cloud environments through least-privilege IAM, encryption, network security, secrets management, and remediation of security findings • Strong CI/CD experience with GitHub Actions across multiple programming languages and repositories, including automated security scanning • Hands-on Docker experience, including multi-stage builds, base image management, runtime configuration, health checks, and image lifecycle management • Solid networking fundamentals, including VPC design, security groups, DNS, load balancing, and CDN configuration • Experience managing relational databases through RDS/Aurora, including performance tuning and backup strategies • Ability to work independently as a primary infrastructure owner while creating standards and automation for future team growth • Availability to work core business hours from 8:00 AM to 5:00 PM Arizona time, Monday through Friday • Eligible to work in the United States • Visa sponsorship is unavailable • Bonus: Experience migrating workloads from Azure or another cloud platform to AWS • Bonus: Familiarity with Azure App Service, Azure SQL, Cosmos DB, Service Bus, Event Grid, Key Vault, or Application Insights • Bonus: Experience with event-driven architectures such as EventBridge, SQS/SNS, or Kinesis • Bonus: Experience with Dagster, AWS Glue, or Step Functions • Bonus: FinOps experience, including cloud cost optimization, reserved capacity, and right-sizing • Bonus: DevSecOps tooling experience with Trivy, Dependabot, Snyk, Checkov, Prowler, Security Hub, or GuardDuty • Bonus: Experience supporting SOC 2 or other compliance programs • Bonus: AWS Security Specialty, CISSP, or another relevant security certification • Bonus: Experience with EdTech platforms or integration standards such as Canvas LMS, PowerSchool, Clever, LTI, or OneRoster • Bonus: Experience building and leading a DevOps or Platform Engineering team

🏖️ Benefits

• A competitive total compensation package, including medical, dental, vision, and voluntary benefits • An on-site gym • Virtual wellness programs • Wellness coaching • Flexible work options for select roles • Unlimited PTO for exempt roles • Additional “life happens” days • A fully paid holiday week off at Christmas • Recognition and rewards for birthdays, meaningful milestones, legendary work, and community service hours • Quarterly Town Halls • Annual social events and traditions, including Halloween celebrations and Wellness Fairs

Apply Now

Similar Jobs

🕒 August 14

Accellor

201 - 500

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Principal FDE designing and deploying governed AI architectures for Accellor’s enterprise clients. Leading customer engagements, production implementations, and the forward-deployed engineering team.

🕒 August 13

Skydio

501 - 1000

🎖️ Defense

🏭 Manufacturing

📦 Logistics

Staff Site Reliability Engineer operating Kubernetes, AWS, and Terraform infrastructure for Skydio’s autonomous drone platform. Ensuring reliable, scalable cloud services through observability, automation, and incident response.

🇺🇸 United States – Remote

💵 $240k - $300k / year

💰 $170M Series E - Skydio on 2024-11

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

🕒 August 12

AlphaSense

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

Principal engineer architecting scalable CI/CD and Kubernetes delivery systems for AlphaSense, an AI-powered market intelligence platform. Defining progressive delivery standards and developer self-service across global engineering teams.

🇺🇸 United States – Remote

💵 $246k - $339k / year

💰 Debt Financing on 2022-06

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

🕒 August 11

Inizio Evoke

1001 - 5000

🏥 Healthcare

💼 Consulting

📣 Marketing

Director, DevOps operating AWS infrastructure for Inizio Evoke’s agentic AI platform. Automating deployments, ensuring reliability, and leading security, compliance, and release operations remotely.

🇺🇸 United States – Remote

💵 $100k - $145k / year

⏰ Full Time

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 August 11

Hexion Inc.

1001 - 5000

🚘 Automotive

🏗️ Construction

🏭 Manufacturing

Reliability Engineer improving asset performance across Hexion’s North American manufacturing plants. Leading failure elimination, maintenance optimization, and cross-site reliability standardization.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

🔴 Lead

⛑ DevOps & Site Reliability Engineer (SRE)