Staff DevOps Engineer

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of StrongMind

StrongMind

51 - 200 employees

Founded 2001

📚 Education

☁️ SaaS

🤝 B2B

Education • SaaS • B2B

StrongMind is a K-12 education company that provides digital learning solutions, curriculum, and services for schools and families. It offers an adaptive, engaging digital curriculum and technology platform aimed at strengthening student outcomes and simplifying K-12 digital learning, along with enrollment marketing and support services. StrongMind also provides a homeschool product for families with built-in curriculum and daily planning guidance.

📋 Description

• Own and evolve AWS infrastructure, including ECS, VPC networking, Route 53, secrets management, and cloud security posture • Improve cluster reliability, load balancing, auto-scaling, health checks, and disaster recovery • Maintain and expand Infrastructure-as-Code using Pulumi or similar technologies • Build and standardize GitHub Actions CI/CD pipelines across Ruby, C#, Python, JavaScript, and TypeScript codebases • Establish blue/green and canary deployments, automated rollback, and secure pipeline gates • Build observability with centralized logging, distributed tracing, metrics, dashboards, alerting, and SLOs/SLIs • Strengthen cloud security through least-privilege IAM, encryption, network security, secrets management, vulnerability scanning, and cloud security posture management • Partner with Security and Compliance on SOC 2 evidence, audit trails, security controls, and remediation • Migrate remaining Azure workloads to AWS, including containerization, traffic cutovers, and database migrations • Create architecture decision records, runbooks, technical documentation, and operational standards • Mentor engineers on cloud-native architecture, Infrastructure-as-Code, secure-by-default design, and operational excellence • Collaborate with Release and engineering teams on deployment quality gates, incident response, and post-incident learning

🎯 Requirements

• 8+ years of experience in DevOps, SRE, Platform Engineering, or a related discipline • Deep hands-on AWS experience, including ECS Fargate, RDS, DynamoDB, SQS/SNS, Lambda, API Gateway, EventBridge, CloudWatch, IAM, VPC networking, and S3 • Proven experience improving observability, uptime, reliability, and operational maturity in production cloud environments • Strong Infrastructure-as-Code experience managing multi-environment infrastructure, including Pulumi or similar tools and policy-as-code/IaC scanning • Experience securing production cloud environments through least-privilege IAM, encryption, network security, secrets management, and remediation of security findings • Strong CI/CD experience with GitHub Actions across multiple programming languages and repositories, including automated security scanning • Hands-on Docker experience, including multi-stage builds, base image management, runtime configuration, health checks, and image lifecycle management • Solid networking fundamentals, including VPC design, security groups, DNS, load balancing, and CDN configuration • Experience managing relational databases through RDS/Aurora, including performance tuning and backup strategies • Ability to work independently as a primary infrastructure owner while creating standards and automation for future team growth • Availability to work core business hours from 8:00 AM to 5:00 PM Arizona time, Monday through Friday • Eligible to work in the United States • Visa sponsorship is unavailable • Bonus: Experience migrating workloads from Azure or another cloud platform to AWS • Bonus: Familiarity with Azure App Service, Azure SQL, Cosmos DB, Service Bus, Event Grid, Key Vault, or Application Insights • Bonus: Experience with event-driven architectures such as EventBridge, SQS/SNS, or Kinesis • Bonus: Experience with Dagster, AWS Glue, or Step Functions • Bonus: FinOps experience, including cloud cost optimization, reserved capacity, and right-sizing • Bonus: DevSecOps tooling experience with Trivy, Dependabot, Snyk, Checkov, Prowler, Security Hub, or GuardDuty • Bonus: Experience supporting SOC 2 or other compliance programs • Bonus: AWS Security Specialty, CISSP, or another relevant security certification • Bonus: Experience with EdTech platforms or integration standards such as Canvas LMS, PowerSchool, Clever, LTI, or OneRoster • Bonus: Experience building and leading a DevOps or Platform Engineering team

🏖️ Benefits

• A competitive total compensation package, including medical, dental, vision, and voluntary benefits • An on-site gym • Virtual wellness programs • Wellness coaching • Flexible work options for select roles • Unlimited PTO for exempt roles • Additional “life happens” days • A fully paid holiday week off at Christmas • Recognition and rewards for birthdays, meaningful milestones, legendary work, and community service hours • Quarterly Town Halls • Annual social events and traditions, including Halloween celebrations and Wellness Fairs

Apply Now

Similar Jobs

🔥 11 hours ago

Expel

201 - 500

🔒 Cybersecurity

☁️ SaaS

Principal SRE securing Expel’s cloud-native cybersecurity platform. Leading Kubernetes, infrastructure, incident response, and reliability initiatives across distributed systems.

🔥 23 hours ago

Accellor

201 - 500

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Principal FDE designing and deploying governed AI architectures for Accellor’s enterprise clients. Leading customer engagements, production implementations, and the forward-deployed engineering team.

🕒 Yesterday

Skydio

501 - 1000

🎖️ Defense

🏭 Manufacturing

📦 Logistics

Staff Site Reliability Engineer operating Kubernetes, AWS, and Terraform infrastructure for Skydio’s autonomous drone platform. Ensuring reliable, scalable cloud services through observability, automation, and incident response.

🕒 Yesterday

Renesas Electronics

10,000+ employees

🏭 Manufacturing

🏥 Healthcare

📦 Logistics

Quality and reliability engineer qualifying AI server power modules at Renesas, a global semiconductor solutions company. Developing reliability tests, component qualification processes, validation plans, and manufacturing prototypes for high-performance computing products.

🕒 Yesterday

General Dynamics Information Technology

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

GDIT DevSecOps Engineer maintaining secure AWS, Kubernetes, and GitLab pipelines. Supporting defense and intelligence technology, video encoding, automated security scanning, and system performance validation.