Site Reliability Engineer

Job not on LinkedIn

🔥 0 minutes ago

🇺🇸 United States – Remote

💵 $87.4k - $123.4k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Empower

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Finance • Fintech • B2C

Empower is a leading provider of financial services focused on helping individuals and organizations achieve financial freedom through retirement planning and investment management. Serving over 19 million Americans, Empower offers a comprehensive suite of finance-related services, including smart planning and investment advice, and tools like the Empower Personal Dashboard™ for a complete financial view. The company is renowned as a top retirement plan provider and works closely with personal investors, workplace plan savers, plan sponsors, and financial professionals. Empower is also recognized for initiatives in Diversity, Equity, Inclusion, and has a social commitment that bolsters community impact.

📋 Description

• Establish key indicators (SLIs) measuring service performance and build proactive monitors and alerts • Support projects of varying complexity and impact across multiple disciplines and teams • Conduct in-depth problem analysis to identify relevant findings and root causes • Ensure high availability, resilience, and scalability of containerized production applications • Document critical systems and create incident runbooks • Lead capacity planning and right-sizing exercises • Maintain and optimize infrastructure as code (IaC) • Troubleshoot and resolve complex system and deployment issues • Manage observability within Kubernetes, specifically EKS • Collaborate with development teams to support releases and create scalable, resilient, maintainable services • Work in a GitOps-driven environment

🎯 Requirements

• Experience maintaining high availability and resiliency within AWS infrastructure components, including EKS, EC2, RDS, S3, VPC, and others • Proficiency with Infrastructure as Code frameworks such as Terraform and CloudFormation • Demonstrated experience with Docker and Kubernetes containerization and orchestration • Experience with technologies, systems, and networks that affect production incident detection and response • Strong problem-solving abilities and desire to learn • Bachelor’s degree in Computer Science, Information Systems, or equivalent experience (listed under “What will set you apart”) • AWS, Kubernetes, or relevant certifications • Experience with observability suites and APM tooling such as DataDog, AppDynamics, or New Relic • Strong programming skills in one or more languages such as shell, Go, or Python • Experience supporting Java Spring Boot applications • Production experience in Kubernetes, especially EKS • Applicants must be authorized to work for any employer in the U.S.; employment visa sponsorship, including CPT/OPT, is unavailable • Reliable high-speed internet with a wired connection and a minimally disruptive home workspace required for remote work

🏖️ Benefits

• Medical, dental, vision and life insurance • Retirement savings – 401(k) plan with generous company matching contributions (up to 6%) • Financial advisory services • Potential company discretionary contribution • Broad investment lineup • Tuition reimbursement up to $5,250/year • Business-casual environment with the option to wear jeans • Generous paid time off upon hire • Ten paid company holidays • Three floating holidays each calendar year • Paid volunteer time — 16 hours per calendar year • Paid parental leave • Paid short- and long-term disability • Family and Medical Leave (FMLA) • Business Resource Groups (BRGs) • Bonus program opportunity for non-sales positions • Reliable high-speed internet and wired connection required for remote work • Necessary computer equipment provided

Apply Now

Similar Jobs

🔥 24 minutes ago

InnoData

2 - 10

🤝 B2B

💼 Consulting

🌍 Social Impact

Application Reliability Engineer supporting GCP and Google App Engine enterprise applications for global data engineering company. Handling incidents, deployments, observability, troubleshooting, and reliability enhancements.

🔥 3 hours ago

True Zero Technologies, LLC

11 - 50

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Cloud/DevSecOps Engineer designing AWS and hybrid-cloud infrastructure for True Zero Technologies, a veteran-owned government technology company. Leading secure, compliant mission-critical cloud operations and migrations.

🔥 4 hours ago

CoLogix Analytics

1 - 10

🏥 Healthcare

💼 Consulting

📦 Logistics

DevOps Engineer building automation and internal systems solutions for Cologix, a North American network-neutral interconnection and hyperscale edge data center company. Supporting reliable deployments, production infrastructure, and operational workflows.

🔥 6 hours ago

Astronomer

201 - 500

🏥 Healthcare

🏭 Manufacturing

💼 Consulting

Customer Reliability Engineer helping Astronomer customers operate its managed Apache Airflow DataOps platform. Troubleshooting reliability issues, meeting SLAs, and contributing to Airflow and monitoring systems.

🔥 8 hours ago

CVS Health

10,000+ employees

🏥 Healthcare

⚕️ Healthcare Insurance

🛒 Retail

Salesforce DevOps Engineer automating enterprise healthcare platform deployments at CVS Health. Managing Salesforce environments, CI/CD pipelines, security scanning, and release governance across engineering teams.