Site Reliability Engineer

🔥 1 hour ago

🇺🇸 United States – Remote

💵 $87.4k - $123.4k / year

⏰ Full Time

🟢 Junior

🟡 Mid-level

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Empower

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Finance • Fintech • B2C

Empower is a leading provider of financial services focused on helping individuals and organizations achieve financial freedom through retirement planning and investment management. Serving over 19 million Americans, Empower offers a comprehensive suite of finance-related services, including smart planning and investment advice, and tools like the Empower Personal Dashboard™ for a complete financial view. The company is renowned as a top retirement plan provider and works closely with personal investors, workplace plan savers, plan sponsors, and financial professionals. Empower is also recognized for initiatives in Diversity, Equity, Inclusion, and has a social commitment that bolsters community impact.

📋 Description

• Own operational excellence for assigned systems and services while supporting projects of varying complexity across teams • Participate in on-call rotations, respond to incidents, troubleshoot complex system and deployment issues, and drive resolution • Lead postmortem processes, conduct root cause analysis, and implement preventive measures • Establish service level indicators, build proactive monitoring and alerting, and manage observability for Kubernetes environments, including EKS • Build, maintain, and optimize infrastructure as code across multiple AWS environments • Manage and optimize EKS clusters to support the availability, resilience, and scalability of containerized applications in production • Collaborate with development teams to support releases and implement scalable, resilient, and maintainable services using GitOps and progressive delivery practices • Maintain and improve CI/CD pipelines and automation tools to reduce toil and improve operational efficiency • Lead capacity planning and right-sizing efforts to support performance optimization and system reliability • Document critical systems, runbooks, architecture decisions, and system behaviors, and mentor entry-level SREs on operational best practices

🎯 Requirements

• Bachelor’s degree in Computer Science, Information Technology, or a related field, or equivalent practical experience • 2 to 4 years of experience in Site Reliability Engineering, DevOps, or Systems Engineering • Experience maintaining high availability and resiliency across AWS infrastructure components, including EKS, EC2, RDS, S3, VPC, and similar services • Production experience with Kubernetes and containerization technologies such as Docker, including deployment, troubleshooting, and optimization • Proficiency with infrastructure as code frameworks such as Terraform and CloudFormation • Experience with observability platforms and the technologies, systems, and networks that affect incident detection and response • Understanding of CI/CD principles and experience with GitLab CI, Jenkins, or equivalent tools • Knowledge of networking fundamentals, troubleshooting, and high-availability architecture patterns • Familiarity with GitOps workflows, incident management, and on-call practices • Strong problem-solving skills, sound judgment, and a desire to learn

🏖️ Benefits

• Medical, dental, vision and life insurance • Retirement savings – 401(k) plan with generous company matching contributions (up to 6%), financial advisory services, potential company discretionary contribution, and a broad investment lineup • Tuition reimbursement up to $5,250/year • Business-casual environment that includes the option to wear jeans • Generous paid time off upon hire – including a paid time off program plus ten paid company holidays and three floating holidays each calendar year • Paid volunteer time — 16 hours per calendar year • Leave of absence programs – including paid parental leave, paid short- and long-term disability, and Family and Medical Leave (FMLA) • Business Resource Groups (BRGs) – BRGs facilitate inclusion and collaboration across our business internally and throughout the communities where we live, work and play. BRGs are open to all.

Apply Now

Similar Jobs

🔥 2 hours ago

MdotM

11 - 50

📣 Marketing

Intermediate DevOps Engineer supporting infrastructure for AI-driven investment solutions platform. Automating CI/CD and managing cloud resources for efficient software delivery.

🔥 4 hours ago

Vynca

51 - 200

💼 Consulting

⚕️ Healthcare Insurance

🏥 Healthcare

Site Reliability Engineer designing and managing scalable AWS infrastructure for Vynca's healthcare tech platform. Collaborating with Software Engineers and improving system reliability through automation and observability practices.

🔥 6 hours ago

Akkadian Labs

51 - 200

☁️ SaaS

🏢 Enterprise

📡 Telecommunications

DevOps Engineer supporting design, implementation, and maintenance of secure infrastructure. Collaborating with teams to enable reliable deployments and improve system observability at Akkadian Labs.

🔥 6 hours ago

Sentara Health

10,000+ employees

🏥 Healthcare

⚕️ Healthcare Insurance

DevOps Engineer focusing on developing Azure cloud infrastructure and Kubernetes environments for Sentara Health. Managing CI/CD pipelines and collaborating with development teams in a remote setting.

🔥 7 hours ago

Pekin Insurance

501 - 1000

🏥 Healthcare

📦 Logistics

🛡️ Insurance

Software Engineer at Pekin Insurance collaborating with product teams to create and support rich applications. Designing, coding, and implementing solutions within an Agile framework.