Senior Engineering Manager – Site Reliability

🕒 July 16

🇺🇸 United States – Remote

💵 $260k - $280k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Horizon3.ai

Horizon3.ai

51 - 200 employees

Founded 2019

🔒 Cybersecurity

🤖 Artificial Intelligence

☁️ SaaS

Cybersecurity • Artificial Intelligence • SaaS

Horizon3. ai is a cybersecurity company that specializes in autonomous penetration testing through its platform, NodeZero. The company's mission is to proactively identify and address attack vectors before they can be exploited, allowing organizations to continuously assess their security posture across various environments including cloud, IoT, and on-premises systems. Founded by veterans from the US Special Operations and National Security sectors, Horizon3. ai offers a self-service SaaS solution that provides organizations with insights into their security vulnerabilities without requiring persistent or credentialed agents.

📋 Description

• Build a SRE team, from scratch. Hire experienced site reliability staff and build a team of 4-6 in year one. • Professionalize incident management. Define and document incident processes and practices for your SRE team and for the application feature teams. • Drive incident professionalism and reliability culture across the engineering organization through training and process adoption. • Use design reviews, code reviews, and blameless retrospectives to drive a culture of quality and excellence in engineering. • Balance incident response while also executing on a roadmap of observability and reliability engineering initiatives. • Hire and directly manage site reliability engineers. • Responsible for recruiting, onboarding, mentoring, coaching, and developing your team. • Recognizing and retaining high performers. • Leading horizontally with peer management & senior leaders.

🎯 Requirements

• Demonstrated experience leading hiring and growing SRE or Infrastructure teams • Experience leading or building SRE functions, including incident management processes, on-call programs, SLO/SLA definition, and operational runbooks • Previous career experience as a Site Reliability Engineer • Deep hands-on experience with observability: application performance management, logs and traces, and golden signals and service-specific metrics • Experience in selecting and deploying incident management tooling (e.g., PagerDuty, FireHydrant,etc.) • Strong working knowledge of at least one major cloud provider (AWS, GCP, or Azure)

🏖️ Benefits

• Health, vision & dental insurance for you and your family • Flexible vacation policy • Generous parental leave • Equity package in the form of stock options

Apply Now

Similar Jobs

🕒 July 16

Accelerant

201 - 500

🛡️ Insurance

☁️ SaaS

🤝 B2B

Senior SRE driving reliability and observability across financial data platform at Accelerant. Partnering with engineering to enhance monitoring, metrics, and incident management processes.

🇺🇸 United States – Remote

💰 $150M Private Equity Round - Accelerant on 2023-06

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info

🕒 July 15

Empower AI

501 - 1000

🎖️ Defense

🏥 Healthcare

📦 Logistics

Sr. DevOps Engineer architecting scalable infrastructure and automating solutions for USCIS in a fully remote role. Mentoring engineers and supporting critical systems with complex challenges.

🇺🇸 United States – Remote

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 15

Encoura

51 - 200

💼 Consulting

📣 Marketing

📚 Education

Azure DevOps Engineer responsible for ensuring reliability of large-scale production systems at Encoura. Leading CI/CD initiatives and collaborating with engineering teams on cloud infrastructure.

🇺🇸 United States – Remote

💵 $116k - $128.8k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 15

Vytalize Health

201 - 500

🏥 Healthcare

☁️ SaaS

⚕️ Healthcare Insurance

Data Reliability Engineer ensuring operational health of healthcare data pipelines at Vytalize Health. Focused on data quality, compliance, and collaboration across teams.

🇺🇸 United States – Remote

💰 $100M Series C - Vytalize Health on 2023-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info

🕒 July 15

11:11 SYSTEMS

201 - 500

💼 Consulting

🏥 Healthcare

📦 Logistics

Infrastructure Deployment Engineer managing deployment projects across global data centers. Leading cross-functional teams to ensure timely and standardized execution of infrastructure projects.

🇺🇸 United States – Remote

💵 $94k - $130.5k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)