Site Reliability Engineer

Job not on LinkedIn

🕒 July 8

🇺🇸 United States – Remote

💵 $120k - $185k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 25%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Prove

Prove

201 - 500 employees

🏥 Healthcare

🛡️ Insurance

💼 Consulting

Healthcare • Insurance • Consulting

Prove is the world’s most accurate identity verification and authentication platform, trusted by over 1,000 leading companies to reduce fraud and improve consumer experiences. The company offers a range of digital identity solutions, including Prove Pre-Fill®, Prove Identity®, and Prove Auth®, which streamline the onboarding process, enhance authentication, and provide real-time identity management. With advanced KYC-compliant identity verification and solutions that enable a seamless user experience, Prove is widely used across various industries such as banking, healthcare, and fintech to mitigate fraud and accelerate digital onboarding. Prove ensures bank-grade security, offering frictionless and secure identity verification across the globe.

📋 Description

• Design, implement, maintain, and deploy highly available, complex, scalable, and reliable systems • Leverage automation, effective monitoring, and infrastructure-as-code • Work closely with application engineering teams to ensure reliability, performance, and security • Design and implement comprehensive observability solutions across infrastructure and applications • Establish metrics, logging, and tracing systems • Create alerting thresholds and automated responses based on SLOs • Provide actionable insights into service-to-service communications • Design, build, and maintain scalable AWS cloud infrastructure • Implement infrastructure-as-code using Terraform and related tools • Automate routine operational tasks to reduce toil and improve efficiency • Ensure infrastructure security compliance and implement least-privilege access controls • Design infrastructure-as-code deployments for container-based applications • Scale containers based on custom metrics • Conduct post-incident reviews and implement preventative measures • Perform root cause analysis and system improvements using observability data • Participate in a 24/7 on-call rotation to achieve 99.999% system availability • Improve reliability, performance, scalability, and cost efficiency • Scale developer experiences through an opinionated platform approach

🎯 Requirements

• 3+ years of experience in Site Reliability or Platform Engineering teams for IC3; 5+ years for Senior level • Deep understanding of cloud platforms, particularly AWS • Strong experience with Kubernetes and container orchestration • Experience with Terraform and infrastructure-as-code tools • Bachelor's degree in Computer Science, Engineering, or equivalent practical experience • Senior level: expert knowledge of observability platforms and practices, including OpenTelemetry, Prometheus, Grafana, Jaeger, ELK stack / Splunk • Senior level: proficiency in at least one programming language, Go or Python • Preferred: experience with distributed systems and microservice architectures • Preferred: experience working in a high compliance environment • Preferred: experience with holistic monitoring and alerting for developing platforms • Preferred: skilled proficiency in Go or Python • Preferred: familiarity with service mesh technologies • Preferred: contributions to open-source projects • Preferred: experience in identity verification or financial technology • Preferred: application development experience

🏖️ Benefits

• Competitive salaries • Bonus Plan (for eligible roles) • Equity Plan • Modern Health for financial, mental, and physical wellness • 401(k) Retirement Plan & Match (US Offices) • Local Country Pension (International Offices) • Unlimited Vacation • Flexible hours • Comprehensive medical benefits for you and your family • Emotional & Physical Wellness – Access to wellness services (EAP & Prove Well-Being Reimbursement) • Bottomless snacks & beverages for certain office locations • Daily GrubHub stipend for lunch if coming into the office (US Offices)

Apply Now

Similar Jobs

🕒 July 8

Cribl

501 - 1000

☁️ SaaS

Senior Site Reliability Engineer unlocking the value of observability data for Cribl. Engaging with teams to improve service delivery and reliability in a remote-first environment.

🕒 July 8

ShorePoint Inc

1 - 10

💼 Consulting

🏥 Healthcare

📦 Logistics

DevSecOps Engineer supporting cloud-based cybersecurity data systems in fast-paced public sector environments. Drive operational excellence through engineering, operating, and monitoring data infrastructure.

🕒 July 8

Health Catalyst

1001 - 5000

🏥 Healthcare

💼 Consulting

📦 Logistics

Site Reliability Engineer on Central AI team supporting AI systems for healthcare organizations at Health Catalyst. Train teams in AI practices and ensure governance and best practices are followed.

🕒 July 7

Nametag

11 - 50

🏥 Healthcare

🛡️ Insurance

📦 Logistics

Software Engineer focusing on infrastructure and reliability at Nametag for secure digital identity. Designing scalable systems and tooling to enhance engineering productivity.

🇺🇸 United States – Remote

💵 $120k - $190k / year

💰 Series unknown on 2021-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🕒 July 6

Infarsight

51 - 200

🤖 Artificial Intelligence

✈️ Travel

📦 Logistics

Senior DevOps & Cloud Infrastructure Engineer optimizing AWS environments for automation and product innovation. Leading deployment strategies and resource management across AWS, Vercel, and RackSpace.