Senior Site Reliability Engineer

đź•’ 3 days ago

🍂 Massachusetts, Pennsylvania – Remote

infoinfo

đź’µ $95.3k - $158.8k / year

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź‘» Ghost score 4%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of RELX

RELX

10,000+ employees

đź’Ľ Consulting

🏥 Healthcare

🛡️ Insurance

Consulting • Healthcare • Insurance

RELX is a global provider of information-based analytics and decision tools for professional and business customers. The company focuses on enabling its clients to make better decisions, improve results, and enhance productivity by leveraging advanced technology and data. RELX serves various sectors, including Risk, Scientific, Technical & Medical, Legal, and Exhibitions, by offering specialized information and analytical tools that facilitate critical decision-making. The company is committed to corporate responsibility and delivering societal benefit through its products by contributing to scientific advancement, legal justice, and effective market transactions.

đź“‹ Description

• Create monitoring queries and establish service level baselines • Support senior engineers during incidents • Contribute to post-mortems and root cause analyses • Participate in disaster recovery tests • Implement automation and execute code in production environments • Contribute to SRE knowledge documentation • Support deployment, monitoring, and reliability of services integrating AI tools • Support architecture and senior engineers in creating infrastructure topology drawings and deployment workflows • Test availability, reliability, and recoverability in non-production environments • Lead complex reliability initiatives and drive automation to reduce operational toil • Design and implement solutions that improve service availability, streamline operations, and enhance system recovery capabilities • Collaborate with engineering teams and host-function stakeholders • Support handover and capability-building so solutions remain owned and operable after the squad moves on

🎯 Requirements

• Expertise in advanced Terraform, including modules, providers, state management, lifecycle controls, drift detection, safe refactoring, remote state, locking, and cross-stack dependencies • Hands-on experience managing production, multi-account, multi-region AWS environments across ECS, RDS, ALB, VPC, IAM, Route53, ECR, S3, Lambda, DynamoDB, SQS, Secrets Manager, KMS, and CloudWatch • Experience building and troubleshooting reusable GitHub Actions CI/CD workflows, OIDC authentication, approval gates, runners, Terraform deployments, application deployments, and migration pipelines • Knowledge of Docker, ECR, ECS task definitions and services, IAM roles, health checks, autoscaling, ALB integration, and deployment rollbacks • Proficiency in AWS networking and security, including VPCs, networking, ALBs, Route53, ACM/TLS, IAM, OIDC, Secrets Manager, KMS, and cloud security best practices • Skilled in incident response and observability using logs, metrics, alarms, deployment history, root cause analysis, rollback decisions, and operational runbooks • Strong Linux and Git fundamentals with Bash/Python scripting for AWS CLI automation, CI/CD, and operational tooling • Hands-on experience integrating and operating AI services and APIs in production, including monitoring, reliability, and security practices for AI-powered features • Ability to support multiple engineering teams, troubleshoot across infrastructure and application layers, document solutions, and enable secure self-service practices

🏖️ Benefits

• Annual incentive bonus • Country-specific benefits • Accommodation or adjustment support during the hiring process

Apply Now

Similar Jobs

đź•’ 4 days ago

Backblaze

201 - 500

🛍️ eCommerce

🏢 Enterprise

Site Reliability Engineer III managing Backblaze’s Vitess and Cassandra production databases. Ensuring reliability, automation, incident response, and operational readiness for its open cloud storage platform.

🇺🇸 United States – Remote

đź’µ $125k - $150k / year

đź’° $5M Series A on 2012-07

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

đź•’ 4 days ago

Shippo

201 - 500

📦 Logistics

📣 Marketing

đź’Ľ Consulting

SRE Manager leading Kubernetes, cloud infrastructure, and reliability platforms for Shippo’s global shipping technology. Enabling product teams to deploy and operate scalable services.

🇺🇸 United States – Remote

đź’µ $175k - $238k / year

đź’° $50M Series E - Shippo on 2021-06

⏰ Full Time

đźź  Senior

đź”´ Lead

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

infoinfo

đź•’ 4 days ago

Climavision

11 - 50

đź’Ľ Consulting

📦 Logistics

🤖 Artificial Intelligence

Senior SRE operating Kubernetes, observability, and automated recovery for Climavision’s weather intelligence platform. Driving high availability across Azure, colocation, and edge infrastructure.

🇺🇸 United States – Remote

đź’µ $130k - $170k / year

đź’° $100M Series A on 2021-06

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź•’ 4 days ago

Synapticure Inc.

11 - 50

🏥 Healthcare

📡 Telecommunications

⚕️ Healthcare Insurance

Senior DevOps and Security Engineer securing AWS, Kubernetes, and software delivery environments. Supporting Synapticure’s virtual neurodegenerative disease care and life sciences research platform.

🇺🇸 United States – Remote

đź’µ $150k - $170k / year

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź•’ 4 days ago

Wursta

51 - 200

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Digital Workplace Deployment Engineer leading Google Workspace migrations and cloud deployment services. Delivering digital transformation, managed services, cybersecurity, and AI solutions at Wursta.

🇺🇸 United States – Remote

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)