Search Remote Jobs

Senior Site Reliability Engineer

🔥 1 minute ago

🍂 Massachusetts, Pennsylvania – Remote

infoinfo

đź’µ $95.3k - $158.8k / year

⏰ Full Time

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź‘» Ghost score 4%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of RELX

RELX

10,000+ employees

đź’Ľ Consulting

🏥 Healthcare

🛡️ Insurance

Consulting • Healthcare • Insurance

RELX is a global provider of information-based analytics and decision tools for professional and business customers. The company focuses on enabling its clients to make better decisions, improve results, and enhance productivity by leveraging advanced technology and data. RELX serves various sectors, including Risk, Scientific, Technical & Medical, Legal, and Exhibitions, by offering specialized information and analytical tools that facilitate critical decision-making. The company is committed to corporate responsibility and delivering societal benefit through its products by contributing to scientific advancement, legal justice, and effective market transactions.

đź“‹ Description

• Create monitoring queries and establish service level baselines • Support senior engineers during incidents • Contribute to post-mortems and root cause analyses • Participate in disaster recovery tests • Implement automation and execute code in production environments • Contribute to SRE knowledge documentation • Support deployment, monitoring, and reliability of services integrating AI tools • Support architecture and senior engineers in creating infrastructure topology drawings and deployment workflows • Test availability, reliability, and recoverability in non-production environments • Lead complex reliability initiatives and drive automation to reduce operational toil • Design and implement solutions that improve service availability, streamline operations, and enhance system recovery capabilities • Collaborate with engineering teams and host-function stakeholders • Support handover and capability-building so solutions remain owned and operable after the squad moves on

🎯 Requirements

• Expertise in advanced Terraform, including modules, providers, state management, lifecycle controls, drift detection, safe refactoring, remote state, locking, and cross-stack dependencies • Hands-on experience managing production, multi-account, multi-region AWS environments across ECS, RDS, ALB, VPC, IAM, Route53, ECR, S3, Lambda, DynamoDB, SQS, Secrets Manager, KMS, and CloudWatch • Experience building and troubleshooting reusable GitHub Actions CI/CD workflows, OIDC authentication, approval gates, runners, Terraform deployments, application deployments, and migration pipelines • Knowledge of Docker, ECR, ECS task definitions and services, IAM roles, health checks, autoscaling, ALB integration, and deployment rollbacks • Proficiency in AWS networking and security, including VPCs, networking, ALBs, Route53, ACM/TLS, IAM, OIDC, Secrets Manager, KMS, and cloud security best practices • Skilled in incident response and observability using logs, metrics, alarms, deployment history, root cause analysis, rollback decisions, and operational runbooks • Strong Linux and Git fundamentals with Bash/Python scripting for AWS CLI automation, CI/CD, and operational tooling • Hands-on experience integrating and operating AI services and APIs in production, including monitoring, reliability, and security practices for AI-powered features • Ability to support multiple engineering teams, troubleshoot across infrastructure and application layers, document solutions, and enable secure self-service practices

🏖️ Benefits

• Annual incentive bonus • Country-specific benefits • Accommodation or adjustment support during the hiring process

Apply Now

Similar Jobs

🔥 10 hours ago

Backblaze

201 - 500

🛍️ eCommerce

🏢 Enterprise

Site Reliability Engineer III managing Backblaze’s Vitess and Cassandra production databases. Ensuring reliability, automation, incident response, and operational readiness for its open cloud storage platform.

🔥 11 hours ago

Shippo

201 - 500

📦 Logistics

📣 Marketing

đź’Ľ Consulting

SRE Manager leading Kubernetes, cloud infrastructure, and reliability platforms for Shippo’s global shipping technology. Enabling product teams to deploy and operate scalable services.

🔥 14 hours ago

Climavision

11 - 50

đź’Ľ Consulting

📦 Logistics

🤖 Artificial Intelligence

Senior SRE operating Kubernetes, observability, and automated recovery for Climavision’s weather intelligence platform. Driving high availability across Azure, colocation, and edge infrastructure.

🔥 15 hours ago

Synapticure Inc.

11 - 50

🏥 Healthcare

📡 Telecommunications

⚕️ Healthcare Insurance

Senior DevOps and Security Engineer securing AWS, Kubernetes, and software delivery environments. Supporting Synapticure’s virtual neurodegenerative disease care and life sciences research platform.

🔥 15 hours ago

Wursta

51 - 200

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Digital Workplace Deployment Engineer leading Google Workspace migrations and cloud deployment services. Delivering digital transformation, managed services, cybersecurity, and AI solutions at Wursta.