Search Remote Jobs

Site Reliability Engineer

Job not on LinkedIn

🔥 0 minutes ago

🇬🇧 United Kingdom – Remote

đź’µ ÂŁ70k - ÂŁ85k / year

⏰ Full Time

🟡 Mid-level

đźź  Senior

⛑ DevOps & Site Reliability Engineer (SRE)

đź‘» Ghost score 17%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Climb Channel Solutions NA

Climb Channel Solutions NA

51 - 200 employees

Founded 1982

đź’Ľ Consulting

📦 Logistics

📣 Marketing

Consulting • Logistics • Marketing

Climb Channel Solutions NA is an IT distribution company focused on providing leading and innovative technology solutions. They support technology resellers by offering expertise in areas such as virtualization, cloud, data management, and cybersecurity, thereby enhancing their partners' success. Climb is dedicated to transforming IT distribution with exceptional service and an extensive vendor marketplace to facilitate business growth for its partners across various sectors, including public and private markets.

đź“‹ Description

• Own the availability and performance of production SaaS applications running on Azure across multiple geographic regions • Lead troubleshooting and resolution of cloud infrastructure and application issues, including AKS failures, deployment rollbacks, ingress and networking issues, and autoscaling problems • Participate in an on-call rotation, including weekends, and drive incident response from detection through resolution • Improve disaster recovery, failover, and incident management processes across multi-region deployments • Build and maintain automation scripts and monitoring tools to reduce manual toil • Author post-incident reviews (RCAs), identify root causes, and drive preventive actions to closure • Partner with senior engineers and cross-functional teams on reliability, observability, and performance best practices • Contribute to continuous improvement across infrastructure, tooling, and processes • Communicate with customer-facing stakeholders during incidents through external status updates and written incident summaries

🎯 Requirements

• 5+ years of relevant experience in Site Reliability Engineering, DevOps, or Cloud Administration • Demonstrated ownership of production systems • Hands-on experience administering Azure environments, including AKS (Kubernetes), core Azure services, cloud networking, and cloud security fundamentals • Solid understanding of monitoring, logging, and alerting practices, including Datadog, Azure Monitor, or ELK stack • Hands-on troubleshooting with log analysis and stack traces using Datadog APM • Familiarity with firewalls, load balancers, VPNs, DNS, and routing • Experience with automation and scripting using PowerShell, Python, or similar • Practical understanding of backup, redundancy, and disaster recovery strategies in cloud environments • Strong ownership across the full incident lifecycle • Customer-first approach • Clear and professional written communication • Experience with AWS Cloud Platform • Experience with CI/CD tools such as Azure DevOps • Experience with infrastructure-as-code tools such as Terraform or ARM templates • Prior experience operating SaaS products with regional tenant architectures

🏖️ Benefits

• Meaningful bonus program • 10% annual bonus • Healthcare insurance • Pension/retirement matching • Comprehensive life insurance • Employee assistance program • Time off plans • Paid company holidays • Career progression opportunities • Meaningful work • Culture of innovation

Apply Now

Similar Jobs

🔥 14 hours ago

Tether.to

11 - 50

₿ Crypto

đź’ł Fintech

đź’¸ Finance

DevOps Engineer building CI/CD, Docker, and IaC infrastructure for Tether’s blockchain-powered digital finance platform. Automating secure releases across web, desktop, and mobile applications.

🔥 17 hours ago

NEC Software Solutions

5001 - 10000

🏥 Healthcare

đź’Ľ Consulting

📦 Logistics

Senior DevOps Engineer operating AWS infrastructure and delivery pipelines for NEC Software Solutions’ mission-critical public-service systems. Securing scalable applications across national programmes.

🔥 18 hours ago

Mozilla

501 - 1000

👥 B2C

đź”’ Cybersecurity

Release engineer optimizing scalable build, test, and deployment pipelines for Mozilla’s Firefox browser. Improving developer experience and supporting reliable releases across platforms.

đź•’ 5 days ago

InfluxData

201 - 500

đź’Ľ Consulting

🏭 Manufacturing

📦 Logistics

DevOps Engineer operating multi-cloud Kubernetes infrastructure for InfluxData’s time-series platform. Automating operations and supporting highly available distributed services.

đź•’ August 26

4Pharma Ltd

11 - 50

đź’Ľ Consulting

🍽️ Food & Beverage

📦 Logistics

DevOps Engineer building scalable, secure AWS infrastructure and CI/CD pipelines for BC Platforms’ healthcare data and analytics. Growing into a technical DevOps leadership role through automation and AI-native practices.