Site Reliability Engineer I

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Backblaze

Backblaze

201 - 500 employees

Founded 2007

🛍️ eCommerce

🏢 Enterprise

💰 $5M Series A on 2012-07

Cloud Storage • eCommerce • Enterprise

Backblaze is a cloud storage company that provides scalable and secure data backup solutions for both businesses and individuals. Their B2 Cloud Storage service offers S3 compatible object storage, allowing users to easily protect and manage their data with transparent pricing. Backblaze specializes in automatic and unlimited backup services for computer systems, ensuring data protection and recovery options for users, while also supporting integration with applications for enhanced functionality.

📋 Description

• Act as first point of contact for all customer-affecting issues • Drive resolution of technical problems • Follow incident management processes and complete incident post-mortems • Communicate consistently with management • Respond to and monitor Zabbix alerts, taking direct action or escalating • Ensure successful handoffs of escalations • Maintain pod health across all sites and define pod alerts in Zabbix • Perform daily filesystem checks for pods • Troubleshoot technical issues for DC Techs, including pod, deployment, migration, and Ansible playbook issues • Identify and escalate potential network issues • Configure and test Vault before deployments • Start and monitor Vault migrations and perform migration pod health checks • Document and automate daily tasks • Document and provide network IPs for upcoming deployments • Monitor server farm releases and updates and escalate issues • Participate in on-call rotation shifts • Assist TechOps team members with tasks • Recommend improvements to organizational productivity • Work outside normal business hours, including weekends, holidays, and evenings, as needed

🎯 Requirements

• Must be located in Bangalore • 2–4 years of relevant experience • Knowledge of system administration and Linux • Willingness to learn and develop necessary technical skills • Strong analytical thinking • Strong collaboration and communication skills across teams • Knowledge of network cabling, network classification, and network topology

🏖️ Benefits

• Equal Opportunity Employer • Diversity, equity, and inclusion commitment • Culture emphasizing learning, developing, and growing

Apply Now

Similar Jobs

🔥 9 hours ago

SailPoint

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

DevSecOps Engineer securing SailPoint’s AWS-based identity security SaaS platform. Implementing security automation, hardening infrastructure, and supporting compliance and on-call operations.

AWS

Azure

Chef

Cloud

Cyber Security

Jenkins

Puppet

Python

Ruby

Terraform

🔥 10 hours ago

Ontrac Solutions

11 - 50

🤖 Artificial Intelligence

💼 Consulting

🤝 B2B

Site Reliability Engineer operating scalable cloud infrastructure for a client's Cloud Operations team. Automating deployments, monitoring, incident response, and Kubernetes migration for high-concurrency production systems.

Ansible

AWS

Cloud

Docker

HAProxy

Java

JavaScript

Kubernetes

Linux

NGINX

Node.js

Prometheus

Puppet

Python

Terraform

Go

🕒 4 days ago

Weekday

501 - 1000

👗 Fashion

🛒 Retail

🛍️ eCommerce

DevOps Engineer building scalable cloud infrastructure and CI/CD systems for a client. Managing containers, automation, monitoring, and reliability using AWS, Docker, and Kubernetes.

Ansible

AWS

Azure

Cloud

Docker

Google Cloud Platform

Grafana

Jenkins

Kafka

Kubernetes

Linux

Microservices

Prometheus

Python

Terraform

Go

🕒 4 days ago

Weekday (YC W21)

11 - 50

💼 Consulting

👥 HR Tech

☁️ SaaS

DevOps Engineer building cloud infrastructure, CI/CD pipelines, and containerized systems for a Weekday client. Automating provisioning, monitoring reliability, and deployment workflows using AWS, Docker, Kubernetes, and IaC tools.

Ansible

AWS

Azure

Cloud

Docker

Google Cloud Platform

Grafana

Jenkins

Kafka

Kubernetes

Linux

Microservices

Prometheus

Python

Terraform

Go

🕒 5 days ago

Signalmash

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

DevOps Engineer owning Kubernetes, CI/CD, PostgreSQL, observability, and security for Signalmash’s cloud communications platform. Improving reliability, deployment speed, recovery, and infrastructure costs from India.

AWS

Azure

Cloud

Docker

Flux

Google Cloud Platform

Grafana

JavaScript

Kubernetes

Linux

Node.js

Postgres

Prometheus

Python

Shell Scripting