
201 - 500 employees
Founded 2007
🛍️ eCommerce
🏢 Enterprise
💰 $5M Series A on 2012-07
Cloud Storage • eCommerce • Enterprise
Backblaze is a cloud storage company that provides scalable and secure data backup solutions for both businesses and individuals. Their B2 Cloud Storage service offers S3 compatible object storage, allowing users to easily protect and manage their data with transparent pricing. Backblaze specializes in automatic and unlimited backup services for computer systems, ensuring data protection and recovery options for users, while also supporting integration with applications for enhanced functionality.
🔥 0 minutes ago
Ansible
AWS
Azure
Cloud
Distributed Systems
Docker
Google Cloud Platform
Grafana
Jenkins
Kubernetes
Linux
Microservices
Prometheus
Python
Terraform
Go
Improve your chances of getting an interview by checking your resume score before you apply.

201 - 500 employees
Founded 2007
🛍️ eCommerce
🏢 Enterprise
💰 $5M Series A on 2012-07
Cloud Storage • eCommerce • Enterprise
Backblaze is a cloud storage company that provides scalable and secure data backup solutions for both businesses and individuals. Their B2 Cloud Storage service offers S3 compatible object storage, allowing users to easily protect and manage their data with transparent pricing. Backblaze specializes in automatic and unlimited backup services for computer systems, ensuring data protection and recovery options for users, while also supporting integration with applications for enhanced functionality.
• Support the availability and durability of critical services across production environments • Monitor service health using SLIs, SLOs, and error budgets, escalating issues when thresholds are at risk • Participate in on-call rotations, incident response, and post-incident reviews • Follow ITIL/OSS processes for incident, change, problem, and capacity management • Develop automation for common operational tasks and reduce manual intervention • Contribute to monitoring, logging, and alerting frameworks such as Prometheus, Grafana, Catchpoint, and ELK • Work with CI/CD pipelines, configuration management, and infrastructure-as-code tools including Terraform, Ansible, and Jenkins • Write Bash, Python, or Go scripts to improve system reliability and efficiency • Partner with engineering, product, and operations teams on resilient system design and operations • Assist with capacity planning and disaster recovery exercises • Work with vendors and service providers to troubleshoot issues and track SLA performance • Document systems, share learnings, and contribute to a reliability-minded engineering culture • Contribute to playbooks, runbooks, and operational documentation • Identify recurring issues and propose long-term improvements • Promote reliability-focused practices within development and operations teams
• Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent experience • 2–4 years of experience in site reliability, systems engineering, or operations • Exposure to large-scale, production-grade systems • Solid Linux systems administration and troubleshooting skills • Familiarity with monitoring, alerting, incident response, and root cause analysis • Proficiency in at least one scripting language: Python, Bash, or Go • Understanding of containers, including Kubernetes and Docker, and microservices concepts • Knowledge of incident response and operational best practices • Experience in a SaaS, service provider, or distributed systems environment (preferred) • Familiarity with ITIL/OSS practices and SLO/SLAs (preferred) • Experience with cloud platforms such as AWS, GCP, or Azure (preferred) • Ability to work independently, take ownership, and drive projects from problem discovery through resolution (preferred)
Apply Now🔥 9 hours ago
DevSecOps Engineer securing SailPoint’s AWS-based identity security SaaS platform. Implementing security automation, hardening infrastructure, and supporting compliance and on-call operations.
AWS
Azure
Chef
Cloud
Cyber Security
Jenkins
Puppet
Python
Ruby
Terraform
🔥 10 hours ago
Site Reliability Engineer operating scalable cloud infrastructure for a client's Cloud Operations team. Automating deployments, monitoring, incident response, and Kubernetes migration for high-concurrency production systems.
Ansible
AWS
Cloud
Docker
HAProxy
Java
JavaScript
Kubernetes
Linux
NGINX
Node.js
Prometheus
Puppet
Python
Terraform
Go
🕒 4 days ago
DevOps Engineer building scalable cloud infrastructure and CI/CD systems for a client. Managing containers, automation, monitoring, and reliability using AWS, Docker, and Kubernetes.
🇮🇳 India – Remote
💵 ₹1.5M - ₹4.5M / year
⏰ Full Time
🟡 Mid-level
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
Ansible
AWS
Azure
Cloud
Docker
Google Cloud Platform
Grafana
Jenkins
Kafka
Kubernetes
Linux
Microservices
Prometheus
Python
Terraform
Go
🕒 4 days ago
DevOps Engineer building cloud infrastructure, CI/CD pipelines, and containerized systems for a Weekday client. Automating provisioning, monitoring reliability, and deployment workflows using AWS, Docker, Kubernetes, and IaC tools.
Ansible
AWS
Azure
Cloud
Docker
Google Cloud Platform
Grafana
Jenkins
Kafka
Kubernetes
Linux
Microservices
Prometheus
Python
Terraform
Go
🕒 5 days ago
DevOps Engineer owning Kubernetes, CI/CD, PostgreSQL, observability, and security for Signalmash’s cloud communications platform. Improving reliability, deployment speed, recovery, and infrastructure costs from India.
AWS
Azure
Cloud
Docker
Flux
Google Cloud Platform
Grafana
JavaScript
Kubernetes
Linux
Node.js
Postgres
Prometheus
Python
Shell Scripting