
51 - 200 employees
📡 Telecommunications
✈️ Travel
Telecommunications • Technology • Travel
Airalo is the world's first eSIM store that provides digital SIM cards (eSIMs) to travelers in over 200 countries and regions globally. The company offers an innovative solution to avoid high roaming charges by allowing users to purchase and activate eSIMs via its app, ensuring instant connectivity without the need for physical SIM cards. Airalo caters to various travelers by providing local, regional, and global eSIMs with transparent, prepaid pricing plans, supported by 24/7 customer service. The platform also supports partnership through APIs and offers incentives such as referral credits. Airalo represents a modern approach to global mobile connectivity, making it an essential tool for frequent travelers.
🕒 June 26
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
📡 Telecommunications
✈️ Travel
Telecommunications • Technology • Travel
Airalo is the world's first eSIM store that provides digital SIM cards (eSIMs) to travelers in over 200 countries and regions globally. The company offers an innovative solution to avoid high roaming charges by allowing users to purchase and activate eSIMs via its app, ensuring instant connectivity without the need for physical SIM cards. Airalo caters to various travelers by providing local, regional, and global eSIMs with transparent, prepaid pricing plans, supported by 24/7 customer service. The platform also supports partnership through APIs and offers incentives such as referral credits. Airalo represents a modern approach to global mobile connectivity, making it an essential tool for frequent travelers.
• Lead the design of scalable, fault-tolerant and self-healing systems in a multi-region AWS environment. • Define and track Service Level Objectives (SLOs) and Service Level Indicators (SLIs) to drive architectural decisions and error budget policies. • Conduct blameless post-incident reviews to uncover systemic root causes and implement long-term preventive measures. • Identify patterns of manual work and lead the development of internal tools/automation to permanently eliminate them. • Develop and maintain automated runbooks and playbooks for common operational tasks and complex incident response. • Shift from simple monitoring to deep observability, ensuring high cardinality data leads to proactive actionable insights. • Proactively identify and mitigate operational risks through chaos engineering and architecture reviews. • Work with software engineers to design systems for reliability, scalability, and maintainability from the early stages of the SDLC. • Continuously evaluate and optimize system performance, capacity, and cost efficiency. • Beyond just participating, you will refine the on-call experience to reduce alert fatigue, improve MTTR, and ensure sustainable rotation health.
• Bachelor’s degree in Computer Engineering or a similar discipline. • 5+ years of experience as a Site Reliability Engineer or in a similar role. • 3+ years of experience with AWS services including strong knowledge of container orchestration. • 2+ years of Kubernetes experience. • Deep understanding of observability principles and tools such as: Prometheus, Datadog, OpenTelemetry and similar. • Experience with leading incident management and complex postmortem analysis. • Experience and interest in managing infrastructure as code (Terraform). • Experience with chaos engineering and other techniques for testing system resilience. • Experience with CI/CD tools such as GitHub Actions for automated delivery. • Proficiency in at least one programming language (Python, Go, Java, etc.) for building automation and internal tooling. • Event-driven architecture experience (SNS, SQS etc). • Ability to work independently and collaboratively in a fast-paced environment. • Team player and open to new ideas. • Good communication skills and fluency in English.
• Remote work • Generous PTO • Wellness allowances • Learning allowances • Annual Airalo Away retreat
Apply Now🕒 June 19
Senior Site Reliability Engineer at Redzone, ensuring reliability and performance of mission-critical services. Evolving SRE practices while driving automation and operational excellence within the team.
Distributed Systems
🕒 June 18
Site Reliability Engineer at Tempo working on infrastructure to support various global engineering products. Collaborating with teams and ensuring high availability and performance standards.
Ansible
AWS
Cloud
Docker
Java
Kotlin
Kubernetes
Linux
Terraform
🕒 June 2
DevOps Engineer at Alkemy leading innovation and collaboration, specializing in digital growth through technological expertise.
🗣️🇪🇸 Spanish Required
AWS
Azure
Google Cloud Platform
Kubernetes
Linux
SDLC
🕒 May 29
DevOps Engineer optimizing software deployments and enhancing collaboration between Software Development and IT Operations at SGS. Focused on automation, reliability, and secure software delivery.
Azure
Cloud
Docker
Jenkins
Kubernetes
Prometheus
Terraform
🕒 May 6
Senior Site Reliability Engineer with Stellar Cyber, enhancing cloud reliability and efficiency through Kubernetes and observability tools. Collaborating in a diverse team for operational excellence.
🇪🇸 Spain – Remote
💰 $38M Series B on 2021-11
⏰ Full Time
🟠 Senior
⛑ DevOps & Site Reliability Engineer (SRE)
AWS
Azure
Cloud
Distributed Systems
ElasticSearch
Google Cloud Platform
Grafana
Kafka
Kubernetes
Linux
MongoDB
Prometheus
Python
Redis
Spark
Terraform