Site Reliability Engineer – II

🕒 Julho 13

🇮🇳 Índia – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 35%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of MRSOOL | مرسول

MRSOOL | مرسول

201 - 500 funcionários

Fundada em 2015

🍽️ Alimentos e Bebidas

✈️ Turismo

💼 Consultoria

Food & Beverage • Travel • Consulting

A MRSOOL é uma das maiores plataformas de entrega na região, oferecendo uma experiência sob demanda com altas classificações de usuários tanto na App Store da Apple quanto no Google Play. A MRSOOL oferece um serviço de "peça qualquer coisa de qualquer lugar" apoiado por uma grande frota de entregadores registrados. Ela permite que empresas acessem uma vasta base de usuários, facilitando a transformação para o e-commerce sob demanda e oferecendo um sistema de lances flexível para preços de serviços. A MRSOOL também oferece uma experiência de entrega personalizada com rastreamento em tempo real e comunicação com os entregadores, tornando-se uma opção de destaque para pedidos de lojas locais, mercados e restaurantes diretamente na sua porta.

Descrição

• Collaborate with development teams to design and implement scalable Infrastructure. • Collaborate with development teams to design and implement automated deployment and testing pipelines. • Develop and maintain monitoring and alerting systems to proactively identify and address issues. • Troubleshoot and escalate production incidents to minimize downtime and improve system reliability. • Continuously improve our infrastructure and processes to optimize scalability and efficiency. • Participate and take ownership for on-call rotations as needed to ensure 24/7 support for our application. • Perform routine maintenance and upgrades as needed to keep our systems up to date. • Contribute to ongoing efforts to improve our security posture and compliance with industry standards. • Communicate complex technical concepts clearly and concisely to both technical and non-technical stakeholders in order to make the right decision. • Mentor and coach junior engineers, fostering their professional growth and enabling them to deliver high-quality work. • Stay up-to-date with the latest advancements and trends in site reliability engineering and share knowledge and insights with the team. • Identify opportunities for organizational enhancements and propose alternatives to optimize team structures and execution.

🎯 Requisitos

• Bachelor’s degree in Computer Engineering, Computer Science, or related field. • 5+ years of experience in a similar role, preferably with experience in a high-traffic, high-availability environment. • Proficiency in at least one programming language (Python, Ruby, Java, Go, etc.). • Strong understanding of cloud infrastructure and related technologies (AWS, GCP, Azure, Kubernetes, Docker, etc.) • Excellent troubleshooting and problem-solving skills. • Experience with one or more automation and configuration management tools (Chef, Ansible, Puppet, Terraform, etc.). • Familiarity with monitoring and alerting tools (Prometheus, Grafana, Nagios, etc.) • Strong communication and interpersonal skills, enabling effective collaboration with cross-functional teams. • Ability to navigate ambiguity, set clear expectations, and thrive in a fast-paced, dynamic environment. • A strong grasp of computer science fundamentals when it comes to dealing with distributed systems and networks.

🏖️ Benefícios

• Inclusive and Diverse Environment: We foster an inclusive and diverse workplace that values innovation and offers remote environments. • Competitive Compensation: Our compensation packages are highly competitive and include potential share options for certain roles. • Personal Growth and Development: We are committed to your personal and professional growth, providing regular training and an annual learning stipend to help you advance your career in a dynamic environment. • Autonomy and Mentorship: You'll enjoy a high degree of autonomy in your role, supported by mentorship and ambitious goals that pave the way for both your success and the company's growth.

Candidatar-se

Vagas Similares

🕒 Julho 9

DBSync

51 - 200

💼 Consultoria

🏥 Saúde

📦 Logística

Forward Deployment Engineer at DBSync solving tasks for Cloud technology users and ensuring customer success through technical expertise.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 8

MariaDB

201 - 500

🏢 Corporativo

QA and release engineer testing MariaDB MaxScale, a database proxy for MariaDB clusters. Managing releases, Linux packages, CI/CD automation, and build infrastructure.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 8

Pythian

201 - 500

💼 Consultoria

🏥 Saúde

📦 Logística

Site Reliability Engineer at Pythian focusing on operating large-scale distributed systems. Responsible for designing, deploying, and operating infrastructure with strong collaboration across teams.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 7

Resilinc

201 - 500

💼 Consultoria

📦 Logística

🏥 Saúde

Site Reliability Engineer responsible for platform availability and automation in cloud environments at Resilinc. Focused on leveraging agentic AI for impactful supply chain solutions.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Julho 3

Moniepoint Inc. (Formerly TeamApt Inc.)

1001 - 5000

💳 Fintech

🏦 Bancário

Site Reliability Engineer ensuring reliability of Moniepoint’s distributed financial platform. Building automation, observability, and self-healing systems for Moniepoint’s high-growth payments and banking services.

🗣️🇺🇸🇬🇧 Inglês obrigatório