Senior Site Reliability Engineer

🕒 Março 12

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $141.000 - $208.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of ClickHouse

ClickHouse

51 - 200 funcionários

Fundada em 2016

☁️ SaaS

🏢 Corporativo

🤖 Inteligência Artificial

SaaS • Enterprise • Artificial Intelligence

ClickHouse é um data warehouse em tempo real rápido e eficiente no uso de recursos e um banco de dados open source, projetado para oferecer desempenho superior de consultas para aplicações de missão crítica e sensíveis ao tempo. Está disponível como serviço em nuvem nas principais plataformas, como AWS, GCP e Azure, com opção de "Bring Your Own Cloud" e uma ampla gama de integrações para operação fluida em diferentes stacks de tecnologia. O ClickHouse se destaca em analytics em tempo real, machine learning, business intelligence e observabilidade, sendo uma escolha ideal para tarefas como serviços financeiros, detecção de fraudes e analytics para jogos. Ele oferece operações SQL amigáveis para desenvolvedores, soluções de armazenamento com ótimo custo-benefício e uma alternativa open source a bancos de dados tradicionais. Empresas como Sony, Lyft, Cisco, GitLab e Twilio utilizam o ClickHouse por sua escalabilidade, eficiência e facilidade de uso.

Descrição

• Collaborate with various engineering teams in ClickHouse to design and implement scalable, secure, and highly available systems for ClickHouse. • Establish and manage service level objectives (SLOs) and service level agreements (SLAs) for ClickHouse Cloud. • Ensure all the infrastructure components in ClickHouse Cloud (including Dataplane, Control Plane and ClickHouse Core) have monitoring and alerting in place to ensure timely detection and resolution of incidents. • Enhance and refine incident response processes and post-mortem analysis for any outages in ClickHouse Cloud including working with the support team to communicate to the impacted customers. • Continuously improve the reliability and performance of our ClickHouse services. • Plan, enable, and drive Chaos initiatives across Engineering teams, based upon internal priorities. • Manage on-call processes to respond to performance and reliability issues, and establish best practices for coordinating escalation to resolve issues and minimize downtime.

🎯 Requisitos

• Bachelor’s or Master’s degree in Computer Science or a related field. • At least 8 years of experience in Site Reliability Engineering or a related field. • Previous experience using ClickHouse in production. • Hands on experience with Go and/or Python. • Strong knowledge of cloud computing platforms such as AWS, Azure, or Google Cloud Platform. • Excellent understanding of distributed databases and SQL, particularly ClickHouse is a major plus. • Hands on experience with container orchestration tools such as Kubernetes or Docker Swarm. • Strong experience with automation and configuration management tools such as Ansible, Terraform, or Puppet. • You are a strong problem solver and have solid production debugging skills. • You are passionate about efficiency, availability, scalability, and data governance. • You thrive in a fast paced environment, and see yourself as a partner with the business with the shared goal of moving the business forward. • You have a high level of responsibility, ownership, and accountability. • Excellent communication and interpersonal skills.

🏖️ Benefícios

• Flexible work environment - ClickHouse is a globally distributed company and remote-friendly. We currently operate in 20 countries. • Healthcare - Employer contributions towards your healthcare. • Equity in the company - Every new team member who joins our company receives stock options. • Time off - Flexible time off in the US, generous entitlement in other countries. • A $500 Home office setup if you’re a remote employee. • Global Gatherings – We believe in the power of in-person connection and offer opportunities to engage with colleagues at company-wide offsites.

Candidatar-se

Vagas Similares

🕒 Março 10

Deepgram

51 - 200

💼 Consultoria

🏥 Saúde

📦 Logística

Site Reliability Engineer managing AI/ML infrastructure for Deepgram. Architecting, building, and optimizing hybrid systems with Kubernetes, AWS, and Terraform.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $150.000 - $220.000 / ano

💰 $47.000.000 Series B em 2022-11

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 7

Inetum

10.000+ funcionários

💼 Consultoria

🏥 Saúde

🛡️ Seguros

Expert DevOps / DevSecOps supporting Generative AI initiatives at Inetum for digital transformation in the United States. Designing high-value GenAI use cases and integrating new tools and practices.

🇺🇸 Estados Unidos – Remoto (EUA)

💰 Post-IPO Equity em 2007-03

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇫🇷 Francês obrigatório

🕒 Março 7

Flywire

1001 - 5000

🏥 Saúde

✈️ Turismo

📦 Logística

Manager II of Site Reliability Engineering at Flywire driving reliability, automation, and performance in cloud infrastructure. Collaborating with Engineering teams to achieve production excellence in a global environment.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $160.000 - $200.000 / ano

💰 $60.000.000 Series F em 2021-03

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 4

ALTEN Technology USA

501 - 1000

💼 Consultoria

🎖️ Defesa

🏥 Saúde

Design and Release Engineer developing vehicle components and systems from concept to production at ALTEN Technology USA.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Março 4

Akamai Technologies

5001 - 10000

🔒 Cibersegurança

🏢 Corporativo

📱 Mídia

Senior II DevOps Engineer developing and maintaining cloud infrastructures and web applications for top-tier security solutions. Engaging with highly skilled colleagues in a dynamic learning environment.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $112.500 - $202.500 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório