Site Reliability Engineer

🕒 Julho 17

🇪🇸 Espanha – Remoto

💵 €58.000 - €97.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 4%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

🗣️🇪🇸 Espanhol obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Tinybird

Tinybird

11 - 50 funcionários

Fundada em 2019

🔌 API

🏢 Corporativo

⚡ Produtividade

API • Enterprise • Productivity

Tinybird é uma plataforma de infraestrutura de dados que permite que equipes de software desenvolvam e lancem rapidamente funcionalidades de análise. Ao abstrair a ingestão de dados, armazenamento, computação e desenvolvimento de API em um único fluxo de trabalho, o Tinybird reduz significativamente as dependências e acelera o processo de desenvolvimento. Ele permite que os usuários criem produtos de dados em tempo real, análises voltadas para o usuário e painéis de controle com facilidade, mantendo alta performance e confiabilidade.

Descrição

• As part of the Platform team, your work will focus on the systems that keep Tinybird reliable, efficient, observable, and scalable. • Enhancing high availability and elasticity so the system can scale automatically and efficiently. • Boosting our observability capabilities, from low-level resource usage to high-level service metrics. • Improving disaster recovery with better tools, incident discovery, and enhanced oncall experiences. • Handling Kubernetes lifecycle tasks, managing cluster infrastructure, autoscaling, and ensuring safe deployments. • Understanding how ClickHouse operates under the hood and extracting the best performance possible from it. • Identifying bottlenecks and improving performance across storage, networking, and compute. • Reducing operational burden by transforming manual or fragile processes into repeatable, well-managed systems. • Assisting with incident prevention, operational reviews, and follow-up tasks after reliability issues. • Strengthening CI/CD foundations to help teams build and deploy changes with greater confidence.

🎯 Requisitos

• You have strong experience in designing, building, and running distributed cloud architectures and large-scale web-based production systems. • You have deep knowledge of Kubernetes, which is essential for this role. • You are skilled in AWS and GCP. • Coding skills are required. • You are comfortable operating close to production: debugging incidents, understanding system behavior, improving observability, and enhancing service reliability. • You think in systems and pay attention to edge cases, failure modes, and specific implementation details. • You care about performance, reliability, cost efficiency, and operational simplicity. • You take ownership, follow through, and are willing to tackle issues that may be broken. • You enjoy data and SQL, and you are curious about how real-time analytical systems work. • Familiarity with Traefik, Varnish, Redis, Terraform, or Ansible is not mandatory, but it's helpful. • You communicate clearly in writing. • You use AI tools such as Claude Code, Cursor, ChatGPT, and others. • You are fluent in English and Spanish. • You are willing to participate in on-call rotations. • You are located in an EU timezone.

🏖️ Benefícios

• 22 days of holiday a year (plus your birthday and public holidays). • Freedom to work from wherever suits you best. • We provide up to €2,800 to help you set up your home workspace.

Candidatar-se

Vagas Similares

🕒 Julho 16

CONVOTIS

501 - 1000

🤝 B2B

☁️ SaaS

DevOps Engineer at CONVOTIS focusing on Cloud Adoption, offering flexible hours and remote work. Requires 3 years of experience, particularly in Azure and DevOps.

🇪🇸 Espanha – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇪🇸 Espanhol obrigatório

🕒 Julho 8

knowmad mood

1001 - 5000

💼 Consultoria

🏥 Saúde

📦 Logística

Automation & DevOps Engineer leading AWS cloud migration and enhancing workplace automation. Join a leader in digital transformation with 3,000+ innovative team members.

🇪🇸 Espanha – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇪🇸 Espanhol obrigatório

🕒 Junho 30

Progress

1001 - 5000

🤖 Inteligência Artificial

Senior Cloud Operations Engineer at Progress, building and managing multi-cloud platforms for next-generation AI products. Leading infrastructure workflows and collaborating with product and engineering teams in a global environment.

🇪🇸 Espanha – Remoto

💰 Post-IPO Equity em 1995-01

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 26

Airalo

51 - 200

📡 Telecomunicações

✈️ Turismo

Senior Site Reliability Engineer building and maintaining reliable systems for Airalo's eSIM platform. Collaborating with software engineers and leading SRE principles for innovative solutions.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Junho 19

QAD

1001 - 5000

🏭 Manufatura

📦 Logística

🍽️ Alimentos e Bebidas

Senior Site Reliability Engineer at Redzone, ensuring reliability and performance of mission-critical services. Evolving SRE practices while driving automation and operational excellence within the team.

🗣️🇺🇸🇬🇧 Inglês obrigatório