Database Reliability Engineer – Core Team

🕒 vor 4 Monaten

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 40%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of ClickHouse

ClickHouse

51 - 200 Mitarbeiter

Gegründet 2016

☁️ SaaS

🏢 Unternehmen

🤖 Künstliche Intelligenz

SaaS • Enterprise • Artificial Intelligence

ClickHouse ist ein schnelles, ressourceneffizientes Echtzeit‑Data‑Warehouse und eine Open‑Source‑Datenbank, die für überragende Abfrageleistung in geschäfts‑ und zeitkritischen Anwendungen entwickelt wurde. Es ist als Cloud‑Service auf führenden Plattformen wie AWS, GCP und Azure verfügbar, bietet eine Option „Bring Your Own Cloud“ und eine breite Palette an Integrationen für den nahtlosen Betrieb in unterschiedlichen Tech‑Stacks. ClickHouse überzeugt bei Echtzeit‑Analysen, Machine Learning, Business Intelligence und Observability und ist damit eine ideale Wahl für Anwendungsfälle wie Financial Services, Fraud Detection und Gaming‑Analytics. Es unterstützt entwicklerfreundliche SQL‑Operationen, bietet kosteneffiziente Storage‑Lösungen und stellt eine Open‑Source‑Alternative zu traditionellen Datenbanken dar. Unternehmen wie Sony, Lyft, Cisco, GitLab und Twilio setzen ClickHouse wegen seiner Skalierbarkeit, Effizienz und Benutzerfreundlichkeit ein.

Beschreibung

• Continuously improve the reliability and performance of ClickHouse core. • Improve and create metrics and alerts for ClickHouse to be able to identify and prevent problems in production before they affect customers. • Dig deeper into the most common problems encountered by customers in ClickHouse Core to identify the root cause of problems and submit bug fixes, issue reports and suggest improvements. • Enhance and refine incident response processes and post-mortem analysis for ClickHouse core related outages including working with support and Cloud teams to communicate to the impacted customers. • Plan, enable, and drive Chaos initiatives across Engineering teams, based upon internal priorities. • Manage on-call processes to respond to performance and reliability issues, and establish best practices for coordinating escalation to resolve issues and minimize customer impact.

🎯 Anforderungen

• Bachelor’s or Master’s degree in Computer Science or a related field. • At least 5 years of experience in Reliability Engineering, QA or customer facing engineering. • Previous experience operating ClickHouse or other SQL databases in production. • Excellent understanding of distributed database internals and SQL, particularly ClickHouse is a major plus. • Scripting experience with Shell or Python, and ability to read and understand C++ code. • Knowledge of cloud computing platforms such as AWS, Azure, or Google Cloud Platform. • You are a strong problem-solver and have solid production debugging skills. • You thrive in a fast-paced environment as part of a global team, and you see yourself as a partner with the business with the shared goal of moving the business forward. • You have a high level of responsibility, ownership, and accountability. • Excellent communication skills.

🏖️ Vorteile

• Flexible work environment - ClickHouse is a globally distributed company and remote-friendly. We currently operate in 20 countries. • Healthcare - Employer contributions towards your healthcare. • Equity in the company - Every new team member who joins our company receives stock options. • Time off - Flexible time off in the US, generous entitlement in other countries. • A $500 Home office setup if you’re a remote employee. • Global Gatherings – We believe in the power of in-person connection and offer opportunities to engage with colleagues at company-wide offsites.

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 4 Monaten

Prima

1001 - 5000

💼 Beratung

📦 Logistik

🛡️ Versicherung

Senior Site Reliability Engineer shaping the future of motor insurance at a leading provider. Collaborating across engineering teams to build reliable and scalable systems.

🇬🇧 Vereinigtes Königreich – Remote

💰 €115.800.000 Series A im 2018-11

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 4 Monaten

Luupli

11 - 50

👥 B2C

🛍️ eCommerce

🌍 Soziale Wirkung

Site Reliability Engineer optimizing reliability, scalability, and performance for Luupli's AWS cloud infrastructure. Collaborating with teams to enhance automation and incident management.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 4 Monaten

RemoteStar

11 - 50

💼 Beratung

📦 Logistik

📣 Marketing

Senior Site Reliability Engineer Manager ensuring infrastructure and service reliability. Leading SRE team and driving operational excellence in a B2B diamond marketplace.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 4 Monaten

Whitespace Software

51 - 200

🔌 API

💸 Finanzen

🛡️ Versicherung

DevOps Engineer at WhiteSpace managing cloud provisioning and high availability systems. Collaborating with development team on user stories and ensuring environment security.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟢 Junior

🟡 Mittelstufe

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 4 Monaten

Whitespace Software

51 - 200

🔌 API

💸 Finanzen

🛡️ Versicherung

Senior DevOps Engineer at WhiteSpace Technology managing cloud provisioning and high availability. Collaborating with developers and implementing CI/CD while ensuring system hardening and security.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich