Senior Site Reliability Engineer

🕒 vor 1 Monat

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

👻 Geisterscore 12%

infoinfo

🗣️🇺🇸🇬🇧 Englisch erforderlich

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Omilia - Conversational Intelligence

Omilia - Conversational Intelligence

201 - 500 Mitarbeiter

Gegründet 2002

💼 Beratung

🛡️ Versicherung

✈️ Reisen

Consulting • Insurance • Travel

Omilia ist ein führendes Unternehmen im Bereich Conversational AI, spezialisiert auf Sprach- und Chatlösungen, die natürliche, umfassende Kundeninteraktionen ermöglichen. Die Omilia Cloud Platform bietet fortschrittliche, KI-gesteuerte Kundenservicetools, darunter Echtzeit-Unterstützung für Agenten, Sprachbiometrie zur Betrugsprävention und Datenanalysen zur Verbesserung der Kundenkenntnisse. Omilia bedient Branchen wie Finanzdienstleistungen, Versicherungen, Einzelhandel, Automobilindustrie und Reisen und konzentriert sich auf die Automatisierung des Kundenservice bei gleichzeitiger Gewährleistung einer sicheren und personalisierten Erfahrung.

Beschreibung

• Ensure platform reliability and availability across production and pre-production environments through proactive monitoring, alerting, and automation. • First response for incidents, contribute to problem management and root cause analysis. • Supporting the development team's effort towards reliability, creating a solid reliability culture within the development lifecycle. • Develop troubleshooting documentation for production support resources. • Collaborate with Engineering teams to develop optimised and productive runbooks, operational documentation and automation of operational tasks. • Collaborate with development and cloud engineering teams to embed reliability and performance into the software delivery lifecycle. • Design, implement, and evolve observability solutions (metrics, logs, traces, dashboards) using tools such as Prometheus, Grafana, and ELK. • Participate in on-call rotations and continuously improve alert quality and response processes. • Champion a culture of reliability, performance, and continuous improvement across teams.

🎯 Anforderungen

• - Bachelor's Degree or MS in Engineering or equivalent. • - Experience in operating at least one container orchestration cluster (Kubernetes, Docker Swarm). • - Experience developing or maintaining software for production services at scale. • - Experience with ELK. • - Experience with AWS. • - Experience with Grafana/Prometheus stack. • - Strong scripting skills (Bash, Python or Go). • - Excellent communication skills. • - Thinking out of the box and anticipating challenges. It is imperative we are not simply reactive; we must expect challenges and question technologies, procedures and thinking already in place. You will be expected to constantly review and challenge at all levels. • - Versatility. We work with agile/lean methods. We'd much rather iterate and learn than assume we know all the answers. • - Being a team player. You don't (always) work in isolation and are excited by the thought of using your team whilst involving product, experience design, engineering, and more in the process. • **Will be considered as a plus:** • - Telephony knowledge (SIP, VoIP); • - Experience in Linux Administration (RedHat, CentOS, AL); • - Working knowledge in Configuration Management tools (Terraform, Ansible); • - Experience with TCP/IP and general networking concepts; • - RDBMS knowledge (MySQL, Postgres); • - NoSQL knowledge (Redis).

🏖️ Vorteile

• - Fixed compensation; • - Long-term employment with the working days vacation; • - Development in professional growth (courses, training, etc); • - Being part of successful cutting-edge technology products that are making a global impact in the service industry; • - Proficient and fun-to-work-with colleagues; • - Apple gear.

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

DeepHealth

11 - 50

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

DevOps Engineer managing AWS infrastructure and enhancing platform reliability at DeepHealth. Collaborating with teams to automate processes and improve software delivery.

🇬🇧 Vereinigtes Königreich – Remote

💵 £60.000 - £70.000 / Jahr

💰 €225.000 Grant im 2019-08

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

DeepHealth

11 - 50

🏥 Gesundheitswesen

💼 Beratung

📦 Logistik

DevOps Engineer responsible for AWS and Kubernetes platform management at DeepHealth, a healthcare SaaS provider. Ensuring cloud infrastructure is secure and reliable for efficient software delivery.

🇬🇧 Vereinigtes Königreich – Remote

💵 £60.000 - £70.000 / Jahr

💰 €225.000 Grant im 2019-08

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

RTX

10.000+ Mitarbeiter

🏭 Fertigung

💼 Beratung

📦 Logistik

Principal Site Reliability Engineer managing complex AWS infrastructures for Collins Aerospace. Delivering B2B solutions and overseeing cloud migration projects in a collaborative team environment.

🇬🇧 Vereinigtes Königreich – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Fuse Energy

11 - 50

💼 Beratung

📦 Logistik

⚡ Energie

Database Reliability Engineer ensuring reliability, performance, and scalability of database infrastructure at Fuse Energy. Responsible for building and maintaining data pipelines and analytical schemas.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Intermedia Cloud Communications

1001 - 5000

💼 Beratung

🏥 Gesundheitswesen

⚖️ Rechtswesen

Team Lead DevOps Engineer leading a small team for a cloud communications provider. Overseeing Kubernetes, CI/CD processes, and collaborating with multiple teams.

🇬🇧 Vereinigtes Königreich – Remote

💰 Venture Round im 2017-02

⏰ Vollzeit

🟠 Senior

⛑ DevOps- und Site Reliability Engineer (SRE)

🗣️🇺🇸🇬🇧 Englisch erforderlich