Site Reliability Engineer 3

Vaga não está no LinkedIn

🕒 Maio 16

🇮🇳 Índia – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 50%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Granicus

Granicus

501 - 1000 funcionários

Fundada em 1999

🏛️ Governo

☁️ SaaS

📋 Conformidade

Government • SaaS • Compliance

A Granicus é uma empresa de tecnologia focada no governo que fornece uma Government Experience Cloud e uma gama de serviços digitais para agências locais, estaduais, federais, de educação e de distritos especiais. Seus produtos incluem plataformas de engajamento e comunicação, nuvens de serviço e operações (para licenças, registros, solicitações de serviço/311), gerenciamento de reuniões e agendas, sites/CMS, ferramentas de conformidade e um Agente de Experiência do Governo com tecnologia AI para oferecer autoatendimento 24 horas por dia, 7 dias por semana. A Granicus ajuda organizações do setor público a modernizar a prestação de serviços, aumentar o engajamento dos cidadãos, automatizar fluxos de trabalho e melhorar a eficiência operacional.

Descrição

• Provide production support according to the team on-call roster • Work on customer and internal engineering/implementation team tickets • Work on SRE backlog items • Monitor the health and performance of services, systems, and infrastructure • Respond promptly to alerts and incidents to maintain high availability • Develop and maintain automation scripts and tools • Troubleshoot and resolve incidents, perform root cause analysis, and implement long-term fixes • Design and implement system improvements for reliability, scalability, and performance • Collaborate with software engineers on application requirements, design, architecture, deployment, and releases • Create and maintain process, procedure, and troubleshooting documentation • Assist with capacity planning • Implement and follow security best practices to protect systems and data • Lead efforts to build and maintain robust infrastructure and guide implementation of site reliability best practices

🎯 Requisitos

• Good understanding of Linux/Unix systems, networking, and cloud services such as AWS, Azure, or Google Cloud • Experience with Python, Bash, or Ruby • Bachelor’s or master’s degree in computer science, Information Technology, or a related field, or equivalent practical experience • 5+ years of experience in site reliability engineering, system administration, or a similar role • Proven track record managing large-scale, high-availability systems • Familiarity with AI/ML operations, including model lifecycle management, vector databases, and inference performance tuning • Expertise in Linux/Unix systems, networking, and cloud services • Proficiency in Python, Bash, Ruby, Go, Java, or C++ • Advanced knowledge of Elastic, Prometheus, Grafana, Splunk, Ansible, Chef, Puppet, and CI/CD pipelines • Strong analytical and problem-solving skills • Excellent verbal and written communication skills • Ability to lead and mentor a team, drive projects to completion, and manage cross-functional initiatives • Relevant certifications such as AWS Certified DevOps Engineer, AWS Certified Machine Learning – Specialty, or Google Cloud Professional DevOps Engineer are a plus

🏖️ Benefícios

• Remote work / remote-first company • Employee Resource Groups • Coffee with Mark sessions with the CEO • Microsoft Teams communities focused on wellness, art, furbabies, family, and parenting • Special guest sessions addressing issues impacting employees

Candidatar-se

Vagas Similares

🕒 Maio 16

Proofpoint

1001 - 5000

🔒 Cibersegurança

🏢 Corporativo

🔐 Segurança

Site Reliability Engineer at Proofpoint managing and operating scalable distributed systems. Focused on Kubernetes infrastructure, CI/CD, and incident response processes across regions.

🇮🇳 Índia – Remoto

💰 $28.000.000 Series F em 2008-02

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 13

Shuru

51 - 200

🤖 Inteligência Artificial

🤝 B2B

🏢 Corporativo

Senior DevOps Engineer at Shuru Technologies enhancing cloud platform infrastructure. Collaborating with teams for scalable solutions and operational readiness in a remote-first environment.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Maio 12

Volvo Cars

10.000+ funcionários

🏭 Manufatura

🚗 Transporte

🚘 Automotivo

Salesforce Release Engineer driving digital innovation at Volvo Cars. Managing Salesforce release lifecycle across global teams and developing cutting-edge technology solutions for the automotive industry.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Abril 29

Tookitaki

51 - 200

🤖 Inteligência Artificial

Site Reliability Engineer maintaining and scaling infrastructure for fintech solutions at Tookitaki. Collaborating with engineering and DevOps teams for high availability and performance.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Abril 26

Avaya

5001 - 10000

💼 Consultoria

📣 Marketing

📦 Logística

Tooling Expert at Avaya serving as a technical liaison in Cloud Operations. Focusing on complex debugging, troubleshooting, and maintaining deployment tooling for CI/CD pipelines.

🇮🇳 Índia – Remoto

💰 Post-IPO Debt em 2022-06

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório