Senior Site Reliability Engineer

Vaga não está no LinkedIn

🕒 Abril 1

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of IO Connect Services

IO Connect Services

51 - 200 funcionários

🏢 Corporativo

Enterprise • Cloud • IT Services

A IO Connect Services é uma empresa especializada em fornecer soluções de tecnologia em nuvem de alta qualidade. A empresa oferece uma gama de serviços que inclui Revisão de Arquitetura de Soluções, Avaliações de Inteligência Artificial, DevOps, Migração e Modernização, e Serviços Gerenciados. A IO Connect Services é especialista em tecnologia de nuvem e trabalha com as principais plataformas de integração do mundo para SOA, SaaS e APIs. Eles oferecem suporte abrangente para desenvolvimento nativo em nuvem e garantem segurança e conformidade para as empresas. Além disso, a empresa atingiu o Status de Competência em Varejo da AWS, sinalizando sua proficiência em fornecer soluções de nuvem especializadas para o setor de varejo.

Descrição

• Responsible for designing, building, maintaining, and scaling production services and server farms across multiple data centers for complex and data-intensive cloud services. • Design and enhance software architecture to improve scalability, service reliability, capacity, and performance. • Write automation code for provisioning and operating infrastructure at massive scale. You are not an operator, you’re an experienced software engineer focused on operations. • Work with development teams to make sure the applications fit nicely within the infrastructure and scalability/reliability is designed and implemented from the grounds up. You will work with QA on building pipelines and automation for delivering and deploying applications to production. • Roll up the sleeves to troubleshoot incidents, formulate theories and test your hypothesis, and narrow down possibilities to find the root cause. • Write postmortem reviews and remediation recommendation. • Identify bad trends before they become problems; respond to automated system alerts, effectively troubleshoot system errors and work incidents to return systems to normal operating conditions • Author and update high-quality documentation of all relevant specifications, systems and procedures • Support and comply with the company’s Quality Management System policies and procedures.

🎯 Requisitos

• Bachelor’s degree (or equivalent) in computer science or related discipline • Knowledge of IaC technologies such as Terraform, Ansible, Puppet, Chef. • Knowledge of Cluster creation and management through Kubernetes • Knowledge of Microsoft Azure, AWS, Google Cloud, Azure services, Virtual Machine in Azure, Virtual Network Configuration. • Knowledge in design patterns such as: Iaas, Paas, and Saas • Knowledge in CI/CD • Scripting knowledge with PowerShell • IPs and Mask knowledge • Ability to program (structured and OOP) using one or more high-level languages, such as Python, Java, C/C++, Ruby, and JavaScript • Experience with distributed storage technologies such as NFS, HDFS, Ceph, and Amazon S3, as well as dynamic resource management frameworks (Apache Mesos, Kubernetes, Yarn) • Proactive approach to identifying problems, performance bottlenecks, and areas for improvement

🏖️ Benefícios

• Base Salary and permanent contract directly with the company • Continuous training plan with paid certifications • Carreer plan according to your development and knowledge • Benefits above the law: 12 days of Paid Time Off, 30 day Christmas Bonus, Medical Insurance, Life Insurance, Savings Fund, Groceries Bonus • Quarterly Performance Bonus • Computer equipment for your work • Optional 100% Home Office

Candidatar-se

Vagas Similares

🕒 Fevereiro 17

Alten México

10.000+ funcionários

🚘 Automotivo

💼 Consultoria

Electronic Design & Release Engineer ensuring compatibility and compliance in automotive standards for electronic components. Lead redesign and release processes in collaboration with suppliers and internal teams.

🇲🇽 México – Remoto

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Janeiro 9

Solera, Inc.

5001 - 10000

🚘 Automotivo

💼 Consultoria

📦 Logística

Manager leading Cloud and DevOps engineering teams to optimize solutions in Public Cloud environments at Solera. Collaborating with cross-functional teams and managing engineering excellence in a global market.

🇲🇽 México – Remoto

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório