Senior Site Reliability Engineer

🕒 4 dias atrás

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Fortress Information Security

Fortress Information Security

201 - 500 funcionários

Fundada em 2015

🔒 Cibersegurança

🏛️ Governo

🤖 Inteligência Artificial

💰 $125.000.000 Series C - Fortress Information Security em 2022-04

Cybersecurity • Government • Artificial Intelligence

A Fortress Information Security é uma empresa de cibersegurança impulsionada por IA que defende infraestruturas críticas, agências governamentais e suas cadeias de suprimentos contra ameaças cibernéticas e riscos de missão. A empresa foca em inteligência de ameaças, gestão de vulnerabilidades, gerenciamento de riscos de fornecedores e terceiros (TPRM), segurança cibernética da cadeia de suprimentos e produtos (C-SCRM), SBOMs e proteção de clientes do setor público e de infraestrutura crítica.

Descrição

• Transition applications from traditional (non-containerized) Ansible deployments to containerized, orchestrated deployments in AWS and on-premises environments • Build upon current CI/CD efforts to support different deployment strategies (Blue/Green, Canary, etc.) • Support Development/QA/UAT efforts by building an environment for anyone in the company to test drive a release anywhere in the development lifecycle • Improve common infrastructure for developers, such as CI/CD pipelines, log/application monitoring, cluster management, and configuration management • Migrate application secrets and configuration from Ansible Vault to Hashicorp Vault • Automate infrastructure provisioning during deployment • Handle code deployments in all environments (cloud, on-premises) • Implement tools to monitor and alert with respect to service level metrics and objectives • Provide technical guidance and educate team members and coworkers on development and operations • Monitor relevant systems for availability and performance • Available to support daytime and after business hours release activities

🎯 Requisitos

• 5-8 years hands-on experience in a SRE/DevOps role supporting production systems • Production experience supporting Linux-based infrastructure and administering services on AWS (RDS, VPC, ECR, CloudWatch, Cloud Formation, Lambda, API Gateway) and on-premises • Demonstrable experience with deployment technologies such as Kubernetes, Ansible, Jenkins and Terraform (or similar technologies) • Excellent documentation skills so anyone on the team can come up-to-speed on changes quickly • Excellent written and verbal communication skills • Experience implementing rolling upgrades (canary, blue/green, etc.) • Strong scripting and tooling skillset (Bash, Python, JS, etc.) • Self-motivated, resourceful and a persistent problem-solving aptitude with advanced time management skills • Ability to independently use and refine prompts to enhance the quality, efficiency, and insight of regular work processes • Must be willing to participate in technical interviews and technical questions, which may be recorded or transcribed for evaluation purposes (required)

🏖️ Benefícios

• Remote and Hybrid working environment • Competitive pay structure • Medical, dental, vision plans with employees covered up to 90% with highly progressive options for dependents and families • Company paid life, short- and long-term disability insurance • Employee Assistance Program • 401(k) match • Flexible Paid Time Off • Parental Leave

Candidatar-se

Vagas Similares

🕒 5 dias atrás

MyFitnessPal

51 - 200

🏥 Saúde

🍽️ Alimentos e Bebidas

💼 Consultoria

Site Reliability Engineer improving MyFitnessPal’s production systems reliability and security. Engaging in incident response, observability, and infrastructure management.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $165.000 / ano

💰 $18.000.000 Series A em 2013-08

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

info

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

CXM

201 - 500

💸 Finanças

💳 Fintech

Application Site Reliability Engineer focusing on .NET/C# services reliability for trading systems. Collaborating with software engineers to enhance service resilience and operational excellence.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

NVIDIA

10.000+ funcionários

🏥 Saúde

🏭 Manufatura

🤖 Inteligência Artificial

DevOps Engineer supporting NVIDIA’s Rapids project for AI and data science initiatives. Collaborating with teams to ensure high-quality software releases and infrastructure maintenance.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

Global Enterprise Services, LLC (GES)

11 - 50

💼 Consultoria

📦 Logística

Reliability Engineer responsible for cloud platform performance and incident response, managing compliance. Requires strong technical expertise and 8 years of experience.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

🔴 Especialista

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 6 dias atrás

Mirantis

501 - 1000

💼 Consultoria

🏥 Saúde

📦 Logística

Senior DevOps Engineer handling high-performance storage for AI platforms at Mirantis. Integrating and operating storage solutions within Kubernetes environments.

🗣️🇺🇸🇬🇧 Inglês obrigatório