Senior Site Reliability Engineer

🕒 4 dias atrás

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $180.000 - $200.000 / ano

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of PayNearMe

PayNearMe

201 - 500 funcionários

Fundada em 2009

💳 Fintech

☁️ SaaS

🤝 B2B

🔥 Investimento no último ano

💰 $50.000.000 Series E - PayNearMe em 2025-09

Fintech • SaaS • B2B

A PayNearMe é uma empresa de tecnologia de pagamentos que oferece uma plataforma moderna e completa para que as empresas aceitem, distribuam e gerenciem pagamentos. A plataforma enfatiza o Gerenciamento da Experiência de Pagamento, oferecendo recursos como processamento de pagamentos moderno, automação e autosserviço, gerenciamento de exceções e uma rede de pagamento em dinheiro no varejo. A PayNearMe atende clientes empresariais em diversos setores (financiamento de automóveis e consumo, pedágios, iGaming, concessionárias "compre aqui, pague aqui", cooperativas de crédito, serviços hipotecários, escritórios de advocacia, etc. ) e divulga métricas incluindo mais de 16 anos de operação, mais de $50 bilhões processados anualmente, mais de 16. 000 empresas na sua plataforma e mais de 62. 000 locais de pagamento em dinheiro no varejo. Os serviços de transmissão de dinheiro são fornecidos por suas subsidiárias PayNearMe MT, Inc. e PayNearMe Financial, Inc.

Descrição

• Infrastructure Management: Design, implement, and maintain scalable and resilient infrastructure using Terraform for infrastructure as code, ensuring high availability and performance • Kubernetes and Containers: Deploy, manage, and optimize Kubernetes clusters and containerized applications using Docker. Implement best practices for container orchestration and management • Systems and Application Monitoring/Observability: Develop and maintain comprehensive monitoring and observability solutions using Datadog. Ensure detailed visibility into system performance and application health • SLOs and SLA Management: Define, monitor, and maintain Service Level Objectives (SLOs) and Service Level Agreements (SLAs) to ensure reliable and consistent service delivery • Incident Response and Troubleshooting: Respond to incidents, perform root cause analysis, and implement solutions to prevent recurrence. Participate in post-incident reviews and contribute to blameless postmortems • Reliability and Production Environment Management: Ensure the reliability and stability of our production environments. Continuously assess and improve system reliability, identifying and addressing potential points of failure • Automation and Scripting: Develop automation scripts and tools to reduce manual intervention and improve system reliability using Python, Bash, or Go. Implement and improve CI/CD pipelines • CI/CD Pipeline Management: Enhance and maintain continuous integration and continuous deployment pipelines using GitLab CI. Ensure seamless and reliable deployment processes • Capacity Planning and Scaling: Assist in capacity planning and ensure that systems are scalable to meet future demands. Implement auto-scaling strategies where applicable • Security and Compliance: Implement security best practices and ensure compliance with industry standards. Regularly review and update security policies and procedures • Collaboration and Support: Work closely with development teams to ensure reliability and scalability of new features and services. Provide technical support and guidance on infrastructure-related issues • Software Engineering for Operations: Develop and maintain internal tools and services that enhance the efficiency and reliability of our operations • On-Call Rotation: Participate in an on-call rotation to address production issues and collaborate in incident response efforts

🎯 Requisitos

• +3 years of experience in SRE, DevOps, or a related role • Cloud Platform Experience: Proficient with cloud platforms such as AWS, GCP, or Azure Experience with EC2, RDS, VPCs, and security groups is essential. • Kubernetes and Containers: Strong experience with Kubernetes and Docker, including deployment, scaling, and management of containerized applications • Infrastructure as Code: Expert in using Terraform for infrastructure as code. Proficient with configuration management tools such as Ansible, Puppet, or Chef • Monitoring and Observability: Extensive experience with monitoring and observability tools like Datadog, Prometheus, Grafana, ELK stack, or Splunk. Skilled in setting up detailed monitoring and logging systems • SLOs and SLA Management: Proven ability to define, monitor, and maintain SLOs and SLAs to ensure reliable service delivery • Scripting and Automation: Strong skills in scripting languages like Python, Bash, or Go. Experience automating repetitive tasks and processes • CI/CD Practices: Familiarity with GitLab CI or similar tool for continuous integration and deployment. Experience in setting up and managing pipelines • Production Environments: Experience supporting production environments running Go or Ruby/Rails applications • Tool Development: Ability to write and update tools to support infrastructure and application management, demonstrating the principle that “SRE is what happens when you ask a software engineer to design an operations team • DevOps Best Practices: Deep understanding of DevOps principles, practices, and tools to drive continuous improvement in the software development lifecycle • Soft Skills: Strong organizational skills, attention to detail, and the ability to work collaboratively in a team environment. Excellent documentation skills to ensure accurate and detailed records • Problem-Solving Ability: Excellent analytical and problem-solving skills to diagnose and resolve complex system issues quickly and effectively.

🏖️ Benefícios

• Competitive salary and benefits with growth-company options grant • Fast- paced and professional work culture • Stock options with standard startup vesting - 1 year cliff; 4 years total • $50 monthly communication expense stipend to go towards your phone/internet bill • $250 stipend to enhance your WFH setup • Reimbursement for peripheral equipment: monitor (up to $400), keyboard and mouse (up to $200) • Premium medical benefits including vision and dental (100% coverage for employees) • Company-sponsored life and disability insurance • Paid parental bonding leave • Paid sick leave, jury duty, bereavement • 401k plan • Flexible Time Off (our team members typically take off ~3-4 weeks per year) • Volunteer Time Off • 13 scheduled holidays

Candidatar-se

Vagas Similares

🕒 4 dias atrás

Made4net

51 - 200

📦 Logística

☁️ SaaS

🏢 Corporativo

Cloud Operations Engineer supporting AWS infrastructure for supply chain software solutions. Monitoring systems and ensuring reliability in a global operations team.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 / ano

💰 Private equity em 2021-02

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

Sycurio

51 - 200

☁️ SaaS

🔐 Segurança

📋 Conformidade

Deployment Engineer for Sycurio solutions deployment and testing. Involves supporting installations and training customer support engineers.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

Otoe Missouria Group

11 - 50

🏛️ Governo

🔒 Cibersegurança

💼 Consultoria

DevSecOps Engineer supporting federal agency application modernization efforts in Washington, DC. Building secure, efficient pipelines for mission-critical systems.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

Imagineeer

11 - 50

🏛️ Governo

🔒 Cibersegurança

💼 Consultoria

DevOps Engineer specializing in SharePoint and Microsoft 365 platform deployment automation. Collaborating with teams to enhance cybersecurity, IT modernization in a federal context.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $130.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 4 dias atrás

Imagineeer

11 - 50

🏛️ Governo

🔒 Cibersegurança

💼 Consultoria

DevOps Engineer building and maintaining cloud infrastructure and CI/CD pipelines. Driving automation and security operations in alignment with federal mission requirements.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $120.000 - $130.000 / ano

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório