Senior DevOps Engineer

🕒 Agosto 28

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

👻 Score fantasma 26%

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

Candidatar-se
Encontrar Vagas Remotas Similares

📊 Verifique sua pontuação de currículo para esta vaga

Melhore suas chances de conseguir uma entrevista verificando sua pontuação de currículo antes de se candidatar.

Logo of Alpaca

Alpaca

201 - 500 funcionários

🔌 API

💳 Fintech

₿ Cripto

API • Fintech • Crypto

Alpaca é uma empresa de fintech que fornece uma plataforma completa de corretagem e negociação por meio de um conjunto de APIs. Essas APIs permitem que desenvolvedores e empresas integrem trading algorítmico, desenvolvimento de apps e investimentos embutidos aos seus serviços. A Alpaca oferece serviços como negociação de ações dos EUA, ETFs e criptomoedas, com opções para transações em moeda local. A empresa é reconhecida por suas práticas de segurança cibernética e é membro da FINRA e da SIPC. A Alpaca é ideal para startups de fintech, broker-dealers, hedge funds e outras empresas de serviços financeiros que buscam criar aplicações e plataformas de negociação sofisticadas, com o mínimo de fricção, por meio de sua Broker API bem documentada.

Descrição

• Design and evolve cloud architecture on GCP, including networking, interconnects, IAM and high-availability topology, expressed as Terraform code following GitOps. • Build and own CI/CD pipelines that plan, review, test and safely apply IaC changes, including Policy-as-Code guardrails, drift detection and progressive rollout. • Advance Platform-as-a-Product by building self-serve capabilities and golden paths for engineers. • Strengthen observability across metrics, logs, traces and alerting using Prometheus, Thanos, Grafana, Loki, Tempo and Alertmanager. • Operate GKE clusters and infrastructure services, including Helm-packaged workloads, RabbitMQ, IBM MQ and data stores. • Participate in the Follow-The-Sun on-call model; triage alerts, join and declare incidents, lead debugging and escalation, and drive blameless post-mortems and follow-up actions. • Embed SRE practices such as SLIs/SLOs, error budgets and capacity planning into Core Infrastructure operations, partnering closely with SRE.

🎯 Requisitos

• 5+ years in a DevOps, Platform/Infrastructure, or SRE role, with a proven track record operating large-scale, high-availability, high-performance systems in production. • Deep hands-on experience designing cloud architecture on Google Cloud Platform (GCP) as the primary cloud - landing zones, networking, IAM and high-availability topology. • Strong Infrastructure-as-Code skills with Terraform, structuring large codebases across multiple environments, with GitOps as a first principle and least-privilege as a default mindset. • Proven experience building CI/CD pipelines for IaC - automated plan/apply, code review, Policy-as-Code, drift detection and safe rollout. • Significant production experience with Kubernetes (ideally GKE) and packaging/deploying workloads with Helm. • Solid cloud and L3/L4-L7 networking fundamentals (VPCs, routing, load balancing, DNS, TLS, interconnects) and comfort debugging cross-service connectivity. • Hands-on experience with a modern observability stack - Prometheus, Thanos, Grafana, Loki, Tempo and Alertmanager - across metrics, logs, traces and alerting. • Operator-level familiarity with data stores such as PostgreSQL and Message Brokers (e.g. RabbitMQ, RedPanda) - able to run and troubleshoot them in production. • A good understanding of SRE practices - SLOs/error budgets, capacity planning - and a Platform-as-a-Product mindset. • Strong grasp of incident management end to end: joining and declaring incidents, structured debugging under pressure, escalation, clear documentation, and post-mortems that drive real change. • Able and willing to take part in a Follow-The-Sun on-call rotation from APAC hours, and to work effectively in a distributed, async-first team with strong written communication. • Bonus: Policy-as-code and IaC quality tooling (OPA/Conftest, Checkov, tflint, Atlantis, or similar). • Bonus: Experience managing Terraform state, module registries and versioning at scale across many teams. • Bonus: Experience building self-serve developer platforms and internal golden paths (e.g. with Backstage, Tilt, or similar). • Bonus: Experience with the Alloy collector and incident tooling such as Rootly. • Bonus: Working proficiency in Go for automation and tooling. • Bonus: Strong Linux (Debian/Ubuntu) and container (Docker/containerd) fundamentals. • Bonus: Security and compliance experience in a regulated environment (SOC 2, secrets management, audit logging). • Bonus: Familiarity with trading, brokerage, or other regulated fintech domains, and with low-latency systems.

🏖️ Benefícios

• Competitive Salary & Stock Options • Health Benefits • New Hire Home-Office Setup: One-time USD $500 • Monthly Stipend: USD $150 per month via a Brex Card

Candidatar-se

Vagas Similares

🕒 Agosto 27

Bitdeer Group

201 - 500

💼 Consultoria

📦 Logística

🏗️ Construção

Cloud validation and release engineer for Bitdeer’s AI and Bitcoin mining infrastructure. Automating GPU compatibility, regression, acceptance, and regional release validation.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $145.000 - $260.000 / ano

💰 Post-IPO Equity em 2023-05

⏰ Tempo Integral

🟡 Pleno

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 27

Koniag Government Services

1001 - 5000

🏛️ Governo

🎖️ Defesa

💼 Consultoria

Senior DevOps Engineer securing AWS/Azure cloud infrastructure for Koniag Government Services. Automating DevSecOps, CI/CD security, compliance, and incident response for federal customers.

🇺🇸 Estados Unidos – Remoto (EUA)

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 27

Guidehouse

10.000+ funcionários

🏥 Saúde

🎖️ Defesa

📦 Logística

Senior DevOps Engineer automating cloud infrastructure, Kubernetes deployments, and CI/CD for Guidehouse government applications. Supporting secure, reliable delivery across development, QA, and operations.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $115.200 - $172.800 / ano

💰 Grant em 2023-02

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 27

Virta Health

201 - 500

🏥 Saúde

⚕️ Seguro de Saúde

🧘 Bem-estar

DevSecOps Engineer securing Virta Health’s cloud-native healthcare platform. Automating application security, IAM, vulnerability management, and compliance across GCP and Kubernetes.

🗣️🇺🇸🇬🇧 Inglês obrigatório

🕒 Agosto 26

Cross River

501 - 1000

🏦 Bancário

💳 Fintech

☁️ SaaS

Senior Site Reliability Engineer building reliable cloud infrastructure for Cross River’s fintech products. Driving DevOps, CI/CD, observability, incident response, and operational excellence.

🇺🇸 Estados Unidos – Remoto (EUA)

💵 $160.000 - $200.000 / ano

💰 $620.000.000 Series D em 2022-03

⏰ Tempo Integral

🟠 Sênior

⛑ DevOps & Engenheiro de Confiabilidade do Site (SRE)

🦅 Patrocina Visto H1B

infoinfo

🗣️🇺🇸🇬🇧 Inglês obrigatório