
201 - 500 employees
đ API
đł Fintech
âż Crypto
API ⢠Fintech ⢠Crypto
Alpaca is a fintech company that provides a comprehensive brokerage and trading platform through a suite of APIs. These APIs enable developers and businesses to integrate algorithmic trading, app development, and embedded investing into their services. Alpaca offers services like trading in US stocks, ETFs, and cryptocurrency with options for local currency transactions. The company is recognized for its cyber security practices and is a member of FINRA and SIPC. Alpaca is ideal for fintech startups, broker-dealers, hedge funds, and other financial services looking to build sophisticated trading applications and platforms with minimal friction through their well-documented Broker API.
đ August 28
đşđ¸ United States â Remote
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đť Ghost score 21%
Improve your chances of getting an interview by checking your resume score before you apply.

201 - 500 employees
đ API
đł Fintech
âż Crypto
API ⢠Fintech ⢠Crypto
Alpaca is a fintech company that provides a comprehensive brokerage and trading platform through a suite of APIs. These APIs enable developers and businesses to integrate algorithmic trading, app development, and embedded investing into their services. Alpaca offers services like trading in US stocks, ETFs, and cryptocurrency with options for local currency transactions. The company is recognized for its cyber security practices and is a member of FINRA and SIPC. Alpaca is ideal for fintech startups, broker-dealers, hedge funds, and other financial services looking to build sophisticated trading applications and platforms with minimal friction through their well-documented Broker API.
⢠Design and evolve cloud architecture on GCP, including networking, interconnects, IAM and high-availability topology, expressed as Terraform code following GitOps. ⢠Build and own CI/CD pipelines that plan, review, test and safely apply IaC changes, including Policy-as-Code guardrails, drift detection and progressive rollout. ⢠Advance Platform-as-a-Product by building self-serve capabilities and golden paths for engineers. ⢠Strengthen observability across metrics, logs, traces and alerting using Prometheus, Thanos, Grafana, Loki, Tempo and Alertmanager. ⢠Operate GKE clusters and infrastructure services, including Helm-packaged workloads, RabbitMQ, IBM MQ and data stores. ⢠Participate in the Follow-The-Sun on-call model; triage alerts, join and declare incidents, lead debugging and escalation, and drive blameless post-mortems and follow-up actions. ⢠Embed SRE practices such as SLIs/SLOs, error budgets and capacity planning into Core Infrastructure operations, partnering closely with SRE.
⢠5+ years in a DevOps, Platform/Infrastructure, or SRE role, with a proven track record operating large-scale, high-availability, high-performance systems in production. ⢠Deep hands-on experience designing cloud architecture on Google Cloud Platform (GCP) as the primary cloud - landing zones, networking, IAM and high-availability topology. ⢠Strong Infrastructure-as-Code skills with Terraform, structuring large codebases across multiple environments, with GitOps as a first principle and least-privilege as a default mindset. ⢠Proven experience building CI/CD pipelines for IaC - automated plan/apply, code review, Policy-as-Code, drift detection and safe rollout. ⢠Significant production experience with Kubernetes (ideally GKE) and packaging/deploying workloads with Helm. ⢠Solid cloud and L3/L4-L7 networking fundamentals (VPCs, routing, load balancing, DNS, TLS, interconnects) and comfort debugging cross-service connectivity. ⢠Hands-on experience with a modern observability stack - Prometheus, Thanos, Grafana, Loki, Tempo and Alertmanager - across metrics, logs, traces and alerting. ⢠Operator-level familiarity with data stores such as PostgreSQL and Message Brokers (e.g. RabbitMQ, RedPanda) - able to run and troubleshoot them in production. ⢠A good understanding of SRE practices - SLOs/error budgets, capacity planning - and a Platform-as-a-Product mindset. ⢠Strong grasp of incident management end to end: joining and declaring incidents, structured debugging under pressure, escalation, clear documentation, and post-mortems that drive real change. ⢠Able and willing to take part in a Follow-The-Sun on-call rotation from APAC hours, and to work effectively in a distributed, async-first team with strong written communication. ⢠Bonus: Policy-as-code and IaC quality tooling (OPA/Conftest, Checkov, tflint, Atlantis, or similar). ⢠Bonus: Experience managing Terraform state, module registries and versioning at scale across many teams. ⢠Bonus: Experience building self-serve developer platforms and internal golden paths (e.g. with Backstage, Tilt, or similar). ⢠Bonus: Experience with the Alloy collector and incident tooling such as Rootly. ⢠Bonus: Working proficiency in Go for automation and tooling. ⢠Bonus: Strong Linux (Debian/Ubuntu) and container (Docker/containerd) fundamentals. ⢠Bonus: Security and compliance experience in a regulated environment (SOC 2, secrets management, audit logging). ⢠Bonus: Familiarity with trading, brokerage, or other regulated fintech domains, and with low-latency systems.
⢠Competitive Salary & Stock Options ⢠Health Benefits ⢠New Hire Home-Office Setup: One-time USD $500 ⢠Monthly Stipend: USD $150 per month via a Brex Card
Apply Nowđ August 27
Cloud validation and release engineer for Bitdeerâs AI and Bitcoin mining infrastructure. Automating GPU compatibility, regression, acceptance, and regional release validation.
đşđ¸ United States â Remote
đľ $145k - $260k / year
đ° Post-IPO Equity on 2023-05
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ August 27
Senior DevOps Engineer securing AWS/Azure cloud infrastructure for Koniag Government Services. Automating DevSecOps, CI/CD security, compliance, and incident response for federal customers.
đşđ¸ United States â Remote
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đ August 27
Senior DevOps Engineer automating cloud infrastructure, Kubernetes deployments, and CI/CD for Guidehouse government applications. Supporting secure, reliable delivery across development, QA, and operations.
đşđ¸ United States â Remote
đľ $115.2k - $172.8k / year
đ° Grant on 2023-02
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŚ H1B Visa Sponsor
đ August 27
DevSecOps Engineer securing Virta Healthâs cloud-native healthcare platform. Automating application security, IAM, vulnerability management, and compliance across GCP and Kubernetes.
đşđ¸ United States â Remote
đľ $179.5k - $187.9k / year
â° Full Time
đĄ Mid-level
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŚ H1B Visa Sponsor
đ August 26
Senior Site Reliability Engineer building reliable cloud infrastructure for Cross Riverâs fintech products. Driving DevOps, CI/CD, observability, incident response, and operational excellence.
đşđ¸ United States â Remote
đľ $160k - $200k / year
đ° $620M Series D on 2022-03
â° Full Time
đ Senior
â DevOps & Site Reliability Engineer (SRE)
đŚ H1B Visa Sponsor