
201 - 500 employees
🔌 API
đź’ł Fintech
₿ Crypto
API • Fintech • Crypto
Alpaca is a fintech company that provides a comprehensive brokerage and trading platform through a suite of APIs. These APIs enable developers and businesses to integrate algorithmic trading, app development, and embedded investing into their services. Alpaca offers services like trading in US stocks, ETFs, and cryptocurrency with options for local currency transactions. The company is recognized for its cyber security practices and is a member of FINRA and SIPC. Alpaca is ideal for fintech startups, broker-dealers, hedge funds, and other financial services looking to build sophisticated trading applications and platforms with minimal friction through their well-documented Broker API.
🔥 0 minutes ago
🇺🇸 United States – Remote
⏰ Full Time
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
đź‘» Ghost score 24%
Improve your chances of getting an interview by checking your resume score before you apply.

201 - 500 employees
🔌 API
đź’ł Fintech
₿ Crypto
API • Fintech • Crypto
Alpaca is a fintech company that provides a comprehensive brokerage and trading platform through a suite of APIs. These APIs enable developers and businesses to integrate algorithmic trading, app development, and embedded investing into their services. Alpaca offers services like trading in US stocks, ETFs, and cryptocurrency with options for local currency transactions. The company is recognized for its cyber security practices and is a member of FINRA and SIPC. Alpaca is ideal for fintech startups, broker-dealers, hedge funds, and other financial services looking to build sophisticated trading applications and platforms with minimal friction through their well-documented Broker API.
• Design and evolve cloud architecture on GCP, including networking, interconnects, IAM and high-availability topology, expressed as Terraform code following GitOps. • Build and own CI/CD pipelines that plan, review, test and safely apply IaC changes, including Policy-as-Code guardrails, drift detection and progressive rollout. • Advance Platform-as-a-Product by building self-serve capabilities and golden paths for engineers. • Strengthen observability across metrics, logs, traces and alerting using Prometheus, Thanos, Grafana, Loki, Tempo and Alertmanager. • Operate GKE clusters and infrastructure services, including Helm-packaged workloads, RabbitMQ, IBM MQ and data stores. • Participate in the Follow-The-Sun on-call model; triage alerts, join and declare incidents, lead debugging and escalation, and drive blameless post-mortems and follow-up actions. • Embed SRE practices such as SLIs/SLOs, error budgets and capacity planning into Core Infrastructure operations, partnering closely with SRE.
• 5+ years in a DevOps, Platform/Infrastructure, or SRE role, with a proven track record operating large-scale, high-availability, high-performance systems in production. • Deep hands-on experience designing cloud architecture on Google Cloud Platform (GCP) as the primary cloud - landing zones, networking, IAM and high-availability topology. • Strong Infrastructure-as-Code skills with Terraform, structuring large codebases across multiple environments, with GitOps as a first principle and least-privilege as a default mindset. • Proven experience building CI/CD pipelines for IaC - automated plan/apply, code review, Policy-as-Code, drift detection and safe rollout. • Significant production experience with Kubernetes (ideally GKE) and packaging/deploying workloads with Helm. • Solid cloud and L3/L4-L7 networking fundamentals (VPCs, routing, load balancing, DNS, TLS, interconnects) and comfort debugging cross-service connectivity. • Hands-on experience with a modern observability stack - Prometheus, Thanos, Grafana, Loki, Tempo and Alertmanager - across metrics, logs, traces and alerting. • Operator-level familiarity with data stores such as PostgreSQL and Message Brokers (e.g. RabbitMQ, RedPanda) - able to run and troubleshoot them in production. • A good understanding of SRE practices - SLOs/error budgets, capacity planning - and a Platform-as-a-Product mindset. • Strong grasp of incident management end to end: joining and declaring incidents, structured debugging under pressure, escalation, clear documentation, and post-mortems that drive real change. • Able and willing to take part in a Follow-The-Sun on-call rotation from APAC hours, and to work effectively in a distributed, async-first team with strong written communication. • Bonus: Policy-as-code and IaC quality tooling (OPA/Conftest, Checkov, tflint, Atlantis, or similar). • Bonus: Experience managing Terraform state, module registries and versioning at scale across many teams. • Bonus: Experience building self-serve developer platforms and internal golden paths (e.g. with Backstage, Tilt, or similar). • Bonus: Experience with the Alloy collector and incident tooling such as Rootly. • Bonus: Working proficiency in Go for automation and tooling. • Bonus: Strong Linux (Debian/Ubuntu) and container (Docker/containerd) fundamentals. • Bonus: Security and compliance experience in a regulated environment (SOC 2, secrets management, audit logging). • Bonus: Familiarity with trading, brokerage, or other regulated fintech domains, and with low-latency systems.
• Competitive Salary & Stock Options • Health Benefits • New Hire Home-Office Setup: One-time USD $500 • Monthly Stipend: USD $150 per month via a Brex Card
Apply Now🔥 9 hours ago
Cloud validation and release engineer for Bitdeer’s AI and Bitcoin mining infrastructure. Automating GPU compatibility, regression, acceptance, and regional release validation.
🇺🇸 United States – Remote
đź’µ $145k - $260k / year
đź’° Post-IPO Equity on 2023-05
⏰ Full Time
🟡 Mid-level
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 14 hours ago
Senior DevOps Engineer supporting Patterson’s cloud-based software applications and infrastructure. Automating deployments, managing Azure environments, and improving reliability, scalability, and production delivery.
🇺🇸 United States – Remote
đź’µ $109.1k - $145.4k / year
⏰ Full Time
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 15 hours ago
Salesforce DevOps Engineer automating CI/CD, releases, and governance for CVS Health’s enterprise healthcare Salesforce platform. Supporting federated engineering teams and improving deployment reliability.
🇺🇸 United States – Remote
đź’µ $83.4k - $166.9k / year
⏰ Full Time
🟡 Mid-level
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 16 hours ago
Data SRE modernizing secure, reliable cloud data platforms for GDIT’s U.S. federal courts program. Improving observability, DevSecOps, incident response, and FinOps optimization.
🇺🇸 United States – Remote
đź’µ $111.2k - $150.4k / year
⏰ Full Time
🟡 Mid-level
đźź Senior
⛑ DevOps & Site Reliability Engineer (SRE)
🔥 17 hours ago
10,000+ employees
đź’Ľ Consulting
🏥 Healthcare
📦 Logistics
Senior DevOps Engineer automating Oracle-based healthcare deployments for GDIT’s CMS program. Advancing secure DevSecOps pipelines, release orchestration, and deployment governance.