Senior Infrastructure Engineer, SRE

🔥 14 hours ago

🏄 California, District of Columbia, +1 more states – Remote

infoinfo

💵 $150k - $185k / year

⏰ Full Time

🟠 Senior

👷 Infrastructure Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Rocket Money (formerly Truebill)

Rocket Money (formerly Truebill)

51 - 200 employees

💸 Finance

💳 Fintech

👥 B2C

Finance • Fintech • B2C

Rocket Money is a comprehensive personal finance management app designed to help users save more, spend less, and take control of their financial lives. The app offers features such as subscription management, spending insights, bill negotiation, net worth tracking, budgeting, and credit score monitoring. With its automated services, Rocket Money aims to simplify money management by finding and tracking subscriptions, helping to negotiate lower bills, and automating savings processes. Users can link all of their financial accounts in one place to get a full financial picture and set financial goals. The app also provides access to financial experts who can assist with personal finance queries.

📋 Description

• Build and improve the reliability and resiliency of systems and services • Establish SLIs, SLOs, and error budgets for critical services and user journeys • Review SLIs, SLOs, and error budgets with owning teams • Own and evolve the disaster recovery strategy, including recovery objectives, failover and restore paths, and regular exercises • Partner with product engineering teams to help them own and operate their services • Evolve observability platforms and standards across metrics, tracing, and logs • Improve instrumentation paved roads, alert quality, and observability cost • Strengthen incident practices by tuning paging thresholds, maintaining runbooks, and following up on postmortem actions • Contribute to Cloud Infrastructure work, including infrastructure build-outs and platform backlog • Participate in a shared on-call rotation one week out of every six weeks • Partner with engineering and internal support teams • Lead or contribute to reliability and observability modernization, internal tooling, game days, chaos experiments, DR exercises, and observability cost optimization

🎯 Requirements

• 5+ years of hands-on cloud or infrastructure engineering experience, with substantial time spent on reliability and production operations at scale • Defined SLIs and SLOs for real production services • Hands-on experience with an observability platform in production; Datadog strongly preferred • Comfortable writing code in Python, Go, TypeScript, or similar • Production Terraform experience • Comfortable working in AWS • Built or operated a disaster recovery plan, including recovery goals, failover and restore steps, and drills • On-call experience for services helped build • Knowledge of effective alerting practices

🏖️ Benefits

• Health, Dental & Vision Plans • Competitive Pay • 401k Matching • Unlimited PTO • Lunch daily (in-office only) • Snacks & Coffee (in-office only) • Commuter benefits (in-office only) • Bonus

Apply Now

Similar Jobs

🔥 22 hours ago

Precise Software Solutions, Inc.

51 - 200

🏛️ Government

🤖 Artificial Intelligence

🤝 B2B

AWS infrastructure engineer designing secure, automated cloud platforms for federal modernization programs. Supporting resilient infrastructure, DevSecOps, compliance, and operations for government organizations.

🕒 Yesterday

Coinbase

1001 - 5000

💼 Consulting

₿ Crypto

💸 Finance

Senior Infrastructure Engineer building Coinbase’s institutional crypto exchange infrastructure. Operating low-latency trading systems and ensuring reliability for high-volume financial data.

🕒 Yesterday

Nava

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Azure Infrastructure Engineer designing secure, governed cloud platforms. Nava PBC helps government agencies modernize digital services and infrastructure.

🕒 Yesterday

OpenTeams

11 - 50

💼 Consulting

📣 Marketing

☁️ SaaS

Senior Infrastructure Engineer building secure Kubernetes platforms for OpenTeams’ government AI test and evaluation teams. Operating GPU workloads, infrastructure automation, observability, and compliance in restricted environments.

🇺🇸 United States – Remote

💵 $145k - $250k / year

💰 $100k Pre Seed Round on 2019-08

⏰ Full Time

🟠 Senior

👷 Infrastructure Engineer

🕒 Yesterday

ElevenLabs

1 - 10

🤖 Artificial Intelligence

📱 Media

HPC Infrastructure Engineer operating NVIDIA GPU clusters for ElevenLabs’ AI voice and media platforms. Automating provisioning, scheduling, storage, networking, and cluster reliability.