Senior Site Reliability Engineer

🔥 0 minutes ago

🇺🇸 United States – Remote

💵 $104.1k - $140k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MeridianLink

MeridianLink

501 - 1000 employees

Founded 1998

💳 Fintech

🏦 Banking

☁️ SaaS

💰 $485M Post-IPO Debt on 2021-11

Fintech • Banking • SaaS

MeridianLink is a leading provider of SaaS solutions for financial institutions, specializing in loan origination systems and digital transformation technologies. Their end-to-end platform enhances digital experiences through integration with mortgage LOS, deposit account opening solutions, and more. MeridianLink's cloud-based systems improve efficiency in loan processing and collections, data-driven decision-making, and account management. The company collaborates with partners to expand market reach and drive growth in the fintech industry. With over 25 years of experience, MeridianLink is dedicated to supporting banks, credit unions, and other financial service providers through technology and business intelligence.

📋 Description

• Administer AWS services, accounts, access, PostgreSQL, and other data stores • Own backup posture across databases, S3 buckets, and queues • Verify restores and maintain a tested disaster recovery plan • Monitor production using CloudWatch dashboards, metric alarms, log-based metrics, and Slack alerting • Lead production debugging and incident response • Build and maintain runbooks and participate in the on-call rotation • Resolve queue and dead-letter-queue failures through retry, redrive, and recovery • Refine infrastructure for deployability and scalability • Maintain infrastructure as code, retire unused infrastructure, and keep costs visible and justified • Share production-operations knowledge with the team

🎯 Requirements

• Bachelor's degree and 4–6 years of related experience or equivalent work experience • 5+ years of experience in DevOps, site reliability, or platform operations, with significant responsibility for production systems • 3+ years of hands-on experience with AWS, emphasizing serverless services including Lambda, SQS, EventBridge, CloudWatch, and S3 • Strong PostgreSQL database administration, backup and recovery, and query performance experience • Comfort administering other data stores • Proficiency in TypeScript, Python, and bash scripting • Strong understanding of Linux, DNS, TLS, Docker, GitHub Actions, and infrastructure as code using SST, Pulumi, or Terraform • Experience with production monitoring and alerting, incident response, and on-call ownership • Must be legally authorized to work in the United States; application asks whether future sponsorship will be required • Must pass comprehensive background, credit, and drug checks as part of the offer process

🏖️ Benefits

• Insurance coverage (medical, dental, vision, life, and disability) • Flexible paid time off • Paid holidays • 401(k) plan with company match • Remote work

Apply Now

Similar Jobs

🔥 1 hour ago

NVIDIA

10,000+ employees

🏥 Healthcare

🏭 Manufacturing

🤖 Artificial Intelligence

Senior SRE improving NVIDIA GeForce NOW’s reliable GPU cloud gaming infrastructure. Building observability, automation, Kubernetes, and incident-response tooling for service SLOs.

🔥 2 hours ago

SimpliGov

11 - 50

🏛️ Government

☁️ SaaS

⚡ Productivity

Senior DevOps/MLOps Engineer operating SimpliGov’s Azure AI platform for government customers. Building secure Kubernetes infrastructure, compliant inference paths, observability, releases, and cost controls.

🔥 2 hours ago

VetsEZ

201 - 500

🏥 Healthcare

💼 Consulting

📦 Logistics

Senior Backend DevOps Engineer operating AWS containerized microservices for the VA’s JLV clinical data viewer. Building CI/CD, observability, security, and disaster recovery capabilities.

🔥 3 hours ago

Group 1001

501 - 1000

💼 Consulting

🏥 Healthcare

💸 Finance

Senior Network Reliability Engineer automating insurance company network platforms at Group 1001. Applying SRE, cloud, Kubernetes, security, and observability practices to improve reliability and reduce operational toil.

🔥 6 hours ago

PathAI

501 - 1000

🏥 Healthcare

💼 Consulting

📦 Logistics

Senior/Staff SRE designing and operating secure on-premises and hybrid-cloud data centers for PathAI’s AI-powered pathology platform. Improving reliability, automation, observability, and incident response for machine-learning infrastructure.