Senior Database Reliability Engineer

🔥 14 hours ago

🇺🇸 United States – Remote

💵 $220k - $240k / year

⏰ Full Time

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

👻 Ghost score 17%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Prompt Therapy Solutions Inc

Prompt Therapy Solutions Inc

11 - 50 employees

🏥 Healthcare

💼 Consulting

📦 Logistics

Healthcare • Consulting • Logistics

Prompt Therapy Solutions Inc. is a company that provides therapy-focused software solutions designed to enhance the efficiency and profitability of clinics. They offer a range of features including AI-powered scheduling, patient management, billing, and reporting tools. Their platform caters to various practice sizes, from startup clinics to large enterprises, aiming to optimize clinic operations and improve patient satisfaction. The company's AI solutions help streamline workflows for roles such as front desk staff, therapists, billers, and clinic owners. Through advanced analytics and automated processes, Prompt Therapy Solutions seeks to modernize practice management and drive growth for therapy providers.

📋 Description

• Own the reliability, performance, and availability of production Aurora MySQL databases on AWS across EHR, Data, and AI teams • Manage database alerting end to end and maintain the observability pipeline feeding metrics and alerts into Datadog • Integrate after-hours database alerts into the DevOps on-call rotation • Educate and train engineering teams on database hygiene and high-performing queries • Implement safeguards such as automatic termination of long-running or runaway queries • Provide fast, always-available visibility into queries executing on any database instance • Author and maintain database runbooks for rapid, consistent incident response • Diagnose, optimize, and re-index problematic queries to reduce latency and scale-out load • Review EMR and application queries before release as a performance gate • Architect and right-size Aurora reader topology and read-routing strategy • Own database cost efficiency through instance right-sizing, reserved-capacity/Savings Plan strategy, and storage tiering • Manage replication health, backups, restore testing, failover, and disaster recovery • Establish safe schema-change and migration practices across teams • Partner with the Platform Architect on database workload architecture • Handle protected health information responsibly in a HIPAA-compliant environment

🎯 Requirements

• 6+ years in database reliability engineering, database administration, or database engineering, including ownership of production systems at scale • Deep MySQL expertise, including query optimization, execution-plan analysis, indexing strategy, and replication • Aurora MySQL experience strongly preferred • Hands-on experience operating large, multi-reader Aurora or RDS clusters at terabyte scale, including read-routing and connection management • Strong command of database observability tooling, including Datadog, Percona Monitoring and Management (PMM), Performance Insights, Prometheus/Grafana • Proficiency in Python, Bash, or a similar scripting language for automation • Comfort with infrastructure-as-code, particularly Terraform • Solid AWS operational depth around RDS/Aurora, including cost optimization, instance sizing, reserved capacity, and storage types • Experience with on-call/incident response and authoring runbooks • Excellent communication skills and ability to collaborate across engineering teams • Experience with a cloud data warehouse such as Snowflake, Databricks, or Redshift preferred • Experience in a HIPAA-regulated or other compliance-heavy environment handling sensitive data preferred • Familiarity with query auto-remediation tooling such as pt-kill, Percona Toolkit, statement timeouts, or custom killers preferred • Application-side familiarity with PHP preferred • Knowledge of security best practices for database and cloud environments preferred • Previous experience with agile methodologies and fast-paced environments preferred • Must be legally authorized to work in the US • Must be located in the US

🏖️ Benefits

• Potential equity compensation for outstanding performance • Flexible PTO • Company-wide sponsored lunches • Company paid disability and life insurance benefits • Company paid family and medical leave • Medical, dental, and vision insurance benefits • Discounted pet insurance • FSA/DCA and commuter benefits • 401k • Credits for online fitness classes/gym memberships • Recovery suite at HQ — includes a cold plunge, sauna, and shower • Remote/hybrid environment • Positive impact helping outpatient rehab organizations treat more patients and deliver better care

Apply Now

Similar Jobs

🔥 15 hours ago

ACT

501 - 1000

📚 Education

🤝 Non-profit

Senior DevOps Engineer building secure AWS, Kubernetes, and GitOps delivery platforms for ACT’s education mission. Improving automation, resiliency, observability, and cloud cost management.

🔥 15 hours ago

ACT

501 - 1000

📚 Education

🤝 Non-profit

Senior DevOps Engineer building secure AWS, Kubernetes, and GitOps delivery platforms for ACT’s education assessment mission. Improving automation, resiliency, observability, and cloud cost management.

🔥 15 hours ago

24-MAG

2 - 10

🤝 B2B

💼 Consulting

DevOps Engineer evaluating Kubernetes, AWS, Terraform, and CI/CD infrastructure. Contributing expert technical assessments and solutions to advanced AI infrastructure workstreams.

🔥 17 hours ago

Kong Inc.

201 - 500

💼 Consulting

📦 Logistics

🔌 API

Site Reliability Engineer scaling Kong Konnect, an API and AI connectivity SaaS platform. Automating Kubernetes infrastructure, observability, and multi-region reliability across major clouds.

🇺🇸 United States – Remote

💵 $123k - $150k / year

💰 $100M Series D on 2021-02

⏰ Full Time

🟡 Mid-level

🟠 Senior

⛑ DevOps & Site Reliability Engineer (SRE)

🔥 18 hours ago

CACI International Inc

10,000+ employees

💼 Consulting

🎖️ Defense

AWS DevSecOps Engineer modernizing CACI’s federal Grant Solutions cloud platform. Automating secure infrastructure, CI/CD, observability, and compliant software delivery.