Search Remote Jobs

Senior Data Engineer

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Healthfirst

Healthfirst

1001 - 5000 employees

Founded 1993

🛡️ Insurance

🏥 Healthcare

⚕️ Healthcare Insurance

Insurance • Healthcare • Healthcare Insurance

Healthfirst is a health insurance provider dedicated to helping New Yorkers access affordable health coverage for individuals and families. With over 30 years of experience, Healthfirst offers a range of plans including Medicaid managed care, Medicare Advantage, long-term care, and essential health plans. The company focuses on providing quality healthcare options, comprehensive benefits, and support to ensure members can maintain their health and well-being.

đź“‹ Description

• Design and implement ELT/ETL solutions for batch and streaming ingestion, integration, refinement, and publishing on the Lakehouse • Develop reusable data processing frameworks and configuration-driven pipelines using Python and PySpark • Build and maintain scalable orchestration workflows for production data delivery, including retries, historical loads, and operational runbooks • Implement data quality checks, validation frameworks, and monitoring to meet data contracts and SLAs • Apply DataOps practices including Git-based development, CI/CD/CT, automated testing, and controlled environment promotion • Contribute to data lifecycle practices including retention, archival, disaster recovery, and resiliency • Support platform modernization and cloud migration of legacy data flows into Lakehouse patterns • Collaborate with stakeholders to map technical designs to business processes, non-functional requirements, and consumption needs • Establish and document standards, naming conventions, and engineering practices; participate in Agile ceremonies and cross-team delivery • Mentor engineers, conduct design and code reviews, and improve reliability, performance, and cost efficiency • Own reliable datasets that enable analytics and reporting and ensure governed access across the Lakehouse

🎯 Requirements

• Bachelor's degree in Computer Science, Information Systems, Engineering, or a related technical field, or equivalent work experience • 8+ years of overall IT experience • 5+ years of hands-on experience designing and developing enterprise-scale data engineering solutions • Strong experience developing scalable data pipelines and reusable frameworks using Python and PySpark • Experience with AWS Glue, dbt, Apache Spark, or comparable technologies • Strong understanding of data warehousing, dimensional modeling, and modern data lake/Lakehouse architectures • Experience with AWS, Microsoft Azure, or Google Cloud Platform; AWS preferred • Strong SQL expertise with relational databases; NoSQL familiarity is a plus • Experience with Apache Spark, Amazon EMR, or Hadoop-based platforms • Experience using Git-based source control and Agile software development methodologies • Strong analytical, problem-solving, and communication skills • Preferred: AWS services including S3, Glue, EMR, Athena, Redshift, Lambda, Lake Formation, IAM, and CloudWatch • Preferred: Apache Iceberg or similar open table formats and governed Lakehouse patterns • Preferred: CI/CD, DataOps, continuous testing, and infrastructure automation such as Terraform • Preferred: Apache Airflow, MWAA, AWS Step Functions, or similar workflow orchestration tools • Preferred: RESTful APIs or other access layers and enterprise application integrations • Preferred: Kafka, Kinesis, or Spark Structured Streaming • Preferred: Data quality, observability, monitoring, and automated validation frameworks • Preferred: Data governance, metadata management, lineage, and enterprise data catalog solutions • Preferred: Data security, encryption, access controls, and healthcare regulatory compliance including HIPAA/PHI • Preferred: Distributed workload optimization for scalability, reliability, and cloud cost efficiency • Preferred: Mentoring engineers, conducting design and code reviews, and establishing engineering best practices

🏖️ Benefits

• Medical coverage • Dental coverage • Vision coverage • Incentive and recognition programs • Life insurance • 401k contributions • Competitive compensation and benefits package

Apply Now

Similar Jobs

🔥 2 hours ago

Cherry

201 - 500

🏥 Healthcare

🍽️ Food & Beverage

🏨 Hospitality

Senior data engineering leader owning Cherry’s BNPL data platform, governance, and analytics engineering. Building a high-performing team enabling secure, self-serve data across the medical-financing business.

🔥 5 hours ago

Manulife

10,000+ employees

🛡️ Insurance

đź’¸ Finance

Data Engineer building cloud data pipelines for Manulife’s insurance and financial services business. Developing analytics, machine learning infrastructure, databases, dashboards, and secure integrations remotely in California.

🔥 5 hours ago

AECOM

10,000+ employees

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Licensed architect leading hyperscale and colocation data center design for AECOM, a global infrastructure consulting firm. Delivering projects from concept through construction administration while coordinating multidisciplinary teams and clients.

🔥 5 hours ago

AECOM

10,000+ employees

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Lead architect designing hyperscale cloud and colocation data centers for AECOM, a global infrastructure consulting firm. Guiding integrated design, construction administration, compliance, and client delivery across complex mission-critical projects.

🔥 5 hours ago

AECOM

10,000+ employees

đź’Ľ Consulting

🏥 Healthcare

📦 Logistics

Licensed architect leading hyperscale cloud and colocation data center design for AECOM, a global infrastructure consulting firm. Guiding projects from concept through construction closeout.