Lead Data Engineer – PySpark, Palantir Foundry

🔥 50 minutes ago

☕ Washington – Remote

infoinfo

💵 $156.3k - $175.2k / year

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Logic20/20, Inc.

Logic20/20, Inc.

201 - 500 employees

Founded 2005

🏥 Healthcare

📦 Logistics

📣 Marketing

Healthcare • Logistics • Marketing

Logic20/20, Inc. is a high-value business and technology consulting firm that helps organizations across utilities, energy & utilities, financial services, government, healthcare, life sciences, and technology & communications modernize and improve performance through digital strategy and transformation. The firm provides services including product, project and change management, compliance, solution assessment and planning, agile transformation, and comprehensive AI & analytics capabilities—covering data strategy, data platforms and engineering, and data visualization—to embed AI and data-driven decisioning into operations, capital planning, and compliance.

📋 Description

• Deliver client value and ensure high client satisfaction • Design, enhance, and maintain production-grade data pipelines supporting model output aggregation and downstream risk analysis • Establish and mature repository governance practices, including branching strategies, pull request standards, merge policies, release tagging, and version control workflows • Own release engineering practices for reproducible, traceable, and auditable production releases • Refactor and improve pipeline code for modularity, maintainability, scalability, and documentation quality • Develop configuration-driven pipeline patterns across environments and releases • Support testing, validation, benchmarking, and change management for critical data pipelines • Partner with data scientists, machine learning engineers, data engineers, product stakeholders, and other technical teams • Align teams on schemas, interfaces, inputs, and delivery expectations • Translate complex technical concepts into clear updates for technical and non-technical stakeholders • Improve structure and governance in codebases, repositories, and engineering workflows • Contribute to engineering best practices in a regulated, audit-sensitive delivery environment

🎯 Requirements

• 10-15+ years of data engineering, data science, machine learning engineering, and/or relevant experience using Python • Experience leading technical teams and overseeing enterprise-scale data initiatives • Strong expertise in PySpark, SQL, and cloud services • Ability to improve, refactor, or stabilize existing codebases and pipeline environments • Experience with cloud-optimized datasets, efficient partitioning strategies, and large-scale spatial operations • Understanding of machine learning model outputs flowing into downstream data pipelines, platforms, or production systems • Experience in highly regulated industries such as utilities, financial services, healthcare, insurance, or similar environments • Experience designing maintainable, scalable, and well-documented cloud-based data infrastructure or modern data platform environments • Experience supporting reproducibility, dataset versioning, release traceability, and audit readiness • Ability to define expected inputs, outputs, schemas, and interfaces across technical teams • Strong communication skills • Detail-oriented, governance-minded approach to engineering • Practical experience with Git-based workflows, code reviews, branching strategies, and release management • Experience with Palantir Foundry is highly preferred • Experience with GIS technologies and geospatial data platforms

🏖️ Benefits

• Competitive base salary • Performance-based bonuses • Other incentives • Training and mentorship opportunities • Project opportunities for career development • Supportive, globally connected work environment

Apply Now

Similar Jobs

🔥 2 hours ago

CVS Health

10,000+ employees

🏥 Healthcare

⚕️ Healthcare Insurance

🛒 Retail

Senior Manager owning CVS Health’s Medicare revenue data engineering foundation. Building auditable healthcare data pipelines for revenue forecasting, risk adjustment, Finance, and Actuarial teams.

🔥 3 hours ago

Humana

10,000+ employees

🏥 Healthcare

🛡️ Insurance

⚕️ Healthcare Insurance

Senior Data Engineer building scalable Azure data pipelines for Humana, a U.S. healthcare company. Transforming large datasets for analytics, AI, and member-focused services.

🔥 3 hours ago

Capital One

10,000+ employees

🏦 Banking

💳 Fintech

💸 Finance

Senior Staff Data Engineer leading Capital One’s enterprise data pipelines and cloud architecture. Driving scalable data platforms, governance, and innovation across banking technology.

🔥 5 hours ago

Republic Services

10,000+ employees

💼 Consulting

📦 Logistics

Data Engineer III building AWS data lakes, pipelines, and analytics systems for Republic Services’ environmental services business. Developing MarkLogic APIs, Snowflake warehouses, and streaming ingestion architectures.

🔥 8 hours ago

NeuraFlash

201 - 500

💼 Consulting

🏭 Manufacturing

📣 Marketing

Salesforce Data Architect migrating CPQ, Billing, and Quote-to-Cash data for NeuraFlash, an AI and Salesforce consultancy. Designing scalable architectures and migration frameworks for client revenue transformations.