Principal Data Engineer

Job not on LinkedIn

🔥 0 minutes ago

🇵🇱 Poland – Remote

⏰ Full Time

🔴 Lead

🚰 Data Engineer

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Humaneva Group

Humaneva Group

1001 - 5000 employees

🏥 Healthcare

💊 Pharmaceuticals

🤝 B2B

💰 $50M Private Equity Round - Humaneva on 2024-09

Healthcare • Pharmaceuticals • B2B

Humaneva Group is a clinical-research-focused organization that integrates patient care, scientific innovation, and data-driven approaches to provide reliable clinical research data to the medical community. The Group comprises three interrelated companies: a research site network, a contract research organization (CRO), and a technology platform that together create a connected data stream between patients, investigators, research centers, and data standards. Humaneva emphasizes integration of clinical operations with technology and appears to support therapeutic areas such as oncology, according to its site navigation.

📋 Description

• Design, build, test, deploy, and operate production-grade pipelines, transformations, integrations, data models, and reusable platform capabilities • Lead complex data-engineering topics and resolve critical engineering problems • Assess the current data warehouse and define the target architecture • Decide which platform components should be retained, modernized, replaced, or redesigned, including Snowflake’s future role • Introduce measurable reliability standards, automated data-quality controls, end-to-end monitoring, logging, alerting, incident handling, and root-cause analysis • Evaluate and select technologies based on business fit, reliability, interoperability, security, maintainability, talent availability, and total cost of ownership • Use proofs of concept where evidence is needed • Enable governed self-service analytics through Power BI, SAS, and other approved platforms • Provide trusted and well-documented data products while maintaining security, quality, and ownership controls • Lead through delivery, design reviews, code reviews, and practical problem-solving • Mentor less-experienced engineers and raise implementation, testing, documentation, and operational standards • Collaborate with Product, Clinical Operations, Finance, Quality, Security, Legal, and business owners • Translate business needs into robust technical designs and explain significant decisions in clear business language • Optimize compute, storage, processing, and operational costs • Introduce visibility into workload behavior and cost drivers and evaluate alternatives using total cost of ownership • Establish source-system integration and batch or near-real-time ingestion • Manage data transformation, orchestration, modeling, and serving • Address metadata, cataloguing, lineage, data-product ownership, CI/CD, infrastructure automation, release controls, identity, access management, privacy, retention, auditability, and interfaces for analytics and future AI workloads • Embed privacy, security, and compliance requirements into architecture and engineering practices • Implement access controls, segregation of duties, traceability, lineage, and audit evidence • Support GDPR requirements and regulated use cases, including applicable GxP and 21 CFR Part 11 expectations • Document context, constraints, alternatives, expected benefits, risks, trade-offs, cost implications, and recommended direction for material decisions • Assess the existing DWH, pipelines, platform risks, and modernization priorities during the first year • Establish a target architecture and roadmap, including a recommendation on Snowflake’s future role • Implement observability and automated quality controls for critical data flows • Reduce pipeline failures and eliminate critical silent failures • Define ownership and measurable reliability expectations for critical data products • Introduce consistent engineering, testing, deployment, and documentation standards • Improve team technical capability through hands-on delivery and practical mentoring • Provide reliable, governed data access for Product, Clinical Operations, Finance, and other functions

🎯 Requirements

• Extensive experience designing, delivering, and operating production-grade data platforms • Strong hands-on expertise in data engineering and software engineering • Proven experience modernizing data warehouses or building new data platforms and migration paths • Advanced SQL • Strong data-modeling skills • Strong programming capability in at least one language used for production data engineering • Practical experience with automated testing, data-quality controls, monitoring, alerting, incident diagnosis, and root-cause analysis • Strong understanding of cloud data architecture, version control, CI/CD, infrastructure automation, and controlled deployment practices • Ability to evaluate vendor-specific and open technologies without being constrained by a predefined stack • Ability to lead complex work through implementation and mentor engineers • Ability to communicate technical decisions to engineering, executive, and business audiences • Experience with Snowflake architecture, engineering, performance optimization, and cost management • Experience with AWS, Azure, or multi-cloud environments • Experience enabling Power BI, SAS, or comparable self-service analytics platforms • Experience with data mesh, domain-oriented data products, or heterogeneous platform architectures • Experience in clinical research, healthcare, or another regulated industry • Familiarity with GDPR, GxP, audit-trail requirements, or 21 CFR Part 11 • Experience with Platform Engineering, Infrastructure as Code, or shared developer capabilities • Experience preparing reliable data foundations for machine learning or generative AI

🏖️ Benefits

• Employee position • Predominantly hands-on individual contributor role with substantial autonomy over data architecture, engineering standards, and technology selection • Potential path toward broader responsibility as Head of Data or an expert technical track • Potential future expansion into Platform Engineering and AI enablement

Apply Now

Similar Jobs

🕒 August 18

RS

5001 - 10000

🤝 B2B

💼 Consulting

Principal Data Architect shaping RS Group’s Snowflake enterprise data platform for industrial customers. Leading architecture, governance, transformation initiatives, and Data Architects to deliver scalable global analytics capabilities.

Cloud

ETL

🕒 July 3

OpenX

201 - 500

💼 Consulting

📣 Marketing

Big Data Engineer developing large-scale data processing systems for OpenX's cloud-based platform. Collaborating with teams to improve efficiency and scalability in processing billions of ad requests.

🇵🇱 Poland – Remote

💵 zł23.4k - zł26.1k / month

💰 Secondary Market on 2015-05

⏰ Full Time

🟠 Senior

🔴 Lead

🚰 Data Engineer

🗣️🇵🇱 Polish Required

Airflow

Apache

AWS

BigQuery

Cloud

Docker

Google Cloud Platform

Java

Kubernetes

NoSQL

Python

RDBMS

Scala

Spark

SQL

🕒 April 16

Infosys

10,000+ employees

🏢 Enterprise

💼 Consulting

🤖 Artificial Intelligence

Data Architect utilizing data management skills and strategic approach in a consulting role. Collaborating with clients to design and implement innovative data solutions.

🇵🇱 Poland – Remote

💰 $200M Post-IPO Equity on 2008-07

⏰ Full Time

🟠 Senior

🔴 Lead

🚰 Data Engineer

Airflow

Apache

AWS

Azure

Cloud

ETL

Hadoop

Kafka

MongoDB

NoSQL

Oracle

Spark

Splunk

SQL

🕒 January 21

Look4IT

11 - 50

💼 Consulting

📣 Marketing

📦 Logistics

SAP Data Migration Consultant implementing data migration for SAP projects by executing extraction and transformation strategies. Collaborating with teams to ensure accurate data migration.

🗣️🇩🇪 German Required

ETL