Senior Data Engineer, Databricks, PySpark

Job not on LinkedIn

🔥 1 hour ago

🇵🇱 Poland – Remote

💵 €38 / hour

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Optiveum

Optiveum

1 - 10 employees

📦 Logistics

📣 Marketing

🏭 Manufacturing

Logistics • Marketing • Manufacturing

Optiveum is a recruitment and consulting company based on over 20 years of experience in HR and IT services. It operates both locally in Poland and internationally, offering candidates project-based or permanent job opportunities, either remotely or on-site. Optiveum specializes in placing professionals across various fields including IT, finance, accounting, and manufacturing, ensuring that both clients and candidates achieve their professional goals.

📋 Description

• Design and implement scalable data products using Databricks, Delta Lake, and PySpark • Build and maintain data pipelines processing complex datasets from ERP, procurement, sustainability, and external systems • Optimize workloads for performance, scalability, and cost efficiency • Build reusable engineering patterns • Implement automated data quality controls throughout the data lifecycle • Identify and resolve data issues before they impact reporting • Implement CI/CD pipelines, automated deployments, testing frameworks, and Infrastructure as Code • Support platform security controls and access management • Partner with sustainability experts, business analysts, and reporting teams on requirements gathering, solution design, and production releases

🎯 Requirements

• 5–7 years of experience designing, developing, and operating data platforms and pipelines • Extensive hands-on experience with Databricks, including Azure Databricks and Databricks Workflows, in production environments • Expert-level proficiency in PySpark, Python, Spark SQL, Data Modelling, and Data Pipeline Design • Experience implementing CI/CD pipelines for data engineering workloads • Git-based development and version control • Solid understanding of data lineage, metadata management, governance, auditing, and validation frameworks • Fluent English, both written and spoken • Proactive, self-driven mindset with strong troubleshooting and root-cause analysis skills • Experience with sustainability reporting, ESG data, and regulatory reporting requirements (nice to have) • Knowledge of procurement, supplier, finance, and ERP data domains (nice to have) • Experience developing Databricks applications, dashboards, or user-facing data tools (nice to have)

🏖️ Benefits

• Fully remote work from Warsaw • Full-time schedule of 40 hours/week • B2B cooperation agreement

Apply Now

Similar Jobs

🔥 10 hours ago

ELEKS

1001 - 5000

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior Data Engineer building Azure, Databricks, and SQL data solutions for a global insurance broker. Optimizing cloud workflows, synchronization, and database performance.

Azure

Cloud

Python

Spark

SQL

Unity

🕒 4 days ago

Solvd, Inc.

501 - 1000

💼 Consulting

🏥 Healthcare

📦 Logistics

Data Engineer maintaining Solvd’s production Databricks Data Lake for AI and technology consulting services. Supporting pipelines, Python transformations, reports, incidents, and documentation.

Azure

Cloud

Python

🕒 6 days ago

XTB online investing

1001 - 5000

💳 Fintech

💸 Finance

👥 B2C

Data Engineer building reliable pipelines and Python services for XTB's tax and accounting systems. Supporting Databricks migration and financial data integration for an online trading fintech.

🇵🇱 Poland – Remote

💰 $194.1M Post-IPO Secondary - Xtb on 2025-05

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

Airflow

Apache

ETL

Flask

MS SQL Server

PySpark

Python

Spark

SQL

🕒 September 22

intive

1001 - 5000

💼 Consulting

🏥 Healthcare

📣 Marketing

Senior Data Engineer building reliable data pipelines and curated datasets for intive’s digital innovation projects. Developing dbt models, Power BI datasets, and governed cloud data platforms.

AWS

Cloud

SQL

🕒 September 21

stermedia.ai

11 - 50

🤖 Artificial Intelligence

🤝 B2B

🏭 Manufacturing

Senior Data Engineer building Databricks lakehouse platforms for pharmaceutical operations. Developing scalable pipelines, governance, and analytics-ready datasets from global facility data.

Apache

AWS

Azure

Cloud

ETL

Google Cloud Platform

PySpark

Python

Spark

SQL

Unity