Senior Data Engineer – Databricks Expert

Job not on LinkedIn

🔥 38 minutes ago

🇵🇱 Poland – Remote

💵 zł20k - zł30k / month

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 4%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of stermedia.ai

stermedia.ai

11 - 50 employees

Founded 2009

🤖 Artificial Intelligence

🤝 B2B

🏭 Manufacturing

Artificial Intelligence • B2B • Manufacturing

Stermedia. ai is a B2B AI and software engineering firm that builds custom artificial intelligence, robotic automation, cloud-based manufacturing software and product design solutions. Since 2009 it has provided digital transformation services including deep learning and data science, robotic/process automation, custom web and mobile application development, UX/product design, and cloud architecture (AWS/IBM). The company shows domain experience across manufacturing and automotive (AndonCloud manufacturing software and industry 4. 0 case studies), healthcare (medical imaging AI), HR/recruitment (CV/job-matching), media and public sector projects, and highlights competition wins and multiple client case studies demonstrating technical expertise.

📋 Description

• Design, develop, and maintain data pipelines in Databricks using Python, PySpark, and SQL • Build and maintain Delta Lake tables and transformation layers following bronze, silver, and gold architecture • Integrate data from multiple source systems into consistent, analytics-ready datasets • Implement incremental processing, change data capture, and schema evolution where required • Develop and orchestrate production workflows, including dependencies, scheduling, retries, and monitoring • Optimize Spark workloads, SQL queries, and compute usage to improve performance and manage costs • Implement automated data quality checks, validation, and pipeline tests • Support data access management, governance, and lineage using Unity Catalog • Maintain reusable code, technical documentation, and Git-based development workflows • Collaborate with business stakeholders, data engineers, analysts, and machine learning specialists to translate requirements into reliable data solutions • Contribute to the full delivery lifecycle, from client discussions and business requirements analysis through solution design, implementation, and production deployment

🎯 Requirements

• Strong commercial experience with Databricks, including developing and operating production data pipelines • Advanced knowledge of Python, PySpark, and SQL • Hands-on experience with Apache Spark, distributed data processing, and performance troubleshooting • Practical experience with Delta Lake, including incremental loads, merge operations, and schema management • Strong understanding of lakehouse architecture, ETL/ELT patterns, and data modeling • Experience orchestrating and monitoring workflows in Databricks • Practical knowledge of Unity Catalog, including permissions, data organization, and lineage • Experience optimizing pipeline performance and compute resource usage • Familiarity with cloud storage and services in at least one major cloud environment: Azure, AWS, or GCP • Experience with Git, code reviews, automated testing, and CI/CD workflows • Strong analytical and communication skills, with the ability to work directly with international stakeholders • Good command of English

🏖️ Benefits

• Opportunities to work with modern data engineering and machine learning technologies • An annual self-development budget • The opportunity to contribute to a variety of interesting projects • Internal workshops and knowledge-sharing sessions • Support for personal branding through articles, conference talks, and leading internal workshops • Flexible working hours • The possibility of remote work • A chillout room, free beverages, and team and company events • A friendly atmosphere • MultiSport • LuxMed

Apply Now

Similar Jobs

🔥 7 hours ago

Commit

501 - 1000

🔒 Cybersecurity

Senior Data Engineer designing scalable GCP data solutions for customer environments. Building pipelines, warehouses, lakehouses, and analytics platforms across the full data lifecycle.

Amazon Redshift

AWS

Azure

BigQuery

Cloud

ETL

Google Cloud Platform

Kafka

Python

SQL

Tableau

🕒 3 days ago

Nord Security

1001 - 5000

🔒 Cybersecurity

☁️ SaaS

🤝 B2B

Senior Data Engineer building scalable pipelines and data systems for NordVPN’s cybersecurity products. Supporting privacy protection through PySpark, AWS, Kafka, and Databricks.

🇵🇱 Poland – Remote

💵 PLN23k - PLN30k / month

💰 $100M Private Equity Round - Nord Security on 2023-09

⏰ Full Time

🟠 Senior

🚰 Data Engineer

AWS

Cassandra

Cloud

Cyber Security

ETL

Java

Kafka

Postgres

PySpark

Python

Scala

SQL

Terraform

Go

🕒 4 days ago

G-P

1001 - 5000

💼 Consulting

🏥 Healthcare

⚖️ Legal

Senior Data Engineer architecting G-P’s AI-native Databricks platform. Building scalable data frameworks and streaming systems for G-P’s global employment SaaS platform.

Cloud

Kafka

Spark

🕒 5 days ago

ABB

10,000+ employees

💼 Consulting

📦 Logistics

🚗 Transport

Enterprise Domain Architect shaping ABB’s enterprise data architecture. Enabling governed, scalable data ecosystems for analytics, AI, and business innovation.

🇵🇱 Poland – Remote

💰 $545.9M Post-IPO Debt - ABB on 2023-11

⏰ Full Time

🟠 Senior

🔴 Lead

🚰 Data Engineer

Cloud

🕒 September 14

Atos

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Azure Data Engineer building Microsoft Azure, Fabric, and Synapse data solutions for Atos's Dutch insurance client. Mostly remote from Poland with annual Netherlands business trips.

Azure

ETL