Data Engineer, Spark

🕒 March 31

🇵🇱 Poland – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 49%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Addepto

Addepto

51 - 200 employees

💼 Consulting

📦 Logistics

🏥 Healthcare

Consulting • Logistics • Healthcare

Addepto is a company specializing in Artificial Intelligence and Machine Learning solutions tailored for various industries. They offer AI consulting, MLOps consulting, and data engineering services, providing comprehensive AI-driven strategies and applications. Addepto focuses on implementing sophisticated AI technologies such as Generative AI, computer vision, and large language models to solve complex business problems. They cater to industries like finance, aviation, retail, and logistics, among others, enhancing operations and decision-making processes.

📋 Description

• Develop and maintain a high-performance data processing platform for automotive data, ensuring scalability and reliability. • Design and implement data pipelines that process large volumes of data in both streaming and batch modes. • Optimize data workflows to ensure efficient data ingestion, processing, and storage using technologies such as Spark, Cloudera, and Airflow. • Work with data lake technologies (e.g., Iceberg) to manage structured and unstructured data efficiently. • Collaborate with cross-functional teams to understand data requirements and ensure seamless integration of data sources. • Monitor and troubleshoot the platform, ensuring high availability, performance, and accuracy of data processing. • Leverage cloud services (AWS) for infrastructure management and scaling of processing workloads. • Write and maintain high-quality Python (or Java/Scala) code for data processing tasks and automation.

🎯 Requirements

• At least 4 years of commercial experience implementing, developing, or maintaining Big Data systems, data governance and data management processes. • Strong programming skills in Python (or Java/Scala): writing clean code, OOP design. • Hands-on with Big Data technologies like Spark, Cloudera, Kafka, Data Platform, Airflow, NiFi, Docker, and Iceberg. • Excellent understanding of dimensional data and data modeling techniques. • Experience implementing and deploying solutions in cloud environments. • Consulting experience with excellent communication and client management skills, including prior experience directly interacting with clients as a consultant. • Ability to work independently and take ownership of project deliverables. • Fluent English (at least C1 level). • Bachelor’s degree in technical or mathematical studies. • Nice to have: Experience with an MLOps framework such as Kubeflow or MLFlow. Familiarity with Databricks and/or dbt.

🏖️ Benefits

• Work in a supportive team of passionate enthusiasts of AI & Big Data. • Engage with top-tier global enterprises and cutting-edge startups on international projects. • Enjoy flexible work arrangements, allowing you to work remotely or from modern offices and coworking spaces. • Accelerate your professional growth through career paths, knowledge-sharing initiatives, language classes, and sponsored training or conferences, including a partnership with Databricks, which offers industry-leading training materials and certifications. • Choose your preferred form of cooperation: B2B or a contract of mandate, and make use of 20 fully paid days off. • Participate in team-building events and utilize the integration budget. • Celebrate work anniversaries, birthdays, and milestones. • Access medical and sports packages, eye care, and well-being support services, including psychotherapy and coaching. • Get full work equipment for optimal productivity, including a laptop and other necessary devices. • Experience a smooth onboarding with a dedicated buddy, and start your journey in our friendly, supportive, and autonomous culture.

Apply Now

Similar Jobs

🕒 March 25

Inetum

10,000+ employees

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Data Engineer at Inetum Polska focusing on GenAI-based mail processing. Responsible for designing, implementing, and optimizing integration workflows ensuring platform reliability and performance.

Django

Docker

Grafana

Kubernetes

Microservices

Postgres

Prometheus

Python

🕒 March 19

airSlate

501 - 1000

💼 Consulting

📦 Logistics

📣 Marketing

Data Engineer II designing and maintaining scalable data pipelines in AWS for airSlate’s SaaS solutions. Collaborating on advanced data analytics and machine learning initiatives.

AWS

Cloud

ETL

SQL

🕒 February 13

Swing Development (SwingDev)

11 - 50

💼 Consulting

🛡️ Insurance

🏥 Healthcare

Senior Data Architect for InsurTech company designing AI-ready data architecture. Leading the establishment of data modeling foundations in a greenfield environment.

🗣️🇵🇱 Polish Required

Amazon Redshift

AWS

Azure

BigQuery

Cloud

Google Cloud Platform

🕒 January 23

Technopride

1 - 10

💼 Consulting

📣 Marketing

📦 Logistics

Sr. Data Engineer supporting Data Ingestion team for an IT client based in Poland. Involves designing and developing data processing solutions and collaborating across teams.

Distributed Systems

🕒 January 21

Look4IT

11 - 50

💼 Consulting

📣 Marketing

📦 Logistics

SAP Data Migration Consultant implementing data migration for SAP projects by executing extraction and transformation strategies. Collaborating with teams to ensure accurate data migration.

🗣️🇩🇪 German Required

ETL