Senior Data Engineer, Databricks Migration

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Sigma Software Group

Sigma Software Group

1001 - 5000 employees

Founded 2002

💼 Consulting

🏥 Healthcare

🚘 Automotive

Consulting • Healthcare • Automotive

Sigma Software Group is a multinational company, established in 2002, that specializes in providing high-quality software development, graphic design, testing, and support services. The company focuses on delivering solutions across various industries such as automotive, telecommunications, aviation, advertising, gaming, banking, real estate, and healthcare. Sigma Software values professional growth, offers remote work opportunities worldwide, and caters to world-renowned clients like AstraZeneca, Scania, and SAS. The company emphasizes a culture of continuous education, mentorship, and flexible work environments, making it a preferred workplace for IT specialists aiming to work on complex solutions utilizing cutting-edge technologies. Sigma Software is committed to innovative solutions and engineering the future while also contributing to social causes such as charitable work in Ukraine.

📋 Description

• Participate in the migration of a large-scale analytical platform from BigQuery to Databricks • Design and implement scalable Lakehouse architectures using Databricks and Delta Lake • Analyze existing ETL / ELT workloads and define migration approaches • Develop and optimize data pipelines processing large volumes of retail and analytical data • Implement incremental processing strategies and scalable transformation frameworks • Build and maintain Spark-based data processing solutions using PySpark • Design and maintain Bronze, Silver, and Gold medallion architecture layers • Implement data governance and security best practices using Unity Catalog • Collaborate with Data Science, Analytics, Product, and Customer Engineering teams • Participate in architecture discussions and technical solution design • Develop reusable data platform components and engineering standards • Conduct code reviews and contribute to platform reliability and maintainability • Troubleshoot and optimize complex SQL and Spark workloads • Support production deployments and platform modernization activities

🎯 Requirements

• 5+ years of professional experience as a Data Engineer • Strong programming skills in Python and advanced SQL • Hands-on commercial experience with Databricks • Strong knowledge of Apache Spark, primarily PySpark • Experience designing and building modern cloud-based data platforms • Experience developing ETL / ELT pipelines and large-scale data processing solutions • Hands-on experience with Delta Lake • Experience with Spark Declarative Pipelines • Experience with cluster monitoring, metrics analysis, and performance optimization • Strong understanding of distributed data processing architectures • Solid understanding of data warehousing concepts and dimensional modeling • Experience with Airflow or similar orchestration tools • Experience optimizing complex analytical SQL workloads • Experience implementing CI / CD practices for data engineering platforms • Strong troubleshooting and performance optimization skills • Ability to work collaboratively in cross-functional international teams • Upper-Intermediate or higher English level • Experience with GCP cloud services, AWS, or Azure is a plus • Experience in retail analytics or pricing optimization domains is a plus • Experience supporting machine learning or AI-related data workloads is a plus • Experience with platform modernization and cloud migration initiatives is a plus • Strong analytical and problem-solving mindset • Proactive and ownership-driven approach • Ability to work independently and collaboratively • Good communication and stakeholder collaboration skills • Passion for scalable data engineering and modern data platforms • Interest in continuous learning and technology innovation

🏖️ Benefits

• Remote work • Opportunities for continuous learning • Technology growth opportunities • Meaningful engineering impact • Work on complex international projects

Apply Now

Similar Jobs

🕒 5 days ago

Expleo Group

10,000+ employees

💼 Consulting

🎖️ Defense

📦 Logistics

Data Platform Engineer building scalable cloud data platforms for Expleo, a global engineering, technology, and consulting provider. Designing production data pipelines and supporting reliable, compliant data solutions with AWS and distributed processing.

🗣️🇧🇷🇵🇹 Portuguese Required

Airflow

Apache

AWS

Cloud

Grafana

Kafka

Python

Scala

Spark

SQL

🕒 July 27

HumanIT Digital Consulting

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

CDP Specialist managing customer data platform for Switzerland's largest telecommunications operator. Working fully remotely from Portugal while collaborating with various teams on integrations.

Cloud

SQL

🕒 July 27

Construo

11 - 50

🏗️ Construction

🏪 Marketplace

☁️ SaaS

Senior Data Engineer focused on data modeling and ETL for a financial multinational client. Responsibilities include pipeline design, algorithm development, and data analysis in a remote setting.

Azure

Cloud

ETL

Java

PySpark

Python

🕒 July 11

SIXT

5001 - 10000

📦 Logistics

🚘 Automotive

💼 Consulting

Senior Data Engineer exploring and implementing AWS and big data technologies for SIXT. Collaborating with team members to enhance the SIXT Data Shop through innovative data solutions.

Airflow

Amazon Redshift

Apache

AWS

Cloud

Python

SQL

🕒 June 25

SIXT

5001 - 10000

📦 Logistics

🚘 Automotive

💼 Consulting

Senior Data Engineer managing AI & Agentic Pipelines for SIXT's Data Platform, leveraging AWS and big data technologies. Collaborating with data teams to optimize analytical solutions and enhance data capabilities.

Airflow

Amazon Redshift

Apache

AWS

Cloud

Distributed Systems

Python

SQL