Senior Data Engineer

🕒 il y a 27 jours

🇵🇱 Pologne – Télétravail

⏰ Temps Plein

🟠 Senior

🚰 Ingénieur Data

👻 Score fantôme 11%

infoinfo

🗣️🇺🇸🇬🇧 Anglais requis

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of Sigma Software Group

Sigma Software Group

1001 - 5000 employés

Fondée en 2002

💼 Conseil

🏥 Santé

🚘 Automobile

Consulting • Healthcare • Automotive

Sigma Software Group est une entreprise multinationale, fondée en 2002, qui se spécialise dans le développement de logiciels de haute qualité, le design graphique, les tests et les services de support. L'entreprise se concentre sur la fourniture de solutions dans divers secteurs tels que l'automobile, les télécommunications, l'aviation, la publicité, le jeu vidéo, la banque, l'immobilier et la santé. Sigma Software valorise la croissance professionnelle, offre des opportunités de travail à distance dans le monde entier et travaille pour des clients renommés tels qu'AstraZeneca, Scania et SAS. L'entreprise met l'accent sur une culture d'éducation continue, de mentorat et d'environnements de travail flexibles, ce qui en fait un lieu de travail privilégié pour les spécialistes IT visant à travailler sur des solutions complexes utilisant des technologies de pointe. Sigma Software s'engage à proposer des solutions innovantes et à concevoir l'avenir tout en contribuant à des causes sociales comme le travail caritatif en Ukraine.

Description

• Design and build scalable, cloud-native data platforms from greenfield to production • Implement near-real-time ingestion pipelines using event-driven patterns • Define and enforce platform standards, including Data Lake / Lakehouse principles, medallion architecture, and data contracts • Refactor and optimise existing Spark and PySpark scripts for performance and maintainability • Introduce best practices for code quality, testing, and CI/CD across data pipelines • Drive adoption of AI tooling and agentic workflows within the data engineering team • Ensure data quality, observability, and reliability across all pipelines and platforms • Develop self-service tooling and microservices to simplify platform usage for other teams • Collaborate with Machine Learning, Data Science, and Product teams • Lead greenfield initiatives, cloud migrations, and R&D around agentic AI architectures, event-driven systems, and LLM-ready data pipelines

🎯 Exigences

• 5+ years of professional experience in Data Engineering • Strong Python and SQL development skills for pipeline development and optimisation • Proficiency in Apache Spark / PySpark, including query optimisation and performance tuning • Hands-on experience with Databricks (preferred) or Snowflake • Experience with at least one major cloud provider: Azure (preferred), AWS, or GCP • Experience with stream processing technologies (Kafka, Spark Structured Streaming) • Solid understanding of ETL/ELT patterns, data modelling (dimensional, Data Vault), and data warehousing • Experience with orchestration tools (Apache Airflow, Azure Data Factory, or equivalent) • Knowledge of Infrastructure as Code (Terraform or equivalent) • Understanding of production-grade system requirements: reliability, scalability, observability, and performance • Upper-Intermediate English level • Familiarity with RAG pipeline design and LLM integration patterns • Knowledge of data governance frameworks and tools (Unity Catalog, Apache Atlas, or similar) • Experience with dbt for data transformation and modelling • Familiarity with MLflow, Feature Stores, or ML platform integration • Self-driven and proactive in identifying improvements • Comfortable working in a fast-paced, innovative environment • Strong problem-solving mindset with attention to detail • Open to experimenting with emerging technologies and approaches

🏖️ Avantages

• Remote work option • Full-time employment

Postuler Maintenant

Emplois Similaires

🕒 il y a 1 mois

Sowelo Consulting sp. z o.o. sp. k.

11 - 50

💼 Conseil

📦 Logistique

📣 Marketing

Senior Data Engineer building Databricks pipelines for an AI and data solutions consultancy. Delivering cloud data solutions powering analytics, enterprise AI, MLOps, and Generative AI initiatives.

🇵🇱 Pologne – Télétravail

⏰ Temps Plein

🟠 Senior

🚰 Ingénieur Data

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 1 mois

Inetum

10 000+ employés

💼 Conseil

🏥 Santé

🛡️ Assurance

Data Engineer developing scalable Azure, Databricks, and Spark data solutions for Inetum Polska’s digital transformation clients. Building pipelines and ensuring data quality, performance, and security.

🇵🇱 Pologne – Télétravail

💰 Post-IPO Equity en 2007-03

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

🚰 Ingénieur Data

🗣️🇵🇱 Polonais requis

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 1 mois

InPost Group

10 000+ employés

🛍️ eCommerce

🚗 Transport

📦 Logistique

Data Engineer improving HR data quality and system integrations at InPost. Collaborating with teams across multiple countries while leveraging Databricks and Python for scalable solutions.

🇵🇱 Pologne – Télétravail

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

🚰 Ingénieur Data

🗣️🇵🇱 Polonais requis

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 1 mois

Intetics

501 - 1000

💼 Conseil

🏥 Santé

📦 Logistique

Senior Data Engineer owning Databricks pipelines for predictive data platform. Shaping AI-driven CRM technology with intelligent data solutions in a global environment.

🇵🇱 Pologne – Télétravail

⏰ Temps Plein

🟠 Senior

🚰 Ingénieur Data

🗣️🇺🇸🇬🇧 Anglais requis

Apache

AWS

Azure

ERP

ETL

Informatica

Postgres

PySpark

Python

Spark

SQL

SSIS

🕒 il y a 1 mois

Vecten

51 - 200

🤖 Intelligence artificielle

🏥 Santé

💸 Finance

Senior Data Engineer joining AI-native company to implement data systems for investment decision-making. Collaborating on data integration and building data pipelines for machine learning.

🇵🇱 Pologne – Télétravail

⏰ Temps Plein

🟠 Senior

🚰 Ingénieur Data

🗣️🇺🇸🇬🇧 Anglais requis