Data Engineer

🔥 19 hours ago

🇵🇱 Poland – Remote

💵 zł16.5k - zł28k / month

⏳ Contract/Temporary

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 0%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of CodiLime

CodiLime

201 - 500 employees

Founded 2011

🤝 B2B

📡 Telecommunications

🔧 Hardware

B2B • Telecommunications • Hardware

CodiLime is a software and network engineering services company that partners with networking hardware vendors, software providers, and telecommunications firms to create proofs-of-concept, develop new products, and support production environments. Founded in 2011 and grown to 300+ employees, the company specializes in network automation, low-level systems programming, observability, DevOps, and cybersecurity, serving clients worldwide (US, Japan, Israel, Europe). CodiLime primarily operates as a B2B service provider to tech startups and large industry players.

📋 Description

• Design, build, and maintain batch and streaming data pipelines in Python, including orchestration, scheduling, and monitoring with Airflow • Write reusable, well-typed Python libraries and internal packages used by engineers and analysts • Build Python services and APIs with FastAPI and integrate with third-party and internal APIs • Develop SQL and dbt data transformations on Snowflake and build data models for fast, reliable querying • Write unit, integration, and data-contract tests and maintain automated CI coverage • Profile and optimize Python code and data-processing jobs for runtime, memory, and cost • Deploy and operate code in the cloud using containers, infrastructure as code, and CI/CD • Enforce access controls, secrets handling, and sensitive-data protections throughout the data lifecycle • Use AI coding assistants effectively while validating generated output before production • Find and fix efficiency, reliability, cost, and correctness issues in existing pipelines • Replace one-off scripts and notebooks with tested, packaged, scheduled code • Implement data quality checks, validation, and monitoring • Create matching logic to deduplicate and connect entities across multiple data sources • Document data processes and system architecture and maintain project documentation

🎯 Requirements

• Strong Python experience: data structures, typing, error handling, generators/iterators, context managers, and the standard library • Software engineering fundamentals: modular design, dependency management, packaging, and API design • Testing discipline: pytest, fixtures, mocking, and writing code that's testable by construction • Hands-on experience building and operating ETL/ELT pipelines in production, not just scripts or notebooks • Strong experience with Snowflake and dbt • Experience with Apache Airflow or similar code-based orchestration tools • Solid working knowledge of SQL and data modeling, sufficient to design robust database schemas and query them effectively • Experience with Docker, Kubernetes, and CI/CD practices • Debugging and profiling skills; able to reason about performance, concurrency, and memory in Python • Experience with at least one public cloud (AWS or Azure) • Experience with version control systems (Git) • Able to explain technical trade-offs to both technical and business audiences • Experience using AI coding assistants such as Claude Code, Cursor, or similar on a daily basis • Strong communication skills and good knowledge of English (minimum C1 level) • Experience with Apache Spark, ideally on Databricks • Experience with Pydantic or similar schema-validation libraries • Python web/API frameworks (FastAPI, Flask) • Async Python, multiprocessing, or other concurrency patterns • Experience with Azure AI Search or AWS OpenSearch • A second language: Go, Rust, Scala, or TypeScript • Familiarity with LLMs, Azure OpenAI, or agentic AI systems

🏖️ Benefits

• Flexible working hours and approach to work: fully remotely, in the office or hybrid • Professional growth supported by internal training sessions and a training budget • Solid onboarding with a hands-on approach to give you an easy start • A great atmosphere among professionals who are passionate about their work • The ability to change the project you work on

Apply Now

Similar Jobs

🕒 3 days ago

InPost Group

10,000+ employees

🛍️ eCommerce

🚗 Transport

📦 Logistics

Product Manager evolving InPost’s Azure and Databricks data platform. Driving FinOps, DataOps, governance and scalable data capabilities for European parcel delivery.

Azure

Cloud

Tableau

🕒 3 days ago

InPost Group

10,000+ employees

🛍️ eCommerce

🚗 Transport

📦 Logistics

Product Manager evolving InPost’s Azure and Databricks data platform. Driving FinOps, DataOps and secure data capabilities for Europe’s leading out-of-home parcel delivery network.

Azure

Cloud

Tableau

🕒 August 13

Sowelo Consulting sp. z o.o. sp. k.

11 - 50

💼 Consulting

📦 Logistics

📣 Marketing

Senior Data Architect designing Databricks lakehouse platforms for an AI and data solutions consultancy. Leading migrations, data pipelines, cloud implementations, and client-facing architecture engagements.

Apache

AWS

Azure

Cloud

ETL

Google Cloud Platform

Python

Scala

Spark

SQL

🕒 July 30

Exerizon

11 - 50

💼 Consulting

🤖 Artificial Intelligence

🤝 B2B

Join Exerizon as a Data Engineer focusing on Data Lake and Data Lakehouse projects. Collaborate on AI-supported data migrations while building a data team from scratch.

🗣️🇵🇱 Polish Required

Airflow

AWS

Azure

Oracle

Postgres

SQL

Vault

🕒 July 29

InPost Group

10,000+ employees

🛍️ eCommerce

🚗 Transport

📦 Logistics

Data Platform Product Manager at InPost overseeing technical roadmap for data platform using Azure and Databricks. Driving cloud cost optimization and collaboration across data and engineering teams.

Azure

Cloud