Data Platform Engineer

🔥 0 minutes ago

🌐 Romania, Poland – Remote

infoinfo

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 12%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Software Mind

Software Mind

1001 - 5000 employees

Founded 1999

🤖 Artificial Intelligence

☁️ SaaS

📡 Telecommunications

💰 Private Equity Round on 2020-12

Artificial Intelligence • SaaS • Telecommunications

Software Mind is a technology company that specializes in software development and digital transformation services. With a focus on AI and cloud solutions, the company offers a wide range of services including custom software development, mobile app development, and cloud consulting. Software Mind serves various industries such as financial services, telecom, biotech, and media, providing tailored solutions to accelerate digital transformations and business growth globally.

📋 Description

• Build and operate change-data-capture pipelines from PostgreSQL into Azure using Kafka Connect and Debezium • Configure, deploy and scale connectors end to end, including connector setup, task management, offsets, schema history and snapshot strategy • Run pipelines as stateful workloads on Kubernetes AKS, covering configuration, secrets, networking and resource tuning • Monitor and troubleshoot the platform in production, including connector failures, task rebalances, restarts, throughput, backpressure, message-size limits, retries and recovery • Automate the platform in Python through configuration-driven data-source onboarding, pipeline orchestration, monitoring, alerting, recovery workflows and automated testing • Integrate CDC streams with Azure Event Hubs, ADLS, Azure PostgreSQL, ADF and Databricks • Manage platform infrastructure as code so environments are reproducible and changes are reviewable • Apply data protection requirements to sensitive data flowing through pipelines, including masking, hashing, access control and retention

🎯 Requirements

• Solid commercial experience as a data or platform engineer, with hands-on work on streaming or CDC pipelines rather than batch reporting alone • Practical Kafka knowledge, including topics, partitions, offsets, consumer groups and delivery semantics • At least one Kafka Connect deployment run independently • Strong SQL and PostgreSQL skills • Working knowledge of WAL, logical replication, replication slots and replication lag • Working understanding of CDC concepts, including initial snapshots, inserts, updates and deletes, event ordering, at-least-once delivery and schema evolution • Confident Python for automation and tooling • Hands-on experience with Azure data services such as Event Hubs, ADLS or Azure PostgreSQL • Comfortable working with Kubernetes as a user, including deploying workloads, handling configuration and secrets, reading logs and debugging failing pods • Ability to debug running pipelines from metrics and logs • Production experience with Debezium specifically • Experience operating stateful workloads on AKS, including StatefulSets, stable worker identity and resource tuning under load • Infrastructure-as-code and CI/CD experience for data platform components, using Terraform, Bicep or similar • Hands-on work with Databricks and ADF at production scale • Experience implementing data protection controls for sensitive data, including masking, hashing, access control and retention policies

🏖️ Benefits

• Flexible and remote cooperation options • International projects with leading global clients • Travels related to international projects • Non-corporate atmosphere • Access to language classes • Access to knowledge-sharing initiatives • Access to private healthcare and life insurance • Access to multisport card

Apply Now

Similar Jobs

🔥 5 hours ago

accesa.eu

1001 - 5000

💼 Consulting

🏭 Manufacturing

🏥 Healthcare

Senior Data Platform Engineer automating secure Azure-Databricks platforms for banking clients. Building Terraform infrastructure, Airflow orchestration, and cloud data synchronization pipelines.

Airflow

Apache

Azure

Cloud

Docker

Kubernetes

OpenShift

Terraform

Unity

🕒 6 days ago

Betfair Romania Development

1001 - 5000

💼 Consulting

📣 Marketing

🏥 Healthcare

Data Engineer building scalable Big Data platforms for Flutter’s global casino gaming studios. Developing streaming, cloud, storage and analytics solutions in an agile Scrum environment.

Amazon Redshift

AWS

Cassandra

Cloud

DynamoDB

ETL

Java

Kafka

NoSQL

Pulsar

PySpark

Python

Scala

Spark

SQL

Tableau

🕒 September 1

AllCloud

201 - 500

💼 Consulting

📦 Logistics

🏥 Healthcare

Data Engineer building AWS lakehouse, ETL, and GenAI platforms. Supporting AllCloud’s cloud migration and managed-services customers through hands-on engineering.

🗣️🇮🇱 Hebrew Required

Airflow

Amazon Redshift

AWS

Cloud

ETL

Kafka

Kubernetes

Microservices

Python

Spark

SQL

🕒 August 21

MWDN

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior Data Engineer building scalable pipelines and warehouses for ad-tech business intelligence. Powering real-time analytics from billions of advertising events for MWDN’s global clients.

Airflow

Amazon Redshift

Apache

AWS

BigQuery

Cloud

ETL

Google Cloud Platform

Kubernetes

Python

Scala

Spark

SQL

Terraform

🕒 August 21

MWDN

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior Data Engineer building scalable real-time and batch pipelines for a fast-growing programmatic advertising company. Optimizing auction analytics, yield models, and ad-tech data infrastructure.

Airflow

Amazon Redshift

AWS

BigQuery

Cloud

ETL

Google Cloud Platform

Kafka

Python

Spark

SQL