Senior Data Engineer – Real Time Streams

🔥 14 hours ago

🌐 Ukraine, Armenia, +1 more countries – Remote

infoinfo

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Globaldev Group

Globaldev Group

201 - 500 employees

Founded 12 years

💼 Consulting

📦 Logistics

☁️ SaaS

Consulting • Logistics • SaaS

Globaldev Group is an international company that specializes in providing IT outsourcing and software development services. They focus on assisting businesses in building and managing their engineering teams, offering expertise in various modern technologies and development practices to ensure efficient and effective project delivery.

📋 Description

• Serve as a senior engineer on the team owning the real-time data platform • Design, implement, deploy, and operate streaming jobs, CDC pipelines, Kafka Connect bridges, and downstream data sinks • Design, implement, and operate stateful Apache Flink streaming pipelines using keyed state, windowing, watermarks, timers, side outputs, and custom sources/sinks • Develop and maintain CDC pipelines with Flink CDC or Debezium, including snapshot/incremental handling, schema evolution, and downstream idempotency • Build and operate Kafka and Kafka Connect pipelines, including topic and partition design, distributed source/sink connectors, schema and converter management, and dead-letter routing • Own the Kafka → Flink → Postgres / TimescaleDB / ClickHouse / BigQuery data path, including schema design, idempotency strategy, batch tuning, and observability • Diagnose and fix production issues such as checkpoint failures, backpressure, state growth, sink slowness, autoscaler oscillation, and restart loops • Harden the platform through delivery guarantees, watermarks, DLQ handling, schema migrations, and alerting coverage • Conduct code reviews and mentor mid-level engineers • Contribute to deployment infrastructure through Helm charts, ArgoCD applications, GKE configuration, Grafana dashboards, and Prometheus alert rules • Influence technical direction, including state backend choices, schema migrations, and pipeline decomposition or rebuilding

🎯 Requirements

• 5+ years of professional software/data engineering experience, with strong Java expertise • Production experience with Apache Flink and stateful real-time streaming pipelines • Strong Apache Kafka experience, including Kafka Connect, consumer groups, partitions, delivery guarantees, and schema management • Hands-on CDC experience using Debezium or Flink CDC • Experience designing and operating Kafka → Flink → database/data warehouse pipelines in production • Strong understanding of PostgreSQL and experience with at least one analytical database such as ClickHouse or BigQuery • Experience with Kubernetes and deploying/operating production data workloads • Proven ability to troubleshoot and optimize production streaming systems—backpressure, checkpoint failures, state growth, latency, lag, and sink performance • Strong understanding of data consistency, idempotency, schema evolution, and observability • Ability to work independently and own a data platform component end-to-end, from design through production operation • Experience with TimescaleDB, hypertables, and time-series data would be a plus • Experience with PostGIS and geospatial/spatial query optimization would be a plus • Deep ClickHouse performance tuning, including MergeTree and partitioning strategies, would be a plus • Strong BigQuery optimization and data modeling experience would be a plus • Experience with Flink Kubernetes Operator and Flink autoscaler tuning would be a plus • Experience with ArgoCD, Helm, GKE, Prometheus, and Grafana would be a plus • Experience designing high-throughput, low-latency real-time systems at scale would be a plus • Experience with MQTT or IoT/event-driven systems would be a plus • Knowledge of Kafka Schema Registry, Avro/Protobuf, and advanced schema evolution would be a plus • Experience mentoring engineers and leading technical decisions around streaming architecture would be a plus • Experience with routing, logistics, delivery, mobility, or location-based platforms would be particularly relevant

🏖️ Benefits

• 20 days of paid vacation • 5 sick days • Public holidays • Flexible schedule with a high level of autonomy • Opportunity to influence product and technical decisions • Professional growth and learning opportunities • Comfortable working environment with a supportive team

Apply Now

Similar Jobs

🕒 Yesterday

OneDome

51 - 200

💼 Consulting

📦 Logistics

⚖️ Legal

Sr. Data Engineer owning OneDome’s AWS data warehouse and pipelines. Enabling analytics for a UK-based property technology platform simplifying home transactions.

🇺🇦 Ukraine – Remote

💵 $4.5k - $5.5k / month

💰 Private Equity Round on 2022-11

⏰ Full Time

🟠 Senior

🚰 Data Engineer

Amazon Redshift

AWS

BigQuery

DynamoDB

NoSQL

Postgres

Python

SQL

Terraform

🕒 4 days ago

MWDN

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior Data Engineer building scalable pipelines and warehouses for MWDN’s ad-tech business intelligence clients. Processing SSP, DSP, GAM, CRM, and real-time advertising data.

Airflow

Amazon Redshift

AWS

BigQuery

ETL

Google Cloud Platform

Python

Scala

Spark

SQL

🕒 4 days ago

MWDN

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior Data Engineer building scalable real-time and batch data systems for programmatic advertising. Optimizing auction analytics, yield, and partner performance for an AI-focused AdTech company.

Airflow

Amazon Redshift

AWS

BigQuery

Cloud

ETL

Google Cloud Platform

Kafka

Python

Spark

SQL

🕒 4 days ago

MWDN

51 - 200

💼 Consulting

📦 Logistics

🏥 Healthcare

Senior Data Engineer building scalable pipelines and warehouses for MWDN’s ad-tech business intelligence projects. Processing massive SSP, DSP, ad server, and CRM datasets for real-time analytics.

Airflow

Amazon Redshift

Apache

AWS

BigQuery

Cloud

ETL

Google Cloud Platform

Kubernetes

Python

Scala

Spark

SQL

Terraform

🕒 4 days ago

Fluent Trade Technologies

51 - 200

💳 Fintech

Big Data Engineer architecting time-series infrastructure for Fluent Trade Technologies' FX trading analytics platform. Optimizing massive market-data datasets and low-latency queries for global banks and hedge funds.

Cassandra

Linux

NoSQL

Python