Data Engineer II

🔥 0 minutes ago

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MRSOOL | مرسول

MRSOOL | مرسول

201 - 500 employees

Founded 2015

🍽️ Food & Beverage

✈️ Travel

💼 Consulting

Food & Beverage • Travel • Consulting

MRSOOL is one of the largest delivery platforms in the region, offering an on-demand experience with high user ratings in both Apple's App Store and Google's Play store. MRSOOL provides a "order anything from anywhere" service backed by a large fleet of registered couriers. It enables businesses to access a vast user base, facilitating transformation to on-demand eCommerce and offering a flexible bidding system for service prices. MRSOOL also offers a personalized delivery experience with real-time tracking and communication with couriers, making it a leading option for ordering from local shops, groceries, and restaurants directly to your door.

📋 Description

• Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt • Develop and optimize Bronze, Silver, and Gold data models using Medallion Architecture • Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery • Design data ingestion frameworks using CDC, Kafka, and event-driven architectures • Create and maintain data warehouses and data marts for reporting and self-service analytics • Translate business requirements into scalable data solutions with Product, Analytics, Engineering, and Business stakeholders • Develop reusable dbt models, testing frameworks, and documentation • Optimize Spark jobs, Trino queries, and storage layouts • Own critical data pipeline lifecycles, monitoring, SLA adherence, and incident resolution • Build reusable data-platform frameworks, automation, CI/CD pipelines, and engineering practices • Ensure data quality through validation, monitoring, lineage, observability, security, and governance • Deliver trusted datasets, semantic models, and Metabase dashboards for decision-making

🎯 Requirements

• 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses • Strong proficiency in Spark (Scala, python) and SQL • Experience building production-grade data pipelines and distributed data processing applications • Hands-on experience with Apache Spark, distributed data processing, performance tuning, and optimization • Experience building batch and streaming data pipelines using Kafka, CDC/Maxwell, or similar event-driven architectures • Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization • Experience with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines • Experience designing dimensional models, star schemas, and reliable data marts • Hands-on experience with dbt, reusable models, automated testing, and documentation • Knowledge of data quality, observability, lineage, and engineering best practices • Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs • Familiarity with CI/CD, Git-based development workflows, infrastructure automation, and modern software engineering best practices • Ability to independently own projects from design through production • Experience collaborating across Product, Engineering, Analytics, and Business teams

🏖️ Benefits

• Inclusive and diverse workplace • Remote work environment • Competitive compensation • Potential share options for certain roles • Regular training • Annual learning stipend • High degree of autonomy • Mentorship

Apply Now

Similar Jobs

🔥 15 hours ago

G-P

1001 - 5000

💼 Consulting

🏥 Healthcare

⚖️ Legal

Data Architect designing Databricks-native architecture, ingestion, governance, and security. Supporting G-P’s SaaS global employment platform serving organizations across 180+ countries.

AWS

Azure

Cloud

Google Cloud Platform

PySpark

Python

SDLC

Spark

SQL

🕒 Yesterday

PORCH 💚

1 - 10

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Data Warehouse Engineer building BigQuery-based pipelines for Porch Group’s SaaS and insurance platform. Optimizing SQL, schemas, and warehouse performance for business intelligence and reporting.

BigQuery

Cloud

ETL

SQL

🕒 2 days ago

PORCH 💚

1 - 10

💼 Consulting

🏥 Healthcare

🏨 Hospitality

Senior Data Warehouse Engineer building BigQuery data pipelines for Porch Group’s SaaS and insurance platform. Optimizing SQL, warehouse performance, and reporting data quality.

BigQuery

Cloud

ETL

SQL

🕒 2 days ago

Shuru

51 - 200

🤖 Artificial Intelligence

🤝 B2B

🏢 Enterprise

Snowflake Data Engineer building scalable pipelines and cloud data solutions for Shuru’s AI-native engineering services company. Optimizing data quality, warehousing, and analytics workflows.

Airflow

AWS

Azure

Cloud

ETL

Google Cloud Platform

Matillion

Python

SQL

🕒 5 days ago

Datavail

1001 - 5000

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Data Engineer building automated, governed Databricks and Microsoft Fabric data platforms. Supporting Datavail’s clients with pipelines, quality frameworks, analytics, and modernization.

Apache

AWS

Azure

Cloud

ETL

Google Cloud Platform

PySpark

Spark

SQL

Terraform