Data Engineer II

🕒 July 8

🇮🇳 India – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 14%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MRSOOL | مرسول

MRSOOL | مرسول

201 - 500 employees

Founded 2015

🍽️ Food & Beverage

✈️ Travel

💼 Consulting

Food & Beverage • Travel • Consulting

MRSOOL is one of the largest delivery platforms in the region, offering an on-demand experience with high user ratings in both Apple's App Store and Google's Play store. MRSOOL provides a "order anything from anywhere" service backed by a large fleet of registered couriers. It enables businesses to access a vast user base, facilitating transformation to on-demand eCommerce and offering a flexible bidding system for service prices. MRSOOL also offers a personalized delivery experience with real-time tracking and communication with couriers, making it a leading option for ordering from local shops, groceries, and restaurants directly to your door.

📋 Description

• **What You Will Do ❓** • - Design, build, and maintain scalable batch and real-time data pipelines using Maxwell, Kafka, Spark, and dbt to power analytics and business-critical applications. • - Develop and optimize data models following Medallion Architecture (Bronze, Silver, Gold) to create reliable, reusable, and high-quality datasets. • - Build and maintain cloud-native data platforms using S3, Spark, Trino, and BigQuery, ensuring scalability, reliability, and cost efficiency. • - Design robust data ingestion frameworks leveraging CDC (Maxwell), Kafka, and event-driven architectures to support near real-time data processing. • - Create, optimize, and maintain data warehouses and data marts that enable fast, reliable reporting and self-service analytics. • - Partner closely with Product Managers, Data Analysts, Backend Engineers, and Business stakeholders to translate business requirements into scalable data solutions. • - Develop reusable dbt models, testing frameworks, and documentation to improve data quality, governance, and developer productivity. • - Optimize Spark jobs, Trino queries, and storage layouts for performance, reliability, and cost efficiency. • - Own the end-to-end lifecycle of critical data pipelines, ensuring high availability, monitoring, SLA adherence, and proactive incident resolution. • - Build and enhance the core data platform by developing reusable frameworks, automation, CI/CD pipelines, and engineering best practices. • - Ensure data quality through validation, monitoring, lineage, and observability while implementing best practices for security and governance. • - Enable analytics teams by delivering trusted datasets, semantic models, and dashboards that power decision-making through Metabase.

🎯 Requirements

• **What We're Looking For 🚀** • - 4+ years of hands-on experience designing and building scalable data platforms, data lakes, and data warehouses. • - Strong proficiency in Spark (Scala, python) and SQL, with experience building production-grade data pipelines and distributed data processing applications. • - Hands-on experience with Apache Spark and a solid understanding of distributed data processing, performance tuning, and optimization. • - Experience building batch and streaming data pipelines using technologies such as Kafka, CDC/Maxwell, or similar event-driven architectures. • - Strong understanding of modern data lake architectures, including Medallion Architecture, data modeling, partitioning, and storage optimization. • - Experience working with cloud-native data platforms and technologies such as Amazon S3, BigQuery, Trino, or similar analytics engines. • - Solid experience designing dimensional models, star schemas, and building reliable data marts that support analytics and business intelligence. • - Hands-on experience with dbt, including developing reusable models, implementing automated testing, and maintaining documentation. • - Strong knowledge of data quality, observability, lineage, and engineering best practices to build reliable and maintainable data products. • - Experience optimizing large-scale data pipelines, SQL queries, and distributed processing jobs for performance, scalability, and cost efficiency. • - Familiarity with CI/CD, Git-based development workflows, infrastructure automation, and modern software engineering best practices. • - Excellent problem-solving skills with the ability to independently own projects from design through production. • - Strong communication and stakeholder management skills, with experience collaborating across Product, Engineering, Analytics, and Business teams. • - A passion for building scalable data platforms and continuously improving developer experience, platform reliability, and operational excellence.

🏖️ Benefits

• **What We Offer You❗**Inclusive and Diverse Environment: We foster an inclusive and diverse workplace that values innovation and offers remote environments. • Competitive Compensation: Our compensation packages are highly competitive and include potential share options for certain roles. • Personal Growth and Development: We are committed to your personal and professional growth, providing regular training and an annual learning stipend to help you advance your career in a dynamic environment. • Autonomy and Mentorship: You'll enjoy a high degree of autonomy in your role, supported by mentorship and ambitious goals that pave the way for both your success and the company's growth.

Apply Now

Similar Jobs

🕒 July 7

3Pillar Global

1001 - 5000

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Lead Data Engineer building enterprise AI-native products with 3Pillar. Involved in data pipelines and ETL processes with MongoDB and Snowflake.

ETL

Informatica

MongoDB

MySQL

Python

SQL

🕒 July 4

EXL

10,000+ employees

🏥 Healthcare

🛡️ Insurance

📦 Logistics

Senior AWS Data Engineer managing data engineering projects ensuring alignment with business objectives and overseeing data engineers. Designing and maintaining scalable data pipelines in AWS cloud environment.

Amazon Redshift

AWS

Cloud

DynamoDB

ETL

Kafka

Python

Shell Scripting

Spark

SQL

🕒 June 30

Greenlight Planet

1001 - 5000

⚡ Energy

🌍 Social Impact

👥 B2C

Data Engineer role focusing on data infrastructure and pipeline development for Sun King. Collaborating with teams to ensure clean, reliable, and accessible data for decision making.

Airflow

Amazon Redshift

Apache

AWS

Cloud

EC2

ETL

Kafka

PySpark

Python

Spark

SQL

🕒 June 26

EVS, Inc.

201 - 500

🏗️ Construction

💼 Consulting

📦 Logistics

Data Engineer developing scalable data solutions that support renewable energy engineering and analytics. Collaborating with software and AI teams to innovate data architecture and pipelines.

Airflow

Amazon Redshift

AWS

BigQuery

Cloud

Docker

EC2

Pandas

PostGIS

Postgres

Python

SQL

Terraform

🕒 June 25

Everest Technologies, Inc

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

Senior Data Engineer at ETech focusing on Snowflake and dbt for data transformation and warehousing. Collaborating with business analysts and data scientists to optimize data pipelines.

Airflow

Apache

AWS

Azure

Cloud

Google Cloud Platform

Oracle

Python

Scala

SQL