Lead Data Engineer

Job not on LinkedIn

🔥 13 hours ago

🇮🇳 India – Remote

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 18%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MFSG

MFSG

11 - 50 employees

Founded 2018

🏭 Manufacturing

🔧 Hardware

🚗 Transport

Manufacturing • Hardware • Transport

MFSG is a provider of automated material handling systems (AMHS) and related control software for semiconductor fabs. The company designs and supplies hardware such as overhead hoist transport (OHT) systems, autonomous mobile robots (AMR), conveyors, FOUP/POD stockers, reticle cabinets, buffers, and purge systems, along with software solutions including material control systems, FAB monitoring & simulation, and mobile robot and transport controllers. MFSG focuses on improving fab efficiency, flexibility, and reliability for semiconductor manufacturers through integrated hardware and software solutions, and is based in Singapore.

📋 Description

• Lead the end-to-end design and delivery of scalable data pipelines and data solutions using Snowflake, Snowflake Openflow (Apache NiFi), and dbt • Translate business and data requirements into technical designs, data flows, integration patterns, data models, and delivery plans • Build and maintain Openflow ingestion and orchestration flows, dbt transformation models, and performant Snowflake workloads • Design reusable batch, change data capture, API, and file-based ingestion patterns with validation, reconciliation, restartability, and recovery • Develop in advanced SQL and Python through code reviews, troubleshooting, performance tuning, and complex issue resolution • Ensure data quality, security, scalability, maintainability, observability, and cost efficiency • Collaborate with architecture, analytics, platform, security, application, and business teams • Maintain technical documentation and reusable engineering standards • Lead Data Operations support and coordinate priorities, handoffs, and issue resolution across India-based data colleagues • Establish practices for source control, peer review, testing, CI/CD, deployment, documentation, and production readiness • Mentor colleagues through design and code reviews and promote cross-training • Own and mature Data Operations, including incident response, escalation paths, runbooks, recovery procedures, and service expectations • Drive automated monitoring, alerting, and recovery across data platforms and pipeline workloads • Track service health and incident trends, lead root-cause analysis, and improve reliability and operational efficiency

🎯 Requirements

• Bachelor's degree in computer science, engineering, information systems, or an equivalent combination of education and relevant experience • 10+ years of hands-on data engineering experience, including at least 2 years leading technical delivery or production operations • Strong production experience with cloud data technologies in Snowflake, AWS, Azure, or GCP • Strong hands-on experience with dbt • Experience with Git-based development, automated testing, CI/CD, Terraform, and repeatable environment-promotion practices • Advanced SQL skills • Experience building maintainable, testable production data pipelines and data products • Experience designing ETL/ELT solutions across batch, CDC, API, and file-based ingestion • Experience with data modeling, quality, validation, and recovery practices • Strong solution-design, communication, and technical-leadership skills • Ability to guide decisions, mentor engineers, and work across business and technology teams

🏖️ Benefits

• Comprehensive medical coverage for employee, spouse, two children, and two parents • ₹6 lakhs annual medical coverage • Optional ₹5 lakhs medical top-up • Employer provident fund contribution of ₹1,800 per month • Annual performance-based bonus of up to 5% of base salary • Bucketlist point-based rewards redeemable for gift cards, experiences, and personalized perks • Recognition of milestones, achievements, and impact • Currently remote work model, with a future transition to hybrid • 15 annual leaves • 6 casual leaves • 12 sick leaves • Upskilling programs, mentorship, and professional development resources • Internal mobility programs • Leadership development, industry certifications, and specialized training • Collaborative knowledge sharing and exposure to global teams

Apply Now

Similar Jobs

🔥 13 hours ago

Momentum

51 - 200

💼 Consulting

💳 Fintech

💸 Finance

Lead Data Engineer building scalable Snowflake, dbt, and AWS data platforms. Leading Data Operations and reliable data products for India-focused financial services.

Apache

AWS

Azure

Cloud

ETL

Google Cloud Platform

Postgres

Python

SQL

Terraform

🕒 Yesterday

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Data Architect shaping scalable cloud and data architectures at Empower, a financial services company. Driving modernization, governance, security, and emerging AI innovation.

Amazon Redshift

AWS

Cloud

ETL

Java

Kafka

Kubernetes

Microservices

Spark

🕒 2 days ago

Firmable

51 - 200

🤝 B2B

☁️ SaaS

🤖 Artificial Intelligence

Lead data engineer architecting AI-native extraction and ETL pipelines for Firmable’s B2B sales intelligence platform. Building agentic workflows and production LLM infrastructure across 13 markets.

Airflow

Amazon Redshift

AWS

Cloud

ETL

JavaScript

Python

SQL

🕒 3 days ago

CrowdStrike

5001 - 10000

🔒 Cybersecurity

☁️ SaaS

🤖 Artificial Intelligence

Data Engineer building pipelines, automation, and production data systems for CrowdStrike’s AI-native cybersecurity platform. Supporting machine-learning products and large-scale event processing.

Airflow

AWS

Docker

Kafka

Kubernetes

Linux

Python

Go

🕒 3 days ago

Firmable

51 - 200

🤝 B2B

☁️ SaaS

🤖 Artificial Intelligence

Lead Data Engineer architecting AI-native extraction and ETL pipelines for Firmable’s B2B sales intelligence platform. Building production LLM systems, agentic workflows, and data quality infrastructure across 13 markets.

Airflow

Amazon Redshift

AWS

ETL

JavaScript

Python

SQL