Lead Data Engineer, Contract, Full-Time

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Smart Working

Smart Working

51 - 200 employees

💼 Consulting

🏥 Healthcare

📣 Marketing

Consulting • Healthcare • Marketing

Smart Working is a recruitment service specializing in sourcing and providing top-tier software developers from around the world to meet the needs of businesses. With a robust vetting process that includes technical assessments and background checks, Smart Working ensures that clients receive highly skilled developers adept in various programming languages and frameworks. The company focuses on flexible and remote hiring solutions, allowing businesses to efficiently scale their development teams while benefiting from significant cost savings.

📋 Description

• Architect and build scalable data pipelines and infrastructure supporting AI and product systems • Design and maintain data ingestion, transformation, and storage architectures • Develop and manage batch and real-time data pipelines • Build and optimize vector search, retrieval, and machine learning data pipeline systems • Ensure data reliability, security, and governance • Collaborate with AI and backend engineering teams on training, inference, and product features • Implement monitoring, observability, and data quality frameworks • Optimize large-scale dataset and query-system performance • Contribute to technical architecture decisions and long-term data strategy • Define culture, standards, and hiring bar as the founding data hire • Partner with founders and product leadership to translate data capabilities into product decisions

🎯 Requirements

• 7+ years of professional experience, primarily in dedicated data engineering roles • Strong experience designing and building data pipelines and distributed data systems • Experience with relational databases; PostgreSQL preferred, with MySQL or similar acceptable • Experience with NoSQL databases • Experience with vector databases for modern AI systems • Strong Python programming experience • Ability to make and justify architectural decisions • Experience building scalable backend systems • Experience designing data models and storage architectures • Strong understanding of data-processing performance and optimization • Highly desirable: Apache Spark, Apache Airflow, Kafka, Elasticsearch or OpenSearch • Highly desirable: PostgreSQL, MongoDB, Qdrant, Milvus, or pgvector • Highly desirable: Pandas or Polars • Nice to have: AI or machine learning platform experience • Nice to have: stream processing and event-driven architecture familiarity • Nice to have: GCP, AWS, or Azure cloud infrastructure experience • Nice to have: high-growth startup or early-stage company experience

🏖️ Benefits

• Genuine community focused on growth and well-being • Remote-first work environment • Full-time, long-term roles • Opportunity to work with global teams and products • Personal and professional growth opportunities

Apply Now

Similar Jobs

🕒 3 days ago

Hypersonix Inc.

51 - 200

💼 Consulting

📣 Marketing

📦 Logistics

Data Engineer building scalable pipelines and data infrastructure for Hypersonix.ai’s AI-powered e-commerce insights platform. Supporting analytics, machine learning, and real-time business applications through reliable data systems.

Airflow

AWS

Azure

Cloud

Distributed Systems

ETL

Google Cloud Platform

PySpark

Python

Spark

SQL

🕒 4 days ago

Greenlight Planet

1001 - 5000

⚡ Energy

🌍 Social Impact

👥 B2C

Senior Data Engineer automating incentive compensation systems for Sun King's solar energy business. Ensuring accurate, auditable payouts across products and countries through data pipelines and quality controls.

Airflow

Amazon Redshift

Python

SQL

Tableau

🕒 4 days ago

thinkbridge

201 - 500

💼 Consulting

🏥 Healthcare

🏭 Manufacturing

Data Engineer analyzing Power BI reports and Alteryx workflows for thinkbridge, a technology strategy and development consultancy. Documenting data lineage, mappings, and reporting architecture using SQL.

Azure

ETL

SQL

🕒 4 days ago

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Senior Data Engineer building Empower’s AWS data infrastructure and ingestion frameworks. Supporting scalable financial reporting through Python, ETL, SQL, and Big Data solutions.

Amazon Redshift

AWS

Cloud

DynamoDB

ETL

Hadoop

MySQL

Postgres

PySpark

Python

RDBMS

Spark

SQL

🕒 5 days ago

StarTree

11 - 50

💼 Consulting

📦 Logistics

📣 Marketing

Senior Software Engineer building Apache Pinot distributed analytics systems at StarTree, a cloud software company enabling real-time data insights. Developing fault-tolerant, low-latency analytics applications at scale.

Apache

Distributed Systems

Java

Linux

Maven

Python

SQL

Subversion

Unix

Go