Senior Data Platform Engineer

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of DigiCert

DigiCert

1001 - 5000 employees

Founded 2003

🏥 Healthcare

📦 Logistics

💼 Consulting

Healthcare • Logistics • Consulting

DigiCert is a global leader in providing high-assurance digital security solutions. It offers a wide range of services including TLS/SSL certificates, PKI (Public Key Infrastructure), and IoT (Internet of Things) security solutions. DigiCert's suite of products and services includes certificate lifecycle management, secure signing for code and documents, and devices across various industries such as healthcare, transportation, and smart cities. The company is focused on ensuring digital trust with advanced encryption and identity verification technologies, preparing organizations for the challenges of the quantum era. DigiCert is trusted by the majority of the Global 2000 companies for managing their digital trust needs.

📋 Description

• Design, build, and maintain scalable data and ML pipelines using Python and SQL across batch and streaming workloads • Own and evolve Databricks platform infrastructure, including Delta Lake architecture, Unity Catalog governance, Databricks Workflows orchestration, and compute optimization • Build and maintain end-to-end ML pipelines covering feature engineering, model training, experiment tracking, and model deployment/serving • Collaborate with data scientists to operationalize models in production-grade ML systems • Define and enforce data platform standards, ingestion patterns, data modeling conventions, medallion architecture, and reliability practices • Implement data quality, observability, and monitoring frameworks • Optimize pipelines for performance, cost, and reliability using Spark and PySpark • Evaluate, integrate, and govern platform tooling and data sources in the Databricks ecosystem • Contribute to architectural decisions and the long-term data platform roadmap • Participate in code reviews, technical design discussions, and engineering standards • Mentor junior engineers and improve platform and data engineering practices • Document platform architecture, pipeline design, and operational runbooks

🎯 Requirements

• 6+ years of experience in data engineering, data platform, or ML engineering roles • Strong proficiency in Python and SQL with production-grade data pipeline experience • Hands-on expertise with Databricks, Delta Lake, Unity Catalog, Databricks Workflows, and PySpark • Experience building and maintaining production ML pipelines, including feature engineering, training, experiment tracking, and model deployment • Familiarity with MLflow or comparable experiment tracking and model registry tools • Experience with cloud data platforms such as AWS, Azure, or GCP • Strong understanding of data modeling, dimensional design, and analytics-friendly data architecture • Experience with batch and incremental/CDC pipeline patterns • Proficiency with Git, version control, and CI/CD practices for data and ML workflows • Strong engineering judgment focused on reliability, maintainability, and cost • Clear communication and comfort working with technical and non-technical stakeholders • Nice-to-have: streaming or near real-time pipelines, feature stores, LLM/RAG or AI/BI tooling, data quality and observability tools, dbt, infrastructure as code, Agile/Scrum, and mentoring or platform standards experience

🏖️ Benefits

• Generous time off policies • Top shelf benefits • Education, wellness and lifestyle support

Apply Now

Similar Jobs

🔥 2 hours ago

Smart Working

51 - 200

💼 Consulting

🏥 Healthcare

📣 Marketing

Lead Data Engineer building scalable pipelines, vector databases, and ML data workflows for an AI assistant serving property-management businesses. Defining architecture and engineering standards as the first senior data hire.

Airflow

Apache

AWS

Azure

Cloud

ElasticSearch

Google Cloud Platform

Kafka

MongoDB

MySQL

NoSQL

Pandas

Postgres

Python

Spark

🕒 3 days ago

Hypersonix Inc.

51 - 200

💼 Consulting

📣 Marketing

📦 Logistics

Data Engineer building scalable pipelines and data infrastructure for Hypersonix.ai’s AI-powered e-commerce insights platform. Supporting analytics, machine learning, and real-time business applications through reliable data systems.

Airflow

AWS

Azure

Cloud

Distributed Systems

ETL

Google Cloud Platform

PySpark

Python

Spark

SQL

🕒 4 days ago

Greenlight Planet

1001 - 5000

⚡ Energy

🌍 Social Impact

👥 B2C

Senior Data Engineer automating incentive compensation systems for Sun King's solar energy business. Ensuring accurate, auditable payouts across products and countries through data pipelines and quality controls.

Airflow

Amazon Redshift

Python

SQL

Tableau

🕒 4 days ago

thinkbridge

201 - 500

💼 Consulting

🏥 Healthcare

🏭 Manufacturing

Data Engineer analyzing Power BI reports and Alteryx workflows for thinkbridge, a technology strategy and development consultancy. Documenting data lineage, mappings, and reporting architecture using SQL.

Azure

ETL

SQL

🕒 4 days ago

Empower

10,000+ employees

💸 Finance

💳 Fintech

👥 B2C

Senior Data Engineer building Empower’s AWS data infrastructure and ingestion frameworks. Supporting scalable financial reporting through Python, ETL, SQL, and Big Data solutions.

Amazon Redshift

AWS

Cloud

DynamoDB

ETL

Hadoop

MySQL

Postgres

PySpark

Python

RDBMS

Spark

SQL