Senior Data Engineer

Job not on LinkedIn

🕒 August 4

🇺🇸 United States – Remote

💵 $158.3k - $180k / year

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🦅 H1B Visa Sponsor

info

Airflow

Apache

Keras

Kubernetes

OpenShift

PySpark

Scikit-Learn

Splunk

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Red Hat

Red Hat

10,000+ employees

Founded 1993

🏢 Enterprise

💰 Corporate Round on 1999-03

Enterprise • Cloud

Red Hat is a leading provider of enterprise open source software solutions, helping companies worldwide to build and deploy applications across hybrid cloud infrastructures. With a strong focus on developing secure, stable, and innovative technologies, Red Hat offers a broad portfolio including products like Red Hat Enterprise Linux, Red Hat OpenShift, and Red Hat Ansible Automation Platform. These products support IT services on any infrastructure efficiently. Trusted by more than 90% of the U. S. Fortune 500, Red Hat empowers organizations to modernize their IT environments, leveraging open source communities to drive technological advancement.

📋 Description

• Architect and implement high-volume data pipelines between Snowflake and Databricks using PySpark and dbt • Orchestrate pipeline scheduling, dependencies, monitoring, and automated failure recovery with Apache Airflow • Administer Databricks platforms, including IAM roles, Amazon S3 access, secrets, governance policies, and cluster configurations • Design and deploy vector-based retrieval architecture and AI-driven marketing decision-automation workflows • Operationalize MLOps with MLflow and Lakehouse monitoring • Build NER, predictive, deep learning, and time-series forecasting models • Manage CI/CD pipelines with Git and Tekton • Build and publish container images using buildah and skopeo • Deploy and maintain containerized models and applications on Red Hat OpenShift/Kubernetes • Manage network routes, TLS termination, and high-availability services • Lead enterprise security compliance assessments across 20+ controls • Perform SAST, vulnerability scanning, Privacy Impact Assessments, and STRIDE-based threat modeling • Collaborate with information security teams to remediate vulnerabilities, support compliance audits, and maintain Splunk logging and monitoring

🎯 Requirements

• Bachelor's degree in Computer Science or related field and five (5) years of experience, or Master's degree and three (3) years of experience • U.S. or foreign equivalent degree • At least three (3) years of experience architecting high-volume Snowflake-to-Databricks data pipelines using PySpark and dbt • At least three (3) years of experience orchestrating workflows with Apache Airflow, including DAG dependencies, failure recovery, and monitoring • At least three (3) years of experience administering Databricks, Amazon S3 IAM access, platform vaults, OpenShift secrets, governance, and cluster policies • At least three (3) years of experience building NER systems with TextBlob, gensim, and fastText • At least three (3) years of experience developing models with XGBoost and Scikit-learn • At least three (3) years of experience constructing Keras deep learning architectures and time-series forecasting models • At least three (3) years of experience implementing information retrieval systems using Transformer architectures and Transfer Learning • At least three (3) years of experience with MLflow, automated model monitoring, Git, Tekton, buildah, skopeo, and Red Hat OpenShift/Kubernetes • At least three (3) years of experience with enterprise security compliance assessments, SonarQube, Qualys, Privacy Impact Assessments, and STRIDE threat modeling

🏖️ Benefits

• Remote telecommuting arrangement • Bonus, commission, and/or equity eligibility • Flexible work environments • Reasonable accommodations for job applicants

Apply Now

Similar Jobs

🕒 August 4

PrizePicks

201 - 500

🎮 Gaming

⚽ Sports

Data Platform Engineer scaling PrizePicks’ batch and real-time data infrastructure for sports betting and daily fantasy. Building low-latency ML services, secure platforms, and highly observable pipelines.

🇺🇸 United States – Remote

💵 $145k - $175k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

🕒 August 4

Premier Inc.

1001 - 5000

🏥 Healthcare

⚕️ Healthcare Insurance

💼 Consulting

Healthcare data architect transforming EHR and other healthcare source data into Databricks warehouses for Premier’s analytics products. Mapping, validating, and troubleshooting customer data integrations.

🇺🇸 United States – Remote

💵 $113k - $188k / year

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🕒 August 4

HIKE2

51 - 200

💼 Consulting

🤖 Artificial Intelligence

📋 Compliance

Salesforce Data 360 Engineer building pipelines, identity resolution, and data activation for HIKE2’s transformation consultancy. Connecting governed customer data across marketing, sales, and service platforms.

🇺🇸 United States – Remote

💵 $110k - $175k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

🕒 August 4

Patriot Software

51 - 200

☁️ SaaS

🤝 B2B

Data Engineer III building scalable PostgreSQL databases, ETL/ELT pipelines, and AWS data infrastructure for Patriot Software’s accounting and payroll products. Optimizing performance, reliability, security, and availability while mentoring engineers.

🇺🇸 United States – Remote

💵 $105k - $125k / year

💰 Series B - Patriot Software on 2023-07

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

🕒 August 4

The Helper Bees

51 - 200

🏥 Healthcare

📦 Logistics

💼 Consulting

Data engineering manager leading governed pipelines, reporting, and predictive analytics for The Helper Bees’ in-home care platform. Managing engineers and analysts while advancing healthcare data strategy and AI adoption.

🇺🇸 United States – Remote

💵 $150k - $160k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer