Big Data Engineer

πŸ”₯ 36 minutes ago

Apply Now
Find Similar Remote Jobs

πŸ“Š Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of AdvanSix

AdvanSix

1001 - 5000 employees

Founded 2016

🏭 Manufacturing

🀝 B2B

πŸ’° $12M Grant - AdvanSix on 2024-09

Manufacturing β€’ B2B

AdvanSix is a U. S. -based chemical manufacturer that produces nylon 6, caprolactam, ammonium sulfate and other chemical intermediates and performance materials for industrial customers. The company operates integrated chemical production facilities supplying polymers, fibers and specialty chemicals to customers in plastics, textiles, agriculture (fertilizers) and other industrial markets. AdvanSix primarily serves business customers (B2B) and focuses on large-scale manufacturing and materials production.

πŸ“‹ Description

β€’ AdvanSix is seeking a Big Data Engineer to build and operate our enterprise Unified Data Layer (UDL). β€’ Engineer batch/CDC/streaming pipelines, model curated/semantic layers, and harden run-state with testing, CI/CD, security, and observability. β€’ Partner closely with the data team and larger IT organization. β€’ Design and deliver scalable, secure data pipelines and data models that safely connect operational systems to analytics. β€’ Build ingestion pipelines (batch, CDC, streaming) from S/4HANA/DataSphere, PHD/historian, LIMS, TMS, HSE, and other sources. β€’ Implement data contracts, schema/versioning, SCD handling, partitioning, and performance tuning. β€’ Develop dimensional/semantic models that back certified Power BI datasets and APIs for apps/agents. β€’ Integrate OT data via OPC UA/MQTT and collaborate with plant controls on change control, signal quality, and downtime windows. β€’ Embed data quality rules, unit/integration tests, and validation checks. β€’ Automate build/test/deploy with Git-based CI/CD.

🎯 Requirements

β€’ Minimum 5 years' in data engineering building production pipelines at scale (batch/CDC/streaming). β€’ Hands-on with Azure data stack: Databricks or Fabric/Synapse, ADF/Pipelines, ADLS/OneLake, Azure SQL/SQL MI, Key Vault. β€’ Strong SQL and Python/PySpark; comfort with Spark Structured Streaming and performance tuning. β€’ Experience implementing tests/observability (freshness, schema, expectations), and Git-based CI/CD. β€’ Familiarity with SAP S/4HANA structures and SAP DataSphere semantic modeling. β€’ OT concepts: historians (PHD/PI), OPC UA/MQTT, event/batch frames, ISA-95/99 basics. β€’ Understanding of Power BI consumption (semantic models, RLS) and APIs for downstream AI/ML apps/agents.

Apply Now

Similar Jobs

πŸ”₯ 1 hour ago

Standvast Fulfillment

51 - 200

πŸ“¦ Logistics

πŸ›οΈ eCommerce

Senior Data Engineer in a fully remote-first organization focusing on engineering best practices with cloud infrastructure. Collaborate with cross-functional teams to deliver innovative solutions and mentor junior developers.

BigQuery

Python

Spark

SQL

πŸ”₯ 2 hours ago

Instacart

1001 - 5000

🍽️ Food & Beverage

πŸ“¦ Logistics

πŸ›οΈ eCommerce

Senior Risk & Compliance Engineer developing automated risk programs for Instacart's governance, risk, and compliance team. Writing production-level code, building data pipelines, and developing risk models at scale.

Postgres

Python

SQL

πŸ”₯ 2 hours ago

mPulse

501 - 1000

πŸ₯ Healthcare

☁️ SaaS

🀝 B2B

Data Engineer II responsible for designing and maintaining scalable data pipelines in healthcare. Collaborating with cross-functional teams to enable data-driven solutions for improved outcomes.

Airflow

Amazon Redshift

Apache

AWS

Cloud

ETL

Jenkins

MS SQL Server

Postgres

Python

SQL

Tableau

πŸ”₯ 2 hours ago

Genesys

5001 - 10000

πŸ’Ό Consulting

πŸ₯ Healthcare

πŸ›‘οΈ Insurance

Senior Data Engineer transforming financial data and improving reporting by migrating pipelines to AWS. Collaborating with stakeholders to enhance data management practices and support business objectives.

AWS

Cloud

Python

SQL

πŸ”₯ 4 hours ago

Swish Analytics

11 - 50

πŸ’Ό Consulting

πŸ“£ Marketing

🎲 Gambling

Data Engineer building predictive sports analytics solutions at Swish Analytics. Collaborating on real-time data and infrastructure for sports betting products.

Airflow

AWS

Cloud

ETL

Kubernetes

MySQL

Python

Shell Scripting

SQL