Senior Data Engineer

🔥 59 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Allstate

Allstate

10,000+ employees

Founded 1931

💸 Finance

💰 Post-IPO Equity on 2014-01

Insurance • Finance

Allstate is an industry leader in providing insurance solutions, focusing on home, auto, device, and identity protection. With a commitment to customer well-being, Allstate aims to instill peace of mind and financial security for its customers. The company also emphasizes community impact and sustainability through various initiatives, showcasing their dedication to social responsibility and positive change.

📋 Description

• Design, build, and maintain scalable batch and streaming data pipelines using Apache Spark and cloud‑native data technologies • Develop and optimize ETL/ELT workflows to ingest, transform, and curate data from diverse source systems into analytics‑ready datasets • Implement data modeling and transformation logic to support reporting, dashboards, and downstream analytical and machine learning workloads • Build and manage data processing workloads within modern lakehouse platforms, including Microsoft Fabric / OneLake (preferred) • Ensure data quality, reliability, and consistency by implementing validation checks, monitoring, and reconciliation processes • Optimize Spark jobs for performance, cost efficiency, and scalability across large and complex datasets • Manage and evolve data schemas while handling schema drift and upstream source changes • Develop reusable frameworks, libraries, and standardized patterns to improve data engineering productivity and consistency • Implement CI/CD pipelines for data workloads to enable automated testing, deployment, and rollback • Monitor data pipelines and jobs, troubleshoot failures, and resolve performance or data quality issues • Partner with analytics engineers, BI developers, and data scientists to understand data requirements and deliver curated datasets • Collaborate with platform, security, and governance teams to ensure data security, compliance, and proper access controls • Contribute to Agile delivery processes, including sprint planning, design reviews, and continuous improvement initiatives

🎯 Requirements

• 4+ years of experience in data engineering or equivalent role (preferred) • Strong experience as a Data Engineer building and operating production data pipelines • Hands‑on experience with Apache Spark for large‑scale data processing • Proficiency in Python, SQL, and data transformation best practices • Experience with cloud‑based data platforms and storage (e.g., Data Lakes, Lakehouse architectures) • Familiarity with Microsoft Fabric, OneLake, or similar analytics platforms (strong plus) • Experience designing and optimizing data models for analytical workloads • Understanding of distributed data processing concepts, performance tuning, and fault tolerance • Experience with CI/CD, version control, and infrastructure‑as‑code concepts • Strong problem‑solving skills and ability to troubleshoot complex data issues • Excellent communication skills and ability to collaborate across technical and non‑technical teams

🏖️ Benefits

• Comprehensive technology setup including laptop, monitors, headset, keyboard, and mouse • Monthly connectivity reimbursement to help offset internet costs • 401(k) matching • Health insurance plans • Employee assistance program • Paid time off and holiday pay

Apply Now

Similar Jobs

🔥 1 hour ago

H&R Block

10,000+ employees

💸 Finance

👥 B2C

🤝 B2B

Senior Software Engineer at H&R Block designing scalable data platforms and automation frameworks. Collaborating to enable reliable, secure, and efficient data operations across the organization.

Azure

Cloud

ETL

Python

SQL

.NET

🔥 1 hour ago

Guidehouse

10,000+ employees

Data Engineer designing and deploying data pipelines and platforms for RCM workflows. Building scalable data systems for AI-enabled analytics and performance management.

AWS

Azure

Cloud

ETL

Google Cloud Platform

Python

Spark

SQL

🔥 1 hour ago

Guidehouse

10,000+ employees

Data Engineer focused on building and maintaining data pipelines, applications, and workflows in the Palantir platform. Collaborating with cross-functional teams for operational improvements and client solutions.

ERP

PySpark

🔥 2 hours ago

Aptive Resources

501 - 1000

🏛️ Government

Health IT Data Engineer managing data engineering and analytics for the ICE Health Service Corps. Collaborating with teams to design, develop, and maintain data pipelines in healthcare environments.

Cloud

ETL

Python

SQL

🔥 2 hours ago

People Data Labs

51 - 200

🔌 API

🤝 B2B

☁️ SaaS

Senior Software Engineer at People Data Labs contributing to data acquisition and processing platforms. Building large-scale distributed systems for data collection and ensuring high quality across datasets.

Airflow

Distributed Systems

DNS

ETL

Kafka

Linux

Python

Rust

Unix

Go