Data Engineer

Job not on LinkedIn

πŸ”₯ 11 minutes ago

Apply Now
Find Similar Remote Jobs

πŸ“Š Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Sparibis

Sparibis

11 - 50 employees

Founded 2018

πŸ’Ό Consulting

πŸ”’ Cybersecurity

🏒 Enterprise

Consulting β€’ Cybersecurity β€’ Enterprise

Sparibis is an Orlando-based IT engineering and consulting firm. It provides program management and project controls support along with SDLC leadership and systems engineering. Its listed capabilities include cyber security engineering, cloud and DevSecOps engineering, mobility/software engineering, CRM and enterprise portals, test and operations engineering, and data management.

πŸ“‹ Description

β€’ Plan, create, and maintain data architectures, ensuring alignment with business requirements. β€’ Obtain data, formulate dataset processes, and store optimized data. β€’ Identify problems and inefficiencies and apply solutions. β€’ Determine tasks where manual participation can be eliminated with automation. β€’ Identify and optimize data bottlenecks, leveraging automation where possible. β€’ Create and manage data lifecycle policies (retention, backups/restore, etc). β€’ In-depth knowledge for creating, maintaining, and managing ETL/ELT pipelines. β€’ Create, maintain, and manage data transformations. β€’ Maintain/update documentation. β€’ Create, maintain, and manage data pipeline schedules. β€’ Monitor data pipelines. β€’ Create, maintain, and manage data quality gates (Great Expectations) to ensure high data quality. β€’ Support AI/ML teams with optimizing feature engineering code. β€’ Expertise in Spark/Python/Databricks, Data Lake and SQL. β€’ Create, maintain, and manage Spark Structured Steaming jobs, including using the newer Delta Live Tables and/or DBT. β€’ Research existing data in the data lake to determine best sources for data. β€’ Create, manage, and maintain ksqlDB and Kafka Streams queries/code Data driven testing for data quality. β€’ Maintain and update Python-based data processing scripts executed on AWS Lambdas. β€’ Unit tests for all the Spark, Python data processing and Lambda codes. β€’ Maintain PCIS Reporting Database data lake with optimizations and maintenance (performance tuning, etc). β€’ Streamlining data processing experience including formalizing concepts of how to handle lake data, defining windows, and how window definitions impact data freshness.

🎯 Requirements

β€’ 5+ years of IT experience focusing on enterprise data architecture and management to include data flow charts, diagrams, and other technical documentation. β€’ Experience with Databricks, Structured Streaming, Delta Lake concepts, and Delta Live Tables required. β€’ Python development experience required. β€’ Experience with ETL and ELT tools such as SSIS, Pentaho, and/or Data Migration Services, and the ability to incorporate Python as required. β€’ Advanced level SQL experience (Joins, Aggregation, Windowing functions, Common Table Expressions, RDBMS schema design, Postgres performance optimization). β€’ Proficiency using Git for version control, including repository management, branching, merging, and pull requests. β€’ Active CompTIA Security+ certification preferred.

πŸ–οΈ Benefits

β€’ Applicants must be able to obtain and maintain a secret security clearance. β€’ United States Citizenship is required as part of the eligibility criteria to be able to obtain this type of security clearance.

Apply Now

Similar Jobs

πŸ”₯ 23 minutes ago

Praescient Analytics

51 - 200

πŸŽ–οΈ Defense

πŸ›οΈ Government

πŸ€– Artificial Intelligence

Senior Data Engineer providing leadership for data engineering under national security applications. Transforming legacy data structures into actionable intelligence within an AWS GovCloud environment.

AWS

Cloud

ETL

SQL

πŸ”₯ 23 minutes ago

Praescient Analytics

51 - 200

πŸŽ–οΈ Defense

πŸ›οΈ Government

πŸ€– Artificial Intelligence

Senior Data Architect providing enterprise-level technical leadership for federal background investigation data architectures at Praescient Analytics. Focusing on secure cloud environments and data management practices.

Amazon Redshift

AWS

Cloud

DynamoDB

Kafka

NoSQL

Oracle

Postgres

Spark

Vault

πŸ”₯ 38 minutes ago

AARP

1001 - 5000

🀝 Non-profit

πŸ₯ Healthcare

πŸ‘₯ B2C

Engineer II at AARP focusing on data platforms and technical solutions. Collaborating with cross-functional teams to deliver valuable technology-based business solutions.

πŸ‡ΊπŸ‡Έ United States – Remote

πŸ’΅ $144k - $160k / year

πŸ’° Grant on 2019-08

⏰ Full Time

🟑 Mid-level

🟠 Senior

🚰 Data Engineer

AWS

Cloud

PySpark

Python

SQL

πŸ”₯ 43 minutes ago

Imagineeer

11 - 50

πŸ›οΈ Government

πŸ”’ Cybersecurity

πŸ’Ό Consulting

Data Migration Specialist at Imagineeer supporting the Department of Health and Human Services HR IT migration. Assessing data quality and collaborating with teams for successful data transition.

Cloud

ETL

Informatica

MySQL

Oracle

Perl

Postgres

Python

SQL

SSIS

πŸ”₯ 50 minutes ago

Blue Cloud

2 - 10

πŸ’Ό Consulting

🀝 B2B

🏒 Enterprise

dbt Data Engineer for BlueCloud, focusing on designing, developing, and maintaining data pipelines and warehouses with dbt tools.

AWS

Azure

Cloud

Google Cloud Platform

SQL