Mid-Level Data Engineer

Stelle nicht auf LinkedIn

🕒 vor 1 Monat

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟡 Mittelstufe

🟠 Senior

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

Airflow

Amazon Redshift

Apache

AWS

ETL

Oracle

Postgres

PySpark

Python

Spark

SQL

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Simple Technology Solutions

Simple Technology Solutions

51 - 200 Mitarbeiter

🏛️ Regierung

🤖 Künstliche Intelligenz

Government • Cloud • Artificial Intelligence

Simple Technology Solutions ist ein kleines Unternehmen in einer HUBZone, das sich auf die IT-Modernisierung und digitale Erfahrungen für Regierungsoperationen spezialisiert hat. Es konzentriert sich darauf, Regierungsprozesse mithilfe von Cloud-nativen Technologien und agilen Praktiken zu digitalisieren und Full-Stack-Digitallösungen bereitzustellen. Das Unternehmen legt Wert auf Sicherheit, Skalierbarkeit und Interoperabilität in seinem Unternehmensansatz. Es arbeitet daran, Cloud-Umgebungen zu verbessern, Altsysteme zu migrieren und DevSecOps-Praktiken zu fördern. Darüber hinaus entwickelt Simple Technology Solutions Unternehmensdatenmanagementstrategien, die maschinelles Lernen und künstliche Intelligenz nutzen, modernisiert Anwendungen und bietet Cloud-Kontaktzentren-Services an. Sie bedienen vor allem Bundesbehörden, insbesondere in den Bereichen Strafverfolgung und öffentliche Sicherheit.

Beschreibung

• Develop new ETL pipelines and data ingestion processes alongside senior engineers using AWS Glue (Spark-based, PySpark), MWAA (Airflow), Lambda, and SNS • Integrate the agency's ETL Common Library into Glue jobs for standardized orchestration, error handling, metadata recording, and SNS notifications for all success and error job events • Ingest structured and semi-structured datasets (CSV, XML, JSON, Avro, pipe-delimited) into S3 landing, raw, and curated zones using Apache Iceberg tables • Configure static ETL metadata in the centralized PostgreSQL metadata store; ensure dynamic metadata records job status and timestamps for all key execution steps • Monitor assigned production jobs and participate in operations support rotations • Ensure ETL Load Reports are populated in real-time and ETL Gap Reports are updated on a weekly basis • Build and maintain materialized views and semantic layer objects in Trino and Athena to ensure optimized query performance and consistent business logic • Produce and maintain required documentation for each assigned dataset: Business Requirements, ETL Design Documents, Data Models, Data Dictionaries, Mapping Documents, Deployment Documents, O&M Guides, and ETL Test Plans • Write unit and integration tests achieving the 90% minimum code coverage threshold; complete security scans at least once per sprint • Deploy ETL resources using CloudFormation templates through the agency CICD pipeline • Support transition of ETL jobs from other agency teams and disaster recovery exercises

🎯 Anforderungen

• US Citizenship is required • Bachelor's Degree is required • minimum of 3-5 years' position related experience is required • Hands-on experience with Python (PEP 8), PySpark, and SQL for ETL pipeline development • Experience with AWS services including Glue, S3, MWAA (Airflow), Lambda, SNS, and SQS • Familiarity with Apache Iceberg, Parquet, and ORC file formats and S3 data lake zone concepts • Experience with PostgreSQL and basic familiarity with Redshift or Oracle • Familiarity with Trino or Athena for query and semantic layer development • Experience with CloudFormation, GitHub branching workflows, and CI/CD-integrated deployments • Ability to produce clear ETL documentation including data models (Mermaid format) and data dictionaries • Understanding of ETL metadata concepts including static and dynamic metadata, load reports, and gap reports • Experience in agile development environments with sprint-based delivery • Experience supporting IV&V and/or User Acceptance Testing (UAT) processes in a federal or technical program environment • Experience with automated testing frameworks; ability to write unit and integration tests achieving defined code coverage thresholds • Familiarity with FISMA, NIST 800-53, and OWASP ASVS Level 2 is a plus • Must be able to work 8am-5pm Eastern Time regardless of home location • Active federal public trust suitability determination or ability to obtain one required

🏖️ Vorteile

• Flexible work arrangements • Continuous learning • Professional development • Special incentives for team members living in qualified HUBZones

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

Texas Windstorm Insurance Association

201 - 500

🤝 Non-Profit

🏛️ Regierung

Data Warehouse Developer leveraging expertise in Data Warehousing and Business Intelligence to design resilient data solutions for TWIA/TFPA. Collaborating across teams to transform data into meaningful insights for decision-making.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟠 Senior

🔴 Experte

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

ETL

Guidewire

Informatica

SDLC

SSIS

Tableau

🕒 vor 1 Monat

Samsara

1001 - 5000

🏢 Unternehmen

🚗 Transport

🔐 Sicherheit

Senior Data Engineer developing scalable data pipelines for IoT systems at Samsara. Designing data models and collaborating with cross-functional teams to enhance data analysis efficiency.

🇺🇸 Vereinigte Staaten – Remote

💵 $119.595 - $201.000 / Jahr

💰 Seed Round im 2014-08

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🦅 H1B-Visum-Sponsor

info

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

AssistRx

501 - 1000

⚕️ Krankenversicherung

💊 Pharmazie

☁️ SaaS

Senior Manager leading teams in data engineering for scalable data solutions at AssistRx. Engaging with stakeholders to ensure successful project delivery and team development.

🗣️🇺🇸🇬🇧 Englisch erforderlich

AWS

Azure

Cloud

ETL

Hadoop

Informatica

Spark

SQL

Vault

🕒 vor 1 Monat

NikoHealth

51 - 200

⚕️ Krankenversicherung

☁️ SaaS

Data Migration Analyst at NikoHealth overseeing transitions from legacy platforms into SaaS solutions. Ensure data accuracy and provide customer support during the migration process.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟢 Junior

🟡 Mittelstufe

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Egen

501 - 1000

🤖 Künstliche Intelligenz

Lead Data Engineer at Egen focusing on building and optimizing cloud-native data platforms on Google Cloud. Mentoring engineers and designing sustainable data architectures with a data-first mindset.

🇺🇸 Vereinigte Staaten – Remote

💵 $143.400 - $168.650 / Jahr

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🦅 H1B-Visum-Sponsor

info

🗣️🇺🇸🇬🇧 Englisch erforderlich