Senior Data Engineer

🕒 August 9

🇺🇸 United States – Remote

💵 $143.1k - $174.6k / year

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🦅 H1B Visa Sponsor

infoinfo

👻 Ghost score 27%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Spokeo

Spokeo

51 - 200 employees

👥 B2C

☁️ SaaS

🔌 API

B2C • SaaS • API

<Spokeo> is a people-search and data aggregation service that lets users search by name, phone, email, or address to find contact information, location history, social profiles, property records, court records, and other public records. It combines billions of records from consumer and industry sources into concise, easy-to-read reports, offers account-based report updates, enterprise search capabilities, and an API for integration. Spokeo is positioned for individual consumers seeking to reconnect, identify callers, or verify sellers, while also offering paid services and tools for businesses; it states it is not an FCRA consumer reporting agency.

📋 Description

• Build infrastructure and data automation pipelines to ingest, process, and load data from various sources • Automate and integrate new components into the data pipeline • Collaborate with stakeholders and data science teams to develop data products, including entity resolution and best selection • Execute product vision and strategy in alignment with organizational goals and priorities • Create unit and stress-test components to monitor technical performance and ensure identified issues are resolved • Develop data analysis tools to provide data insights and capture key metrics • Research solutions and maintain technical documentation • Follow best practices for data governance, quality, cleansing, and other ETL-related activities • Develop, optimize, and improve data systems, including ETL pipelines, storage, and entity resolution • Build and improve data products, automation platform features, analytical software packages, and data pipeline orchestration tools

🎯 Requirements

• 7+ years of development experience in data engineering within a production environment (internships and academic settings excluded) • Experience working with large datasets exceeding 100M+ records or multiple terabytes • 5+ years of development experience in highly scalable, distributed systems and cluster architectures using AWS and EMR • 5+ years of hands-on programming experience with Python • 5+ years of professional experience working in big data ecosystems • Spark required; PySpark preferable • 5+ years of experience with SQL, schema design, and dimensional data modeling • 5+ years of professional experience working with dataflow orchestration tools such as Airflow • 2+ years of experience with non-relational databases such as DynamoDB and Elasticsearch • Bachelor's degree in Computer Science, Information Systems, Mathematics, or a related field required

🏖️ Benefits

• Bonus program • Equity plans • 401(k) • Discretionary, merit-based annual salary increase • 100% medical, dental, and vision coverage • Unlimited employee PTO • Remote-first team • Equal opportunity employment

Apply Now

Similar Jobs

🕒 August 8

Oddball

51 - 200

💼 Consulting

📦 Logistics

🎖️ Defense

Cloud Data Architect designing Azure and Databricks platforms for Oddball’s federal Veteran data systems. Leading modernization, governance, security, and AI/ML enablement.

🕒 August 8

Snowflake

5001 - 10000

☁️ SaaS

🏢 Enterprise

🤝 B2B

Data Engineering sales specialist driving Snowflake platform adoption and consumption revenue. Advising enterprise customers on Openflow, Iceberg/Polaris, ingestion, and Data Lakehouse capabilities.

🕒 August 8

Redwood Logistics

1001 - 5000

🚗 Transport

📦 Logistics

Lead Data Engineer leading scalable data pipelines, Data Vault, warehouses, and marts for Redwood Logistics’ supply-chain technology platform. Mentoring engineers and delivering operational metrics.

🕒 August 7

Cortex by Palo Alto Networks

51 - 200

🔒 Cybersecurity

🤖 Artificial Intelligence

Senior Data Engineer owning Cortex’s ingestion, transformation, warehousing, and product analytics. Building a trusted data foundation for AI-driven engineering operations and reporting.

🕒 August 7

General Dynamics Information Technology

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Lead Data Engineer building secure cloud data platforms and pipelines for GDIT’s federal court modernization program. Leading migration, governance, analytics, and engineering delivery.