Senior Data Engineer

🕒 Yesterday

🇺🇸 United States – Remote

💵 $143.1k - $174.6k / year

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🦅 H1B Visa Sponsor

info
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Spokeo

Spokeo

51 - 200 employees

👥 B2C

☁️ SaaS

🔌 API

B2C • SaaS • API

<Spokeo> is a people-search and data aggregation service that lets users search by name, phone, email, or address to find contact information, location history, social profiles, property records, court records, and other public records. It combines billions of records from consumer and industry sources into concise, easy-to-read reports, offers account-based report updates, enterprise search capabilities, and an API for integration. Spokeo is positioned for individual consumers seeking to reconnect, identify callers, or verify sellers, while also offering paid services and tools for businesses; it states it is not an FCRA consumer reporting agency.

📋 Description

• Build infrastructure and data automation pipelines to ingest, process, and load data from various sources • Automate and integrate new components into the data pipeline • Collaborate with stakeholders and data science teams to develop data products, including entity resolution and best selection • Create unit and stress-test components to monitor technical performance and resolve identified issues • Develop data analysis tools to provide insights and capture key metrics • Research solutions and maintain technical documentation • Follow best practices for data governance, quality, cleansing, and other ETL-related activities • Develop, optimize, and improve data systems, including ETL pipelines, storage, and entity resolution • Build and improve data products, automation platform features, analytical software packages, and data pipeline orchestration tools

🎯 Requirements

• 7+ years of development experience in data engineering within a production environment, excluding internships and academic settings • Experience with large datasets exceeding 100M+ records or multiple terabytes • 5+ years of development experience in highly scalable, distributed systems and cluster architectures using AWS and EMR • 5+ years of hands-on programming experience with Python • 5+ years of professional experience in big data ecosystems • Spark required; PySpark preferred • 5+ years of experience with SQL, schema design, and dimensional data modeling • 5+ years of professional experience with dataflow orchestration tools such as Airflow • 2+ years of experience with non-relational databases such as DynamoDB or Elasticsearch • Bachelor’s degree in Computer Science, Information Systems, Mathematics, or a related field required

🏖️ Benefits

• Bonus program • Equity plans • 401(k) • Discretionary, merit-based annual salary increase • 100% medical coverage • 100% dental coverage • 100% vision coverage • Unlimited employee PTO • Remote-first work environment

Apply Now

Similar Jobs

🕒 Yesterday

SmartLight Analytics

1 - 10

🏥 Healthcare

💼 Consulting

⚕️ Healthcare Insurance

Senior Data Engineer modernizing Snowflake and cloud data infrastructure for SmartLight Analytics. Building secure healthcare-data pipelines, warehouse models, and migrations supporting lower-cost self-funded employer healthcare.

🕒 Yesterday

AssuranceAmerica

201 - 500

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Data Warehouse Engineer III modernizing AssuranceAmerica’s insurance data systems with Azure, SQL Server, ETL, and SSAS. Maintaining data pipelines, OLAP cubes, production jobs, and warehouse reliability.

🕒 Yesterday

ICF

5001 - 10000

🏥 Healthcare

📦 Logistics

📣 Marketing

Senior Data Engineer building scalable Spark, Hive, Airflow, and Databricks pipelines for ICF’s consulting and technology services. Developing secure AWS data infrastructure, APIs, testing, and visualizations.

🕒 Yesterday

ConnectWise

501 - 1000

☁️ SaaS

🔒 Cybersecurity

🏢 Enterprise

Enterprise Data Architect governing ConnectWise’s enterprise data ecosystem. Defining models, standards, integrations, and semantic layers for analytics, AI, and business operations.

🕒 Yesterday

Newsela

201 - 500

📚 Education

🛍️ eCommerce

👥 B2C

Data Engineer building K-12 integrations and pipelines for Newsela, an AI-powered education technology company. Ensuring accurate educational data flows across student information systems and classroom analytics platforms.