Data Engineer

Job not on LinkedIn

🔥 0 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of SumerSports

SumerSports

11 - 50 employees

Founded 2022

⚽ Sports

🤖 Artificial Intelligence

☁️ SaaS

Sports • Artificial Intelligence • SaaS

SumerSports is an AI-powered sports analytics and technology company focused on football (NFL and NCAA). Combining over 500 years of NFL experience with machine learning, SumerSports offers products such as SūmerBrain for film retrieval and multi-layered data, SūmerLive for game tracking, SūmerNFL and SūmerNCAA for roster building and team optimization, and a player-verified metrics and talent exposure platform. The company produces draft guides, analytics-driven content with former scouts and Hall of Famers, and tools that serve players, teams, and fans to improve scouting, roster decisions, and performance evaluation.

📋 Description

• Build and operate robust data pipelines for ingestion, cleaning, and transformation using Databricks, Airflow, or Kubernetes • Develop efficient ETL/ELT workflows in Python and SQL for batch and streaming workloads • Partner with ML/AI teams to make datasets and tools discoverable and safe for autonomous agents, including evaluation and guardrails for AI-generated queries • Develop retrieval pipelines (RAG, vector search) over structured statistics and unstructured sources such as scouting notes and video metadata • Model and maintain structured data assets in Delta, Parquet, and Iceberg for reliability, versioning, and lineage tracking • Implement orchestration and monitoring by scheduling jobs, tracking dependencies, and automating failure recovery • Ensure data quality and compliance through validation frameworks, schema enforcement, and audit logging • Contribute to data platform evolution by evaluating tools, standardizing best practices, and improving developer experience • Support performance and cost optimization across compute, storage, and orchestration systems • Collaborate with MLOps and Sports Data teams to integrate data and AI systems

🎯 Requirements

• 3–8 years of experience as a Data Engineer or ETL Developer in a production environment • Proficiency in Python and SQL • Strong familiarity with Databricks, Spark, or equivalent big-data frameworks • Experience with workflow orchestration tools such as Airflow, Dagster, Luigi, or Prefect • Deep understanding of data modeling, data warehousing, and distributed data processing • Knowledge of modern data lakehouse architectures • Familiarity with CI/CD, GitHub Actions, Infrastructure as Code, and data pipeline testing frameworks • Comfort working cross-functionally with ML, product, and analytics teams • Exposure to LLM-powered data tools, including text-to-SQL, RAG, agent/tool interfaces such as MCP, or natural-language analytics • Previous work with cloud infrastructure such as AWS, GCP, or Azure • Experience with container orchestration using Docker or Kubernetes • Preferred: previous experience with sports, telemetry, or sensor data pipelines • Preferred: familiarity with streaming frameworks and event-driven data processing such as Kafka, Spark Structured Streaming, or Flink • Preferred: general knowledge of American football, the NFL, and college football • Preferred: background in data governance, lineage, and observability tools such as Monte Carlo, Great Expectations, Unity Catalog, or OpenLineage • Preferred: experience designing semantic layers or metric definitions consumed by AI and BI tools • Preferred: exposure to machine-learning model management and MLOps best practices

🏖️ Benefits

• Competitive Salary and Bonus Plan • Comprehensive health insurance plan • Retirement savings plan (401k) with company match • Remote working environment • A flexible, unlimited time off policy • Generous paid holiday schedule - 13 in total including Monday after the Super Bowl • Annual performance bonus • Benefits and/or other applicable incentive compensation plans

Apply Now

Similar Jobs

🔥 13 minutes ago

Runpod

51 - 200

🤖 Artificial Intelligence

☁️ SaaS

🤝 B2B

Senior Data Engineer building scalable data pipelines and warehousing for Runpod’s AI Developer Cloud. Partnering across analytics, finance, data science, and engineering to power secure, data-driven decisions.

🇺🇸 United States – Remote

💵 $175k - $220k / year

💰 $20M Seed on 2024-06

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🔥 1 hour ago

TeamWorx Security

11 - 50

🎖️ Defense

🔒 Cybersecurity

☁️ SaaS

Senior Data Engineer building secure Databricks data platforms and pipelines for TeamWorx Security’s federal government clients. Delivering governed data products supporting analytics, automation, AI/ML, and operational decisions.

🇺🇸 United States – Remote

💵 $135k - $165k / year

💰 $500k Venture Round - TeamWorx Security on 2025-06

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🔥 3 hours ago

YPO

201 - 500

🤝 B2B

☁️ SaaS

📣 Marketing

Data Engineer building YPO’s data architecture, pipelines, and analytics for its global executive membership community. Supporting member-facing services, data products, and insights across international operations.

🔥 3 hours ago

8th Light

51 - 200

💼 Consulting

📣 Marketing

Lead Data Engineer building scalable data pipelines and platforms for 8th Light, a technology solutions consultancy. Guiding client data architecture, modeling, and delivery across cloud technologies.

🔥 3 hours ago

Quisitive

501 - 1000

🏥 Healthcare

🏭 Manufacturing

📦 Logistics

Data Engineer building scalable Azure pipelines and Microsoft Fabric solutions for Quisitive, a global Microsoft cloud partner. Automating infrastructure, CI/CD deployments, and secure modern data platform operations.