Data Engineer

Job not on LinkedIn

🕒 July 10

Airflow

Amazon Redshift

Apache

BigQuery

Cloud

ETL

Python

SQL

Tableau

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Troveo AI

Troveo AI

11 - 50 employees

Founded 2024

🤖 Artificial Intelligence

📱 Media

🤝 B2B

Artificial Intelligence • Media • B2B

Troveo AI is a company that curates, licenses, and delivers large-scale, rights-cleared video datasets engineered for training computer vision and multimodal AI models. They work directly with content creators and global licensors to produce annotated, training-ready clips with custom metadata and legal compliance, serving enterprise AI labs and platforms that need high-quality video for model development (e. g. , avatars, action recognition, and camera-motion-aware models). Troveo operates as a B2B data provider and platform focused on media-originated training data.

📋 Description

• Design, build, and maintain robust, scalable ELT/ETL data pipelines (batch and streaming) from various source systems into cloud data platforms and warehouses. • Optimize pipelines for performance, cost, reliability, and scalability. • Design and implement conceptual, logical, and physical data models (including dimensional modeling, star/snowflake schemas). • Build and maintain transformation layers using modern tools (e.g., dbt) to create clean, well-documented, analytics-ready datasets. • Write optimal SQL queries for data exploration, ad-hoc analysis, and troubleshooting. • Support the creation of reports, dashboards, and self-service analytics assets in collaboration with data analysts and business teams. • Translate business questions into data requirements and deliver actionable insights or datasets. • Monitor data pipelines and data delivery processes to ensure SLAs for timeliness, freshness, and accuracy are consistently met. • Proactively identify, troubleshoot, and resolve data issues impacting downstream consumers or business operations. • Manage incidents related to data availability and quality; participate in on-call rotations as needed. • Document data pipelines, models, lineage, and processes.

🎯 Requirements

• 7+ years of professional experience in data engineering or a closely related role (analytics engineering experience is highly relevant). • Strong proficiency in SQL and Python. • Hands-on experience building and maintaining data pipelines and working with cloud data platforms/warehouses. (Snowflake, BigQuery, Redshift, Databricks, etc.). • Experience with data orchestration tools (Apache Airflow, Dagster, Prefect, or similar). • Solid understanding of data modeling techniques and dimensional modeling. • Experience performing data analysis and working with BI/visualization tools (Looker, Tableau, Power BI, or similar). • Proven ability to troubleshoot data issues and support operational reliability/SLAs. • Strong communication skills and ability to collaborate with both technical and non-technical stakeholders.

🏖️ Benefits

• Comprehensive Health Benefits: Medical, dental, and vision coverage (100% employer-paid for employees) • Flexible PTO & Paid Holidays: Unlimited PTO with encouragement to actually use it • Remote First Policy: Work from anywhere in the US (with occasional team offsites) • Learning & Growth: Annual learning stipend, access to top conferences, and direct mentorship from experienced founders • Equity Ownership: Competitive equity package with clear growth potential as we scale • Modern Tech Stack & Tools: Budget for the best equipment and software • Strong Culture: High-trust, low-ego environment focused on impact, transparency, and work-life balance. We believe great work happens when people are supported, challenged, and given ownership.

Apply Now

Similar Jobs

🕒 July 10

Unisys

10,000+ employees

🤖 Artificial Intelligence

🔒 Cybersecurity

Sr. Data Lake Support Engineer supporting database operations, ETL processes, and Azure services. Collaborating with teams to ensure systematic reliability and performance monitoring.

🇺🇸 United States – Remote

💰 Post-IPO Debt on 2020-10

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🦅 H1B Visa Sponsor

info

🕒 July 10

BetMGM

501 - 1000

🎲 Gambling

🎮 Gaming

👥 B2C

Senior Data Engineer responsible for building and maintaining data pipelines for analytics at BetMGM. Collaborating with various teams while leveraging modern tech stacks for operational success.

🇺🇸 United States – Remote

💵 $135k - $170k / year

💰 $25.1k Seed Round on 2022-10

⏰ Full Time

🟠 Senior

🚰 Data Engineer

🦅 H1B Visa Sponsor

info

🕒 July 10

Foodsmart

51 - 200

⚕️ Healthcare Insurance

🧘 Wellness

🏪 Marketplace

Staff Software Engineer with strong backend focus at Foodsmart, leading data platform architecture and initiatives. Driving personalized nutrition journeys and ensuring data reliability and accessibility.

🕒 July 10

Data Engineer designing and building scalable data pipelines for Vibrant Planet's platform. Collaborating with engineering and science teams to manage and deliver geospatial data solutions.

🇺🇸 United States – Remote

💵 $100k - $200k / year

💰 $17M Seed Round on 2022-06

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

🕒 July 10

Sayari

1 - 10

🔬 Science

🤝 B2B

Data Engineer at Sayari building scalable data pipelines using Python and Spark. Collaborating with AI/ML teams to enhance features and optimize data processes.

🇺🇸 United States – Remote

💵 $90k - $120k / year

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer