Fullstack Data Engineer

🕒 il y a 3 mois

🇺🇸 États-Unis – Télétravail

⏰ Temps Plein

🟠 Senior

🔴 Expert

🚰 Ingénieur Data

🗣️🇺🇸🇬🇧 Anglais requis

AWS

ETL

Flask

PySpark

Python

Spark

Unity

Postuler Maintenant
Trouver des Emplois à Distance Similaires

📊 Vérifiez votre score de CV pour ce poste

Améliorez vos chances d'obtenir un entretien en vérifiant votre score de CV avant de postuler.

Logo of Codvo.ai

Codvo.ai

51 - 200 employés

Fondée en 2019

🤖 Intelligence artificielle

🔒 Cybersecurity

☁️ SaaS

Artificial Intelligence • Cybersecurity • SaaS

Codvo. ai est une entreprise technologique spécialisée dans la fourniture de solutions d'entreprise stratégiques grâce à une innovation avancée pilotée par l'IA. Elles se concentrent sur la transformation des données d'entreprise en valeur mesurable en aidant les entreprises à accélérer leur croissance avec des implémentations IA sur mesure adaptées aux défis spécifiques de diverses industries. Leurs offres de services étendues incluent l'automatisation IA/ML, le développement d'applications, l'analyse de données, la cybersécurité et la transformation numérique, garantissant que les organisations peuvent prospérer dans un paysage numérique en évolution rapide.

Description

• Design, build, and maintain Databricks data pipelines (ETL/ELT) for ingestion, transformation, and orchestration using Spark/Delta Lake/Databricks Workflows. • Operationalize machine learning models by building inference pipelines that invoke models authored by data scientists (batch or real-time), ensuring consistency between training and inference environments. • Ensure data reliability, quality, and observability through robust validation, monitoring, alerting, and automated recovery mechanisms. • Collaborate closely with data scientists to productionize models, manage model deployment lifecycles, and optimize inference performance and cost. • Implement best-practice DevOps/MLOps processes such as CI/CD for pipelines, model versioning, environment promotion, and infrastructure-as-code. • Optimize performance and cost across compute clusters, jobs, and storage layers. • Implement and manage the enterprise data catalog, including schema design, table ownership, lineage, governance, and documentation using Unity Catalog. • Experience with some Databricks infrastructure. • Experience with building BI dashboards and visualization. • Experience with coding agents and best practices (spec-driven development, etc.)

🎯 Exigences

• 8+ yrs experience • Databricks platform experience • Python development for data processing and ETL pipelines • Unity Catalog knowledge • AWS data services (S3, IAM, VPC, potentially Glue/Lambda) • Data lake/lakehouse architecture patterns • Dashboard building experience • RESTful API design and development (Flask, FastAPI, or similar) • Authentication/authorization patterns (OAuth, API keys, IAM roles) • Query optimization and performance tuning • PySpark optimization experience • ML/AI pipeline experience • Databricks AI/BI

Postuler Maintenant

Emplois Similaires

🕒 il y a 3 mois

Blend360

501 - 1000

🏥 Santé

🏨 Hôtellerie

✈️ Tourisme

Director of AI Engineering leading AI model development and team management at Blend. Overseeing technical strategy and cross-functional collaboration for impactful AI solutions.

🇺🇸 États-Unis – Télétravail

💵 $180 000 - $240 000 / an

💰 €100 000 000 Private Equity Round en 2022-08

⏰ Temps Plein

🔴 Expert

🚰 Ingénieur Data

🦅 Parrain de Visa H1B

info

🗣️🇺🇸🇬🇧 Anglais requis

Azure

BigQuery

Cloud

Python

PyTorch

Scikit-Learn

Spark

SQL

Tensorflow

🕒 il y a 3 mois

Blend360

501 - 1000

🏥 Santé

🏨 Hôtellerie

✈️ Tourisme

Director of AI Engineering at a leading AI services company. Overseeing AI model development, team management, and client engagements.

🇺🇸 États-Unis – Télétravail

💰 €100 000 000 Private Equity Round en 2022-08

⏰ Temps Plein

🔴 Expert

🚰 Ingénieur Data

🦅 Parrain de Visa H1B

info

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 3 mois

Cross Screen Media

11 - 50

💼 Conseil

🏥 Santé

🚘 Automobile

Data Engineer responsible for designing and developing data pipelines for video advertising campaigns. Collaborating with a senior data engineer and a high-output team to shape data architecture.

🇺🇸 États-Unis – Télétravail

💰 Venture Round en 2018-01

⏰ Temps Plein

🟡 Intermédiaire

🟠 Senior

🚰 Ingénieur Data

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 3 mois

Pinterest

1001 - 5000

📱 Médias

👥 B2C

Staff Data Engineer at Pinterest leading design and implementation of identity services and data governance. Collaborating across teams to ensure trusted and privacy-safe data handling.

🇺🇸 États-Unis – Télétravail

💵 $177 185 - $364 795 / an

💰 Post IPO equity en 2022-08

⏰ Temps Plein

🔴 Expert

🚰 Ingénieur Data

🦅 Parrain de Visa H1B

info

🗣️🇺🇸🇬🇧 Anglais requis

🕒 il y a 3 mois

EITACIES Inc.

51 - 200

💼 Conseil

🏥 Santé

🏭 Fabrication

Experienced Data Engineer supporting large-scale Oracle to Snowflake migration at Eitacies Inc. Collaborating with onsite and offshore teams in a fully remote role.

🇺🇸 États-Unis – Télétravail

💵 $60 / heure

⏰ Temps Plein

🟠 Senior

🚰 Ingénieur Data

🗣️🇺🇸🇬🇧 Anglais requis