Lead Data Architect

🕒 vor 1 Monat

🇺🇸 Vereinigte Staaten – Remote

💵 $181.026 - $259.094 / Jahr

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🦅 H1B-Visum-Sponsor

info

🗣️🇺🇸🇬🇧 Englisch erforderlich

Airflow

Amazon Redshift

Apache

AWS

Azure

BigQuery

Cassandra

Cloud

ETL

Google Cloud Platform

GraphQL

Informatica

Kafka

MongoDB

MySQL

Neo4j

NoSQL

Numpy

Pandas

Postgres

Pulsar

PySpark

Python

Spark

SQL

Terraform

Jetzt Bewerben
Ähnliche Remote-Jobs finden

📊 Überprüfen Sie Ihre Lebenslauf-Bewertung für diese Stelle

Verbessern Sie Ihre Chancen auf ein Vorstellungsgespräch, indem Sie Ihre Lebenslauf-Bewertung vor der Bewerbung überprüfen.

Logo of Henry Schein

Henry Schein

10.000+ Mitarbeiter

Gegründet 1932

⚕️ Krankenversicherung

💊 Pharmazie

🤝 B2B

Healthcare Insurance • Pharmaceuticals • B2B

Henry Schein, Inc. (Nasdaq: HSIC) ist ein Lösungsanbieter für Gesundheitsfachkräfte, der durch ein Netzwerk von Menschen und Technologie unterstützt wird. Mit mehr als 25.000 Team Schein Mitgliedern weltweit bietet das Unternehmensnetzwerk von vertrauenswürdigen Beratern über 1 Million Kunden weltweit mehr als 300 bewährte Lösungen, die den operativen Erfolg und die klinischen Ergebnisse verbessern. Unsere Geschäfts-, klinischen, Technologie- und Lieferkettenlösungen helfen Zahnärzten und niedergelassenen Medizinern, effizienter zu arbeiten, sodass sie qualitativ hochwertige Pflege wirksamer bereitstellen können. Diese Lösungen unterstützen auch Dentallabore, staatliche und institutionelle Gesundheitskliniken sowie andere alternative Pflegeeinrichtungen.

Beschreibung

• Define and implement a scalable, enterprise-wide data architecture aligned with business and technology goals • Develop a data strategy roadmap, ensuring long-term sustainability, scalability, and efficiency • Partner with executive leadership, product teams, and engineering to ensure data initiatives drive business value • Establish enterprise data governance, security, and compliance frameworks leveraging tools like Collibra or Alation • Oversee the design and evolution of data lakes, data warehouses, and cloud-based analytics platforms using Databricks, Snowflake, BigQuery, or Redshift • Lead the adoption of modern data architecture patterns, including event-driven architectures, real-time data streaming (Kafka, Pulsar), and AI-driven analytics • Provide guidance on database optimization, indexing, partitioning, and storage strategies for tools like PostgreSQL, MySQL, and NoSQL solutions like MongoDB or Cassandra • Evaluate emerging technologies, making recommendations for tools and platforms that enhance data capabilities • Direct ETL/ELT strategies, ensuring seamless data flow across systems with Python, Apache Airflow, dbt, or Informatica • Architect cloud-based solutions (AWS, Azure, or GCP) using services such as AWS Glue, Azure Synapse, and Google Cloud Dataflow to support analytics, AI, and operational use cases • Ensure API-first design for data integration using GraphQL, RESTful APIs, or event-driven architectures (Kafka, AWS Kinesis, Pub/Sub) • Define and oversee data quality, lineage, and cataloging efforts using Great Expectations, Monte Carlo, or DataHub • Develop policies for data privacy, access control, and encryption, ensuring compliance with GDPR, CCPA, HIPAA, or other relevant regulations • Implement enterprise-wide metadata management and data lineage tracking using Collibra, Alation, or Data Catalog solutions • Drive best practices for data security and compliance audits, leveraging IAM tools and cloud security solutions • Lead a team of data architects, engineers, and analysts, mentoring them on best practices • Act as a liaison between business and technical teams, translating business needs into scalable data solutions • Champion a culture of innovation, ensuring the data team is adopting cutting-edge methodologies • Conduct data architecture reviews, ensuring alignment with organizational standards

🎯 Anforderungen

• 10+ years of experience in data architecture, data engineering, or related fields • Bachelor’s degree (Master’s preferred) in Computer Science, Applied Mathematics, Statistics, Machine Learning, or a closely related field (or foreign equivalent) • Proven track record in designing large-scale, enterprise data architectures • Expertise in SQL, NoSQL, and distributed database technologies such as Snowflake, Databricks, BigQuery, Redshift, PostgreSQL, MongoDB, and Cassandra • Strong experience with cloud-based data platforms (AWS, Azure, GCP) and services like AWS Glue, Azure Data Factory, and Google Dataflow • Deep understanding of data modeling, ETL/ELT processes, and data pipeline optimization using dbt, Apache Airflow, Informatica, or Talend • Experience with real-time streaming technologies (Kafka, Spark Streaming, Apache Flink, AWS Kinesis) • Strong knowledge of data security, governance, and compliance frameworks • Excellent verbal and written communication skills and ability to resolve disputes effectively and efficiently • Outstanding presentation and public speaking skills • Mastery independent decision making, analysis and problem-solving skills • Ability to quickly understand and assess complex projects, systems and ecosystems and identify relevant relationships and connections between them • Mastery planning and organizational skills and techniques • Communicate effectively with senior management and key stakeholders • Ability to influence, build relationships, understand organizational complexities, manage conflict and navigate politics • Familiarity with the healthcare data domain with previous experience working with healthcare datasets is a plus • Strong Python programming skills, with expertise in data manipulation and pipeline development using Pandas, PySpark, NumPy, and SQLAlchemy • Experience with AI/ML-driven analytics architectures and MLOps frameworks like MLflow or SageMaker • Hands-on experience with Infrastructure as Code (Terraform, CloudFormation) • Familiarity with Graph databases and knowledge graphs (Neo4j, Amazon Neptune) • Certifications in cloud data services (AWS Certified Data Analytics, Google Professional Data Engineer, Databricks Certified Data Engineer)

🏖️ Vorteile

• Medical, Dental and Vision Coverage • 401K Plan with Company Match • PTO • Paid Parental Leave • Income Protection • Work Life Assistance Program • Flexible Spending Accounts • Educational Benefits • Worldwide Scholarship Program • Volunteer Opportunities

Jetzt Bewerben

Ähnliche Jobs

🕒 vor 1 Monat

Lamb Weston

5001 - 10000

🛒 Einzelhandel

🌾 Landwirtschaft

Lead Data Architect responsible for Lamb Weston’s enterprise data ecosystem leveraging Snowflake. Collaborate with teams to optimize data integration and deliver analytics solutions.

🇺🇸 Vereinigte Staaten – Remote

💵 $146.840 - $220.250 / Jahr

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🦅 H1B-Visum-Sponsor

info

🗣️🇺🇸🇬🇧 Englisch erforderlich

AWS

Informatica

Matillion

SQL

🕒 vor 1 Monat

Kobie

201 - 500

Senior Data Engineer specializing in Snowflake and data pipelines at Kobie, enhancing loyalty solutions. Collaborate on event-driven and AI-powered workflows for top brands.

🇺🇸 Vereinigte Staaten – Remote

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Medecision

201 - 500

⚕️ Krankenversicherung

Senior Software Engineer responsible for building and evolving cloud-native data services at Medecision. Focusing on data pipelines within a regulated healthcare environment to enhance clinical analytics.

🇺🇸 Vereinigte Staaten – Remote

💵 $130.000 - $155.000 / Jahr

💰 Series C im 2003-03

⏰ Vollzeit

🟠 Senior

🚰 Dateningenieur

🦅 H1B-Visum-Sponsor

info

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Oyster

501 - 1000

🤝 B2B

👥 HR Tech

☁️ SaaS

Senior Director of Data Platform and AI driving AI initiatives and infrastructure transformation at Oyster. Overseeing data systems to optimize workflows and process efficiency across the organization.

🗣️🇺🇸🇬🇧 Englisch erforderlich

🕒 vor 1 Monat

Sedgwick

10.000+ Mitarbeiter

🏢 Unternehmen

📋 Compliance

Senior Data Engineer at Sedgwick designing and implementing ETL pipelines for AI. Collaborating with Data Science teams to ensure data integrity and availability for advanced analytics.

🗣️🇺🇸🇬🇧 Englisch erforderlich

Airflow

AWS

Azure

ETL

PySpark

Python

SQL