Data Architect, AI

Job not on LinkedIn

🔥 1 hour ago

🇲🇽 Mexico – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 25%

infoinfo

🗣️🇪🇸 Spanish Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of VALCE Talent Solutions

VALCE Talent Solutions

11 - 50 employees

Founded 2016

🤝 B2B

🎯 Recruiter

💼 Consulting

B2B • Recruitment • Consulting

VALCE Talent Solutions is a company specializing in nearshoring, IT talent acquisition, and consultancy aimed at helping businesses expand globally, particularly in Mexico, LATAM, and the United States. They offer customized solutions that include talent recruitment, workforce management, and the integration of technology and artificial intelligence to enhance business processes. With a strong focus on strategic consulting, VALCE aims to connect technology, talent, and tangible results, boasting a robust network of IT professionals and a proven track record of successful placements and project scaling.

📋 Description

• Design end-to-end data architectures on the Databricks Lakehouse Platform • Build reliable, production-ready Bronze, Silver, and Gold data pipelines • Troubleshoot and tune large-scale Apache Spark jobs • Design governance models for data and AI assets • Implement role-based, row-level, and column-level access controls and data lineage • Work with AWS, Azure, or GCP cloud infrastructure • Optimize Databricks clusters, DBU usage, performance, and cloud costs • Potentially build, deploy, and monitor GenAI applications and machine learning models • Automate Databricks workflows and support streaming data ingestion and downstream BI integration

🎯 Requirements

• Expertise in designing end-to-end data architectures using Bronze, Silver, and Gold layers • Deep understanding of distributed computing and Databricks/Spark internals • Ability to troubleshoot and tune large-scale Spark jobs using caching, partitioning, and broadcast joins • Understanding of Delta Lake ACID transactions, schema enforcement, time travel, and Z-ordering • Ability to design unified governance models for data and AI assets, including RBAC, row-level security, column-level security, and data lineage • Strong grasp of AWS, Azure, or GCP native cloud ecosystems, including VNet setups and IAM roles • Advanced SQL skills • Fluency in Python or Scala • Ability to monitor DBUs, right-size serverless and multi-node clusters, and optimize cloud costs • Advanced English • Databricks certifications, GenAI/MLflow, CI/CD and DevOps, streaming data, and BI experience are nice to have, not required

Apply Now

Similar Jobs

🕒 August 29

Enroute

51 - 200

💼 Consulting

📦 Logistics

🏭 Manufacturing

Data Engineer building Snowflake pipelines, data models, and Tableau dashboards for Enroute. Enabling reliable business reporting through ETL/ELT, SQL, and data-quality practices.

Cloud

ETL

SQL

Tableau

🕒 August 28

Stefanini LATAM

10,000+ employees

💼 Consulting

📦 Logistics

📣 Marketing

Senior Graph Data Engineer building Neo4j graph data layers for Stefanini Latam. Designing models, integrating sources, and optimizing production query performance.

ETL

Neo4j

🕒 August 27

3Pillar Global

1001 - 5000

💼 Consulting

🏥 Healthcare

🛡️ Insurance

Senior Python Data Engineer building Snowflake-Azure migration solutions for 3Pillar, an enterprise AI transformation partner. Migrating on-premise data assets and developing scalable cloud software.

Airflow

Azure

Cloud

Django

Docker

Flask

Jenkins

Kubernetes

Oracle

Python

React

SQL

SSIS

🕒 August 26

Sequoia Connect

11 - 50

💼 Consulting

📣 Marketing

📦 Logistics

Senior Data Engineer building Databricks, Spark, and Hadoop data pipelines for a global IT powerhouse. Optimizing large-scale processing and CI/CD workflows for digital transformation projects.

🗣️🇪🇸 Spanish Required

Airflow

Apache

Cloud

ETL

Hadoop

Java

PySpark

Python

Spark

SQL

🕒 August 26

Sequoia Connect

11 - 50

💼 Consulting

📣 Marketing

📦 Logistics

Lead Data Engineer building scalable cloud and on-premises data platforms for a global IT powerhouse. Developing ETL/ELT pipelines and guiding engineers across high-impact digital transformation projects.

🗣️🇪🇸 Spanish Required

Airflow

Apache

Cloud

ETL

Hadoop

Java

Python

Spark

SQL