Data Engineer, Databricks Lakehouse – Senior

🔥 0 minutes ago

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Compass

Compass

10,000+ employees

🏠 Real Estate

📱 Media

Real Estate • Media

Compass is a real-estate-focused content and services site that provides detailed market analysis, buying/selling/renting guides, mortgage and financing information, and home improvement and renovation advice. The site offers resources for homebuyers, sellers, renters, agents, and real estate investors — including articles on appraisals, affordable housing, investment strategies, staging and property maintenance. Compass aims to help users make informed decisions across the housing lifecycle through timely market updates and practical how-to content.

📋 Description

• Collaborate in building the Enterprise Lakehouse platform on Databricks, organized into data domains according to Data Mesh principles. • Implement and manage Unity Catalog as the central governance layer and enable data sharing between domains using Delta Sharing. • Develop data ingestion patterns (streaming and batch) and the medallion architecture template (bronze/silver/gold) for domain use. • Develop the platform's data contract framework and Data Quality framework. • Establish CI/CD pipelines for the Databricks platform and automate infrastructure provisioning. • Integrate the Databricks platform with the existing Data Mesh product on AWS, ensuring catalog interoperability and adherence to the data contract model. • Contribute to architecture decisions, technical documentation, and team mentoring.

🎯 Requirements

• Databricks: strong experience with the platform — workspace administration, clusters, jobs and workflows management; • Unity Catalog: data governance, permission models, lineage, external locations and storage integration; • Delta Lake and medallion architecture (bronze, silver, gold) in production environments; • Data ingestion patterns: streaming and batch at scale; • CI/CD for Databricks (Databricks Asset Bundles, Repos, integration with pipelines such as GitHub Actions, Azure DevOps or similar); • AWS ecosystem: S3, Glue, Lake Formation, EMR (EC2 and Serverless), Athena, Lambda, IAM, DMS, Kinesis, Step Functions, SNS, SQS and EventBridge; • IaC (Terraform, including the Databricks provider); • Advanced Python and PySpark for pipeline development and platform automations; • Advanced SQL; • Data Contracts and metadata tools like OpenMetadata; • Datadog: for observability controls and monitoring; • Desirable: • Experience with Data Mesh projects (domains, data products, federated governance) — significant plus; • Experience in the financial or credit industry — significant plus; • Observability and monitoring of data platforms (system tables, job metrics, data quality); • FinOps: monitoring, allocation and cost optimization in Databricks (DBUs, cluster sizing, compute policies); • Delta Sharing in cross-company, cross-cloud or cross-region scenarios; • AWS certifications (e.g., Data Engineer Associate, Solutions Architect); • Databricks certifications (Data Engineer Associate, Data Engineer Professional, Platform Administrator); • Open data contract specifications (Open Data Contract Standard, ODPS).

🏖️ Benefits

• Role also open to candidates with disabilities (PWD)

Apply Now

Similar Jobs

🔥 1 hour ago

Sigma Software Group

1001 - 5000

💼 Consulting

🏥 Healthcare

🚘 Automotive

Data Engineer building scalable data solutions for high-traffic E-commerce platforms in Brazil. Collaborating with cross-functional teams on modern cloud technologies and large-scale data architectures.

Apache

AWS

Cloud

ETL

Kafka

Python

Spark

🔥 6 hours ago

CI&T

5001 - 10000

💼 Consulting

🏥 Healthcare

📣 Marketing

Senior Data Developer (Azure) managing data systems development in Brazil at CI&T. Collaborating in a multicultural environment to solve complex data challenges.

🇧🇷 Brazil – Remote

💰 $5.5M Venture Round on 2014-04

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

PySpark

Python

SQL

🔥 22 hours ago

EVT

501 - 1000

💼 Consulting

📦 Logistics

🎯 Recruiter

Data Engineer integrating AI & Data team for data solutions. Responsible for developing and sustaining data processes ensuring quality and reliability.

🗣️🇧🇷🇵🇹 Portuguese Required

Airflow

Apache

BigQuery

Cloud

Google Cloud Platform

Python

SQL

🔥 22 hours ago

EVT

501 - 1000

💼 Consulting

📦 Logistics

🎯 Recruiter

Senior Data Engineer with strong GCP experience at EVT, focusing on scalable architecture and data pipeline development. Ensuring quality and availability of data solutions in a collaborative environment.

🗣️🇧🇷🇵🇹 Portuguese Required

Airflow

Apache

BigQuery

Cloud

Google Cloud Platform

Python

SQL

🕒 Yesterday

FCamara Consulting & Training

1001 - 5000

🏥 Healthcare

🛡️ Insurance

📦 Logistics

Senior Data Engineer at FCamara, modernizing and migrating legacy data platforms to Azure. Collaborating with teams to create scalable and efficient data solutions.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

ETL

Informatica

Oracle

PySpark

Spark

SQL

Unity

Vault