Mid-Level Data Engineer – GCP, DBT

Job not on LinkedIn

🔥 1 minute ago

🇧🇷 Brazil – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 14%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Leega

Leega

201 - 500 employees

Founded 2010

💼 Consulting

📣 Marketing

🔌 API

Consulting • Marketing • API

Leega is a leading technology solutions provider in Latin America, specializing in data analytics and cloud solutions. As the first company in the region certified by Google Cloud for Data Analytics, Leega offers a range of services including application development, machine learning, and risk management analytics. The firm partners with major cloud services such as AWS and Microsoft Azure to help businesses enhance their data management and transition effectively to the cloud, ultimately driving digital transformation and innovation.

📋 Description

• Assess data warehouse architecture and requirements • Map data, transformations, and processes across GCP services, including Cloud Storage, BigQuery, and Dataproc • Define data migration strategies: full load, incremental load, and CDC • Develop a GCP data architecture plan • Design BigQuery table schemas with performance, cost, and scalability in mind • Define BigQuery partitioning and clustering strategies • Model Bronze, Silver, and Gold data zones in Cloud Storage • Create transformation routines using Dataproc/Spark or Dataflow to load data into BigQuery • Translate business logic and existing transformations into GCP • Implement data validation and quality mechanisms • Optimize BigQuery queries, Spark jobs in Dataproc, and the use of GCP resources • Implement data security in transit and at rest • Define and enforce IAM policies • Ensure compliance with data governance policies • Troubleshoot performance and functionality issues in pipelines and GCP resources • Document architecture, pipelines, data models, and operational procedures • Communicate with team members, stakeholders, and other departments • Ensure clear communication regarding architecture, software components, development progress, and development quality • Apply Agile methodologies and use Jira

🎯 Requirements

• Proven experience with DBT for at least 3 years • Strong knowledge of models (staging, intermediate, and marts) • Knowledge of ref() and source() • Knowledge of macros (Jinja) • Knowledge of seeds and snapshots • Knowledge of not null, unique, and custom tests • Layered organization: Staging → Transform → Mart • Deep knowledge of BigQuery, including data modeling, query optimization, partitioning, clustering, streaming and batch loads, security, and governance • Experience with Cloud Storage, including buckets, storage classes, lifecycle policies, IAM, and security • Ability to provision, configure, and manage Spark/Hadoop clusters in Dataproc • Knowledge of job optimization and integration with other GCP services • Knowledge of Dataflow, Composer, and DBT for data orchestration and processing • Knowledge of Cloud IAM and granular access control • Understanding of VPCs, networking, subnets, firewall rules, and cloud security • Python and PySpark • Advanced SQL • Shell scripting • Git/GitHub/Bitbucket • Knowledge of Agile methodologies, ceremonies, and proficiency with Jira

🏖️ Benefits

• Porto Seguro health insurance, with the option to add a spouse and children • Porto Seguro dental insurance for employees and dependents • Profit Sharing and Results (PLR) • Childcare assistance • Alelo food and meal vouchers • Home office allowance • Partnerships with educational institutions, offering discounts and incentives for courses and degree programs • Certification incentives, including cloud certifications (GCP, Azure, AWS, among others) • Livelo points • TotalPass • Mindself, with incentives for meditation and mindfulness • Ongoing professional development

Apply Now

Similar Jobs

🔥 4 hours ago

CI&T

5001 - 10000

💼 Consulting

🏥 Healthcare

📣 Marketing

Data Engineer Specialist leading end-to-end Azure data pipelines, medallion architecture, and Power BI semantic models. Building AI-driven technology solutions that transform large enterprises at CI&T.

🇧🇷 Brazil – Remote

💰 $5.5M Venture Round on 2014-04

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

Azure

Cloud

ERP

ETL

Terraform

🔥 6 hours ago

Grupo CVLB

5001 - 10000

Especialista de Engenharia de Dados construindo pipelines, data warehouses e soluções analíticas para o Grupo CVLB, varejista brasileiro. Atuando com BigQuery, Airflow, Python e processamento distribuído.

🗣️🇧🇷🇵🇹 Portuguese Required

Airflow

Amazon Redshift

Apache

BigQuery

Cloud

ETL

MySQL

Oracle

Postgres

PySpark

Python

SQL

🔥 23 hours ago

SysMap Solutions

1001 - 5000

💼 Consulting

📣 Marketing

🤖 Artificial Intelligence

Engenheiro de Dados PL/SR construindo modelos de Machine Learning e arquiteturas escaláveis na Triggo.ai. Aplicação de MLOps, GCP, Databricks e Analytics para gerar impacto nos negócios.

🗣️🇧🇷🇵🇹 Portuguese Required

BigQuery

Cloud

Google Cloud Platform

Python

Scikit-Learn

Spark

SQL

🕒 Yesterday

UltraCon Consultoria

11 - 50

💼 Consulting

☁️ SaaS

📚 Education

SAP Data Engineer Sênior projetando e otimizando modelos de dados SAP HANA para a UltraCon Consultoria em TI. Integração, relatórios, planejamento e gestão de incidentes em ambiente SAP BTP.

🗣️🇧🇷🇵🇹 Portuguese Required

Cloud

SQL

🕒 5 days ago

GFT Technologies

10,000+ employees

💼 Consulting

🛡️ Insurance

🔒 Cybersecurity

Engenheiro de Dados aplicando IA, automação e grafos no Cadastro de Clientes da GFT. Desenvolvendo soluções analíticas com Python, Databricks, PySpark e plataformas cloud.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

Cloud

ETL

Google Cloud Platform

Neo4j

PySpark

Python

SQL