Data Engineer

🔥 3 hours ago

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of GFT Technologies

GFT Technologies

10,000+ employees

Founded 1987

🔒 Cybersecurity

📋 Compliance

☁️ SaaS

Cybersecurity • Compliance • SaaS

GFT Technologies is a leading technology and digital transformation company that specializes in providing advanced solutions for consent management and data privacy compliance. Their flagship product, Cookiebot, enables businesses to automate user consent collection efficiently, ensuring adherence to complex privacy regulations such as GDPR and CCPA. GFT's solutions facilitate seamless integration into existing digital infrastructures, helping over 600,000 customers globally manage user data transparently and securely.

📋 Description

• Rebuild legacy Data Warehouse pipelines in Databricks using PySpark and Spark SQL; • Implement bronze, silver and gold layers following the medallion architecture patterns; • Apply CDC and batch ingestion patterns according to project guidelines; • Execute reconstruction waves by domain, running alongside the legacy DW until cutover; • Implement business rules and transformations with automated tests; • Perform reconciliation and parity validation between legacy environment data and the Lakehouse; • Optimize pipeline performance and costs through partitioning, OPTIMIZE, Z-ORDER and workload sizing; • Contribute to technical documentation for migrated rules; • Support the prioritization and execution of migration waves;

🎯 Requirements

• Advanced knowledge of PySpark and Python; • Experience developing large-scale batch pipelines; • Advanced SQL skills; • Experience with relational and dimensional data modeling; • Production experience with Databricks and Delta Lake; • Knowledge of the medallion architecture (bronze, silver and gold); • Experience with Delta Live Tables, Lakeflow and Asset Bundles; • Experience migrating or rebuilding ETL pipelines; • Experience translating business rules from legacy platforms to Spark; • Knowledge of CDC and batch ingestion from relational databases; • Experience with AWS (S3, Glue, EMR, Athena, Lambda, DMS and Step Functions); • Knowledge of Azure Synapse (SQL Pools and Pipelines); • Experience with data quality and reconciliation; • Knowledge of automated testing and Data Quality controls; • Experience with Git and CI/CD applied to data engineering;

🏖️ Benefits

• Multi-benefits card – you choose how and where to use it. • Study scholarships for Undergraduate, Graduate, MBA and Language courses. • Certification incentive programs. • Flexible working hours. • Competitive salaries. • Annual performance review with a structured career plan. • Opportunity for international career development. • Wellhub and TotalPass. • Private pension plan. • Childcare assistance. • Medical insurance. • Dental insurance. • Life insurance.

Apply Now

Similar Jobs

🕒 2 days ago

DOMVS iT

51 - 200

🤝 B2B

🏢 Enterprise

Data Engineer developing and optimizing cloud data solutions at DOMVS iT. Collaborate with teams to handle large data volumes and build efficient pipelines.

🗣️🇧🇷🇵🇹 Portuguese Required

Amazon Redshift

AWS

Cloud

ETL

NoSQL

Python

SQL

🕒 3 days ago

SysMap Solutions

1001 - 5000

Senior Data Engineer maintaining productive data environments on Azure and Databricks. Supporting incident resolution and continuous operations across various data processes.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

Cloud

🕒 3 days ago

Rox Partner

51 - 200

🔒 Cybersecurity

🤖 Artificial Intelligence

🏢 Enterprise

Senior Data Engineer at Rox, a rapidly growing data consultancy in São Paulo. Collaborating on scalable data solutions aligning with business objectives.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

SQL

Vault

🕒 3 days ago

Sinqia

1001 - 5000

💳 Fintech

🏦 Banking

🛍️ eCommerce

DBA Sênior managing data migration of SAP SQL Anywhere at Evertec. Ensuring database performance and data integrity while supporting critical database environments in Brazil.

🗣️🇧🇷🇵🇹 Portuguese Required

SQL

🕒 3 days ago

Dadosfera

51 - 200

🤖 Artificial Intelligence

☁️ SaaS

Data Engineer designing and developing scalable data platforms and pipelines for AI-powered Data Apps. Collaborating with cross-functional teams and leveraging AWS technologies.

🇧🇷 Brazil – Remote

💰 $1.8M Seed Round on 2022-06

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

AWS

DynamoDB

ETL

Python

SQL