Master Data Developer

🔥 10 minutes ago

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of CI&T

CI&T

5001 - 10000 employees

Founded 1995

💼 Consulting

🏥 Healthcare

📣 Marketing

💰 $5.5M Venture Round on 2014-04

Consulting • Healthcare • Marketing

CI&T is a global tech transformation specialist focusing on helping organizations navigate their technology journey. With services spanning from application modernization and cloud solutions to AI-driven data analytics and customer experience, CI&T empowers businesses to accelerate their growth and maximize operational efficiency. The company emphasizes digital product design, strategy consulting, and immersive experiences, ensuring a robust support system for enterprises in various industries.

📋 Description

• Design, build, and maintain robust ETL/ELT processes to ingest, transform, and deliver data across a modern Data Lake architecture • Develop and optimize distributed data processing workflows using Python and PySpark to handle large-scale datasets efficiently • Implement and refine partitioning strategies for data lake storage frameworks (such as Delta Lake or Apache Iceberg) to balance query performance with storage costs • Write, optimize, and translate complex SQL queries involving CTEs, window functions, conditional expressions, and aggregations • Migrate and modernize data pipelines from legacy RDBMS platforms to cloud-native analytics environments • Work confidently with AWS-native services including Glue (Jobs, Catalog, Triggers, Workflows), Athena, Redshift, S3, Lambda, EventBridge, and related data services • Collaborate with infrastructure and DevOps teams to provision and manage data resources using Infrastructure as Code (IaC) tools such as CloudFormation, CDK, or Terraform • Monitor data pipeline health and performance using CloudWatch and other observability tools, proactively addressing issues and improving reliability • Ensure data integrity, consistency, and compliance across pipelines and storage layers • Partner with data analysts, scientists, and business stakeholders to understand requirements and translate them into scalable technical solutions • Stay current with emerging data engineering practices, tools, and cloud-native innovations

🎯 Requirements

• Solid experience working with ETL processes and data pipeline development with AWS • Strong proficiency in Python as the primary programming language, with demonstrated experience writing and optimizing PySpark code for distributed data processing • Thorough understanding of SQL, including complex queries (CTEs, window functions, aggregations, conditional expressions) and experience translating workloads from legacy RDBMS platforms • Hands-on experience with AWS Glue (Jobs, Catalog, Triggers, Workflows), Athena, and Redshift • Solid understanding of Data Lake architectures and partitioning strategies to optimize performance and cost • Good understanding of object-oriented programming (OOP) principles and experience working with reusable code libraries • Comfortable working with Git, Shell scripts, and Linux environments • Familiarity with observability, monitoring, and metric tracking practices • English Advanced/Fluent

🏖️ Benefits

• Health and dental insurance • Meal and food allowance • Childcare assistance • Extended paternity leave • Partnership with gyms and health and wellness professionals via Wellhub (Gympass) TotalPass; • Profit Sharing and Results Participation (PLR); • Life insurance • Continuous learning platform (CI&T University); • Discount club • Free online platform dedicated to physical, mental, and overall well-being • Pregnancy and responsible parenting course • Partnerships with online learning platforms • Language learning platform • And many more!

Apply Now

Similar Jobs

🔥 3 hours ago

UltraCon Consultoria

11 - 50

💼 Consulting

☁️ SaaS

📚 Education

Desenvolvedor(a) Oracle Retail XStore Sênior para desenvolver e manter soluçþes de ponto de venda. Necessårio 5 anos de experiência com Java e SQL, e espanhol intermediårio.

🗣️🇪🇸 Spanish Required

🗣️🇧🇷🇵🇹 Portuguese Required

Java

Oracle

SQL

🔥 3 hours ago

Spassu

1001 - 5000

💼 Consulting

📦 Logistics

📣 Marketing

Senior Developer at Spassu working remote, focusing on software development with Outsystems and Agile methodologies. Leading software architecture and team support for development projects.

🗣️🇧🇷🇵🇹 Portuguese Required

Microservices

SQL

🔥 15 hours ago

Eteg

51 - 200

☁️ SaaS

Full Stack Developer at Eteg Tecnologia Da Informação S/a, integrating in a collaborative squad for end-to-end solutions development. Focusing on systems, APIs, automations, and integrations with bots.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Docker

DynamoDB

EC2

MongoDB

Postgres

React

Redis

TypeScript

Webpack

🔥 17 hours ago

Spread Tecnologia

1001 - 5000

💼 Consulting

📣 Marketing

🏥 Healthcare

Desenvolvedor React Native criando aplicativos mobile e integrando com APIs REST em uma empresa que valoriza tecnologia e diversidade.

🗣️🇧🇷🇵🇹 Portuguese Required

JavaScript

Jest

React

React Native

Redux

TypeScript

🔥 22 hours ago

Sicredi

10,000+ employees

🛡️ Insurance

📦 Logistics

💼 Consulting

AI Governance Engineer at Sicredi responsible for AI governance architecture and ensuring compliance. Focused on data-driven culture and responsible AI within the organization.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Azure

Cloud

Google Cloud Platform