Lead Data Engineer

🔥 21 hours ago

🇧🇷 Brazil – Remote

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 10%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Experian

Experian

10,000+ employees

Founded 1996

💼 Consulting

📣 Marketing

📦 Logistics

Consulting • Marketing • Logistics

Experian is a global leader in digital experience, technology, and transformation. They partner with recognized brands to enhance customer understanding, innovate product strategies, and implement agile technology solutions. With a focus on delivering superior customer experiences through AI, cloud architecture, and project management, Experian helps businesses streamline their operations and achieve their objectives effectively.

📋 Description

• Design, implement, and evolve a petabyte-scale AWS data platform • Build and optimize scalable data pipelines using JVM languages, Spark, Python, and cloud-native services • Act as a principal-level technical leader by mentoring engineers, conducting code reviews, and promoting best practices • Contribute to architectural and design decisions with architects, product owners, and engineering leadership • Plan and deliver complex technical initiatives within agile value-stream teams • Ensure operational readiness through coding standards, testing practices, monitoring strategies, and release procedures • Support production deployments and maintain platform stability • Explore and adopt GraphQL integrations, vector databases, and AI/LLM-based techniques

🎯 Requirements

• Expert-level software and data engineering experience in large-scale data platforms • Deep experience building petabyte-scale systems using Java-based languages such as Scala and Apache Spark • Strong expertise across the AWS data ecosystem, including Glue, S3, Athena, Managed Airflow, and Iceberg • Advanced understanding of distributed systems and highly parallelized workloads • Strong skills in query optimization, data partitioning, and efficient storage patterns • AWS cloud ecosystem mastery; Azure/GCP are pluses • Experience influencing long-term architecture and platform design • Excellent code review, mentorship, and technical leadership capabilities • Hands-on experience working in agile environments and collaborating across multiple engineering teams • Strong version control and multi-repo collaboration skills using Git, GitHub, and Bitbucket • Bachelor’s degree in Computer Science, Engineering, or related field, or equivalent experience • Relevant experience • Advanced English • Availability to travel to São Carlos/SP when needed • Experience with DBT or modern transformation frameworks • Knowledge of concurrent and parallel programming • Familiarity with Python, Angular, and TypeScript • Experience with vector databases such as pgvector and Redis vector fields • Practical experience applying AI/LLM techniques to data platforms, data access, or software quality

🏖️ Benefits

• Inclusive recruitment initiatives • Professional development initiatives • Affinity groups supporting underrepresented groups: ExperianPride, Ubuntu, Women in Experian, Aspire, and Connecting Generations

Apply Now

Similar Jobs

🕒 Yesterday

GFT Technologies

10,000+ employees

💼 Consulting

🛡️ Insurance

🔒 Cybersecurity

Engenheiro de Dados desenvolvendo pipelines escaláveis com Databricks, Azure Data Factory e PySpark. Construindo soluções Cloud para o projeto de Reforma Tributária da GFT.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

Cloud

ETL

PySpark

Python

Spark

SQL

🕒 Yesterday

Sicredi

10,000+ employees

🛡️ Insurance

📦 Logistics

💼 Consulting

Engenheiro de Dados de Auditoria Interna na Sicredi, primeira instituição financeira cooperativa do Brasil. Desenvolvendo pipelines auditáveis e ambientes analíticos para apoiar avaliações de auditoria.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Cloud

ETL

Kafka

Python

Spark

SQL

🕒 Yesterday

Ambev

10,000+ employees

🏭 Manufacturing

📦 Logistics

📣 Marketing

Engenheiro de Dados Pleno desenvolvendo pipelines, integrações e modelos analíticos para a Ambev. Garantindo qualidade, observabilidade e autonomia de dados no Zé Labs.

🗣️🇧🇷🇵🇹 Portuguese Required

Airflow

AWS

Cloud

PySpark

Python

Spark

SQL

Unity

🕒 Yesterday

Sicredi

10,000+ employees

🛡️ Insurance

📦 Logistics

💼 Consulting

Engenharia de Dados de Auditoria Interna no Sicredi, instituição financeira cooperativa brasileira. Desenvolvendo pipelines auditáveis e ambientes analíticos para apoiar avaliações de auditoria e gestão de riscos.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Cloud

ETL

Kafka

Python

Spark

SQL

🕒 Yesterday

Dadoteca

51 - 200

💼 Consulting

🏥 Healthcare

📣 Marketing

Engenheiro(a) de Dados Sênior na Dadoteca, empresa de Dados e Inteligência Artificial. Implementação de Databricks AI/BI, Genie Spaces e soluções analíticas.

🗣️🇧🇷🇵🇹 Portuguese Required

ETL

PySpark

Python

SQL

Unity