Data Engineer

🔥 11 minutes ago

🇧🇷 Brazil – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 10%

infoinfo

🗣️🇧🇷🇵🇹 Portuguese Required

Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Distrito

Distrito

51 - 200 employees

Founded 2014

🤖 Artificial Intelligence

💼 Consulting

📚 Education

💰 Seed on 2020-12

Artificial Intelligence • Consulting • Education

Distrito is a Brazil-based enterprise focused on end-to-end AI transformation and adoption. It offers AI strategy and governance, executive and technical AI education (Mastering AI for Business / Tech), and implementation services through an AI Factory and AI Squads that build custom, scalable solutions such as predictive models and autonomous agents. Distrito also runs an AI ecosystem (GenAI Lab, Plataforma ÍON), publishes reports and runs events, and operates a physical AI Transformation Hub in São Paulo. The company serves large corporations and startups across sectors (examples in consumer goods, healthcare, infrastructure/utilities and insurance) and emphasizes governance, data engineering, and measurable outcomes; it reports 80+ implemented AI projects and 3k professionals trained.

📋 Description

• Build batch and streaming ingestion pipelines • Design and structure data lakes and data warehouses • Create optimized datasets for machine learning • Implement embedding pipelines • Build vector indexing for RAG • Ensure data quality, governance, and security • Optimize storage and processing costs • Collaborate with AI Engineers to structure feature stores

🎯 Requirements

• Programming language: Python • Advanced SQL • Scala (optional) • Apache Airflow • dbt • Prefect • Spark • Pandas • PySpark • Data lakes (S3, GCS, Azure Blob) • Data warehouses (BigQuery, Snowflake, Redshift) • NoSQL databases • Vector databases (Pinecone, Weaviate, FAISS) • Kafka • Pub/Sub • Docker • Kubernetes • Cloud platforms (AWS, GCP, or Azure) • Data modeling • ETL / ELT • Distributed processing • Scalable data architecture • Experience with unstructured data (text, logs, PDFs) • DataOps concepts • Data versioning and quality

🏖️ Benefits

• Meal and/or food allowance of R$950.00 per month on an iFood card • SulAmérica health insurance (100% covered by the company) • MetLife dental plan • Pet health insurance • Life insurance • TotalPass and Wellhub memberships to support physical activity anytime, anywhere • Birthday day off to celebrate your birthday as you wish • Partnership with Sesc • Partnership with OnHappy (discounts on hotels and flights) • Partnerships with educational institutions • Childcare reimbursement • Work equipment • A culture of feedback and professional development • Dynamic environment with significant growth opportunities • Payroll-deducted loan with some of the lowest interest rates • Variable compensation in recognition of achievement, performance, and cultural alignment

Apply Now

Similar Jobs

🔥 8 hours ago

Leega

201 - 500

💼 Consulting

📣 Marketing

🔌 API

Arquiteto de Dados sênior projetando arquiteturas Lakehouse com GCP e Databricks. Consultoria Leega transforma desafios tecnológicos em soluções de dados e IA.

🗣️🇧🇷🇵🇹 Portuguese Required

Apache

BigQuery

Cloud

Google Cloud Platform

NoSQL

PySpark

Python

Spark

SQL

Terraform

Unity

🔥 20 hours ago

Stefanini Brasil

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Engenheiro de Dados Sênior desenvolvendo pipelines escaláveis com Python, Spark e Databricks. Criando soluções de dados para a Stefanini, ecossistema global de tecnologia.

🗣️🇧🇷🇵🇹 Portuguese Required

Apache

Azure

Cloud

ETL

Kafka

NoSQL

PySpark

Python

Spark

SQL

Terraform

Unity

🕒 Yesterday

Stefanini Brasil

10,000+ employees

💼 Consulting

🏥 Healthcare

📦 Logistics

Engenheiro de Dados Sênior projetando pipelines e sistemas de dados para a Stefanini. Mantendo dados acessíveis, confiáveis e prontos para uso.

🗣️🇧🇷🇵🇹 Portuguese Required

Azure

Cloud

Informatica

🕒 Yesterday

GFT Technologies

10,000+ employees

💼 Consulting

🛡️ Insurance

🔒 Cybersecurity

Engenheiro(a) Sênior de Dados modernizando processos IBM DataStage para Python e PySpark na AWS. Desenvolvendo pipelines escaláveis e apoiando a evolução da plataforma de dados da GFT.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

ETL

PySpark

Python

SQL

🕒 Yesterday

Quality Digital

1001 - 5000

💼 Consulting

📣 Marketing

📦 Logistics

Engenheiro de Dados desenvolvendo pipelines e soluções de IA na Quality Digital, especialista em soluções de TI. Apoio à arquitetura de dados, modelos generativos e processos de negócio com IA.

🗣️🇧🇷🇵🇹 Portuguese Required

AWS

Azure

Cloud

Google Cloud Platform

PySpark

Python

SQL