
51 - 200 employees
Founded 2014
🤖 Artificial Intelligence
💼 Consulting
📚 Education
💰 Seed on 2020-12
Artificial Intelligence • Consulting • Education
Distrito is a Brazil-based enterprise focused on end-to-end AI transformation and adoption. It offers AI strategy and governance, executive and technical AI education (Mastering AI for Business / Tech), and implementation services through an AI Factory and AI Squads that build custom, scalable solutions such as predictive models and autonomous agents. Distrito also runs an AI ecosystem (GenAI Lab, Plataforma ÍON), publishes reports and runs events, and operates a physical AI Transformation Hub in São Paulo. The company serves large corporations and startups across sectors (examples in consumer goods, healthcare, infrastructure/utilities and insurance) and emphasizes governance, data engineering, and measurable outcomes; it reports 80+ implemented AI projects and 3k professionals trained.
🔥 11 minutes ago
🗣️🇧🇷🇵🇹 Portuguese Required
Airflow
Amazon Redshift
Apache
AWS
Azure
BigQuery
Cloud
Docker
ETL
Google Cloud Platform
Kafka
Kubernetes
NoSQL
Pandas
PySpark
Python
Scala
Spark
SQL
Improve your chances of getting an interview by checking your resume score before you apply.

51 - 200 employees
Founded 2014
🤖 Artificial Intelligence
💼 Consulting
📚 Education
💰 Seed on 2020-12
Artificial Intelligence • Consulting • Education
Distrito is a Brazil-based enterprise focused on end-to-end AI transformation and adoption. It offers AI strategy and governance, executive and technical AI education (Mastering AI for Business / Tech), and implementation services through an AI Factory and AI Squads that build custom, scalable solutions such as predictive models and autonomous agents. Distrito also runs an AI ecosystem (GenAI Lab, Plataforma ÍON), publishes reports and runs events, and operates a physical AI Transformation Hub in São Paulo. The company serves large corporations and startups across sectors (examples in consumer goods, healthcare, infrastructure/utilities and insurance) and emphasizes governance, data engineering, and measurable outcomes; it reports 80+ implemented AI projects and 3k professionals trained.
• Build batch and streaming ingestion pipelines • Design and structure data lakes and data warehouses • Create optimized datasets for machine learning • Implement embedding pipelines • Build vector indexing for RAG • Ensure data quality, governance, and security • Optimize storage and processing costs • Collaborate with AI Engineers to structure feature stores
• Programming language: Python • Advanced SQL • Scala (optional) • Apache Airflow • dbt • Prefect • Spark • Pandas • PySpark • Data lakes (S3, GCS, Azure Blob) • Data warehouses (BigQuery, Snowflake, Redshift) • NoSQL databases • Vector databases (Pinecone, Weaviate, FAISS) • Kafka • Pub/Sub • Docker • Kubernetes • Cloud platforms (AWS, GCP, or Azure) • Data modeling • ETL / ELT • Distributed processing • Scalable data architecture • Experience with unstructured data (text, logs, PDFs) • DataOps concepts • Data versioning and quality
• Meal and/or food allowance of R$950.00 per month on an iFood card • SulAmérica health insurance (100% covered by the company) • MetLife dental plan • Pet health insurance • Life insurance • TotalPass and Wellhub memberships to support physical activity anytime, anywhere • Birthday day off to celebrate your birthday as you wish • Partnership with Sesc • Partnership with OnHappy (discounts on hotels and flights) • Partnerships with educational institutions • Childcare reimbursement • Work equipment • A culture of feedback and professional development • Dynamic environment with significant growth opportunities • Payroll-deducted loan with some of the lowest interest rates • Variable compensation in recognition of achievement, performance, and cultural alignment
Apply Now🔥 8 hours ago
Arquiteto de Dados sênior projetando arquiteturas Lakehouse com GCP e Databricks. Consultoria Leega transforma desafios tecnológicos em soluções de dados e IA.
🗣️🇧🇷🇵🇹 Portuguese Required
Apache
BigQuery
Cloud
Google Cloud Platform
NoSQL
PySpark
Python
Spark
SQL
Terraform
Unity
🔥 20 hours ago
Engenheiro de Dados Sênior desenvolvendo pipelines escaláveis com Python, Spark e Databricks. Criando soluções de dados para a Stefanini, ecossistema global de tecnologia.
🗣️🇧🇷🇵🇹 Portuguese Required
Apache
Azure
Cloud
ETL
Kafka
NoSQL
PySpark
Python
Spark
SQL
Terraform
Unity
🕒 Yesterday
Engenheiro de Dados Sênior projetando pipelines e sistemas de dados para a Stefanini. Mantendo dados acessíveis, confiáveis e prontos para uso.
🗣️🇧🇷🇵🇹 Portuguese Required
Azure
Cloud
Informatica
🕒 Yesterday
Engenheiro(a) Sênior de Dados modernizando processos IBM DataStage para Python e PySpark na AWS. Desenvolvendo pipelines escaláveis e apoiando a evolução da plataforma de dados da GFT.
🗣️🇧🇷🇵🇹 Portuguese Required
AWS
ETL
PySpark
Python
SQL
🕒 Yesterday
Engenheiro de Dados desenvolvendo pipelines e soluções de IA na Quality Digital, especialista em soluções de TI. Apoio à arquitetura de dados, modelos generativos e processos de negócio com IA.
🗣️🇧🇷🇵🇹 Portuguese Required
AWS
Azure
Cloud
Google Cloud Platform
PySpark
Python
SQL