Data Engineer

🔥 17 hours ago

🇦🇷 Argentina – Remote

⏰ Full Time

🟡 Mid-level

🟠 Senior

🚰 Data Engineer

👻 Ghost score 25%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of Ryz Labs

Ryz Labs

11 - 50 employees

💼 Consulting

📦 Logistics

📣 Marketing

Consulting • Logistics • Marketing

Ryz Labs is a company focused on building startups from the ground up and assisting other startups in scaling their operations. They provide top-tier technical talent solutions to ensure that emerging businesses can thrive in a competitive market.

📋 Description

• Own and evolve data ingestion pipelines, including source ingestion/scraping, data cleaning and curation using AWS Glue, entity resolution, and event-driven data publishing • Design and maintain reliable batch and event-driven data workflows • Diagnose and resolve identity-resolution issues, including deduplication and entity matching • Develop and maintain data-quality and validation frameworks • Monitor data at the ingestion boundary and identify issues before downstream impact • Trace data and events end-to-end with events bridge and downstream identity-service teams • Collaborate with Product, Assessments, Analytics, engineering, and other stakeholders • Automate infrastructure and pipeline changes using Infrastructure as Code • Build and operate scalable AWS data infrastructure using Glue, Athena, S3, and event-driven services • Participate in production incident response and implement remediation • Improve pipeline reliability, observability, maintainability, and operational processes

🎯 Requirements

• 5+ years of experience building and operating production-grade data pipelines • Strong experience with batch data processing and event-driven architectures • Hands-on experience with AWS Glue, Athena, and S3-based data lakes, including layered/medallion-style data transformations • Strong proficiency in Python and SQL for data transformation, processing, and analysis • Experience with identity resolution, entity matching, or deduplication, including exact and fuzzy matching approaches • Understanding of durable identifiers and first-seen/locked identity mappings • Experience with event schemas and schema-registry-backed contracts such as Protobuf • Strong troubleshooting and production incident-response skills • Ability to understand and debug code written in a functional or concurrent programming language such as Elixir • Strong communication and collaboration skills across technical teams

Apply Now

Similar Jobs

🕒 Yesterday

Azumo

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Data Engineer building production pipelines, warehouses, and RAG retrieval infrastructure for Azumo’s client AI systems. Fully remote across Latin America, using modern cloud and data platforms.

Airflow

Amazon Redshift

AWS

Azure

BigQuery

Cloud

Docker

Kafka

Python

Spark

SQL

Terraform

🕒 2 days ago

MaxIT Consulting - Max Corporate Group

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

Senior Data Architect diseĂąando arquitecturas modernas con Snowflake, dbt y AI. Liderando modelado dimensional, Semantic Layers y soluciones escalables para una multinacional de alimentos.

🗣️🇪🇸 Spanish Required

🕒 2 days ago

MaxIT Consulting - Max Corporate Group

51 - 200

💼 Consulting

📦 Logistics

📣 Marketing

Senior Data Architect modernizing a global food and consumer goods company’s data ecosystem with Snowflake, dbt, dimensional modeling, and AI. Designing scalable solutions across international teams.

🕒 2 days ago

BforeAI

51 - 200

🔒 Cybersecurity

🤖 Artificial Intelligence

💸 Finance

Senior Data Engineer building event-driven threat-intelligence pipelines and production data services. Supporting BforeAI’s predictive cybersecurity SaaS product with reliable, explainable customer protection.

AWS

Azure

Cloud

Cyber Security

Java

Kafka

Neo4j

Python

Rust

Scala

SQL

Go

🕒 5 days ago

Blend360

501 - 1000

🏥 Healthcare

🏨 Hospitality

✈️ Travel

Lead Data Engineer leading Snowflake ingestion, ELT, and data-quality platforms for Blend, an AI services provider. Designing scalable AWS architectures and governed pipelines for enterprise clients.

AWS

SQL

Terraform