Senior Data Engineer – Clinical Platforms, Databricks

Job not on LinkedIn

🔥 19 minutes ago

🇦🇷 Argentina – Remote

⏰ Full Time

🟠 Senior

🚰 Data Engineer

👻 Ghost score 20%

infoinfo
Apply Now
Find Similar Remote Jobs

📊 Check your resume score for this job

Improve your chances of getting an interview by checking your resume score before you apply.

Logo of MUTT DATA

MUTT DATA

51 - 200 employees

📣 Marketing

🛡️ Insurance

📦 Logistics

Marketing • Insurance • Logistics

MUTT DATA is a consulting firm that specializes in leveraging AI and machine learning to create automated systems that drive revenue for their clients. They offer solutions in sectors such as Adtech, Martech, Fintech, and Telecommunication, focusing on optimizing advertising platforms, enhancing data-driven marketing strategies, and providing real-time insights in financial and telecommunications networks. MUTT DATA provides expert team augmentation, seamless integration of data architectures in the cloud, and tailor-made solutions to refine product strategies. Their services include implementing modern data stacks, marketing mix modeling, and generative AI. MUTT DATA partners with leading companies like Amazon Web Services, Google Cloud, and others to deliver advanced data and AI solutions. They are recognized in LATAM for their AWS competencies and expertise in data science and MLOps, ensuring robust, scalable, and cost-effective data systems for their clients.

📋 Description

• Design, build, and optimize enterprise data pipelines, lakehouse storage layers, and data models using Databricks (PySpark, Spark SQL, Delta Lake) to power custom clinical application backends. • Collaborate with frontend developers, software architects, and clinical research teams to build API-driven endpoints, data ingestion engines, and query layers for proprietary clinical trial software. • Build performant, standards-compliant data structures to store EDC outputs, audit trails, device telemetry, and patient-reported outcomes, enabling rapid querying and downstream analytics. • Partner with Clinical QA and Validation teams to ensure database structures, data pipelines, and clinical data repositories comply with GxP, 21 CFR Part 11, HIPAA, and GDPR. • Implement real-time and batch ingestion jobs connecting legacy clinical systems, central labs, EHRs, and wearable devices into a unified Databricks Lakehouse architecture. • Monitor, troubleshoot, and optimize Spark jobs, Delta Lake tables, and query execution times to support high-throughput, low-latency clinical platform workflows.

🎯 Requirements

• 4+ years of hands-on experience building production data pipelines and lakehouse architectures using Databricks, Delta Lake, and Apache Spark (PySpark or Scala). • Demonstrated experience building, extending, or maintaining custom software applications for clinical trials (e.g., custom EDC, CTMS, Clinical Data Repositories, or eCOA/ePRO platforms). • Deep understanding of clinical data standards and regulatory environments, including CDISC (SDTM, ADaM, CDASH), 21 CFR Part 11, GxP validation, and ICH-GCP guidelines. • Strong experience with relational schema design, dimensional modeling, and unstructured data handling within Delta Lake environments. • Proficiency in Python, SQL, RESTful API integrations, CI/CD pipelines, Git, and automated testing frameworks. • Experience working in cloud environments (AWS preferred, Azure or GCP). • Advanced English to discuss technical requirements and solutions with clients in the United States

🏖️ Benefits

• Remote-first culture – work from anywhere! • AWS, DBT, Google Cloud, Azure & Databricks certifications fully covered • In-Company English Lessons. • Birthday off + an extra vacation week (Mutt Week! 🏖️) • Referral bonuses – help us grow the team & get rewarded! • Maslow: Monthly credits to spend in our benefits marketplace. • Annual Mutters' Trip – an unforgettable getaway with the team! • Monthly Childcare Reimbursement – Because supporting families matters too

Apply Now

Similar Jobs

🔥 18 hours ago

SunnyData

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Senior Data Engineer designing scalable data pipelines and analytics solutions for SunnyData customers. Supporting AI, machine learning, and business intelligence initiatives.

AWS

Azure

Google Cloud Platform

Hadoop

Kafka

Pandas

Scikit-Learn

Spark

SQL

🕒 2 days ago

Netrix Global

501 - 1000

💼 Consulting

📣 Marketing

AWS Data Engineer developing cloud data platforms, pipelines, and AI-ready datasets for Netrix Global’s data-driven clients. Remote role based in Argentina with occasional travel.

Airflow

Apache

AWS

Cloud

ETL

Hadoop

Python

Spark

Terraform

🕒 September 23

Jampp

51 - 200

📣 Marketing

🎮 Gaming

Senior Data Engineer architecting distributed data infrastructure for Jampp’s mobile advertising platform. Building real-time pipelines and guiding technical direction across high-throughput systems.

AWS

Distributed Systems

Docker

ETL

Kubernetes

Open Source

PySpark

Python

SQL

🕒 September 23

Ryz Labs

11 - 50

💼 Consulting

📦 Logistics

📣 Marketing

Senior Data Engineer building AWS data-ingestion pipelines and identity-resolution systems for RYZ Labs. Improving data quality, event-driven workflows, and downstream platform reliability.

AWS

Elixir

Python

SQL

🕒 September 23

Azumo

51 - 200

💼 Consulting

🏥 Healthcare

📦 Logistics

Data Engineer building production pipelines, warehouses, and RAG retrieval infrastructure for Azumo’s client AI systems. Fully remote across Latin America, using modern cloud and data platforms.

Airflow

Amazon Redshift

AWS

Azure

BigQuery

Cloud

Docker

Kafka

Python

Spark

SQL

Terraform