MUTT DATA
MUTT DATA
51 – 200 Employees
InsuranceLogisticsMarketing
MUTT DATA is an AI and machine learning consulting firm that helps organizations build automated, revenue-focused systems. Its work spans Adtech, Martech, Fintech, and telecommunications, including advertising optimization, data-driven marketing, financial insights, and network intelligence. The company supports clients through expert team augmentation, cloud-based data architecture, modern data stack implementation, marketing mix modeling, generative AI, product strategy, data science, and MLOps. MUTT DATA is recognized in LATAM for its AWS competencies and collaborates with major cloud and technology providers to deliver scalable data and AI solutions.

Senior Data Engineer, Clinical Platforms (Databricks) - Argentina Remote

MUTT DATA is hiring a Senior Data Engineer to build Databricks-powered platforms for clinical trials in a regulated environment. The role focuses on compliant lakehouse pipelines, integrations, and data systems for clinical applications.

Description

  • Build and optimize Databricks data pipelines, lakehouse layers, and data models with PySpark, Spark SQL, and Delta Lake for clinical application backends.
  • Work with frontend engineers, software architects, and clinical research teams to develop APIs, ingestion systems, and query layers for clinical trial software.
  • Create efficient, standards-compliant structures for EDC data, audit trails, device telemetry, and patient-reported outcomes to support analytics and fast querying.
  • Coordinate with Clinical QA and Validation teams to align databases, pipelines, and clinical repositories with GxP, 21 CFR Part 11, HIPAA, and GDPR requirements.
  • Develop batch and real-time ingestion workflows that connect legacy clinical systems, central laboratories, EHRs, and wearable devices to a unified Databricks Lakehouse.
  • Observe, diagnose, and improve Spark jobs, Delta Lake tables, and query performance for high-volume, low-latency clinical platform workloads.

Requirements

  • At least four years of practical experience developing production data pipelines and lakehouse architectures with Databricks, Delta Lake, and Apache Spark using PySpark or Scala.
  • Experience developing, enhancing, or supporting custom clinical-trial software such as EDC, CTMS, clinical data repositories, or eCOA/ePRO platforms.
  • Strong knowledge of clinical data standards and regulated environments, including CDISC standards, 21 CFR Part 11, GxP validation, and ICH-GCP guidelines.
  • Proven ability to design relational schemas, build dimensional models, and manage unstructured data in Delta Lake environments.
  • Proficiency in Python, SQL, REST API integrations, CI/CD, Git, and automated testing.
  • Professional experience with cloud platforms; AWS is preferred, with Azure or GCP also relevant.
  • Advanced English communication skills for discussing technical requirements and solutions with clients in the United States.

Benefits

  • Remote-first work model with the flexibility to work from anywhere.
  • Full coverage for AWS, dbt, Google Cloud, Azure, and Databricks certification programs.
  • English language lessons provided in the company.
  • Birthday leave plus an additional annual vacation week known as Mutt Week.
  • Employee referral bonuses for successful team referrals.
  • Monthly Maslow credits for use in the benefits marketplace.
  • Annual Mutters' Trip with the team.
  • Monthly childcare reimbursement to help support employees with families.

Related Jobs