Senior Data Engineer – Clinical Platforms, Databricks

Posted 4 days ago

This is a fully remote position, open to applicants in Argentina.

📋 Description

• Architect, construct, and enhance enterprise data pipelines, lakehouse storage layers, and data models utilizing Databricks (PySpark, Spark SQL, Delta Lake) to drive the backend of custom clinical applications.

• Collaborate with frontend developers, software architects, and clinical research teams to create API-driven endpoints, data ingestion engines, and query layers for proprietary clinical trial software.

• Develop efficient, standards-compliant data structures to store EDC outputs, audit trails, device telemetry, and patient-reported outcomes, facilitating rapid querying and downstream analytics.

• Work alongside Clinical QA and Validation teams to ensure that database structures, data pipelines, and clinical data repositories adhere to GxP, 21 CFR Part 11, HIPAA, and GDPR regulations.

• Establish real-time and batch ingestion jobs that integrate legacy clinical systems, central labs, EHRs, and wearable devices into a cohesive Databricks Lakehouse architecture.

• Oversee, diagnose, and enhance Spark jobs, Delta Lake tables, and query execution times to maintain efficient, high-throughput, low-latency clinical platform workflows.


⛳️ Requirements

• A minimum of 4 years of practical experience in developing production data pipelines and lakehouse architectures using Databricks, Delta Lake, and Apache Spark (PySpark or Scala).

• Proven experience in building, extending, or maintaining custom software applications for clinical trials (e.g., custom EDC, CTMS, Clinical Data Repositories, or eCOA/ePRO platforms).

• Extensive knowledge of clinical data standards and regulatory frameworks, including CDISC (SDTM, ADaM, CDASH), 21 CFR Part 11, GxP validation, and ICH-GCP guidelines.

• Strong background in relational schema design, dimensional modeling, and handling unstructured data within Delta Lake environments.

• Expertise in Python, SQL, RESTful API integrations, CI/CD pipelines, Git, and automated testing frameworks.

• Experience in cloud environments (AWS preferred; Azure or GCP are also acceptable).

• Proficient in English to effectively discuss technical requirements and solutions with clients based in the United States.


🏝️ Benefits

• Remote-first culture – work from any location!

• Full coverage for AWS, DBT, Google Cloud, Azure, and Databricks certifications.

• In-House English Lessons.

• Birthday off + an additional vacation week (Mutt Week! 🏖️)

• Referral bonuses – assist us in expanding the team & receive rewards!

• Maslow: Monthly credits available to use in our benefits marketplace.

• Annual Mutters' Trip – an unforgettable retreat with the team!

• Monthly Childcare Reimbursement – Because supporting families is important too.

People also viewed

QAVION GROUP18 hours ago

Senior AI / ML / Data Engineer

US flagUnited States OnlyFull-timeData Engineer€80k – €150k/year
ApplyView job
ShippyPro1 day ago

Data Platform Engineer – Mid-level

US flagUnited States OnlyFull-timeData Engineer€33k – €43k/year
ApplyView job
Blood Cancer United2 days ago

Data Engineer

US flagUnited States OnlyFull-timeData Engineer$89k – $128k/year
ApplyView job
RevoData2 days ago

Senior - Principal Data Engineer

CZ flagCzechia OnlyFull-timeData Engineer$120k – $180k/year
ApplyView job
RevoData2 days ago

Senior / Principal Data Engineer

HU flagHungary OnlyFull-timeData Engineer
ApplyView job
Modash2 days ago

Senior Software Engineer – Data Search

FR flagFrance, +5 more countriesFull-timeData Engineer€100k – €130k/year
ApplyView job

Never miss a great job!

Get handpicked remote jobs straight to your inbox weekly.

Trusted by 7,400+ designers