Remote
Senior Data Engineer Backend
About this role
FLAGLER HEALTH IS BUILDING THE CLINICAL OPERATING SYSTEM FOR MODERN MUSCULOSKELETAL CARE. We partner with MSK provider groups and specialty clinics to help them grow, operate more efficiently, and deliver better longitudinal care across patient acquisition, clinical workflows, and ongoing patient engagement. Our platform sits at the intersection of care delivery and clinic operations, helping providers capture more value across the full patient lifecycle.
We’ve recently raised our Series B https://www.businesswire.com/news/home/20260810011327/en/Flagler-Health-Raises-%2450-Million-Series-B-to-Build-the-AI-Operating-System-for-Musculoskeletal-Care and are entering our next phase of growth. THE ROLE: Flagler Health works with outpatient clinics to run care management programs and identify patients who would benefit from them. That work depends on clinical and claims data pulled from dozens of different EHR systems, each exporting it differently: SFTP drops, portal downloads, vendor APIs, bulk exports.
Formats drift, files arrive late or twice, and everything is protected health information. As a Senior Data Engineer, you will own the pipelines that bring this data in, keep it correct, and get it back out to the product and to research partners. You will make messy input dependable, and you will be the one who notices when it quietly stops being dependable. This is a hands-on role on a small team: you write and operate production code.
WHAT YOU WILL DO - Build and run ingestion pipelines from EHR systems into Databricks: landing, validation, normalization, and the silver and gold tables that analytics and the product read. - Onboard new clinics, which usually means understanding a new export format and writing the orchestration to fetch it on a schedule. - Maintain change-data-capture from our MongoDB application database into the analytics layer and keep the two provably consistent.
- Build de-identified data exports for research partners, with controls that keep real identifiers from ever leaving. - Own data quality: freshness checks, reconciliation against source, and alerts that fire before anyone downstream notices. - Investigate data incidents to the root cause and backfill safely without breaking downstream readers. - Write orchestration code in TypeScript alongside the backend team. Required Qualifications - Five or more years of experience building and operating production data pipelines.