11 Sep
|
Flagler Health+
|
Vancouver
11 Sep
Flagler Health+
Vancouver
Flagler Health is building the clinical operating system for modern musculoskeletal care.
We partner with MSK provider groups and specialty clinics to help them grow, operate more efficiently, and deliver better longitudinal care across patient acquisition, clinical workflows, and ongoing patient engagement. Our platform sits at the intersection of care delivery and clinic operations, helping providers capture more value across the full patient lifecycle. We’ve recently raised our Series B and are entering our next phase of growth.
As a Data Engineer, you will own the pipelines that bring this data in, keep it correct, and get it back out to the product and to research partners. You will make messy input dependable, and you will be the one who notices when it quietly stops being dependable. This is a hands-on role on a small team: you write and operate production code.
What You Will Do
- Build and run ingestion pipelines from EHR systems into Databricks: landing, validation, normalization, and the silver and gold tables that analytics and the product read.
- Onboard new clinics, which usually means understanding a new export format and writing the orchestration to fetch it on a schedule.
- Maintain change-data-capture from our MongoDB application database into the analytics layer and keep the two provably consistent.
- Build de-identified data exports for research partners, with controls that keep real identifiers from ever leaving.
- Own data quality: freshness checks, reconciliation against source, and alerts that fire before anyone downstream notices.
- Investigate data incidents to the root cause and backfill safely without breaking downstream readers.
- Write orchestration code in TypeScript alongside the backend team.
Required Qualifications
- 5+ years building and operating production data pipelines.
- Strong SQL and Python.
- Hands-on experience with Spark or a comparable engine, and with Delta Lake or an equivalent table format.
- Working knowledge of idempotency, deduplication keys, event ordering, and schema evolution in real pipelines.
- Experience with change-data-capture or event-stream processing from an operational database.
- Willingness to write TypeScript for orchestration code.
- Understanding of how sensitive data leaks in practice, including through file names, object keys, and logs.
- Comfort with ambiguous, unglamorous problems, such as why one clinic's spreadsheet has a different header this week.
Preferred Qualifications
- Databricks and Unity Catalog, including governance and access controls.
- Temporal or another durable-execution engine.
- MongoDB.
- Healthcare data: claims, CPT and ICD-10 codes, HIPAA de-identification rules.
- Prior TypeScript or Node.js experience.
- Infrastructure as code (Terraform) and CI/CD for data pipelines.
Work Environment
- Small engineering team. You will work directly with backend engineers, the data scientist, and operations, with no layers in between.
- Our stack: Databricks (Delta Lake, Unity Catalog) on AWS, MongoDB, S3, Temporal with TypeScript for orchestration, Python and SQL for transformation, ClickHouse for analytics.
- All data is PHI. HIPAA compliance is part of everyday work, not a separate team's job.
Our Values:
This is what you can expect of your teammates at Flagler:
- Owner: We own our work like founders. We don't wait to be asked, and we pick up the problem nobody else has.
- Builder: We like making things from scratch. We don't need a playbook to start, and we'd rather ship a rough first version than wait on a perfect spec.
- Fast: We make quick decisions. We favor progress and respect process.
- Transparent: We communicate directly. We value merit and honesty, and we challenge assumptions, including our own.
- Curious: We explore recent approaches and ask "why" often.
📌 Senior Data Engineer Backend (Vancouver)
🏢 Flagler Health+
📍 Vancouver