As a key member of the Pulse Engineering Team, you will combine deep experience managing complex operational data sets while contributing to the implementation of recent product features in our flagship products
Must-have
Solid production ETL/ELT background — Python and PySpark
Real comfort with semi-structured data: parsing and flattening nested JSON, exploding arrays, handling EAV / Question-Answer or survey-response style models, and the data-quality headaches that come with them.
Solid SQL and relational modeling on MySQL
Hands-on AWS Glue: job authoring, the Data Catalog, crawlers, bookmarks for incremental loads, triggers/workflows.
Core AWS data stack around it: S3, IAM, Athena, and an orchestration story (Step Functions, EventBridge, or Glue workflows).
Strong assets
Healthcare data background: OMOP/OHDSI CDM specifically, plus exposure to other standards we publish to. Familiarity with clinical vocabularies (SNOMED CT, LOINC, ICD-10, RxNorm)
Experience building configurable/declarative transformation pipelines (mapping rules as data) rather than bespoke scripts per customer.
Regulated-setting experience — HIPAA, and 21 CFR Part 11 awareness
Prior work with EAV, survey/forms platforms, or clinical registry data
📌 Database Engineer London
🏢 Pulse Infoframe
📍 London
Reply to this offer
Impress this employer describing Your skills and abilities, fill out the form below and leave Your personal touch in the presentation letter.