- Lead product vision and roadmap to transition data ingestion modules into scalable, standalone micro-services.
- Productize event-driven and streaming ingestion pipelines (Kafka, Spark) for high-velocity record processing.
- Define API contracts and integration specifications for HL7, C-CDA, and FHIR data sources.
- Develop analytics strategy using Databricks and Apache Spark, focusing on Medallion Architecture (Bronze/Silver/Gold).
- Define product requirements for FHIR-native MDM engines, including deterministic and probabilistic patient matching.
- Build end-to-end exception lifecycle management, including Dead-Letter Queues (DLQs) and idempotent replay mechanisms.
- Design interfaces for data stewards to manage record merges, unmerges, and data corrections.
PythonSQLSpark+2 more