About the Role
Spectrum Health Care is building a centralized Cloud Data Lakehouse to support clinical analytics, AI initiatives, and business intelligence. This role is key to the company's digital transformation, migrating data from various systems into their cloud stack and Salesforce Health Cloud. The position involves designing, building, and optimizing data pipelines and models to deliver secure, compliant, and high-quality datasets for clinical decision-making, operational excellence, analytics, and AI.
Responsibilities
- Design, deploy and maintain a centralized Azure ADLS Gen2/Databricks/Synapse (or Snowflake) data lake.
- Create batch & real‑time ETL/ELT flows to ingest EHR data and operational telemetry.
- Write and run complex migration scripts using Informatica PowerCenter/IDMC or Azure Data Factory to move data from Epic, other EHRs and legacy OLTP stores into the lakehouse and Salesforce Health Cloud.
- Build automated validation, deduplication, masking, encryption and column‑level security.
- Build high‑throughput pipelines that sync Health Cloud data with the central repository.
- Transform HL7, FHIR and EDI messages into clean canonical schemas.
- Design Robust Data Models – Perform ER/3NF modeling for OLTP and dimensional modeling for OLAP analytics.
Requirements
- Bachelor’s degree in Computer Science, Software Engineering, Data Analytics, or a related technical field.
- 3-7 years of hands-on data engineering experience.
- Proven track record of executing complex data migration projects and building cloud data lake architectures.
- Expertise in ER diagramming, relational database modeling (3NF for OLTP), and dimensional modeling for OLAP systems.
- Experience with enterprise ETL platforms such as Informatica, Azure Data Factory, SSIS, or Talend.
- Experience with cloud data architectures (Azure data services: Azure Data Factory, ADLS Gen2, Azure Synapse, Databricks).
- Familiarity with streaming/real-time pipelines (Event Hubs, Kafka, Spark Streaming).
- Experience working with Salesforce data structures, SOQL, Salesforce Bulk API, and data loading tools.
- Expert-level SQL skills (complex joins, CTEs, window functions, schema design, index tuning, and database optimization).
- Strong Python skills for data manipulation, scripting, and pipeline execution using libraries like Pandas, PySpark, and SQLAlchemy.
Skills
- Azure ADLS Gen2
- Databricks
- Synapse
- Snowflake
- Azure Event Hubs
- Kafka
- Spark Streaming
- Informatica PowerCenter
- IDMC
- Azure Data Factory
- Epic
- Salesforce Health Cloud
- SOQL
- Bulk API 2.0
- MuleSoft
- dbt
- Python
- HL7
- FHIR
- EDI
- ER modeling
- 3NF modeling
- Dimensional modeling
- Star Schema
- Snowflake modeling
- Medallion architecture
- SSIS
- Talend
- Pandas
- PySpark
- SQLAlchemy
Location
- Canada
Work Type
- Full-time
Experience Level
- 3-7 years
Education Level
- Bachelor’s degree in Computer Science, Software Engineering, Data Analytics, or a related technical field
Salary/Compensations
- $95,000 - $105,000
Benefits
- Accommodations available throughout the recruitment process
About the Company
- Spectrum Health Care is a recipient of Canada’s Best Managed Companies award, recognizing top performance, sustained growth, strategy, capabilities, innovation, culture, commitment, and leadership.
- Committed to fostering, cultivating and building a culture of diversity, equity and inclusion.
- Strives to attract, engage and develop a workforce that reflects the diverse communities served.
Equal Opportunity
- In accordance with the Accessibility for Ontarians with Disabilities Act 2005, upon request, support will be provided for accommodations throughout the recruitment process.
- Spectrum Health Care is committed to fostering, cultivating and building a culture of diversity, equity and inclusion within our organization.
