About the Role
We are seeking a talented Data Engineer to design, build, and maintain the data pipelines and software infrastructure that power our real-world evidence research. Working primarily in R, Python, and SQL, you'll create clean, scalable solutions that transform healthcare data into research-ready resources - managing collaborative development through Git and GitLab, maintaining CI/CD tooling that keeps code tested and reproducible, and supporting cloud-based infrastructure (AWS preferred). You'll partner closely with operations and product teams on deployment, testing, and troubleshooting, and help drive ongoing improvements to our systems that support Headwater Science’s mission.
Responsibilities
- Design and implement data pipelines and software solutions that promote operational efficiency and scalability.
- Write clean, maintainable code primarily in R and Python, with a strong emphasis on SQL for data manipulation and transformation.
- Manage collaborative development through Git and GitLab, and maintain the CI/CD tooling to ensure tested, validated and reproducible code.
- Support and optimize cloud-based data infrastructure (AWS preferred).
- Develop and modify databases to support internal applications.
- Collaborate with operations and product teams to support deployment, testing, and maintenance.
- Help troubleshoot production issues and contribute to ongoing improvement efforts.
- Stay up to date on relevant tools, technologies, and best practices.
Requirements
- Bachelor’s degree in computer science, engineering or a related field (or equivalent practical experience).
- Background in healthcare, life sciences, or clinical data.
- 3+ years of software development experience, ideally in data engineering, data platform development, or backend systems.
- Experience with cloud platforms (AWS, GCP, or Azure) and working in cloud-native environments.
- Strong problem-solving and communication skills.
- Ability to work independently and manage tasks with moderate supervision.
Skills
- R
- Python
- SQL
- Git
- GitLab
- AWS
- GCP
- Azure
- R package development
- version control
- CI/CD tools
- GitHub Actions
- software validation practices
- machine learning pipelines
- machine learning models
- Generative AI
- LLM frameworks
- Langchain
- LangGraph
- distributed computing concepts
- Spark
- Dask
- command-line workflows
- UNIX/Linux environments
Location
- Hybrid
Work Type
- Hybrid
Experience Level
- 3+ years of software development experience
Education Level
- Bachelor’s degree in computer science, engineering or a related field (or equivalent practical experience)
Benefits
- Comprehensive health, dental, and vision coverage for you and your family.
- 401(k) with company match.
- Generous PTO and company holidays.
- Paid parental leave.
About the Company
- At Headwater Science, we believe better evidence leads to better medicine.
- We are a data science and methods company dedicated to generating principled, reproducible insights that help life sciences organizations tackle their most complex clinical and regulatory challenges.
- Our work spans causal inference, comparative effectiveness, healthcare utilization research, and regulatory-grade analytical software, all grounded in a shared conviction that evidence should be continuous, cumulative, and built to improve patient outcomes.
- We partner with life sciences organizations as a long-term scientific ally, and as part of the Highlander Health family, we’re helping to build a fundamentally more modern approach to evidence generation.
