
Sr Data Engineer
About Headwater Science
Headwater Science (formerly NoviSci) is a data science and methods company specializing in principled, reproducible evidence generation for complex clinical and regulatory challenges. With deep expertise in comparative effectiveness, causal inference, healthcare utilization and expenditure research, and regulatory-grade analytical software, Headwater Science provides the methodological foundation that delivers reproducible analytic pipelines, novel epidemiologic and statistical methods, and regulatory-grade software validated to hold up under the most demanding scrutiny. The company works with life sciences organizations as a long-term scientific partner. Headwater Science is a Highlander Health company. Learn more at headwaterscience.com.
The Role
We are seeking a talented Sr Data Engineer to design, build, and maintain the data pipelines and software infrastructure that power our real-world evidence research. Working primarily in R, Python, and SQL, you'll create clean, scalable solutions that transform healthcare data into research-ready resources - managing collaborative development through Git and GitLab, maintaining CI/CD tooling that keeps code tested and reproducible, and supporting cloud-based infrastructure (AWS preferred). You'll partner closely with operations and product teams on deployment, testing, and troubleshooting, and help drive ongoing improvements to our systems that support Headwater Science’s mission.
What You’ll Do
- Design and implement data pipelines and software solutions that promote operational efficiency and scalability.
- Write clean, maintainable code primarily in R and Python, with a strong emphasis on SQL for data manipulation and transformation.
- Manage collaborative development through Git and GitLab, and maintain the CI/CD tooling to ensure tested, validated and reproducible code.
- Support and optimize cloud-based data infrastructure (AWS preferred).
- Develop and modify databases to support internal applications.
- Collaborate with operations and product teams to support deployment, testing, and maintenance.
- Help troubleshoot production issues and contribute to ongoing improvement efforts.
- Stay up to date on relevant tools, technologies, and best practices.
What You’ll Bring
- 5+ years of software development experience, ideally in data engineering, data platform development, or backend systems.
- Bachelor’s degree in computer science, engineering or a related field (or equivalent practical experience).
- Background in healthcare, life sciences, or clinical data.
- Experience with cloud platforms (AWS, GCP, or Azure) and working in cloud-native environments.
- Strong problem-solving and communication skills.
- Ability to work independently and manage tasks with moderate supervision.
Nice to Have
- Experience in R package development, SQL, Python, version control and CI/CD tools (e.g., GitLab, GitHub Actions).
- Familiarity with software validation practices for regulated or regulatory-grade environments.
- Experience working with or developing machine learning pipelines or models.
- Exposure to Generative AI or LLM frameworks (e.g., Langchain, LangGraph).
- Knowledge of distributed computing concepts or tools (e.g., Spark, Dask).
- Knowledge of command-line workflows in UNIX/Linux environments.
What We Offer
- Hybrid work environment – 3 days onsite per week.
- Comprehensive health, dental, and vision coverage for you and your family.
- 401(k) with company match.
- Generous PTO and company holidays.
- Paid parental leave.
If you are ready to be part of a team where your work truly matters- where your expertise is valued, your growth is supported, and your contributions help shape the future of healthcare- Headwater Science is the place for you. We’re building something meaningful together, and we’d love for you to be a part of it.
Apply for this job
*
indicates a required field