Staff Data Scientist, Machine Learning in Epidemiology and Patient Data Products
About the role
About Us
Valo Health is a human-centric, AI-enabled biotechnology company working to make new drugs for patients faster. The company’s Opal Computational Platform transforms drug discovery and development through a unique combination of real-world data, AI, human translational models and predictive chemistry.
Our talented team of biologists, chemists and engineers, armed with advanced AI/ML tools, work together to break down traditional R&D silos and accelerate the speed and scale of drug discovery and development.
About the Role
As a Staff Data Scientist, Machine Learning in Epidemiology and Patient Data Products, you will be a core member on a team of data scientists building a powerful computational platform for advancing the discovery and development of new medicines. In this role, you will develop machine learning tools for patient data and drive their adoption across teams, under the guidance of epidemiology and biology program leads. Successful candidates will work with a diverse group of scientists and domain experts, in ways that cut across traditional industry boundaries in an innovative startup environment.
What You’ll Do
- As a senior member of our team, you will lead the development of machine learning (ML) methods and analyses of patient data with diverse stakeholders. For example, integrate clinical insights into supervised and unsupervised learning approaches and generate patient profiles.
- Perform project-specific hands-on analysis and modeling of high-dimensional longitudinal real-world data, spanning electronic medical records (EHRs), clinical notes, sequencing data, and multi-omics, using modern data science tools in cloud environments.
- Contribute to the design, implementation, and evaluation of innovative machine learning approaches for patient data to provide novel clinical insights.
- Be comfortable with scientific uncertainty and embrace curiosity and creative solutions. Many of the challenges we tackle don’t have known solutions or established pathways.
- Use your technical knowledge and intuition to articulate and break down large problems into solvable pieces. There are a lot of problems to solve; you’ll need to prioritize which of these are critical-path today from those that can wait.
- Be a dynamic and active team member, championing shared coding standards, participating in code reviews, and providing regular updates on your work and input into the work of your colleagues.
What You Bring
- MS, MPH, or PhD in health data science, biostatistics, or a related quantitative field, with 5 years of experience developing and applying ML methods, including at least 3 years working directly with real-world patient data. Experience in a biopharmaceutical, epidemiological or biostatistical setting is a plus.
- Extensive experience developing and implementing machine learning solutions in healthcare databases, including EHRs, administrative claim