Snorkelai
Snorkelai

Director, Research - Evaluation & Training

datafull-timeNew York City, NY (Hybrid); San Francisco, CA (Hybrid); United States (Remote)
SALARY
Not listed
WORK TYPE
remote
JOB TYPE
full-time
INDUSTRY
ai
Apply for this position
✦ AutoApply Sick of applying? We apply to roles like this for you, up to 20 a month.
Learn more

About the role

ABOUT THE ROLE

We're looking for a manager to lead a team of researchers to focus on data evaluation, error analysis and data valuation methods to predict model performance. This team is responsible for showcasing the value and quality of Snorkel’s data for model training and evaluation, understanding where today's frontier models fall short, and turning that understanding into a point of view on what benchmarks and datasets these models will benefit from.

You and your team will be responsible for Snorkel’s data design flywheel by analyzing model failures, finding capability and skill gaps in current models, suggesting the next benchmarks to invest in and then proving the value of this data for our customers.

MAIN RESPONSIBILITIES

  • Own a multi-quarter roadmap centered on novel evaluation, error analysis, and data valuation techniques
  • Synthesize and share trends from model-failure analysis and benchmarking into recommendations on the datasets the community should focus on and the ones Snorkel should invest in — making this team a primary input to the company's data strategy.
  • Focus on data valuation techniques that quantify how Snorkel data meaningfully improves model performance
  • Lead and grow a team of researchers, setting a high bar for quality, rigor and speed of execution
  • Act as the primary bridge between the team's findings and Product, GTM, and our customers

PREFERRED QUALIFICATIONS

  • 7+ years in applied AI, ML, or research roles, with 4+ years managing technical teams.
  • A leader who has repeatedly turned research and analysis into business outcomes, and who instinctively connects technical findings to market and customer needs.
  • Strong business and market judgment in the AI/ML space — you understand the competitive and frontier-lab landscape and can prioritize accordingly.
  • Technically conversant and credible: enough depth in LLM evaluation, benchmarking, and model behavior analysis to set direction, judge experimental quality, and pressure-test results — without needing to be the deepest technical expert in the room.
  • A nose for trends: able to look across many evaluation results and failure cases and extract the signal that should drive what gets built next.
  • Excellent communication and storytelling skills, with the ability to make technical results legible and persuasive to non-research audiences.
  • Familiarity with data valuation or data attribution research is a strong plus.
  • Bonus: experience working with frontier labs, public benchmarks, or commercial AI data/eval products.

Actual compensation will be determined based on factors including skills, qualifications, experience, and geographic location.

Salary range(s) for this role

$275,000

$425,000 USD

✦ Sick of applying to 40 jobs a month?
I rewrite your resume for ATS by hand first. Once you sign off on it, AutoApply applies to up to 20 roles like this a month, cover letter in your own voice each time. From $14.99/mo, cancel anytime.
Get AutoApply
Apply now
Director, Research - Evaluation & Training at Snorkelai — Remote