# Research Scientist - Frontier Benchmarks at Snorkel AI

- Company: Snorkel AI
- What the company does: Snorkel AI builds specialized training data, benchmarks, and evaluation environments that help frontier models and agents perform in high-stakes domains. Backed by Greylock, Lightspeed and GV.
- Company website: https://www.snorkel.ai
- Type: Startups (AI role)
- Level: Mid level
- Location: New York City, NY (Hybrid); San Francisco, CA (Hybrid); United States (Remote)
- Work setup: Remote
- Posted: 2026-05-29
- Apply by: 2026-10-08
- Apply: https://job-boards.greenhouse.io/snorkelai/jobs/6009489004
- Page: https://www.1752.vc/careers/jobs/snorkel-ai-research-scientist-frontier-benchmarks/

## About the role

We're looking for a Research Scientist to lead the design of next-generation benchmarks and datasets that push the boundaries of frontier model evaluation. You'll define what "good" looks like across a range of hard tasks, drawing on conversations with customers and academic partners to ground your datasets in real performance gaps.

## What they're looking for

- Strong research background in AI/ML evaluation, NLP, or related fields, with a track record of rigorous experimental design — especially around measuring the impact of training and evaluation data on model behavior
- Exceptional communication skills — able to present complex technical findings clearly to both technical and non-technical audiences
- Comfort operating in a fast-moving, cross-functional environment with ambiguous problem spaces
- Genuine interest in GTM strategy, startup dynamics, and the commercial side of AI data services
- Ph.D. in machine learning, NLP, or a related field preferred, equivalent industry or research lab experience considered

Tags: Research
