Startups · AI

Member of Technical Staff, Data Flywheel

Reflection AI · New York, NY · On-site

← All jobs
About Reflection AI

Make intelligence open and accessible to all. Backed by Battery, Lightspeed and Sequoia.

About the role

The Data Flywheel team closes the gap between benchmark performance and useful performance in the real world. We identify and build the signals, data, and feedback loops that turn model usage into rigorous evaluations, targeted training data, and measurable improvements in future generations of models.

What they're looking for

  • Degree (BS, MS, or PhD) in Computer Science, Machine Learning, or related discipline, or equivalent practical experience
  • Deep technical understanding of LLM training and evaluation, with hands-on experience in areas such as evaluation design, data curation, reinforcement learning, or reward design
  • Strong software engineering skills and experience building automated data/evaluation pipelines or large-scale ML systems
  • A track record of owning high-impact projects end to end, navigating ambiguity, and adapting quickly as priorities change
  • A highly collaborative, action-oriented approach and excitement about defining how a new frontier lab measures and accelerates model progress
  • High agency and thrive in a fast-paced startup environment, bias for impact over process
More about this role

Reflection is a research lab making intelligence open and accessible for everyone to use, customize, and build on. We build open models that let anyone control their intelligence and help shape the future of AI. Our mission: make intelligence open and accessible to all.

The Data Flywheel team closes the gap between benchmark performance and useful performance in the real world. We identify and build the signals, data, and feedback loops that turn model usage into rigorous evaluations, targeted training data, and measurable improvements in future generations of models.

This is a hands-on technical role at the intersection of research and deployment. You'll take ambiguous model behaviors from first observation through measurement, intervention, and validated improvement, working across evaluation, human and synthetic data, infrastructure, post-training, and live deployments. You'll collaborate closely with researchers and engineers across the company, as well as customers, partners, vendors, and the open-source community.

Identify high-value data sources and partnership opportunities, deeply understand the underlying use cases, and translate them into representative...

Read the full posting on Reflection AI's site ↗

Research

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.