# Applied AI Researcher, Benchmarking at Distyl

- Company: Distyl
- What the company does: Architecting the AI-Native Enterprise. Distyl partners with the most ambitious enterprises to design and operationalize their AI transformations. Backed by Khosla, Lightspeed and Peak XV.
- Company website: https://distyl.ai/
- Type: Startups (AI role)
- Level: Mid level
- Location: San Francisco
- Work setup: Remote
- Pay: $150K to $250K base salary per year (USD)
- Posted: 2025-10-16
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/distyl/cf166cd0-c30f-41c8-8e43-fa38e323014d
- Page: https://www.1752.vc/careers/jobs/distyl-applied-ai-researcher-benchmarking/

## About the role

The Benchmarking team defines how progress is measured. Researchers design evaluation frameworks that capture reasoning depth, interaction quality, reliability, and operational impact. They construct benchmarks that reflect real-world complexity. Their systems become the standard by which new architectures, techniques, and releases are judged.

## What they're looking for

- Our researchers come from many academic backgrounds but have strong research track records, operate in an AI-native way, and would be bored staying on the rails of a traditional research org
- Experience Designing and Running Evaluations: You’ve built or maintained benchmarks, test suites, or experimental frameworks to measure model or system performance
- Statistical and Analytical Rigor: You design fair, reproducible experiments and can extract signal from noisy empirical results
- Proven Track Record of Research Results: Whether you’ve published in top journals, posted amazing work on twitter, or somewhere else we want to see what you've done
- Uses AI Every Day: Before you can revolutionize someone else’s workflow, you need to revolutionize yours. You should be using tools like ChatGPT, Cursor, and Perplexity to accelerate your workflow
- Strong Programming and Data Analysis Skills: While you might not consider yourself a software engineer you need to be able to build prototypes of your ideas and then perform the experiments to prove the effectiveness to a F500 Head of AI

Tags: Research
