# Evaluations Engineer at Vals AI

- Company: Vals AI
- What the company does: Private, domain-specific benchmarks in legal, tax, and finance. Backed by a16z and Pear VC.
- Company website: https://www.vals.ai
- Type: Startups (AI role)
- Level: Mid level
- Location: San Francisco
- Work setup: On-site
- Pay: $140K to $185K base salary per year (USD)
- Posted: 2026-06-27
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/vals-ai/13d02313-e987-46ee-8b1c-426ea58e8bd6
- Page: https://www.1752.vc/careers/jobs/vals-ai-evaluations-engineer/

## About the role

We are looking for strong engineers to join our team and own the leaderboards that appear on Vals AI. You'll be responsible for testing new models against our benchmarks as they're released; covering tasks in law, tax, coding, finance, social mobility, and more. You will analyze error modes of models, evaluate their strengths and weaknesses, and work with our communications team to release results.

## What they're looking for

- Familiarity with the LLMs: You should already be familiar with the space - the current leading models, relative performance across them, how to use large language models in practice
- Strong engineering fundamentals : You can build and ship quickly with high quality. You should have a track record of building things of significant scope (at jobs, side projects, open source, etc.)
- Python expertise : Significant experience in Python, especially in a professional setting
- Team collaboration : Experience working in development sprints, Git workflows, and pull request reviews
- Strong work ethic: Willingness to work long hours during model releases and get high-quality results out under tight deadlines
- Location : We are an in-person team based in San Francisco. We will support your relocation or transportation as needed

Tags: Engineering & Research
