Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems. Backed by Accel, Bessemer and General Catalyst.
About the role
Anthropic's Safeguards organization builds the policies, evaluations, and enforcement systems that keep our models from contributing to catastrophic harm. We are hiring a manager to lead the research engineering team responsible for biological safety: the evaluations, datasets, and classifiers that govern how our models handle biological knowledge.
What they're looking for
- Experience managing a technical team, including hiring, coaching, and performance management
- A record of setting technical direction for a team and making prioritization calls under uncertainty
- Proficiency in Python, with a background in scientific programming and data analysis
- A solid grasp of ML fundamentals, sufficient to critically review evaluation design and classifier development
- Knowledge of modern biology across both measurement and engineering: high-throughput assays and functional characterization, as well as gene synthesis, genome editing, strain construction, and protein engineering
- Experience designing quantitative experiments or evaluations and drawing defensible conclusions from noisy results
More about this role
Anthropic’s mission is to create reliable, interpretable, and steerable AI systems. We want AI to be safe and beneficial for our users and for society as a whole. Our team is a quickly growing group of committed researchers, engineers, policy experts, and business leaders working together to build beneficial AI systems.
Anthropic's Safeguards organization builds the policies, evaluations, and enforcement systems that keep our models from contributing to catastrophic harm. We are hiring a manager to lead the research engineering team responsible for biological safety: the evaluations, datasets, and classifiers that govern how our models handle biological knowledge.
You will lead a team of research scientists and engineers who design and run capability evaluations against frontier models, curate training data for our safety classifiers, train and iterate on those classifiers alongside our ML engineers, and measure how they hold up against adversarial pressure in production traffic. You will set the technical direction for that work, decide where the team invests, and own the results.
This is a hands-on management role. Most of your time goes to growing and directing the team, but...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area