# Research Operations, Reinforcement Learning at Anthropic

- Company: Anthropic
- What the company does: Anthropic is an AI safety and research company that's working to build reliable, interpretable, and steerable AI systems. Backed by Accel, Bessemer and General Catalyst.
- Company website: https://www.anthropic.com/
- Type: Startups (AI role)
- Level: Mid level
- Location: San Francisco, CA
- Work setup: On-site
- Posted: 2026-09-22
- Apply by: 2026-11-06
- Apply: https://job-boards.greenhouse.io/anthropic/jobs/5422117008
- Page: https://www.1752.vc/careers/jobs/anthropic-research-operations-reinforcement-learning/

## About the role

Frontier research runs on more than research alone. Behind every model launch, review, and cross-team program is a web of decisions, priorities, and follow-through that keeps hundreds of researchers moving in the same direction. Research Operations owns that work: we bring clarity to ambiguity, provide stability during change, and continually deepen our domain knowledge so that researchers can spend their time on the problems only they can solve. It is one of the highest-leverage functions at Anthropic.

## What they're looking for

- Care deeply about Anthropic's mission and let it underpin all judgment calls you make across the org
- Clear writing: you turn dense, jargon-heavy material into summaries and plans that technical collaborators across the RL org and beyond can act on
- An eye for great dashboards: you can look at data visualizations and scope improvements that make lessons at-a-glance clearer and grounded
- Strong follow-through and attention to detail, nothing falls through the cracks on your watch
- Ability to push back on senior stakeholders to offer strategic guidance as needed, disagree without causing discord, and influence without direct authority
- Comfort with ambiguity and a fast-changing environment, you stay steady when the pace is relentless and the stakes are as high as they can possibly be

Tags: AI Research & Engineering
