dmodel: steerable and explainable AI. Backed by Y Combinator.
About the role
As a Technical Intern, you'll work directly under a supervising Member of Technical Staff. You'll help build the reinforcement learning environments and evaluations we use to study how AI agents approach alignment problems. This role is a strong fit for someone early in their research career who wants hands-on experience in AI safety, interpretability, and reinforcement learning. Run experiments under guidance and document observed patterns in model behavior and failure modes.
More about this role
d_model is a fundamental AI research lab partnering with frontier labs to turn their models into capable interpretability and alignment researchers. Alongside our partnerships, we aim to use the agents we build for independent research.
Our team brings experience from places including OpenAI, Google, Anthropic, EleutherAI, and MATS.
As a Technical Intern, you'll work directly under a supervising Member of Technical Staff. You'll help build the reinforcement learning environments and evaluations we use to study how AI agents approach alignment problems. This role is a strong fit for someone early in their research career who wants hands-on experience in AI safety, interpretability, and reinforcement learning.
Run experiments under guidance and document observed patterns in model behavior and failure modes.
Assist in exploring AI interpretability techniques within our reinforcement learning environments.
Support the development of reinforcement learning environments and evals that test how AI agents approach alignment problems, including implementing scoped components, writing test cases, and helping validate results.
Help test graders for robustness to specification gaming,...
Browse similar: AI jobs · Startup internships · AI startup jobs · Startup jobs · San Francisco Bay Area