# Research Scientist, Reinforcement Learning at Pramaana Labs

- Company: Pramaana Labs
- What the company does: Pramaana Labs is a frontier AI lab building the verification layer for high-stakes AI. We turn complex domain knowledge into machine-checkable systems that return proofs, counterexamples, and traceable explanations. Backed by Accel.
- Company website: https://pramaanalabs.ai/
- Type: Startups (AI role)
- Level: Mid level
- Location: Palo Alto
- Work setup: On-site
- Posted: 2026-08-24
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/pramaana-labs/725f2412-b0cc-4aeb-b5bd-bb225356f3e4
- Page: https://www.1752.vc/careers/jobs/pramaana-labs-research-scientist-reinforcement-learning/

## About the role

We're training foundation models to natively interact with symbolic world models. As an RL Post-Training Researcher, you'll take foundation models and scale their reasoning capabilities: applying RLVR to new domains using verified rewards from the Lean kernel, pushing the frontier of autoformalization and proving, and innovating on RL algorithms, data, and evals. Your work will also define how our models leverage test-time compute to solve long-horizon logical tasks.

## What they're looking for

- Deep, hands-on research experience in reinforcement learning applied to reasoning models at scale
- Experience working with RLVR, test-time RL, or exact deterministic reward signals
- Strong algorithmic and systems intuition — comfortable writing custom RL loops, managing data pipelines, and building robust evals from scratch
- High autonomy: the ability to take a fuzzy problem area and independently drive it to state-of-the-art results without day-to-day direction

Tags: Research & Development
