# Applied Researcher, Audio Post-Training at Cartesia

- Company: Cartesia
- What the company does: Integrate real-time text-to-speech with Sonic-3.6, Cartesia. Backed by General Catalyst, Index and Kleiner Perkins.
- Company website: https://www.cartesia.ai/
- Type: Startups (AI role)
- Level: Mid level
- Location: *HQ - San Francisco, CA
- Work setup: On-site
- Pay: $200K to $350K base salary per year (USD)
- Posted: 2026-07-20
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/cartesia/eeac6a77-eac5-4d72-b3d4-97d53cac4c89
- Page: https://www.1752.vc/careers/jobs/cartesia-applied-researcher-audio-post-training/

## About the role

On the Audio Post-Training team, you’ll be building and improving the capabilities that define how the rest of the world interacts with our generative audio models. This team is where customer needs meet research, and covers the full spectrum of modeling from ideation through productionization.

## What they're looking for

- Strong fundamentals in software engineering, machine learning, debugging complex systems, and the ability + desire to learn quickly
- Experience building and ensuring quality of large multilingual datasets
- Experience training and debugging generative models (speech, text, or multimodal), especially SFT, RL, synthetic data, and evaluation (both human and automated)
- Excitement about solving problems grounded in real customer needs, not just benchmarks
- Bonus points if you have native proficiency in other languages!
- Note: Cartesia participates in E-Verify and will provide the federal government with Form I-9 information to confirm employment eligibility after hire

Tags: Research
