Startups

Research Engineer - Audio & Speech Models

Zyphra · San Francisco · On-site

← All jobs
About Zyphra

The future of intelligence is open.

About the role

As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel architectures and training methodologies.

What they're looking for

  • Strong research taste and intuition. The ability to work through a research project from conception to execution to write-up
  • Strong implementation and prototyping ability (can take an idea from conception to experimentation quickly)
  • The ability to work well with others in a high-paced research setting
  • Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale
  • Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models
  • Experience in training audio autoencoders
More about this role

As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel architectures and training methodologies.

Large-scale audio training runs

Performance optimization of our training stack

Audio dataset collection, processing, and evaluation

Architecture and training methodology ablations and improvements

Strong research taste and intuition. The ability to work through a research project from conception to execution to write-up.

Strong implementation and prototyping ability (can take an idea from conception to experimentation quickly)

The ability to work well with others in a high-paced research setting

Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale

Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models

Experience in training audio...

Read the full posting on Zyphra's site ↗

R&D - Engineering

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.