The future of intelligence is open.
About the role
As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel architectures and training methodologies.
What they're looking for
- Strong research taste and intuition. The ability to work through a research project from conception to execution to write-up
- Strong implementation and prototyping ability (can take an idea from conception to experimentation quickly)
- The ability to work well with others in a high-paced research setting
- Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale
- Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models
- Experience in training audio autoencoders
More about this role
As a Research Engineer - Audio & Speech Models , you will be a core contributor on Zyphra’s Audio Team, building the next generation of open-source autoencoders, ASR, TTS, SSL, and speech-to-speech models. You will be deeply involved in the entire model training process, from data gathering and processing to designing novel architectures and training methodologies.
Large-scale audio training runs
Performance optimization of our training stack
Audio dataset collection, processing, and evaluation
Architecture and training methodology ablations and improvements
Strong research taste and intuition. The ability to work through a research project from conception to execution to write-up.
Strong implementation and prototyping ability (can take an idea from conception to experimentation quickly)
The ability to work well with others in a high-paced research setting
Excellent communication and collaboration skills, and can work effectively on both research and engineering implementation at scale
Expertise and intuition for training models in the audio domain, including text-to-speech, ASR, speech-to-speech, speech-emotion-recognition, or other models
Experience in training audio...
Browse similar: Startup jobs · San Francisco Bay Area