Hippocratic AI builds the safest generative AI healthcare agent for health systems, payors, and pharma. Over 180 million clinical interactions across 1,000+ use cases with 60+ partners worldwide. Backed by General Catalyst, Kleiner Perkins and a16z.
About the role
Design and develop data-driven ASR models for both streaming and non-streaming conversational speech applications, architecting end-to-end speech recognition systems purpose-built for medical accuracy, latency, and robustness Research and implement state-of-the-art speech recognition architectures tailored to the medical domain, addressing problems that off-the-shelf ASR cannot solve—medical terminology, diverse patient populations, real-world acoustic conditions
What they're looking for
- PhD with 3+ years of experience in Speech Recognition or related field or Masters with 5+ years of hands on experience with ASR
- Experience Designing and developing algorithms for accurate and efficient speech recognition for both Streaming and Non-Streaming use cases
- Experience with Training, evaluating, and optimizing ASR models for various factors including accuracy, latency, and resource utilization
- Experience with Preprocessing and curating large speech datasets for training models
- Strong programming skills with working knowledge of Python & C++
- Comfort working in a Linux/ Unix command-line environment
More about this role
As Research Scientist in Speech Technologies, you will lead the research and engineering that makes Hippocratic AI's conversational platform not just intelligent, but genuinely conversational—accurate, fast, and trustworthy in the highest-stakes environments imaginable. You'll define the ASR foundation for healthcare's most advanced conversational AI, ensuring every patient interaction is understood with clinical precision. This role exists because accurate speech recognition in clinical contexts is a frontier problem—no off-the-shelf solution exists, and the work directly determines whether our platform can reliably serve the millions of patients who need it.
Own your first major outcome: By day 90, you will have shipped measurable improvements to our production ASR system (improved accuracy on medical terminology, reduced latency, or expanded robustness to diverse patient populations), validated performance gains on clinically relevant benchmarks, and established the data infrastructure roadmap that will compound our advantage in medical speech recognition.
Drive lasting impact: At 12 months, you will have designed and deployed a next-generation ASR architecture purpose-built...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area