Fast, accurate, and private speech-to-text for Mac, Windows, iPhone, and AI workflows. Backed by Y Combinator.
About the role
We're a small team and growing. You'll touch every part of the system.
What they're looking for
- - Show us what you've built
- - Comfortable switching between languages and domains
- - Can own projects from idea to production
- - Writes readable, maintainable code
- - Experience in production systems
- - You are a winner overall (will bring good luck to team etc.)
More about this role
Aqua Voice is building the voice input layer for the age of AI. We train our own models and build deep OS integrations because doing voice well requires controlling the entire stack.
The way people work is changing. IC work is over; you manage AI agents now. This type of work is wonderfully suited for voice.
We aren't building conversational agents. We believe the most natural way to interact with your computer is voice in text out (VITO). We believe that voice belongs at a level above the application, and that a small company can win by pushing the envelope.
We are applying relentless energy to this opportunity and the results so far have been good. We hope you will join us.
We're a small team and growing. You'll touch every part of the system.
- Real-time transcription server handling thousands of concurrent audio streams
- Custom speech recognition model training and deployment
- Native macOS and Windows integrations using deep system APIs
- Frontend : TypeScript, React, Next.js, Electron
- Backend : Python, real-time server (Bun/Node.js), WebSockets
- Native : Swift (macOS), C# (Windows)
- ML : Custom speech recognition models, inference pipelines
- Infra : Terraform,...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area