Fireworks’ state of the art training and inference platform take you beyond the frontier, transforming open models into your specialized intelligence. Backed by Bessemer, Index and Lightspeed.
About the role
Inference is at the core of what Fireworks does. It’s how specialized intelligence reaches production, and the platform around it is what makes it fast to adopt, easy to trust, and economical at scale. We’re hiring a PM to work on our core inference product and general platform, spanning inference performance and control, accounts and fraud reduction.
What they're looking for
- 2 – 8+ years of product management experience building technical or developer-facing products (we are hiring at multiple levels for this role)
- Strong technical background — CS/EE degree, production engineering experience, or equivalent depth earned on the job
- Familiarity with the inference lifecycle: model serving, latency and throughput tradeoffs, and how these connect to cost in production
- Demonstrated ownership of a product area end to end from strategy, spec, launch, to metrics
- Excellent written communication. You can write a spec, a launch post, and a customer-facing explanation of a tradeoff, and all three will be clear
- Comfort with ambiguity, and a bias toward shipping and learning over waiting for certainty
More about this role
Fireworks is the platform for specialized intelligence, enabling companies to build, train, and serve AI models tailored to their own data, workflows, and products. Founded by the team behind PyTorch and backed by AMD, Atreides, Benchmark Capital, Index Ventures, Lightspeed, NVIDIA, Sequoia Capital, and TCV, Fireworks powers production AI with hundreds of state-of-the-art open models across text, image, embedding, audio, and multimodal workloads. Today, Fireworks is a Series D company valued at $17.5 billion, bringing together an ambitious, collaborative team that's building the future of enterprise AI.
Inference is at the core of what Fireworks does. It’s how specialized intelligence reaches production, and the platform around it is what makes it fast to adopt, easy to trust, and economical at scale. We’re hiring a PM to work on our core inference product and general platform, spanning inference performance and control, accounts and fraud reduction.
As a PM working on our platform, you'll help set strategy, write specs, sit with customers running production traffic, and work across our inference, infrastructure, and go-to-market teams to deliver real customer impact. Example...
Browse similar: AI jobs · AI startup jobs · Startup jobs · Remote jobs · San Francisco Bay Area