# Member of Technical Staff - Reliability Engineering at Fireworks

- Company: Fireworks
- What the company does: Fireworks’ state of the art training and inference platform take you beyond the frontier, transforming open models into your specialized intelligence. Backed by Bessemer, Index and Lightspeed.
- Company website: https://fireworks.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: San Mateo
- Work setup: Remote
- Pay: $200K to $290K base salary per year (USD)
- Posted: 2026-08-17
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/fireworks/ff11ca2c-d95f-4802-8370-09c2cf394842
- Page: https://www.1752.vc/careers/jobs/fireworks-member-of-technical-staff-reliability-engineering/

## About the role

Fireworks AI is one of the industry leaders in inference and training for open models. Open models are how the rest of the world gets to build on frontier AI without handing the keys to a single vendor, and our job is to make them fast, cheap, and dependable enough that this is a real choice. That work is systems work: GPU scheduling, kernel and runtime performance, networking, storage, Linux. We serve over 40 trillion tokens a day doing it.

## What they're looking for

- Systems fundamentals: 5+ years with Linux internals, system performance troubleshooting, and networking fundamentals (TCP/IP, HTTP, gRPC)
- Software engineering: 5+ years in Python, Go, C++, or Rust, writing production-grade tools and systems code
- Cloud-native operations: Operating and debugging Kubernetes, Terraform, and Docker in high-throughput production
- Distributed systems: High-throughput control planes, microservices, or multi-region setups
- Reliability fundamentals: Fault-tolerant design, SLO/SLA management, automated failover, high-availability architecture
- Influence without authority: You can get other teams to adopt a standard through credibility and useful tooling rather than mandate

Tags: Engineering
