Startups

Fleet Reliability Engineer

Specter · San Francisco · On-site

← All jobs
About Specter

Search 55M companies, 550M people and every funding or M&A event in real time—surface hidden startups, size markets and act first. Backed by Entrepreneur First.

About the role

Own the fleet’s reliability data pipeline end to end: telemetry aggregation, storage, and instrumentation. Drive down observability cost — own the tooling spend and cut what we pay for but don’t use.

What they're looking for

  • Strong data and software skills — Python (or Go) and SQL — and the ability to own a data pipeline end to end
  • Hands-on building and tuning observability stacks (OpenTelemetry, Grafana, Prometheus, Datadog, or similar), including their cost
  • Experience operating physical or embedded device fleets at scale, and reasoning about how hardware fails in the field
  • Comfortable turning messy field telemetry into trends, failure modes, and forecasts
  • Fluency with databases and data modeling (PostgreSQL or equivalent), infrastructure-as-code familiarity (Terraform or similar) a plus
  • Bias toward building mechanisms over doing manual work
More about this role

Company Background

Specter's mission is to help automate the physical world.

Today, we build video sensors with state-of-the-art AI agents that answer any question, anywhere in their environments. Our systems can automatically detect and reason about any physical activity captured on camera, from security incidents (e.g. perimeter intrusion, theft, LPR), to safety monitoring (e.g. PPE detection, injured people), to operational efficiency (e.g. material tracking, congestion monitoring). We offer both long range wireless (1km range) and wired sensor variants to suit any deployment.

Our co-founders Xerxes and Philip are passionate about empowering our partners in the fast approaching world of physical AI and robotics. We are a small, fast growing team who hail from Anduril, Tesla, Uber, and the U.S. Special Forces.

The Role

We’re hiring a Fleet Reliability Engineer to keep our sensor fleet running in the field by building the data, analytics, and recovery mechanisms that prevent failures from becoming incidents. As we scale toward thousands of sensors, fleet health becomes a data-and-systems problem.

This is the proactive, highest-leverage side of reliability: own the telemetry...

Read the full posting on Specter's site ↗

Engineering

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.