# Site Reliability Engineer at HappyRobot

- Company: HappyRobot
- What the company does: Autonomous AI workers that communicate, coordinate, and operate at scale across the real economy. Backed by Y Combinator and a16z.
- Company website: https://happyrobot.ai
- Type: Startups (AI role)
- Level: Mid level
- Location: San Francisco
- Work setup: Remote
- Pay: $200K to $240K base salary per year (USD)
- Posted: 2026-06-04
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/happyrobot.ai/5726534e-7958-4b4e-8e3c-553025d38cc2
- Page: https://www.1752.vc/careers/jobs/happyrobot-site-reliability-engineer/

## About the role

We're looking for a Site Reliability Engineer to take the lead on scaling our operational resilience as we grow. You’ll own the stability, observability, and debugging workflows that keep our systems running smoothly. You'll be the go-to person for untangling complex failures in real time, designing tools that turn chaos into clarity, and helping us shift from reactive to proactive operations.

## What they're looking for

- 3+ years of hands-on experience debugging production systems (logs, traces, incidents, etc.)
- Strong problem-solving skills and ability to dive into unfamiliar backend codebases
- Strong Go and Kubernetes experience
- Familiarity with observability and monitoring tools (e.g., Grafana, Prometheus, Sentry)
- Clear, calm communication under pressure — especially during live incidents

Tags: Engineering
