# Site Reliability Engineer at Picogrid

- Company: Picogrid
- What the company does: Picogrid delivers integrated systems for modern missions, connecting sensors, platforms, and operators across land, sea, air, and space. Backed by Bessemer.
- Company website: https://picogrid.com/
- Type: Startups
- Level: Mid level
- Location: El Segundo, CA
- Work setup: On-site
- Pay: $170K to $195K base salary per year (USD)
- Posted: 2026-07-30
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/picogrid/e952a0e7-841c-4fb9-ac2e-a9bb2f35bce5
- Page: https://www.1752.vc/careers/jobs/picogrid-site-reliability-engineer/

## About the role

As Picogrid's first Site Reliability Engineer you will own production reliability across cloud and edge, from observability and incident response through node lifecycle, stateful workloads, and a fleet of hardware edge devices in the field. You will help build and define the systems, processes and best practices that ensure Picogrid's systems can be relied upon by our warfighters in even the toughest battlefield conditions.

## What they're looking for

- 3+ years of experience as an SRE or related roles
- Deep Kubernetes operations experience: node lifecycle, workload scheduling, StatefulSets, graceful drains, and live cluster debugging
- Experience designing comprehensive observability dashboards and high signal-to-noise ratio alerting rules
- You are a competent and experienced incident responder practicing methodical evidence-first triage, blameless postmortems, and turning incidents into durable guardrails
- Production Terraform or OpenTofu experience
- Fluent in AWS including IAM, networking, multi-account environments, and account and workload hardening

Tags: Engineering
