# Staff Site Reliability Engineer at Anduril

- Company: Anduril
- What the company does: Anduril Industries builds advanced autonomous systems and defense technology to protect US and allied forces. Creating the future of national security through AI, robotics, and cutting-edge engineering. Backed by General Catalyst, Lightspeed and a16z.
- Company website: https://www.anduril.com/
- Type: Startups
- Level: Senior
- Location: Costa Mesa, California, United States
- Work setup: On-site
- Posted: 2026-09-24
- Apply by: 2026-11-08
- Apply: https://boards.greenhouse.io/andurilindustries/jobs/5210335007?gh_jid=5210335007
- Page: https://www.1752.vc/careers/jobs/anduril-staff-site-reliability-engineer/

## About the role

This Staff SRE role sets the reliability architecture for the systems that run Anduril’s business and manufacturing operations. You are not responding to tickets. You are the person who decides how production works: how services are observed, how deployments are safe, how incidents are managed, and how reliability scales as the platform grows. You own the infrastructure and mechanisms that make reliability a property of the platform rather than a function of individual heroics.

## What they're looking for

- 10+ years of experience in site reliability engineering, production engineering, infrastructure engineering, or a closely related discipline, including experience operating at architecture or platform-wide scope
- Deep technical fluency across distributed systems, container orchestration (Kubernetes), cloud platforms (AWS, GCP, or Azure), networking, storage, and the failure modes specific to each layer
- Proficiency in systems programming languages (Go, Python, Rust, or equivalent) used for building production infrastructure, tooling, and automation
- Demonstrated experience defining SRE standards, production-readiness frameworks, or operational maturity models and influencing adoption across engineering teams without formal authority
- Track record of leading incident response for complex, multi-system failures and converting post-incident findings into infrastructure investments that prevent recurrence
- Proven ability to work independently through ambiguity, define reliability strategy without top-down direction, and sequence investments against competing organizational demands

Tags: TESTING

Source: 1752vc Careers, https://www.1752.vc/careers/jobs/anduril-staff-site-reliability-engineer/
