Startups

Site Reliability Engineer (SRE)

Mithril Technologies · Palo Alto / San Francisco Bay Area · On-site

← All jobs
About Mithril Technologies

Mithril builds electrostatically actuated antennas that give space-based radar the adaptability of a phased array at a fraction of the cost of either a phased array or a traditional deployable reflector — monitoring, scanning, and opening access to the... Backed by Techstars.

About the role

You will be a core contributor to the stability and performance of Mithril's global GPU orchestration platform. This is not a 'keep the lights on' role — you will build the automation, observability, and tooling that allows Mithril to coordinate advanced compute across multiple cloud providers at scale, ensuring customers have fast, reliable access to the infrastructure they need.

What they're looking for

  • 3+ years of experience in SRE, Production Engineering, or Infrastructure roles at a high-growth technology company
  • Hands-on Kubernetes experience: comfortable managing clusters, deployments, and troubleshooting production incidents in a multi-tenant environment
  • Cloud proficiency in at least one major provider (AWS, GCP, or Azure), including practical understanding of cloud networking fundamentals (VPC, DNS, load balancing, security groups)
  • Coding ability: proficiency in Python or equivalent (Go, Rust, etc.) — you build tools and services, not just scripts. Willing to pick up new languages as needed
  • Linux fundamentals: strong command of Linux systems, TCP/IP networking, and security best practices
  • Disciplined troubleshooter: calm under pressure during production incidents, with a rigorous approach to RCA and long-term remediation
More about this role

Mithril is an AI infrastructure platform built to make GPU compute more accessible and affordable for the world’s leading enterprises, AI startups, and the AI research community, including LG AI Research, Saronic, and the Broad Institute (among many others).

Founded by a former Google DeepMind research scientist and Stanford CS PhD, Mithril has raised $80M across seed and Series A funding led by Sequoia Capital & Lightspeed Venture Partners. Platform revenue has grown >6x over the past year, and we were recently recognized by Fast Company as the 8th Most Innovative Company in Artificial Intelligence for 2026.

Our engineering team is lean and high-impact. This role is a core infrastructure hire that will shape how Mithril scales its platform across a heterogeneous, multi-cloud environment.

You will be a core contributor to the stability and performance of Mithril's global GPU orchestration platform. This is not a 'keep the lights on' role — you will build the automation, observability, and tooling that allows Mithril to coordinate advanced compute across multiple cloud providers at scale, ensuring customers have fast, reliable access to the infrastructure they need.

You will work...

Read the full posting on Mithril Technologies's site ↗

Engineering

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.