Startups

Site Reliability Engineer, Kubernetes Platform (Top Secret Clearance)

SpaceX · Hawthorne, CA · On-site

← All jobs
About SpaceX

SpaceX designs, manufactures and launches advanced rockets and spacecraft. The company was founded in 2002 to revolutionize space technology, with the ultimate goal of enabling people to live on other planets. Backed by Kleiner Perkins, a16z and Founders Fund.

About the role

As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved in designing scalable systems capable of supporting a growing volume of data products being generated in mass. We build tools that enable us to work more efficiently, and that help us build software systems that are secure, reliable, and autonomous. Our engineers are responsible for the life cycle of the systems they create, including development, testing, and operational support.

What they're looking for

  • Bachelor’s degree in computer science, information systems, or an engineering discipline, OR 2+ years of professional experience in software, DevOps, or site reliability engineering in lieu of a degree
  • 1+ year of experience with Kubernetes
  • 1+ year of experience with Linux operating systems
  • Experience in Bash, Python, and/or other scripting languages
  • Experience building, maintaining, and scaling on-premises and/or cloud systems
  • Active Top Secret, Top Secret SCI, or DOE Level Q clearance
More about this role

SpaceX was founded under the belief that a future where humanity is out exploring the stars is fundamentally more exciting than one where we are not. Today SpaceX is actively developing the technologies to make this possible, with the ultimate goal of enabling human life on Mars.

As a member of the Classified IT Systems Engineering team, the Site Reliability Engineer is involved in designing scalable systems capable of supporting a growing volume of data products being generated in mass. We build tools that enable us to work more efficiently, and that help us build software systems that are secure, reliable, and autonomous. Our engineers are responsible for the life cycle of the systems they create, including development, testing, and operational support.

  • Develop automation to deploy and manage compute resources both on-premises and in the cloud
  • Build, maintain, and scale on-premises hardware systems designed to host GPU-accelerated machine learning workloads
  • Deploy and manage core infrastructure such as databases, monitoring and storage
  • Closely collaborate with software engineers to create highly scalable, operable and maintainable products
  • Engage in and improve the...

Read the full posting on SpaceX's site ↗

Information Technology - Corporate

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.