# Member of Technical Staff, Cluster Administration at Inferact

- Company: Inferact
- What the company does: Inferact is a startup founded by creators and core maintainers of vLLM, the most popular open-source LLM inference engine. Our mission is to grow vLLM as the world. Backed by Sequoia and Redpoint.
- Company website: https://inferact.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: San Francisco
- Work setup: On-site
- Pay: $200K to $400K base salary per year (USD)
- Posted: 2026-08-21
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/inferact/595cbea0-7099-4416-a87c-efcc2876e654
- Page: https://www.1752.vc/careers/jobs/inferact-member-of-technical-staff-cluster-administration/

## About the role

We're looking for a hands-on cluster administration engineer to own and operate the high-performance GPU compute infrastructure that keeps Inferact engineering productive. Inferact runs on expensive, high-performance GPU and HPC clusters across neo-cloud and dedicated compute providers. Your job is to make sure that infrastructure is healthy, available, observable, and usable around the clock.

## What they're looking for

- Bachelor's degree or equivalent experience in computer science, engineering, systems administration, or similar
- Hands-on experience administering large compute clusters, HPC environments, university or research clusters, supercomputing systems, or production GPU clusters
- Strong Linux systems administration fundamentals across networking, processes, storage, package management, shell scripting, logs, access control, and system debugging
- Experience operating GPU servers, including driver management, GPU health monitoring, node failures, memory errors, scheduler issues, and hardware diagnostics
- Experience with cluster scheduling and resource allocation using SLURM, Kubernetes, or equivalent tooling
- Ability to own urgent infrastructure incidents end-to-end when compute issues are blocking engineering teams

Tags: Research & Engineering
