# Infrastructure, Large-scale Training at Hark

- Company: Hark
- What the company does: At Hark, we are building the most advanced personal intelligence in the world.
- Company website: https://hark.com
- Type: Startups
- Level: Mid level
- Location: San Jose
- Work setup: On-site
- Pay: $180K to $450K base salary per year (USD)
- Posted: 2026-03-23
- Apply by: 2026-10-08
- Apply: https://job-boards.greenhouse.io/hark/jobs/4193756009
- Page: https://www.1752.vc/careers/jobs/hark-infrastructure-large-scale-training/

## About the role

We are looking for a Member of Technical Staff, Infrastructure Compute to lead and manage large-scale GPU computing clusters powering our AI training and deployment workloads. You'll work at the intersection of systems engineering and machine learning infrastructure, owning the reliability, scalability, and efficiency of the compute platform that our research and engineering teams depend on.

## What they're looking for

- 5+ years of experience in infrastructure, systems, or platform engineering, with at least 2 years working in ML or HPC environments
- Demonstrated experience managing GPU clusters or large-scale distributed compute infrastructure
- Strong proficiency in at least one systems or infrastructure programming language
- Deep understanding of networking fundamentals (RDMA, InfiniBand, or RoCE a plus) relevant to high-throughput training workloads
- Experience with container orchestration, job scheduling, and multi-tenant resource management
- Proven track record owning production systems with high reliability requirements

Tags: AI Infrastructure
