Startups · AI

Staff HPC Systems Architect

Lambda Labs · San Jose Office (First St) · Remote

← All jobs
About Lambda Labs

Train and scale AI on NVIDIA VR200 NVL 72, GB300 NVL 72, B300, B200, H200, H100, and and more GPUs. Launch on-demand instances or reserve a cluster. Get started.

About the role

Act as a technical lead during new platform introductions, guiding validation and performance characterization efforts.

More about this role

Lambda, The Superintelligence Cloud, is a leader in AI cloud infrastructure serving tens of thousands of customers. Our customers range from AI researchers to enterprises and hyperscalers. Lambda's mission is to make compute as ubiquitous as electricity and give everyone the power of superintelligence. One person, one GPU.

If you'd like to build the world's best AI cloud, join us.

*Note: This position requires presence in our San Jose, San Francisco, or Bellevue office location 4 days per week; Lambda’s designated work from home day is currently Tuesday.

What You’ll Do

Architect and define scalable compute platforms optimized for AI/ML, simulation, and high-throughput workloads.

Develop compute system standards and design patterns to ensure consistency, performance, and maintainability across infrastructure.

Evaluate emerging CPU, GPU, and accelerator technologies, owning architectural tradeoff decisions that impact compute density, power, cooling, and total cost.

Collaborate with product and engineering teams to map workload requirements to compute platform capabilities across bare metal and cloud deployments.

Experience converting ambiguous business or customer needs into...

Read the full posting on Lambda Labs's site ↗

Data Center Business

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.