# Cluster Operations Software Engineer at Cerebras Systems

- Company: Cerebras Systems
- What the company does: Cerebras powers the world's fastest AI inference on the biggest wafer chip. Cerebras CS-4 delivers up to 30x faster inference than GPUs.
- Company website: https://www.cerebras.ai
- Type: Startups (AI role)
- Level: Mid level
- Location: Sunnyvale, CA
- Work setup: Remote
- Posted: 2026-08-13
- Apply by: 2026-10-12
- Apply: https://jobs.ashbyhq.com/cerebras/f25b9677-32fa-41ed-84d1-0ed704a98533
- Page: https://www.1752.vc/careers/jobs/cerebras-systems-cluster-operations-software-engineer/

## About the role

We are seeking a highly skilled and experienced AI Cluster Operations Engineer to manage and operate our cutting-edge machine learning compute clusters. These clusters would provide the candidate with an opportunity to work with the world's largest computer chip, the Wafer-Scale Engine (WSE), and the systems that harness its unparalleled power.

## What they're looking for

- 6-8 years of relevant experience in managing and operating complex compute infrastructure, preferably in the context of machine learning or high-performance computing
- Proficient in Python and Go, with experience building operational platforms, workflow automation systems, and reliability tooling for large-scale infrastructure environments
- Experience and Expertise in distributed systems is a must
- Deep understanding of Linux-based compute systems and command-line tools
- Extensive knowledge of Docker containers and container orchestration platforms like k8s
- Proven ability to troubleshoot and resolve complex technical issues in a timely and efficient manner

Tags: Software Engineering
