Cerebras powers the world's fastest AI inference on the biggest wafer chip. Cerebras CS-4 delivers up to 30x faster inference than GPUs.
About the role
Own end-to-end technical execution for server systems and network equipment in Cerebras clusters, including NPIs, platform refreshes, and major component or configuration changes. Drive requirements gathering and technical trade-off decisions, converting inputs into executable plans with clear milestones, readiness gates, and cross-functional deliverables.
What they're looking for
- B.S. or M.S. in Computer Science, Electrical/Computer Engineering, or equivalent experience
- 8+ years in technical leadership, systems engineering, or technical program leadership for server, network, or infrastructure platforms from concept through production
- Experience technically leading complex server and/or datacenter network programs across OEM/ODMs, switch vendors, component suppliers, and internal engineering teams
- Strong knowledge of server architecture—including CPU/NUMA, memory bandwidth, PCIe, NIC, and storage I/O—and networking fundamentals including leaf-spine fabrics, switch platforms, optics, and high-performance interconnects
- Familiarity with Linux server fleet management, including provisioning, firmware/BIOS, drivers, and field triage
- Demonstrated ability to make sound technical decisions, guide cross-functional teams, and drive issues to closure
More about this role
Cerebras Systems builds the world's largest AI chip, 56 times larger than GPUs. This architecture allows Cerebras to deliver industry-leading training and inference speeds; over 10 times faster than GPU-based hyperscale cloud inference services.
This order of magnitude increase in speed is transforming the user experience of AI applications, unlocking real-time iteration and increasing intelligence via additional agentic computation.
Cerebras works with the leading model labs, global enterprises, and cutting-edge AI-native startups. OpenAI recently announced a multi-year partnership with Cerebras, to deploy 750 megawatts of scale, transforming key workloads with ultra high-speed inference.
As a Technical Lead Manager, Infrastructure Hardware (Server and Network Systems) on the Cluster Architecture Team, you will provide technical leadership and drive end-to-end execution of server and network platform programs—including new product introductions (NPIs)—across Cerebras CS-3–based AI clusters. You will lead programs from requirements and technical trade-offs through vendor selection, lab bring-up, qualification, and production rollout. You will be the technical execution owner...
Browse similar: AI jobs · AI startup jobs · Startup jobs