# Member of Technical Staff - Compute Platform at Reflection AI

- Company: Reflection AI
- What the company does: Make intelligence open and accessible to all. Backed by Battery, Lightspeed and Sequoia.
- Company website: https://reflection.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: New York, NY
- Work setup: On-site
- Posted: 2026-03-20
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/reflectionai/d9ead368-c050-40ef-a9dc-34a84a2df829
- Page: https://www.1752.vc/careers/jobs/reflection-ai-member-of-technical-staff-compute-platform/

## About the role

Reflection’s Compute Platform team specializes in keeping our compute layer healthy and highly available. We run a K8s-based platform distributed across multiple neo-clouds. We manage multi-cloud scheduling, node health, and performance debugging at this scale presents genuinely hard systems problems. More broadly, you will work closely with Reflection's training teams to co-design fault tolerance, node health checks, and remediation strategies.

## What they're looking for

- • Systems-level engineering experience with a focus on cluster-wide behavior and maintenance
- • Strong coding ability and a demonstrated focus on systems or GPU infrastructure
- • Deep GPU hardware knowledge beyond standard Kubernetes,e.g., familiarity with NCCL
- • Alignment with a K8s-first architecture
- • Cloud storage expertise, specifically managing high-performance data products (like VAST) across multiple data centers, connecting those storage environments together and handling datasets and checkpointing at scale

Tags: Engineering
