# Principal Engineer, AI Platform & Infrastructure at SpreeAI

- Company: SpreeAI
- What the company does: SPREEAI is a fast-growing, innovative AI company at the forefront of fashion and e-commerce, revolutionizing how consumers engage with fashion through lifelike photorealistic try-on technology and hyper-personalized shopping experiences.
- Company website: https://spreeai.com
- Type: Startups (AI role)
- Level: Principal and up
- Location: Hybrid (San Francisco, California, US)
- Work setup: Hybrid
- Posted: 2026-04-24
- Apply by: 2026-10-08
- Apply: https://ats.rippling.com/spreeai/jobs/72b22613-9b51-4ad6-b19d-24be942c57b1
- Page: https://www.1752.vc/careers/jobs/spreeai-principal-engineer-ai-platform-and-infrastructure/

## About the role

This role spans ML platform engineering, deployment systems, GPU infrastructure, and observability. You will partner closely with Applied Science, AI Platform, Product, and Partner Engineering to enable rapid research iteration and reliable model delivery at scale.

## What they're looking for

- Build and operate SPREEAI’s end-to-end ML platform spanning training, evaluation, deployment, and monitoring
- Enable scalable and reliable training workflows through orchestration, infrastructure, and resource management systems
- Define platform standards for model packaging, model registry, dataset lineage, experiment tracking, checkpointing, and deployment automation
- Enable reliable and scalable inference deployments through standardized serving, orchestration, and monitoring frameworks
- Build and operate model deployment pipelines with versioning, reproducibility, rollback, approval gates, evaluation gates, and production observability
- Establish production SLOs for latency, availability, error rate, GPU saturation, cold-start time, cost per inference, and model quality drift

Tags: Engineering
