# Founding Mid-Training/RL Infrastructure Engineer at Model AI

- Company: Model AI
- What the company does: ModelOp's Enterprise AI Command Center is the system of record that unifies every AI asset so you bring ML, GenAI, and agentic AI to production 10× faster. Backed by a16z.
- Company website: https://www.modelop.com
- Type: Startups (AI role)
- Level: Mid level
- Location: Palo Alto
- Work setup: Remote
- Posted: 2026-09-28
- Apply by: 2026-11-12
- Apply: https://jobs.ashbyhq.com/modelai/5830d9da-5d01-46b3-9f63-d63ea66a1ca3
- Page: https://www.1752.vc/careers/jobs/model-ai-founding-mid-training-rl-infrastructure-engineer/

## About the role

We are looking for a Large-Scale Mid-Training/RL Infrastructure Engineer to help build, optimize, and scale out our in-house foundation model training stack, with an emphasis on both core pretraining and RL-enhanced methods. This role is deeply technical and directly impacts Peano AI's core model product.

## What they're looking for

- Significant experience with large-scale deep learning model training and distributed system design, including large GPU/TPU clusters (GB300/VR200, TPU v7x or similar)
- Proven track record of pretraining and/or RL-based fine-tuning of large models (LLMs or comparable scale)
- Deep familiarity and hands-on experience with frameworks such as Megatron, Transformer-Engine, verl, slime, or similar large-scale/foundation-model toolkits
- Experience with data pipeline design, model/data/optimizer sharding, checkpointing, rollout/reward design, and training operations at scale
- Strong accelerator (NVIDIA GB300/VR200 GPUs, TPU v7x) and memory optimization skills, including at cluster scale
- Excellent debugging, profiling, and performance-tuning abilities in distributed environments

Tags: Technical Staff

Source: 1752vc Careers, https://www.1752.vc/careers/jobs/model-ai-founding-mid-training-rl-infrastructure-engineer/
