# Inference Engineering and Product Lead at Modal

- Company: Modal
- What the company does: Bring your own code, and run CPU, GPU, and data-intensive compute at scale. The serverless platform for AI and data teams. Backed by Accel, General Catalyst and AI Grant.
- Company website: https://modal.com/
- Type: Startups (AI role)
- Level: Senior
- Location: San Francisco
- Work setup: On-site
- Pay: $300K to $350K base salary per year (USD)
- Posted: 2026-09-16
- Apply by: 2026-10-31
- Apply: https://jobs.ashbyhq.com/modal/ead55e1a-873d-4837-8450-95fe8f4c8931
- Page: https://www.1752.vc/careers/jobs/modal-inference-engineering-and-product-lead/

## About the role

Modal's LLM inference platform delivers frontier performance for open-source models with best-in-class elasticity and developer experience, made in part possible by our custom runtime with GPU memory snapshots and multi-cloud substrate .

## What they're looking for

- 10+ years of industry experience, including 3+ years in a leadership role
- Track record building high-performance systems at scale
- Strong background in cloud infrastructure
- Deep knowledge of low-level OS foundations (Linux kernel, file systems, containers, etc.)
- Nice to have: Experience working with LLM inference in production and familiarity with underlying concepts like engines, kernels, routing, KV cache management and speculative decoding

Tags: Engineering
