# Member of Technical Staff - ML Performance at Modal

- Company: Modal
- What the company does: Bring your own code, and run CPU, GPU, and data-intensive compute at scale. The serverless platform for AI and data teams. Backed by Accel, General Catalyst and AI Grant.
- Company website: https://modal.com/
- Type: Startups (AI role)
- Level: Senior
- Location: New York
- Work setup: On-site
- Pay: $200K to $350K base salary per year (USD)
- Posted: 2026-04-21
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/modal/af17da5e-23ca-4802-854d-5f0546e1ed32
- Page: https://www.1752.vc/careers/jobs/modal-member-of-technical-staff-ml-performance/

## About the role

We are looking for strong engineers with experience in making ML systems performant at scale. If you are interested in contributing to open-source projects and Modal’s container runtime to push language and diffusion models towards higher throughput and lower latency, we’d love to hear from you!

## What they're looking for

- 5+ years of experience writing high-quality, high-performance code
- Experience working with torch, high-level ML frameworks, and inference engines (vLLM or TensorRT)
- Familiarity with Nvidia GPU architecture and CUDA
- Experience with ML performance engineering (tell us a story about boosting GPU performance — debugging SM occupancy issues, rewriting an algorithm to be compute-bound, eliminating host overhead, etc)
- Nice-to-have: familiarity with low-level operating system foundations (Linux kernel, file systems, containers, etc)

Tags: Engineering
