# Member of Technical Staff - Research, Inference at Modal

- Company: Modal
- What the company does: Bring your own code, and run CPU, GPU, and data-intensive compute at scale. The serverless platform for AI and data teams. Backed by Accel, General Catalyst and AI Grant.
- Company website: https://modal.com/
- Type: Startups (AI role)
- Level: Senior
- Location: New York
- Work setup: On-site
- Pay: $150K to $350K base salary per year (USD)
- Posted: 2026-07-05
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/modal/73c97bbc-8e27-4c5d-b38b-90b3afdb0d93
- Page: https://www.1752.vc/careers/jobs/modal-member-of-technical-staff-research-inference/

## About the role

Most of the value of owning a model shows up at serving time. We're building a platform that covers the whole life of an LLM -- train it, deploy it, observe it -- and inference is where teams feel the difference every day. We already run elastic inference, sandboxes, distributed volumes, and multi-node training, and we control the infrastructure underneath, so the serving stack is ours to shape rather than something we resell.

## What they're looking for

- A research-leaning or systems background in LLM inference, with work you can point to
- Fluency in the LLM serving stack, from kernels and quantization up to schedulers and autoscaling
- A record of shipping research or systems that other people build on, whether in a lab or in industry
- The drive to independently take a research bet from idea to result, working in the open with the rest of the team
- Ability to work in-person, in our NYC or San Francisco office

Tags: Engineering
