# Inference Engineering, Co-op at Inferact

- Company: Inferact
- What the company does: Inferact is a startup founded by creators and core maintainers of vLLM, the most popular open-source LLM inference engine. Our mission is to grow vLLM as the world. Backed by Sequoia and Redpoint.
- Company website: https://inferact.ai/
- Type: Startups (AI role)
- Level: Internship
- Location: San Francisco
- Work setup: On-site
- Posted: 2026-09-15
- Apply by: 2026-10-30
- Apply: https://jobs.ashbyhq.com/inferact/2b6032f9-12a5-4083-b5e6-4bec5376cbaa
- Page: https://www.1752.vc/careers/jobs/inferact-inference-engineering-co-op/

## About the role

Inferact's mission is to grow vLLM as the world's AI inference engine and accelerate AI progress by making inference cheaper and faster. Founded by the creators and core maintainers of vLLM, we sit at the intersection of models and hardware, a position that took years to build.

## What they're looking for

- Currently pursuing a bachelor's, master's, or PhD degree in computer science, engineering, mathematics, or a related technical field, and eligible for a University of Waterloo co-op work term
- Strong programming ability in Python, C++, Rust, Go, or another systems-oriented language
- Strong computer science fundamentals and the ability to learn from research papers, technical documentation, and complex systems code
- Evidence that you have built and debugged nontrivial software through coursework, research, a prior internship, open-source work, or an ambitious side project
- A builder mindset: you define what success means, measure results, validate correctness, communicate clearly, and keep iterating until the system works
- Depth in at least one relevant area such as ML or inference systems, distributed systems, GPU or accelerator programming, compilers, high-performance computing, operating systems, networking, Kubernetes, or cloud infrastructure

Tags: Research & Engineering
