# Research Engineer, Infrastructure, Inference at Thinking Machines

- Company: Thinking Machines
- What the company does: Connectionism: Research Blog by Thinking Machines Lab. Backed by a16z, Accel and GV.
- Company website: https://thinkingmachines.ai/
- Type: Startups (AI role)
- Level: Mid level
- Location: San Francisco
- Work setup: Remote
- Pay: $350K to $475K base salary per year (USD)
- Posted: 2026-08-04
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/thinkingmachines/4087e0f6-4295-419a-ba02-08e95a74ceea
- Page: https://www.1752.vc/careers/jobs/thinking-machines-research-engineer-infrastructure-inference/

## About the role

We’re looking for an infrastructure research engineer to design, optimize, and scale the systems that power large AI models. Your work will make inference faster, more cost-effective, more reliable, and more reproducible to enable our teams to focus on advancing model capabilities rather than managing bottlenecks.

## What they're looking for

- Bachelor’s degree or equivalent experience in computer science, engineering, or similar
- Understanding of deep learning frameworks (e.g., PyTorch, JAX) and their underlying system architectures
- Experience with inference serving systems optimized for throughput and latency (e.g., SGLang, vLLM)
- Thrive in a highly collaborative environment involving many, different cross-functional partners and subject matter experts
- A bias for action with a mindset to take initiative to work across different stacks and different teams where you spot the opportunity to make sure something ships
- Strong engineering skills, ability to contribute performant, maintainable code and debug in complex codebases

Tags: Research Infrastructure (ML...
