# Software Engineer, Inference at Luma AI

- Company: Luma AI
- What the company does: Luma AI is the creative AI platform for video generation and image creation. Powered by the world's leading video generation models, Ray and Uni, and creative agents handling end-to-end workflows. Trusted by leading agencies and brands. Try it free. Backed by General Catalyst, a16z and CRV.
- Company website: https://lumalabs.ai/
- Type: Startups (AI role)
- Level: Mid level
- Location: Redwood City, CA
- Work setup: Remote
- Posted: 2026-07-24
- Apply by: 2026-10-08
- Apply: https://jobs.ashbyhq.com/lumaai/c4b9ff5f-40b2-40d4-9f6f-8b7aad9b8860
- Page: https://www.1752.vc/careers/jobs/luma-ai-software-engineer-inference/

## About the role

You'll own how Luma's models get served — integrating new architectures into the inference engine, scaling deployments across thousands of machines, and keeping expensive GPU fleets busy while meeting internal SLOs.

## What they're looking for

- Strong Python and system-architecture skills
- Experience deploying models with PyTorch, Hugging Face, vLLM, SGLang, TensorRT-LLM, or similar
- Experience with queues, scheduling, traffic control, and fleet management at scale
- Experience with Linux, Docker, and Kubernetes, and with orchestration, deployment, and scheduling
- Familiarity with Redis and S3-compatible storage

Tags: Research & AI
