# Member of Technical Staff, Backend, LLM Applications at Inception Labs

- Company: Inception Labs
- What the company does: We are leveraging diffusion technology to develop a new generation of LLMs. Our dLLMs are much faster and more efficient than traditional autoregressive LLMs. Backed by AI Grant and Amplify.
- Company website: https://www.inceptionlabs.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: San Mateo, United States
- Work setup: On-site
- Pay: $200K to $350K base salary per year (USD)
- Posted: 2026-03-10
- Apply by: 2026-10-12
- Apply: https://jobs.gem.com/inception/am9icG9zdDr59IN5LAF_rI9KiFhT8kI7
- Page: https://www.1752.vc/careers/jobs/inception-labs-member-of-technical-staff-backend-llm-applications/

## About the role

We seek experienced backend engineers to own the systems that serve our diffusion LLMs in production. You'll build and operate infrastructure that handles billions of inference requests — optimizing for latency, throughput, cost, and reliability. This role sits at the intersection of ML systems and backend infrastructure.

## What they're looking for

- Design, build, and operate scalable backend services and model serving infrastructure for our diffusion LLMs
- Implement and manage load balancing, autoscaling, and traffic routing for model endpoints
- Build systems for model versioning, canary deployments, and zero-downtime rollouts
- Develop monitoring, alerting, and observability tooling to ensure SLA compliance and rapid incident response
- Benchmark and evaluate serving frameworks and hardware configurations to inform infrastructure decisions
- BS/MS/PhD in Computer Science or a related field (or equivalent experience)

