# Member of Technical Staff, Kernels at Inception Labs

- Company: Inception Labs
- What the company does: We are leveraging diffusion technology to develop a new generation of LLMs. Our dLLMs are much faster and more efficient than traditional autoregressive LLMs. Backed by AI Grant and Amplify.
- Company website: https://www.inceptionlabs.ai/
- Type: Startups (AI role)
- Level: Senior
- Location: San Mateo, United States
- Work setup: On-site
- Pay: $200K to $350K base salary per year (USD)
- Posted: 2026-03-10
- Apply by: 2026-10-12
- Apply: https://jobs.gem.com/inception/am9icG9zdDoQh4DVhIRtu9dWOfwZ04fs
- Page: https://www.1752.vc/careers/jobs/inception-labs-member-of-technical-staff-kernels/

## About the role

We're looking for engineers and scientists to design, optimize, and maintain the compute foundations that power large-scale language model training and inference. You will develop high-performance ML kernels, enable efficient low-precision arithmetic, and improve the distributed compute stack that makes training and serving large models possible.

## What they're looking for

- Design and implement custom ML kernels (CUDA, CuTe, Triton) for core dLLM operations such as attention, matrix multiplication, gating, and normalization, optimized for modern GPU architectures
- Design compute primitives to reduce memory bandwidth bottlenecks and improve kernel efficiency
- Contribute to infrastructure stability and scalability, ensuring reproducibility, consistency across precision formats, and high utilization of compute resources
- BS/MS/PhD in Computer Science, Engineering, or a related field (or equivalent experience)
- Proficiency in CUDA, CuTe, Triton, or other GPU programming frameworks
- Understanding of ML frameworks (PyTorch, TensorFlow) from a systems perspective

