We are leveraging diffusion technology to develop a new generation of LLMs. Our dLLMs are much faster and more efficient than traditional autoregressive LLMs. Backed by AI Grant and Amplify.
About the role
We seek experienced engineers and scientists to bridge the gap between research and real-world applications by training and deploying our diffusion large language models. You'll build our core product offerings, partner with customers, and ensure our models perform reliably at scale in production environments.
What they're looking for
- Design, develop, and optimize our models for production use cases
- Partner with customers to understand their requirements and translate them into technical solutions
- Implement innovative approaches for post-training generative AI models, including agentic workflows
- Work on data preprocessing pipelines, model evaluation, and alignment to enterprise use cases
- Contribute to the deployment and maintenance of models in production environments
- Collaborate with product teams to design and implement customer-facing ML features
More about this role
Inception creates the world’s fastest, most efficient AI models. Our Mercury model is the world’s fastest reasoning LLM and first commercially available diffusion LLM, delivering 5x greater speed and efficiency than today’s LLMs, with best-in-class quality.
We are the AI researchers and engineers behind such breakthrough AI technologies as diffusion models, flash attention, and DPO.
The Role
We seek experienced engineers and scientists to bridge the gap between research and real-world applications by training and deploying our diffusion large language models. You'll build our core product offerings, partner with customers, and ensure our models perform reliably at scale in production environments.
Key Responsibilities
- Design, develop, and optimize our models for production use cases.
- Partner with customers to understand their requirements and translate them into technical solutions.
- Implement innovative approaches for post-training generative AI models, including agentic workflows.
- Work on data preprocessing pipelines, model evaluation, and alignment to enterprise use cases.
- Contribute to the deployment and maintenance of models in production environments.
- Collaborate...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area