Startups · AI

Member of Technical Staff, Forward Deployed AI Engineer

Inception Labs · San Mateo, United States · On-site

← All jobs
About Inception Labs

We are leveraging diffusion technology to develop a new generation of LLMs. Our dLLMs are much faster and more efficient than traditional autoregressive LLMs. Backed by AI Grant and Amplify.

About the role

This role sits at the intersection of product engineering, customer implementation, evals, data collection, model optimization, and enterprise deployment ownership. You will work directly with enterprise customers to identify high-value AI workflows, collect and structure customer data, build LLM-as-judge evaluation systems, tune model and product behavior for customer-specific goals, and turn fast proof-of-concepts into production deployments.

What they're looking for

  • Enterprise customer deployments: Work directly with strategic enterprise customers to identify high-value AI workflows and turn them into production deployments
  • Rapid prototyping: Build and run fast proof-of-concepts, iterating on customer requirements and technical constraints on 2-week cycles
  • Production AI applications: Build full-stack AI applications, agentic workflows, integrations, internal tools, and customer-facing systems that bring Inception models into real enterprise environments
  • Data collection & feedback loops: Collect, structure, and operationalize customer data to improve model and product performance on customer use cases
  • Measurement and Evaluation: Define success metrics for customer deployments and design LLM-as-judge workflows, evaluation harnesses, and feedback loops for customer-specific use cases
  • Model and product optimization: Tune and customize Mercury models, prompts, workflows, and system architecture to meet customer-specific performance goals
More about this role

Inception creates the world’s fastest, most efficient AI models. Our Mercury model is the world’s fastest reasoning LLM and first commercially available diffusion LLM, delivering 5x greater speed and efficiency than today’s LLMs, with best-in-class quality.

We are the AI researchers and engineers behind such breakthrough AI technologies as diffusion models, flash attention, and DPO.

The Role

Inception is hiring Forward Deployed AI Engineers to help enterprise customers deliver the highest quality AI experiences using our diffusion-based language models.

This role sits at the intersection of product engineering, customer implementation, evals, data collection, model optimization, and enterprise deployment ownership. You will work directly with enterprise customers to identify high-value AI workflows, collect and structure customer data, build LLM-as-judge evaluation systems, tune model and product behavior for customer-specific goals, and turn fast proof-of-concepts into production deployments.

This is not a traditional solutions engineering role, a pure research role, or a long-cycle consulting implementation role. We are looking for full-stack engineers who can operate close to...

Read the full posting on Inception Labs's site ↗

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.