Build and scale faster on the purpose-built AI cloud, engineered from silicon to API. Backed by Accel.
About the role
We offer competitive compensation and benefits packages. Actual compensation will be determined based on job-related factors, including experience, skills, qualifications, the level at which the candidate is hired, and geographic location, consistent with applicable law.
What they're looking for
- Leading template and model delivery — from weights through validation to production endpoint
- Coordinating model onboarding and production delivery, from infrastructure readiness to successful customer availability
- Running the Applied AI / Customer Delivery sync — driving open issues to resolution and unblocking teams between meetings
- Coordinating the bring-up of new GPU platforms for serving — engine and kernel readiness, benchmarking against current hardware, and template migration once validated
- Partnering with Token Factory Product on the serverless launch path — ensuring validated templates hand off cleanly into catalogue, model card, and public availability
- Building and maintaining execution plans, identifying risks early, and ensuring blockers are resolved before they impact delivery
More about this role
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.
Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
Role
As a Technical Project Manager in Token Factory, your primary focus will be coordinating complex cross-functional projects, including new model launches, bringing new GPU platforms into production serving, engineering quality and platform-wide improvement projects, and the end-to-end delivery pipeline for new AI models. You will bring together multiple engineering teams, understand the critical path, proactively manage dependencies and risks, and...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area