Build and scale faster on the purpose-built AI cloud, engineered from silicon to API. Backed by Accel.
About the role
While many team members joined without extensive AI/ML experience, we have rapidly developed strong expertise in large-scale model serving and AI infrastructure.
What they're looking for
- Performance, quality, and smoke-testing frameworks
- Hyperparameter optimization for inference framework configurations
- Gibberish detection systems
- Automated rollout pipelines for inference framework upgrades
- Diagnostics and observability tooling
- Traffic replay systems
More about this role
Nebius is leading a new era in cloud infrastructure for the global AI economy. We are building a full-stack AI cloud platform that supports developers and enterprises from data and model training through to production deployment, without the cost and complexity of building large in-house AI/ML infrastructure.
Built by engineers, for engineers. From large-scale GPU orchestration to inference optimization, we own the hard problems across compute, storage, networking and applied AI.
Listed on Nasdaq (NBIS) and headquartered in Amsterdam, we have a global footprint with R&D hubs across Europe, the UK, North America and Israel. Our team of 1,500+ includes hundreds of engineers with deep expertise across hardware, software and AI R&D.
About the Product
Token Factory is focused on building a next-generation platform that enables companies to seamlessly integrate AI into their products and workflows. Our vision is to create a powerful, open, and scalable alternative for deploying and managing AI systems—making advanced AI infrastructure more accessible to both fast-growing startups and large enterprises.
We work with a wide range of customers, from AI-first companies to established...
Browse similar: AI jobs · AI startup jobs · Startup jobs · Remote jobs