Develop, deploy, and grow with Unity, the world’s leading 3D game engine. Build for all major platforms from mobile, to PC and console as well as XR, acquire players, monetize your game, and power industrial applications. Backed by Sequoia.
About the role
We are building the next generation of AI-driven game experiences — generative world models, neural rendering, and multi-modal understanding that turn images, text, and 3D primitives into interactive worlds. As our Staff Machine Learning Engineer, you will be a core technical leader bringing state-of-the-art computer vision and multi-modal models — transformers, diffusion networks, vision-language models (VLMs), and JEPA-style architectures — from research into robust, production-grade systems.
What they're looking for
- 6+ years in ML engineering, with significant depth in computer vision and/or multi-modal modeling
- Proven production experience with transformer-based and diffusion-based vision models (e.g., ViT, CLIP/SigLIP-style encoders, Stable Diffusion, DETR/SAM-style architectures)
- Strong command of the full model lifecycle: data curation, training and fine-tuning, evaluation, and serving at scale
- Familiarity with efficient attention, diffusion samplers, multi-modal fusion, and vision-language alignment techniques
- Strong Python and modern deep-learning tooling (PyTorch), solid software
- engineering fundamentals
More about this role
The opportunity
We are building the next generation of AI-driven game experiences — generative world models, neural rendering, and multi-modal understanding that turn images, text, and 3D primitives into interactive worlds. As our Staff Machine Learning Engineer, you will be a core technical leader bringing state-of-the-art computer vision and multi-modal models — transformers, diffusion networks, vision-language models (VLMs), and JEPA-style architectures — from research into robust, production-grade systems.
This is a deeply hands-on, high-impact role. You will help define the modeling and\ deployment strategy, drive architectural decisions across the ML stack, and mentor a team\ of senior and mid-level engineers. Your work will directly shape the quality, capability, and\ performance of AI features experienced by billions of players — across cloud, server, and\ on-device targets.
Technical Leadership
- Help set the technical vision and roadmap for computer vision and multi-modal AI models, spanning transformers, diffusion models, vision-language models, and JEPA-style generative architectures.
- Drive design and implementation of models for image and video understanding,...
Browse similar: AI jobs · AI startup jobs · Startup jobs · San Francisco Bay Area