Startups · AI

Research Member of Technical Staff- Video Generation Modeling

Rhoda · Mountain View · On-site

← All jobs
About Rhoda

Redefining Robotic Intelligence. Backed by Khosla.

About the role

Design and train large-scale causal video generation models on web-scale video data Develop and validate training objectives, model architectures, and data mixtures for video prediction at scale

What they're looking for

  • Strong background in large-scale generative modeling — either video generation (autoregressive video models, diffusion transformers, causal video architectures) or language model pretraining (LLMs, autoregressive transformers at scale)
  • Hands-on experience training large generative models from scratch at scale
  • Deep understanding of autoregressive modeling, causal architectures, and scaling behavior
  • Fluency with modern ML frameworks (PyTorch required, JAX a plus)
  • Ability to design experiments, interpret results, and iterate quickly
  • Strong research taste: ability to identify high-leverage questions and cut through noise
More about this role

At Rhoda AI, we’re building the next generation of generalist intelligent robots. We own the full robotics stack from high-performance hardware and robot systems to the infrastructure and state-of-the-art foundation world models that control our robots. Our robots are designed to be generalists capable of operating in complex, real-world environments and handling long-tail edge cases, made possible by our cutting edge research and end-to-end system design. We've raised over $450M and are investing aggressively in model research, infrastructure, hardware development, and manufacturing scale-up to make generalist robotics a reality.

We're looking for Research Scientists and Research Engineers to push the frontier of large-scale pre-training for our video action model. Our approach formulates robot control as video prediction — we pre-train causal video generation models on web-scale video data, then adapt them to predict robot actions from real-world demonstrations. You'll work on the core architectures, training objectives, and scaling strategies that determine how well our models learn from internet-scale video. We hire across levels — from senior to staff — and welcome both...

Read the full posting on Rhoda's site ↗

Research

Build your edge while you search

Free tools for founders and investors, plus VC Unfiltered, our take on startups, venture and the people who build them.