Research Scientist - World Model

Lumalabs Ai

Apply to this job
SF Bay Area, CA Remote, International remote Until 8/21/2026 0+ years exp First posted June 1, 2026 Last posted June 1, 2026
Job description
THE ROLE
This is the role at the center of the thesis. Luma already trains the strongest generative video models in the industry; the next step is turning those models into world models — interactive, controllable, physically faithful, and useful as a substrate for embodied reasoning. As a Research Scientist on the World Models team, you'll work on the next generation of generative models that can be rolled out as worlds.

WHAT YOU'LL DO
- Invent next-generation world model architectures — diffusion, transformer, autoregressive, or hybrid — with a particular focus on controllability and physical consistency.
- Develop controllability mechanisms that let an agent step into the world: action conditioning, view conditioning, long-horizon rollouts.
- Define and own the metrics: physical fidelity, long-horizon coherence, action-following, and downstream usefulness for policy training.
- Run scaling studies that tell us where compute, data, and architecture pay off.
- Publish at the frontier; contribute to the open-source release that is the long-term deliverable.

MINIMUM QUALIFICATIONS
- PhD or equivalent research record in ML, computer vision, robotics, or related fields.
- Deep expertise in at least one of: large-scale generative modeling (video/3D/world), self-supervised representation learning, model-based RL.
- Strong PyTorch and large-scale training experience — you've trained models that hit the limits of a multi-node cluster.
- A research record the field knows (top-venue publications and/or widely-used open releases).

PREFERRED
- Prior work on world models, model-based RL, generative video, neural simulation, or 4D scene representations.
- Experience using generative models for downstream embodied tasks (planning, control, evaluation).
- Excitement about open-sourcing frontier models.

About Luma


Luma’s mission is to build unified general intelligence that can generate, understand, and operate in the physical world.

We believe that multimodality is critical for intelligence. To go beyond language models and build more aware, capable and useful systems, the next step function change will come from vision. So, we are working on training and scaling up multimodal foundation models for systems that can see and understand, show and explain, and eventually interact with our world to effect change.
About this role

Summary

Research and develop next-generation generative world models for embodied reasoning

Job title

Research Scientist - World Model

Experience level

PhD or equivalent

Minimum experience

0+ years exp

Industry

software

Location requirements

Remote and international candidates in SF Bay Area or Singapore allowed

Salary

Not specified

Management role

No

Skills & keywords

Required skills

PyTorchlarge-scale traininggenerative modelingself-supervised learningmodel-based RL

Preferred skills

world modelsmodel-based RLgenerative videoneural simulation4D scene representations

Specializations

generative modelingcomputer visionRLdeep learningtransformer
Locations

Structured locations inferred from the posting.

San Francisco, CA, USA

Remote City

Singapore

Remote City