# Senior Research Engineer, Interactive World Models

**Company**: NVIDIA
**Location**: Santa Clara, CA
**Experience**: senior
**Job type**: full-time
**Category**: Engineering
**Industry**: Technology

**Apply**: https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Research-Engineer--Interactive-World-Models_JR2023826?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply
**Canonical**: https://yubhub.co/jobs/job_a1286cc5-861

## Description

FlashDreams and FastGen are NVIDIA's core technologies for turning video models into real-time world simulations. The stack spans model adaptation for faster generation and richer control, plus the execution layer that runs those models as responsive experiences.

As a Senior Research Engineer, you will lead engineering across model development and runtime systems, building capabilities and turning research into systems that work in real applications.

**Responsibilities:**

- Build and optimize the continuous autoregressive serving loop, including per-step control inputs, model and KV-cache state management, GPU inference, frame streaming, and model integrations to speed-of-light.

- Advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency.

- Lead end-to-end delivery of capabilities such as multi-user experiences and simulation workflows, from prototype through evaluation, integration, and release.

**Requirements:**

- Experience in one or more areas such as video or world models, diffusion and generative modeling, model distillation and adaptation, simulation, robotics, computer vision, or real-time, stateful ML systems.

- MS or PhD in Computer Science, Electrical Engineering, or a related field (or equivalent experience).

- 5+ years of equivalent experience in applied ML or research engineering.

- A record of advancing applied ML or ML systems through research, open-source software, patents, or deployed technology.

- Hands-on experience with Python, PyTorch and GPU-accelerated training, inference, performance optimization, or serving.

**Nice to Have:**

- Experience with post-training generative video models, including distillation, self-forcing, action conditioning, or long-horizon memory.

- Experience building and optimizing real-time, stateful generative inference systems.

- Technical stewardship of an open-source ML project used by researchers or developers.

We offer competitive salaries and a generous benefits package.

## Skills

### Required
- Python
- PyTorch
- GPU-accelerated training
- inference
- performance optimization
- serving
- video models
- world models
- diffusion and generative modeling
- model distillation and adaptation
- simulation
- robotics
- computer vision
- real-time ML systems

### Nice to have
- post-training generative video models
- distillation
- self-forcing
- action conditioning
- long-horizon memory
- real-time stateful generative inference systems
- open-source ML project

---

Source: [Apply at nvidia.wd5.myworkdayjobs.com](https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Research-Engineer--Interactive-World-Models_JR2023826?utm_source=yubhub.co&utm_medium=jobs_feed&utm_campaign=apply)
