New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
NVIDIA

Senior Research Engineer, Interactive World Models

NVIDIA
Apply →
senior full-time Santa Clara, CA

First indexed 21 Aug 2026

Description

FlashDreams and FastGen are NVIDIA's core technologies for turning video models into real-time world simulations. The stack spans model adaptation for faster generation and richer control, plus the execution layer that runs those models as responsive experiences.

As a Senior Research Engineer, you will lead engineering across model development and runtime systems, building capabilities and turning research into systems that work in real applications.

Responsibilities:

  • Build and optimize the continuous autoregressive serving loop, including per-step control inputs, model and KV-cache state management, GPU inference, frame streaming, and model integrations to speed-of-light.
  • Advance the production-ready world model frontier by working with researchers on few-step distillation, causal or autoregressive generation, reward fine-tuning, action conditioning, and long-horizon spatiotemporal memory and consistency.
  • Lead end-to-end delivery of capabilities such as multi-user experiences and simulation workflows, from prototype through evaluation, integration, and release.

Requirements:

  • Experience in one or more areas such as video or world models, diffusion and generative modeling, model distillation and adaptation, simulation, robotics, computer vision, or real-time, stateful ML systems.
  • MS or PhD in Computer Science, Electrical Engineering, or a related field (or equivalent experience).
  • 5+ years of equivalent experience in applied ML or research engineering.
  • A record of advancing applied ML or ML systems through research, open-source software, patents, or deployed technology.
  • Hands-on experience with Python, PyTorch and GPU-accelerated training, inference, performance optimization, or serving.

Nice to Have:

  • Experience with post-training generative video models, including distillation, self-forcing, action conditioning, or long-horizon memory.
  • Experience building and optimizing real-time, stateful generative inference systems.
  • Technical stewardship of an open-source ML project used by researchers or developers.

We offer competitive salaries and a generous benefits package.