New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
NVIDIA

Senior System Software Engineer - Scientific Computing PaaS

NVIDIA
Apply →
remote senior full-time Santa Clara, CA

First indexed 2 Sept 2026

Description

We are seeking a Senior System Software Engineer to help build out our scientific computing platform workflows on Cloud.

This Cloud-based scientific computing cloud platform enables physics-based numerical simulation solvers, AI-based training, inference, and visualization workflows for physical science and engineering problems.

Those applications include weather prediction, climate modeling, industrial design, and digital twins simulation in various domains, such as aerospace, automotive, sports, renewable energy, and biomedical.

Responsibilities:

  • Design services and take ownership of underlying cloud infrastructure for physics-informed and data-driven scientific workflows.
  • Design novel algorithms and actively engage with operations to increase overall system performance across the stack, including deep understanding of application code, such as DL frameworks, numerical solvers, microservices, APIs, and heterogeneous accelerated computing with CPUs and GPUs.
  • Design, build, deploy, and operate scalable I/O infrastructure for checkpointing, data loading, pre- and post-processing of data.
  • Optimize compute, storage, and network architecture specific to physics and simulation-driven applications.

Requirements:

  • BS/MS degree in Computer Science or related areas or equivalent experience.
  • 10+ years of experience working on building and operating distributed compute and data-intensive platforms as a service on the cloud.
  • Proven skill in a compiled language (Go, Rust, C++, or otherwise).
  • Strong foundational knowledge in cloud computing, such as "The Datacenter is a Computer" architecture, cloud security architecture, virtualization, resource pooling, and elasticity.
  • Proven skills in distributed systems and parallel processing, such as system models of distributed computation, topology abstraction, logical time, synchronization, and deadlock detection in distributed systems.
  • Hands-on debugging skills with processes, threads, deadlocks, and synchronization.
  • Strong evidence of algorithmic thinking and system design skills.
  • Be self-motivated, have strong interpersonal skills, and be able to work independently with multiple teams with minimal direction.

Benefits:

  • Equity
  • Benefits package