New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
NVIDIA

AI Computing Development Engineer, TensorRT-LLM

NVIDIA
Apply →
onsite mid full-time Shanghai

First indexed 18 May 2026

Description

Join our fast-paced delivery-focused team as an AI Computing Development Engineer, working on the inferencing software used across our product lines. You will craft and develop robust inferencing software that can be scaled to multiple platforms for functionality and performance.

Key responsibilities include:

  • Performance analysis, optimisation and tuning
  • Closely following academic developments in the field of artificial intelligence and featuring updates to TensorRT-LLM
  • Providing feedback into the architecture and hardware design and development
  • Collaborating across the company to guide the direction of machine learning inferencing, working with software, research and product teams
  • Publishing key results in scientific conferences

We are looking for a candidate with a strong background in computer engineering, computer science, applied mathematics or a related computing-focused degree. Relevant software development experience is also essential. Excellent C/C++ or Python programming and software design skills, including debugging, performance analysis, and test design, are required. A strong curiosity about artificial intelligence and awareness of the latest developments in deep learning are also necessary.

As a member of our team, you will have the opportunity to work on cutting-edge projects, collaborate with talented individuals, and contribute to the development of innovative AI solutions.