Description
NVIDIA is seeking a Senior Systems Software Engineer to develop efficient on-device AI software for RTX and DGX-class systems.
The role focuses on delivering high-performance local inference with low latency, optimized memory utilization, robust infrastructure, and practical deployment on resource-constrained platforms.
Responsibilities:
- Partner with NVIDIA’s software, research, architecture, and product teams to align technical requirements and strategic priorities.
- Build and optimize the local AI inference stack for RTX, RTX Pro, and DGX GPUs.
- Design and develop modern inference runtimes and execution stacks using frameworks such as llama.cpp, vLLM, PyTorch, WinML, DXCGC, and TensorRT-RTX.
- Perform end-to-end optimization of AI models, data pipelines, and inference runtimes.
- Conduct system-level debugging, performance tuning, and performance-accuracy trade-off analysis.
Requirements:
- 5+ Years of experience with Bachelor’s, Master’s, or PhD in Computer Science, Software Engineering, Mathematics, or a related field.
- Excellent C++ programming and debugging skills.
- Proven experience developing and optimizing AI inference pipelines and applications using ML/DL frameworks.
- Deep understanding of inference backends and runtime internals.
- Strong analytical and problem-solving skills.
- Excellent written and verbal communication skills.
Benefits:
- Competitive salaries
- Generous benefits package
- Opportunity to work alongside talented professionals
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/India-Pune/Senior-System-Software-Engineer---LocalAI_JR2021897