Description
NVIDIA is seeking a Systems Software Engineer to build efficient on-device AI software for RTX and DGX-class systems. The role focuses on high-performance local inference, low latency, efficient memory use, infrastructure, and practical deployment on resource-constrained platforms.
Responsibilities:
- Partner with NVIDIA's software, research, architecture, and product teams to align strategies and technical needs.
- Build and optimize local AI inference stack for RTX, RTX Pro, and DGX GPUs.
- Develop modern inference runtimes and execution stacks for frameworks like Llama.cpp, vLLM, PyTorch, WinML, DXCGC, and TensorRT-RTX.
- Perform end-to-end optimization of AI models, data pipelines, and inference runtimes.
- Debug, optimize, and analyze system-level performance.
Requirements:
- 2+ years of experience with Bachelor's, Master's, or PhD in Computer Science, Software Engineering, Mathematics, or a related field.
- Excellent C++ programming and debugging skills.
- Proven experience with AI inferencing pipelines and applications using ML/DL frameworks.
- Strong analytical and problem-solving abilities.
- Outstanding written and oral communication skills.
Benefits:
- Competitive salaries
- Generous benefits package
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/India-Pune/System-Software-Engineer---Local-AI_JR2022875