Description
NVIDIA's Local AI team is building the software stack for large language models and generative AI applications to run efficiently on NVIDIA edge AI hardware.
The Senior Software Engineer will track and evaluate innovations in leading open-source LLM inference frameworks, analyze new model architectures and inference algorithms, characterize multi-node inference behavior, produce performance analysis reports, own the model validation workflow, develop and maintain developer-facing inference recipes, and engage with the community and partners.
The ideal candidate will have a BS, MS, or PhD in Computer Science, Computer Engineering, Electrical Engineering, or equivalent experience, 12+ years of software engineering experience with depth in GPU computing, ML systems, or high-performance inference, strong Python or C++ programming skills, hands-on experience with GPU kernel development or optimization, working knowledge of LLM inference internals, container engineering expertise, and strong analytical skills.
NVIDIA offers highly competitive salaries and a comprehensive benefits package, including equity and benefits.