Description
NVIDIA is seeking a Senior Software Engineer to help develop distributed storage services for AI/ML. You will work closely with the broader NVIDIA team to design and build a reliable, scalable, and efficient storage-as-a-service tailored to AI applications that can be deployed anywhere and scale without limitations.
Responsibilities:
- Lead the overall architecture and design of the distributed storage service optimized for AI/ML
- Develop and maintain distributed, robust, and scalable Go programs deployed to state-of-the-art open-source ecosystems, including Kubernetes
- Develop and maintain user-space applications, containers, Go-bindings, and CLI tools
- Build features for a distributed storage service to enhance availability and reliability for large-scale deployments
- Engage and collaborate with NVIDIA Research, Computing, Product teams, cross-functional teams, and external customers to deliver Cloud services
- Automate distributed storage service end-to-end, including deployment, management, and monitoring
Requirements:
- Bachelor’s of Science in Computer Science, or related field (or equivalent experience) with 8+ years of industry experience
- Strong background in developing distributed systems involving Golang, Kubernetes, and Cloud Service Provider integrations
- Strong track record of delivering distributed services in a variety of distributed computing environments
- Experience in implementing storage services and interfaces to ensure scalable, high-performance, and reliable solutions
- History of ownership of product delivery from inception to support
- Experience developing and maintaining enterprise software
- Great communication and presentation skills
Preferred Qualifications:
- Experience architecting, building, and deploying a distributed service that runs on large-scale clusters, multi-petabyte to exabyte in size, with millions of users
- Ownership of all lifecycle stages of software development and delivery
- Passion for innovating and investing in groundbreaking technologies, particularly in accelerated Computing environments such as GPU Direct Storage, DPU, and RDMA
You will also be eligible for equity and benefits.
This listing is enriched and indexed by YubHub. To apply, use the employer's original posting:
https://nvidia.wd5.myworkdayjobs.com/en-US/NVIDIAExternalCareerSite/job/US-CA-Santa-Clara/Senior-Distributed-Software-Engineer--Golang---DGX-Cloud_JR2024570-1