New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
NVIDIA

Senior Solutions Architect, NVIDIA Cloud Partners

NVIDIA
Apply →
remote senior full-time Santa Clara, CA

First indexed 8 Jul 2026

Description

NVIDIA is seeking an experienced Solutions Architect to be a trusted technical advisor, bridging design to deployment of large-scale AI/HPC GPU infrastructure.

The role involves driving outtake and consumption by integrating libraries, frameworks, models, and software applications. You will deliver GenAI, AI, and ML hardware/software to production with key customers and partners.

Responsibilities:

  • Collaborate with NVIDIA Cloud Partners to create, implement, and deliver on NVIDIA's innovative hardware and software solutions.
  • Partner with SAs, Account Managers, Engineering, Product, and business leaders to align on strategies, assess technical needs, secure business opportunities for NVIDIA.
  • Become the primary technical driver for customers during the design, development, construction, integration, and production of GPU Cloud infrastructure and applications throughout the entire customer lifecycle.
  • Conduct regular technical customer meetings for project/product details, feature discussions, intro to new technologies, and debugging sessions.
  • Work closely with customers to build and adopt NVIDIA solutions including PoCs to address critical business needs covering infrastructure, libraries, and applications.
  • Prepare and deliver technical content to customers including presentations, workshops, reference architectures, tutorials, publications.

Requirements:

  • BS/MS/PhD in Electrical/Computer Engineering, Computer Science, Physics, Mathematics, or other Engineering fields or equivalent experience.
  • 10+ years of Solution Engineering (or similar Sales Engineering, Cloud Engineering, Solution Architecture) including experience working directly with partners and customers.
  • Experience crafting and deploying large-scale cluster environments, hands-on experience designing, developing, delivering distributed Cloud architectures.
  • Strong fundamentals in programming, optimizations and software design, especially in Python and Deep Learning frameworks such as PyTorch and TensorFlow.
  • Practical expertise fine-tuning and deploying models, integrating software application stacks, libraries, and frameworks to drive consumption from GPU platforms.
  • Motivation and skills to own and drive complex multi-disciplinary technical engagements with customers throughout the full customer lifecycle and cross-functional teams.
  • Efficient time management and capable of balancing multiple tasks. Excellent presentation, communication and collaboration skills.
  • Self-starter with a passion for growth, continuous learning, and sharing insights.

Preferred Qualifications:

  • Practical experience with NVIDIA GPUs, software libraries, frameworks, and foundation models, such as NVIDIA Nemotron, NVIDIA NeMo Framework, NVIDIA Dynamo, NeMo Retriever, NVIDIA Triton Inference Server, TensorRT, TensorRT-LLM, NVIDIA CUDA-X.
  • Hands-on expertise with scaled AI cloud environments (e.g., AWS, Azure, GCP) and on-premises/hybrid infrastructure, in particular inference and training workloads.
  • Familiarity with NVIDIA hardware (such as GPUs, networking, storage) and systems technology such as NCCL, DCGM, UFM, Mission Control, Base Command Manager.
  • Proficiency with large-scale AI model training/deployment encompassing GPU systems, performance testing, AI benchmarking, fine-tuning, strong focus on MLOps and cluster orchestration (SLURM, K8s, orchestrator, load balancing, cloud architecture).
  • Experience working with enterprise developers and strong customer-facing skills.

Benefits:

  • Equity
  • Comprehensive benefits package