New The Skills of Tomorrow: how AI-exposed is every skill in 2026? See the data →
NVIDIA

Senior Developer Technology Engineer - Edge Agentic AI

NVIDIA
Apply →
senior full-time Santa Clara, CA

First indexed 23 Jul 2026

Description

At NVIDIA, we're tapping into the unlimited potential of AI to define the next era of computing.

As a Senior Developer Technology Engineer - Edge Agentic AI, you will work closely across internal engineering and product teams as well as external app developers and enterprise ISVs on solving local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX & DGX.

Responsibilities:

  • Work on solving local end-to-end agentic AI GPU deployment challenges on NVIDIA RTX & DGX.
  • Apply powerful profiling and debugging tools for analyzing most demanding accelerated end-to-end agentic AI workflows to detect insufficient system utilization resulting in suboptimal runtime performance.
  • Conduct hands-on trainings, develop sample code and host presentations to give good guidance on efficient end-to-end agentic AI deployment targeting optimal runtime performance.
  • Improve LLM & GenAI user experience by working on feature and performance enhancements of OSS software, including but not limited to projects like GGML, Llama.cpp, Ollama, vLLM, ONNX Runtime.
  • Collaborate with GPU driver and architecture teams as well as NVIDIA research to influence next generation GPU features by providing real-world workflows and giving feedback on partner and customer needs.
  • Provide technical leadership and mentorship to junior engineers, encouraging an inclusive and high-performing team environment.

Requirements:

  • A proven track record 5+ years of professional experience in local GPU deployment, profiling and optimization.
  • A Bachelor's or Master's degree or equivalent experience in Computer Science, Engineering, or a related field.
  • Strong proficiency in C/C++, Python, software design, programming techniques.
  • Familiarity with and development experience on Windows and Linux.
  • Experience with CUDA and NVIDIA's Nsight GPU profiling and debugging suite.
  • Some travel is required for conferences and for on-site visits with external partners.
  • Strong problem-solving skills and the ability to work both independently and collaboratively in a fast-paced environment.
  • Excellent interpersonal and communication skills and a passion for keeping track with the latest advancements in AI technology.

Nice to Have:

  • Experience with GPU-accelerated AI inference driven by NVIDIA APIs and SDKs, specifically TensorRT-RTX, cuDNN, NVIDIA Model Optimizer.
  • Expertise with professional agentic AI use cases, i.e., digital content creation and productivity workflows.
  • Experience working with open-source LLM and GenAI software.
  • Detailed knowledge of the latest generation GPU architectures.
  • Experience with AI deployment on NPUs and ARM architectures.

We offer highly competitive salaries and a comprehensive benefits package, including equity.