NVIDIA
NVIDIA
NVIDIA develops accelerated computing and artificial intelligence technologies used across gaming, data centers, cloud computing, healthcare, manufacturing, automotive, and robotics. Its work spans graphics processing units, AI platforms, simulation, and industry-focused solutions, including NVIDIA Omniverse for collaborative 3D workflows, NVIDIA DRIVE for autonomous vehicle development, and NVIDIA Clara for healthcare applications. The company brings together large technical teams working on the infrastructure and software that support advanced computing, simulation, and AI-driven analytics.

Senior Developer Relations Manager, NVIDIA Dynamo AI Inference

Grow the developer ecosystem for NVIDIA Dynamo, an open-source framework for distributed AI inference. Create technical resources, support developers, and shape inference product priorities through community feedback.

Description

  • Grow the developer ecosystem for NVIDIA Dynamo, NVIDIA’s open-source framework for distributed AI inference.
  • Work with vLLM, SGLang, Mooncake, llm-d, and other inference projects through joint demos, integration guides, upstream contributions, and community programs.
  • Create runnable demos, sample code, technical documentation, and videos that explain Dynamo’s modular ecosystem and data center performance value.
  • Bring developer feedback into NVIDIA by translating it into engineering and product requirements, influencing priorities, and sharing progress with developers.
  • Represent NVIDIA externally through technical talks, live demos, workshops, and office hours.
  • Help developers troubleshoot real deployments and contribute to projects.
  • Measure repeat participation, contributor growth, and issues resolved.
  • Create technical resources, recurring meetups, and hands-on workshops with inference communities.
  • Turn developer feedback, serving-stack choices, and emerging inference workloads into recommendations for product roadmaps, ecosystem investments, and partner priorities.

Requirements

  • Degree in computer science, engineering, or a related field, or equivalent experience.
  • At least 8 years of experience in the technology industry, including 4 or more years in AI/ML infrastructure, model serving, or distributed systems.
  • Experience leading developer or partner programs across engineering, product, sales, marketing, and legal teams.
  • Hands-on experience deploying and managing model-serving systems, including configuring deployments, monitoring latency and throughput, scaling capacity, and resolving reliability or performance issues.
  • Familiarity with inference engines such as vLLM, SGLang, or TensorRT-LLM.
  • Familiarity with distributed serving stacks such as Dynamo or llm-d.
  • Experience deploying and operating inference workloads on Kubernetes.
  • Experience working with open-source maintainers through issues, pull requests, and design discussions.
  • Experience turning developer feedback into technical priorities.
  • A portfolio of technical guides, working demos, talks, or workshops personally delivered to help developers solve practical problems.
  • Ability to explain inference architecture and performance tradeoffs to audiences ranging from engineers to executives.
  • A record of earning developers’ trust externally and influencing decisions internally.
  • Experience working directly with the vLLM, SGLang, or Mooncake communities.
  • Ability to identify patterns in developer feedback, serving-stack choices, and emerging inference workloads and turn them into recommendations.

Benefits

  • Equity.
  • Employee benefits.
  • Information about NVIDIA employee and family benefits is available at www.nvidiabenefits.com/.

Related Jobs

Knowtion Health

Remote Talent Acquisition Manager

Knowtion Health

Oversee recruiting systems, requisitions, analytics, and contingent workforce operations at Knowtion Health. Help support scalable hiring processes for a growing healthcare company.

Open