Designworks Talent LLC
Designworks Talent LLC
Designworks Talent LLC este o firmă de recrutare și consultanță în domeniul talentelor, care deservește startup-uri cu creștere accelerată și întreprinderi globale. Fondată în 2009, compania combină recrutori cu experiență și instrumente bazate pe inteligență artificială pentru a sprijini procese de angajare mai rapide și mai bine direcționate. Activitatea sa acoperă căutări complete de candidați, recrutare bazată pe proiecte și modele flexibile la cerere, alături de strategii de recrutare scalabile, operațiuni de recrutare, coordonarea achiziției de talente la nivel de întreprindere, identificarea candidaților pe baza datelor și experiența candidaților. Abordarea echipei este concepută pentru a oferi organizațiilor un sprijin adaptabil, pe măsură ce nevoile lor de recrutare evoluează.

Senior GPU Performance Kernel Engineer - Bellevue Hybrid

Optimize GPU kernels, utilization, latency, and throughput for a next-generation AI cloud platform. Improve performance across large-scale training and inference infrastructure.

Descriere

  • Profile, analyze, and optimize GPU kernels to improve latency, throughput, and overall utilization
  • Locate and remove data-plane bottlenecks that limit GPU performance across large-scale AI workloads
  • Tune performance-sensitive workloads for both training and inference environments
  • Partner with AI infrastructure, machine learning, and platform engineering teams to understand workloads and improve system behavior
  • Create benchmarking methods and performance measurement practices for GPU infrastructure
  • Assess emerging GPU technologies, profiling tools, and optimization methods as hardware platforms develop
  • Advance engineering practices that increase GPU efficiency, scalability, and fleet reliability

Cerințe

  • Substantial experience developing and optimizing GPU kernels with CUDA, ROCm, or comparable GPU programming frameworks
  • Proven ability to improve GPU utilization, lower latency, or raise throughput for production AI workloads
  • Deep understanding of GPU architecture, memory hierarchy, parallel computing, and the path from application code to hardware execution
  • Experience profiling and debugging performance problems in complex AI or distributed computing systems
  • Ability to take ownership of technically complex problems and deliver solutions in a fast-paced engineering environment
  • Strong systems programming foundation and performance engineering approach
  • Preferred: experience optimizing workloads for both NVIDIA and AMD GPU architectures
  • Preferred: experience with GPU compilers, runtime optimization, or low-level systems performance
  • Preferred: contributions to open-source GPU performance projects, compiler tools, or AI systems optimization
  • Preferred: experience with large-scale AI training, inference platforms, HPC environments, or cloud GPU infrastructure
  • Preferred: familiarity with Nsight Systems, Nsight Compute, ROCm profiling tools, or comparable technologies
  • U.S. work authorization is required
  • Visa sponsorship is not currently available
  • Must be willing to work in a hybrid arrangement in downtown Bellevue, Washington, with at least three in-office days per week once the permanent office is established

Beneficii

  • Some roles qualify for merit increases, annual bonuses, and long-term incentives tied to individual performance
  • Medical, dental, and vision coverage for U.S.-based employees
  • 401(k) plan with company matching
  • Paid holidays each calendar year
  • Hybrid work arrangement

Locuri de muncă similare

Kreato Global | BPO and Language Solutions

Remote English-Spanish OPI/VRI Interpreter

Kreato Global | BPO and Language Solutions

Interpret remotely between English- and Spanish-speaking people in medical, financial, social service, and customer care settings. Provide language support for Kreato Global across Latin America.

Deschide