
Staff AI Training Infrastructure Engineer
Designworks Talent LLCLead the development of distributed GPU training infrastructure for large AI models. Improve reliability, efficiency, fault tolerance, and production training operations at scale.


Lead the development of distributed GPU training infrastructure for large AI models. Improve reliability, efficiency, fault tolerance, and production training operations at scale.

Optimize GPU kernels, utilization, latency, and throughput for a next-generation AI cloud platform. Improve performance across large-scale training and inference infrastructure.

Senior engineer responsible for Kubernetes orchestration and virtualization supporting GPU-intensive AI and HPC workloads. Builds secure, scalable, multi-tenant infrastructure for an AI cloud platform.

Lead GPU kernel optimization for an AI infrastructure cloud platform, improving utilization, latency, throughput, and scalability. Work across training and inference workloads to strengthen performance across the GPU fleet.

Lead the development of production model-serving systems for an AI infrastructure company. Improve GPU utilization, latency, throughput, scalability, and reliability across large-scale AI workloads.

Vadīt lietišķos pētījumus par mākslīgā intelekta infrastruktūru, secināšanas sistēmām un paātrinātāju stratēģiju. Veidot inženierijas, produktu, finanšu un komerciālus lēmumus mērogojamām mākslīgā intelekta slodzēm.