Mirantis
Mirantis
Mirantis develops cloud-native infrastructure and container management software for organizations building, securing, and operating Kubernetes environments at scale. Its portfolio includes Mirantis Kubernetes Engine, Mirantis OpenStack for Kubernetes, Mirantis Container Cloud, Mirantis Container Runtime, and Mirantis Secure Registry, supporting cluster operations and software supply-chain security. The company also contributes to the open-source ecosystem through projects and tools such as Lens Desktop, a Kubernetes IDE, alongside enterprise technical support. Mirantis works with organizations in consulting, healthcare, logistics, public services, financial services, and technology-focused industries adopting modern cloud infrastructure.

Senior Site Reliability Engineer (SRE)

Senior SRE role focused on deploying NVIDIA-backed Kubernetes and cloud AI infrastructure for Mirantis. The position covers reliability, security, automation, and customer delivery across distributed international teams.

Description

  • Design, develop, deploy, operate, maintain, and troubleshoot open-source cloud and AI infrastructure solutions
  • Deploy AI infrastructure on NVIDIA-certified hardware in line with approved architecture and implementation designs
  • Maintain the reliability, security, and performance of container infrastructure
  • Collaborate with geographically distributed international teams on technical challenges and process improvements
  • Gather and refine technical requirements with stakeholders
  • Improve system performance, reliability, and scalability
  • Investigate, debug, and resolve complex technical issues
  • Contribute to code reviews
  • Design and implement AI-driven automation throughout the DevOps lifecycle
  • Transfer knowledge to customers during delivery phases
  • Mentor team members and Mirantis customers
  • Define technical strategies and integrate cloud and software services seamlessly

Requirements

  • At least five years of professional DevOps experience focused on cloud and infrastructure technologies such as Kubernetes or OpenStack
  • Experience with high-performance data center processing, networking, and storage
  • Exposure to Golang plus working knowledge of Python and JavaScript
  • Strong understanding of distributed systems, microservices architecture, and CI/CD pipelines
  • Advanced problem-solving and debugging skills across networking, storage, Linux, and Kubernetes
  • Knowledge of performance optimization and security
  • Ability to lead technical tasks and work effectively with diverse teams
  • Ability to exercise independent judgment when working directly with customers
  • Excellent written and spoken English
  • Strong customer-facing communication skills
  • Commitment to innovation, continuous learning, and high-quality delivery
  • Willingness to travel internationally and domestically up to 25%
  • Bachelor’s degree in Computer Science or a related field, or equivalent experience
  • At least five years of experience in DevOps, software development, or a comparable role
  • Preferred: experience with network or storage architecture
  • Preferred: experience with high-performance computing or GPU infrastructure, including GPU scheduling, MIG/vGPU, RDMA/RoCE or InfiniBand, NVLink, DCGM health checks, GPU driver and firmware lifecycle, or NVIDIA AI Enterprise
  • Preferred: participation in open-source communities through upstream contributions or conference presentations
  • Preferred: experience with Rancher, OpenShift, and VMware

Benefits

  • Professional development and training opportunities
  • Attendance at conferences and working groups
  • Company outings, happy hours, hackathons, and technology talks
  • Competitive compensation package with a comprehensive benefits plan
  • Collaboration with passionate, talented, and engaging colleagues
  • Work with Fortune 500 and Global 2000 customers on next-generation cloud technologies
  • Participation in open-source innovation
  • High-energy culture centered on openness, collaboration, calculated risk-taking, and continuous growth

Related Jobs

Telix Pharmaceuticals Limited

Senior Clinical Project Manager, Remote (United States)

Telix Pharmaceuticals Limited
501 – 1,000 Employees
BiotechnologyConsultingHealthcare

Lead global clinical trials for Telix Pharmaceuticals’ precision radiopharmaceuticals, overseeing study delivery from protocol finalization through closeout. Coordinate vendors, budgets, risks, and cross-functional teams across complex multinational research projects.

Open
Autodesk

Remote Analytics Engineer, Forecasting and Revenue Intelligence at Autodesk

Autodesk

Build Autodesk’s Snowflake and dbt models for sales forecasting and revenue intelligence. Deliver trusted dashboards and reporting that inform sales leadership decisions.

Open