NetApp
NetApp

NetApp develops intelligent data infrastructure solutions that help organizations store, manage, protect, and use information across on-premises and cloud environments. Its portfolio includes unified data storage, cloud management, and data protection tools designed to simplify complex data operations and support more efficient workflows.

The company works at the intersection of cloud software, artificial intelligence, APIs, and enterprise data management, helping businesses modernize their storage environments and make data more accessible across platforms. NetApp’s large global organization supports customers across a range of industries and offers opportunities related to building and operating technologies for modern data infrastructure.

Site Reliability Engineering Manager, OpenSearch — NetApp, Bangalore Onsite

Lead the technical operations team responsible for OpenSearch and Kafka clusters within NetApp Instaclustr’s open-source cloud platform. Oversee incidents, migrations, upgrades, service levels, and ongoing operational improvements.

Description

  • Stay current with NetApp Instaclustr offerings, including OpenSearch, Apache Kafka, and supporting systems
  • Lead engineers in delivering strong customer outcomes and service quality
  • Conduct performance reviews and related people-management responsibilities
  • Assess work demands and support requests, set priorities, and ensure prompt customer responses
  • Keep customer issues within SLA targets through training, queue oversight, and resource planning
  • Lead major operational initiatives, including cluster migrations and fleetwide upgrades
  • Plan operational project schedules and assign resources
  • Join customer calls across multiple geographic regions
  • Detect and coordinate conflicting fleetwide activities to reduce operational risk
  • Ensure established processes and procedures are consistently followed
  • Manage team shift rosters
  • Serve on the Level 3 on-call rotation as Major Incident Manager, including rotating weekend coverage
  • Support post-incident review and follow-up work
  • Train new team members and customers on TechOps operating procedures
  • Advance continuous improvement through better tools, processes, and partnership with product engineering

Requirements

  • Demonstrated technical leadership experience with a clear interest in technical management, or recent experience as a Technical Manager
  • At least 8 years of overall professional experience
  • Excellent written and spoken English communication skills
  • Capability to serve as a Major Incident Manager
  • Strong record of managing teams in demanding environments
  • Ability to sustain team morale and foster a positive workplace culture
  • System administration or programming experience is preferred
  • Proven ability to manage multiple tasks simultaneously
  • Maintain the professional qualifications required for the position

Benefits

  • Benefits information was not specified.

Related Jobs