CrowdStrike
CrowdStrike
CrowdStrike is a cybersecurity and SaaS company founded in 2011 that develops cloud-based tools for protecting endpoints, identities, networks, and cloud environments. Its Falcon platform combines artificial intelligence with real-time visibility, threat detection, and response capabilities to help organizations identify and stop sophisticated cyberattacks. CrowdStrike’s work spans cybersecurity, managed detection and response, and applied AI, supporting businesses seeking to strengthen the security of their systems and data.

Engineer II, Site Reliability - Bangalore Hybrid

Join CrowdStrike as a Site Reliability Engineer supporting Temporal workflow infrastructure on Kubernetes. Automate deployments, observability, incident response, and capacity planning for cybersecurity engineering teams.

Description

  • Run Temporal infrastructure in production across multiple environments with Helm, Kubernetes, and FluxCD.
  • Release updates, monitor cluster health, handle alerts, and maintain availability across regions.
  • Automate deployments, upgrades, scaling, and troubleshooting to minimize manual operational work.
  • Support capacity planning and performance optimization by analyzing resource usage, finding bottlenecks, and tuning configurations.
  • Develop observability with metrics, logs, dashboards, and alerting.
  • Take part in on-call coverage and incident response.
  • Create operational runbooks and incident summaries.
  • Apply FluxCD to GitOps and infrastructure-as-code processes.
  • Investigate deployment failures, connectivity issues, and performance degradation.
  • Work with internal engineering teams to onboard them to Temporal and resolve integration problems.
  • Contribute to documentation while building expertise in PostgreSQL, AWS/GCP, Kubernetes networking, Helm, certificate rotation, secret management, and distributed systems operations.

Requirements

  • At least 3 years of experience in DevOps, SRE, platform engineering, or infrastructure roles.
  • Working knowledge of Kubernetes fundamentals, including pods, deployments, services, kubectl, and YAML.
  • Experience deploying applications with Helm charts and values files.
  • Some infrastructure-as-code experience with Terraform, Ansible, FluxCD, or ArgoCD.
  • Exposure to AWS or GCP compute, networking, and storage fundamentals.
  • Ability to create automation scripts in Bash, Python, or Go.
  • Basic understanding of stateful systems and databases, preferably PostgreSQL, including backups, schema management, and connection handling.
  • A willingness to learn and seek assistance when needed.
  • Demonstrated use of AI technologies to support decision-making, streamline workflows and processes, improve efficiency, and drive business outcomes.

Benefits

  • Compensation and equity awards aligned with market standards.
  • Comprehensive physical and mental wellness programs.
  • Competitive vacation and holiday time for rest and recharge.
  • Paid parental and adoption leave.
  • Professional development opportunities available across levels and roles.
  • Employee networks, local neighborhood groups, and volunteer programs that support connection.
  • An engaging office environment with world-class amenities.
  • Great Place to Work Certified™ workplaces worldwide.

Related Jobs