RELX
RELX
RELX dezvoltă instrumente de analiză și de luare a deciziilor bazate pe informații pentru clienți profesioniști și companii din întreaga lume. Produsele sale combină conținut specializat, date și tehnologie pentru a sprijini decizii mai bune și rezultate mai solide în domeniile riscurilor, științei, tehnologiei și medicinei, juridic și expozițiilor. RELX activează și pe piețe asociate consultanței, serviciilor medicale și asigurărilor, ajutând organizațiile să își îmbunătățească productivitatea, să gestioneze riscurile, să avanseze cercetarea, să susțină activitatea juridică și să desfășoare tranzacții comerciale mai eficiente.

Site Reliability Engineering Lead at RELX (Remote, United States)

Lead SRE teams and reliability programs for LexisNexis Risk Solutions’ cloud-based risk platforms. Drive Kubernetes, Terraform, Azure, automation, incident response, and operational excellence.

Descriere

  • Lead, mentor, and develop SREs through regular 1:1s, performance reviews, and career planning
  • Manage hiring, onboarding, team capacity, and resourcing decisions
  • Establish team objectives, prioritize the backlog, and lead planning activities
  • Promote blameless incident learning and collaboration across Development, Security, and Product
  • Direct reliability improvements across infrastructure and services
  • Coordinate incident response and ongoing service enhancements
  • Advance platform-wide automation and operational excellence
  • Help build scalable, secure, and resilient cloud-native environments
  • Manage a small to medium-sized team, including performance, compensation, and recruitment responsibilities
  • Lead post-incident reviews and ensure root-cause analyses are completed promptly
  • Evaluate and resolve issues whose effects extend beyond the immediate team

Cerințe

  • Expertise in Kubernetes cluster architecture, upgrades, autoscaling, security hardening, and large-scale troubleshooting
  • Advanced Terraform experience covering modular infrastructure as code, state management, multi-environment provisioning, and policy as code
  • In-depth Azure knowledge spanning compute, networking, Azure AD identity, storage, and cost optimization
  • Experience designing and scaling CI/CD pipelines with GitHub Actions, release strategies, and automated rollback
  • Experience with Prometheus, Grafana, OpenTelemetry, and SLO, SLA, and error-budget practices
  • Strong automation capabilities aimed at reducing toil through self-healing systems and automated infrastructure
  • Advanced Python, Bash, and/or PowerShell skills for tooling and automation
  • Deep understanding of TCP/IP, DNS, load balancing, VPNs, and cloud-native networking
  • Background in SRE, DevOps, or infrastructure engineering, including engineering team leadership
  • Demonstrated success leading incident response and improving service reliability

Beneficii

  • Annual incentive bonus
  • Benefits tailored to the employee’s country
  • Support for accommodations or adjustments during the hiring process

Locuri de muncă similare

Kreato Global | BPO and Language Solutions

Remote English-Spanish OPI/VRI Interpreter

Kreato Global | BPO and Language Solutions

Interpret remotely between English- and Spanish-speaking people in medical, financial, social service, and customer care settings. Provide language support for Kreato Global across Latin America.

Deschide