Veeam Software
Veeam Software
Veeam Software izstrādā datu noturības un aizsardzības programmatūru organizācijām, kas pārvalda hibrīdus, vairāku mākoņu, virtuālus, fiziskus un SaaS darba slodzes veidus. Veeam Data Platform apvieno dublēšanas, atjaunošanas, drošības, krātuves un datu analītikas iespējas, kā arī atbalsta tādus pakalpojumus kā Microsoft 365, AWS un Google Cloud. Uzņēmuma produkti ir paredzēti, lai palīdzētu uzņēmumiem aizsargāt kritiski svarīgu informāciju, reaģēt uz traucējumiem, tostarp izspiedējprogrammatūras uzbrukumiem, un pārvaldīt datus mainīgā infrastruktūras vidē. Veeam darbība aptver uzņēmumu programmatūru un uz SaaS orientētu datu pārvaldību, radot iespējas profesionāļiem, kurus interesē dublēšana, mākoņtehnoloģijas, kiberdrošība un noturīgas IT darbības.

Staff Site Reliability Engineer at Veeam Software (Onsite, United States)

Lead the development of reliable-by-default platforms for Veeam Data Cloud, a SaaS data resilience and security offering. Drive observability, resilience automation, and distributed-systems reliability across global engineering teams.

Apraksts

  • Build productized reliability capabilities, including libraries, services, controllers, deployment safeguards, rate limiters, circuit breakers, load-shedding adapters, and back-pressure controls.
  • Design and implement observability platforms with telemetry pipelines, SLIs and SLOs, error-budget policies, SDKs, CLIs, plugins, and release gates.
  • Develop change-safety tooling for progressive delivery, automated rollback, release validation, reusable services and operators, and CI/CD integration.
  • Create resilience automation through fault-injection APIs, chaos experiments, traffic shadowing, and load and performance harnesses.
  • Build golden-path platform components with Terraform or Pulumi modules, Kubernetes operators, Helm charts, and reference microservice templates.
  • Automate incident learning by capturing context and timelines, tracking actions, and enabling systemic code fixes.
  • Write production-quality code in Go, TypeScript/Node.js, C#, or Java while designing APIs, testing thoroughly, and delivering iteratively.
  • Lead the design of distributed, multi-region services initially deployed on Azure.
  • Work with Staff and Principal engineers across product and platform teams to establish reliability standards and encourage adoption.
  • Instrument systems, automate detection and response, and maintain actionable alerts.
  • Lead complex incidents, foster blameless learning, and deliver systemic improvements.
  • Mentor senior engineers through design reviews, architecture decision records, and pair programming.
  • Lead strategic initiatives and establish architectural practices for Veeam’s global SRE organization.

Prasības

  • At least 8 years of software engineering experience developing cloud-based products.
  • Substantial experience designing and operating distributed systems at scale.
  • Strong proficiency in one or more backend languages, including C#, Java, Go, or TypeScript/Node.js.
  • Experience building production-grade services and libraries.
  • Hands-on experience working with Kubernetes.
  • Hands-on infrastructure-as-code experience with Terraform or Pulumi.
  • Experience with CI/CD tools such as GitHub Actions, GitLab, or ArgoCD.
  • Practical expertise in metrics, tracing, and logging.
  • Experience translating SLOs and error budgets into engineering workflows.
  • Ability to lead cross-team initiatives, influence architecture, and deliver measurable reliability improvements.
  • Willingness to participate in a follow-the-sun on-call model with 8-hour daytime rotations and coverage.

Priekšrocības

  • Unlimited paid time off.
  • 12 paid holidays, including four global VeeaMe Days for self-care.
  • 24 paid volunteer hours each year through Veeam Cares.
  • Paid parental leave of 8 weeks for all parents and 16 weeks for birthing parents.
  • Medical, dental, and vision coverage beginning on the first day.
  • Mental health support, therapy sessions, and digital wellness tools through the Employee Assistance Program.
  • 401(k) retirement plan with company matching contributions.
  • Fertility, adoption, and surrogacy assistance through Maven.
  • Free 24/7 virtual veterinary care through AirVet.
  • Legal services, identity protection, and supplemental health insurance options.
  • Tax-advantaged spending accounts for healthcare, dependent care, and commuting.
  • Learning and development resources including LinkedIn Learning, O’Reilly, mentoring, workshops, and learning events.
  • Flexible work arrangements.
  • Competitive pay and benefits.
  • Compensatory benefits for on-call rotations according to local policy.

Saistītās vakances

Knowtion Health

Remote Talent Acquisition Manager

Knowtion Health

Oversee recruiting systems, requisitions, analytics, and contingent workforce operations at Knowtion Health. Help support scalable hiring processes for a growing healthcare company.

Atvērt
Napco National

Purchasing Coordinator, Saudi Arabia (Remote)

Napco National

Coordinate purchasing operations for a manufacturing business in Saudi Arabia, from purchase orders and supplier deliveries to customs paperwork and material transfers. Support shipment clearance, invoice processing, supplier claims, and product certificate renewals.

Atvērt
Terumo Medical Corporation

Region Manager, Terumo Interventional Systems Sales

Terumo Medical Corporation

Lead medical device sales across a North Central New Jersey region, managing field teams and hospital relationships. Drive regional revenue, sales performance, and compliant promotion of Terumo Interventional Systems products.

Atvērt