Seismic
Seismic
Seismic ir digitālā mārketinga platforma, kas palīdz uzņēmumiem veidot ciešāku saikni ar savu auditoriju, izmantojot QR kodus, saīsinātus URL, mobilajām ierīcēm pielāgotas galvenās lapas un saišu pārvaldības rīkus. Tās analītikas iespējas sniedz komandām pārskatāmību par iesaisti, savukārt pielāgojami satura rīki atbalsta kampaņas mazumtirdzniecības, veselības aprūpes, finanšu pakalpojumu un citās nozarēs. Seismic vienā platformā apvieno saziņu ar auditoriju, zīmolam atbilstošu pieredzi un veiktspējas izsekošanu.

Senior Site Reliability Engineer in Hyderabad (On-site)

Senior SRE role focused on improving reliability across AWS, Azure, IBM Cloud, and OCI for Seismic’s AI-powered revenue execution platform. The position covers automation, observability, incident response, and proactive risk reduction.

Apraksts

  • Create and maintain automation and operational tooling that minimizes repetitive work
  • Advance reliability practices alongside Product and Engineering leadership
  • Help improve the full incident-management lifecycle, from detection and escalation through mitigation, communication, review, and corrective actions
  • Join the Global SRE team’s 12-hour follow-the-sun on-call rotation
  • Keep incident-management practices customer-focused, evidence-based, blameless, and consistent across teams
  • Analyze alert, incident, support, and SLO trends to replace reactive response with proactive risk reduction
  • Connect incident learnings with engineering standards, service maturity, product priorities, and vendor actions
  • Collaborate with application engineering teams to enhance developer experience and reduce operational toil
  • Work with service owners to establish production-readiness standards before releases
  • Advise on capacity planning, resilience testing, game days, disaster recovery, and modernization of fragile or legacy systems
  • Map critical customer workflows, define service health expectations, identify dependencies, and align reliability work with business priorities
  • Contribute to cross-team reliability initiatives and influence decisions without direct authority
  • Develop strategic vendor partnerships across observability, incident response, and cloud infrastructure
  • Apply AI-assisted and agentic workflows to alert triage, incident mitigation, postmortems, trend analysis, capacity planning, SLO analysis, and self-service knowledge
  • Maintain qualified human oversight of actions that may affect production
  • Strengthen reliability-workflow context using service metadata, observability data, incident records, runbooks, architecture documentation, and corrective-action quality

Prasības

  • Professional experience in a production-facing SRE role supporting a complex SaaS environment
  • Sound technical judgment across distributed systems, multi-cloud platforms, Kubernetes, networking, infrastructure technologies, GitOps, and CI/CD
  • Experience creating and maturing SRE practices involving SLOs, error budgets, observability, capacity planning, incident response, and toil reduction
  • Ability to use observability data to resolve severe incidents and conduct root-cause analysis
  • Experience leading critical incidents and communicating with engineers, executives, customer-facing teams, and external vendors
  • Bachelor’s or master’s degree in computer science or a related discipline
  • At least 6 years of software engineering experience
  • At least 4 years in DevOps roles focused on developing CI/CD pipelines
  • Advanced knowledge of Kubernetes, Docker, and orchestration platforms
  • Strong hands-on experience with AWS, Azure, or GCP
  • Proficiency with infrastructure-as-code tools such as Terraform, Chef, or Ansible
  • Practical experience with New Relic, Prometheus, Grafana, or comparable observability platforms
  • Proficiency in Python, Go, Bash, or similar programming and scripting languages
  • Knowledge of event-driven autoscaling and advanced Kubernetes configuration
  • Familiarity with Buildkite, Spinnaker, GitHub Actions, or comparable CI/CD platforms
  • Experience with microservices, containerization, and DevOps operating practices
  • Strong understanding of distributed systems, scalability, high availability, and performance optimization
  • Excellent problem-solving ability in a fast-paced work environment

Priekšrocības

  • An inclusive workplace culture that supports employee growth and belonging
  • Participation in a 12-hour follow-the-sun on-call rotation

Saistītās vakances

Deutsche Windtechnik

Site Coordinator, Wind Energy Operations

Deutsche Windtechnik

Coordinate project documents and maintenance reporting for Deutsche Windtechnik’s wind energy operations in Taiwan. Work with engineers, clients, and other stakeholders to manage permits, safety records, and document control.

Atvērt
Ambush

Senior Machine Learning and AI Engineer

Ambush

Build production-grade generative AI and agent workflows for Ambush’s financial services clients. Develop Python and FastAPI services, SQL data workflows, retrieval-augmented generation, and cloud-based AI systems.

Atvērt
GFN

Senior Software Engineer, Machine Learning

GFN

Build backend infrastructure and AI-enabled learning products for GFN, a German education provider. Develop scalable educational platforms with a focus on robust software architecture and practical machine learning.

Atvērt
Upwind Security

AI Engineer, LLM Agents

Upwind Security
51 – 200 Darbinieki
DrošībaMākoņdrošībaSaaS

Develop and deploy LLM agents for Upwind Security’s cloud security platform. Research, evaluate, and build autonomous workflows for investigating and remediating runtime threats.

Atvērt
TEKsystems

Remote Data Program Project Manager — Washington

TEKsystems

Coordinate enterprise data initiatives, report migrations, and adoption of new processes and reporting solutions for TEKsystems. Track project progress, risks, deliverables, and communications with business and technical stakeholders.

Atvērt