PROS
PROS
PROS develops AI-powered software for pricing, revenue management, and digital selling. Its platform helps businesses make more informed pricing decisions, configure and quote complex offerings, and improve sales efficiency across industries such as airlines, automotive parts, chemicals, consumer goods, manufacturing, and travel. Products include Smart Configure Price Quote (CPQ), Smart Price Optimization & Management (POM), and airline-focused tools for revenue management and dynamic pricing. PROS supports B2B organizations seeking to strengthen customer engagement, increase conversion, and turn pricing and selling data into practical commercial decisions.

Site Reliability Engineer II at PROS in Sofia, Bulgaria (Hybrid)

Improve reliability, automation, and security for PROS airline revenue optimization technology. Monitor cloud systems, resolve incidents, and strengthen deployment infrastructure.

Description

  • Administer, support, troubleshoot, and resolve issues across complex systems and services
  • Monitor service performance, reliability metrics, and infrastructure stability
  • Analyze system performance and identify opportunities for improvement
  • Participate in disaster recovery testing and implement reliability improvements
  • Define and maintain service-level objectives, visualizations, and alerts
  • Partner with product and development teams to address performance bottlenecks and reliability issues
  • Implement and maintain automated deployment processes and self-service tools
  • Develop and troubleshoot automation scripts for operational activities
  • Apply automation to improve system scalability and efficiency
  • Participate in follow-the-sun on-call rotations and respond to incidents
  • Investigate production incidents, determine root causes, and prepare post-incident reports
  • Maintain documentation, user stories, and operational procedures
  • Share knowledge in team sessions and support continuous improvement
  • Develop automation for security audits and vulnerability mitigation
  • Work with security teams to strengthen cloud security posture
  • Conduct thorough post-incident analysis and maintain the resulting documentation

Requirements

  • Working knowledge of operating systems, networking, and database administration
  • Advanced scripting and automation skills for deployment, scaling, and maintenance
  • Proficiency in at least one high-level language: Ruby, Go, or Java
  • Knowledge of automating infrastructure and configuration management
  • Advanced ability to create monitoring and alerting rules with Prometheus and Grafana
  • Ability to implement and optimize cloud environments
  • Knowledge of RESTful API design and development
  • Familiarity with API testing tools such as Postman
  • University degree in computer science or a related field
  • Knowledge of IT security best practices and procedures
  • Excellent command of English
  • Relevant IT certifications are preferred
  • System administration experience is preferred
  • Previous cloud services experience is preferred, including open-source technologies, software development, systems engineering, scripting languages, and multiple cloud provider environments
  • Ability to work both collaboratively and independently
  • Willingness to innovate, learn, and share knowledge
  • Excellent communication, time management, organizational, crisis management, and problem-solving skills

Benefits

  • Flexible working arrangements
  • Ongoing opportunities for learning
  • Opportunities to grow, innovate, and develop professionally

Related Jobs

Kreato Global | BPO and Language Solutions

Remote English-Spanish OPI/VRI Interpreter

Kreato Global | BPO and Language Solutions
201 – 500 Employees
HealthcareHospitalityLogistics

Interpret remotely between English- and Spanish-speaking people in medical, financial, social service, and customer care settings. Provide language support for Kreato Global across Latin America.

Open