SabreSabre

Site Reliability Engineer IV

Bengaluru, Karnataka, IndiaOn-site or hybridFULL_TIMEtodayFresh

Powering the agentic revolution in travel. Sabre is an AI native technology leader, backed by one of the world’s largest travel data clouds.

Java

**Powering the agentic revolution in travel.**Sabre is an AI-native technology leader, backed by one of the world’s largest travel data clouds. Built on an open, modular, cloud-native architecture, Sabre serves as the backbone for both established leaders and bold, new disruptors, guiding them to the next age of travel retailing through intelligent, connected, and personalized experiences. With AI at its core and operating at unparalleled scale, Sabre transforms insights into innovation, empowering airlines, hoteliers, agencies and other partners to retail, distribute and fulfill travel worldwide.

This role is for Hotels platform where the Sabre Hotels organization builds and operates the technology platform that powers hotel distribution for global travel partners. Our systems connect 170,000+ hotels to the world’s largest online and offline travel agencies. The applications processes high-volume transactions, is highly available/reliable, multi-instance, and business-critical, requiring strong site reliability engineering discipline and ownership.

Role and Responsibilities:

  • Provide application on-call support (8–12-hour shifts, one week at a time), responding to system alerts/pages, troubleshooting issues, and taking action when needed.
  • Participate on incident bridges and lead investigation of severe issues affecting customer experience.
  • Review, execute, validate and monitor changes to ensure stability and risk are addressed
  • Build alerting and monitoring solutions for new or existing services in GCP, and other observability tools.
  • Continuously improve reliability by engaging in SRE practices such as blameless postmortems, building SLIs for CUJs, reducing toil and increasing automation.
  • Support development partners with infrastructure deployments, ci-cd pipelines, and other non-functional requirements to be prod-ready
  • Take ownership of services, addressing risks, and keeping up with other KTLO tasks including managing infrastructure currency, PCI audits, capacity monitoring, and cost optimization.
  • Provides technical mentorship and cultural/competency-based guidance to teams.

Qualifications and Education Requirements:

  • 5-6 years of relevant experience.
  • Google Cloud Platform knowledge
  • Analysis, debugging and troubleshooting skills, persistence in problem solving
  • Good collaboration within the team and with teams up and across the organization
  • Understanding of CI/CD concepts
  • Familiarity with Jenkins
  • Understanding of container orchestration technologies (Docker, Kubernetes)
  • Linux/UNIX knowledge, Terraform, shell scripting, networking
  • Familiarity with monitoring and alerting tools: AppDynamics, Google Cloud ops, Data dog, Prometheus, Grafana, Elastic search
  • Experience with Change Management process
  • Very good written and verbal English communication skills
  • Willingness to learn
  • Self-disciplined and commitment oriented

NICE TO HAVE SKILLS:

  • Understanding of databases (relational, Oracle, Couchbase, Datastore, Spanner)
  • Java development experience
  • Experience with Sabre Change Management process, Service Now, runbooks
  • Google SRE knowledge

We will give careful consideration to your application and review your details against the position criteria. You will receive separate notification as your application progresses.

Please note that only candidates who meet the minimum criteria for the role will proceed in the selection process.

#LI-Hybrid#LI-NG1

Frequently asked

Is this Site Reliability Engineer IV role remote?

This role is based in Bengaluru, Karnataka, India.

Related