Senior Site Reliability Engineer

Open 1d

Join CellPoint Digital: Shape the Future of Payments with Us!

At CellPoint Digital, we’re revolutionizing the way businesses in the air, travel, and hospitality sectors manage their payments.

Our leading payment orchestration platform helps some of the world’s leading brands make payments simpler, smarter and profitable- from improving payment experiences and maximizing approval rates, to reducing costs and unlocking new opportunities.

But technology is only part of the story. It’s our people who make the difference.

We’re a global team of innovators, problem-solvers and ambitious professionals who care about the work we do and the impact we have. We work across borders, functions and cultures, bringing different perspectives together to solve complex problems and deliver meaningful results for our customers.

Join us as a Senior Site Reliability Engineer on our mission to turn payments into possibilities!

Senior Site Reliability Engineer

The Site Reliability Engineering (SRE) team ensures the reliability, availability, scalability, and performance of a mission-critical payment orchestration platform. The platform operates in a high-volume, API-driven environment and supports integrations with multiple payment service providers.

SREs collaborate closely with the engineering, product, security, and operations teams to maintain resilient systems, lead incident response, implement release management processes, and continuously improve platform stability through automation, observability, and infrastructure best practices.

The Senior SRE provides technical leadership across the SRE function, owns reliability strategy, release management, major incident leadership, and security posture initiatives.

All SRE roles include participation in a 24×7 on-call rotation.

Please note: We are currently seeking candidates who are available to join immediately or within a short notice period. Applications from candidates with extended notice periods will not be considered at this time.

How You Will Make an Impact:

  • Lead major (P0/P1) production incidents, including cross-team coordination

  • Define and evolve reliability standards, SLIs, SLOs, and alerting practices

  • Architect and review highly available, fault-tolerant systems on GCP

  • Drive large-scale automation initiatives to reduce operational risk and on-call load

  • Lead release management strategy for production deployments

  • Own complex Terraform designs and infrastructure reviews

  • Partner with Security and Platform Engineering to define and strengthen security posture

  • Mentor Junior SREs and SREs and review operational documentation

  • Communicate clearly with senior leadership during critical incidents

Skills you will have fine-tuned:

  • Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field

  • 6+ years of experience in SRE, DevOps, or platform reliability roles

  • Deep expertise in GCP, including Kubernetes (GKE), CloudSQL, Spanner, networking, and IAM

  • Experience supporting high-volume, mission-critical payment or financial systems

  • Advanced hands-on experience with Terraform and infrastructure automation

  • Proven ability to lead major incidents and influence cross-functional teams

  • Excellent communication, documentation, and stakeholder management skills

What's in it for you:

  • We offer you the opportunity to be an innovator, challenge the status quo, and redefine the payments category

  • Competitive salary in a fast-growing start-up

  • Medical insurance with coverage for dependents (parents, spouse, children)

  • Opportunity for personal and professional growth in a dynamic industry

Please note, we retain applicant data for 12 months for recruitment purposes and future opportunities, after which it is securely deleted. At CellPoint Digital we are committed to be globally compliant with robust data protection.

See also

要針對這個職缺調整履歷嗎?

目前無法檢查您與這個職缺的符合程度;請先將履歷加入個人檔案,下次即可查看。

A new version of freehire is available