Senior Site Reliability Engineer
Join CellPoint Digital: Shape the Future of Payments with Us!
At CellPoint Digital, we’re revolutionizing the way businesses in the air, travel, and hospitality sectors manage their payments.
Our leading payment orchestration platform helps some of the world’s leading brands make payments simpler, smarter and profitable- from improving payment experiences and maximizing approval rates, to reducing costs and unlocking new opportunities.
But technology is only part of the story. It’s our people who make the difference.
We’re a global team of innovators, problem-solvers and ambitious professionals who care about the work we do and the impact we have. We work across borders, functions and cultures, bringing different perspectives together to solve complex problems and deliver meaningful results for our customers.
Join us as a Senior Site Reliability Engineer on our mission to turn payments into possibilities!
Senior Site Reliability Engineer
The Site Reliability Engineering (SRE) team ensures the reliability, availability, scalability, and performance of a mission-critical payment orchestration platform. The platform operates in a high-volume, API-driven environment and supports integrations with multiple payment service providers.
SREs collaborate closely with the engineering, product, security, and operations teams to maintain resilient systems, lead incident response, implement release management processes, and continuously improve platform stability through automation, observability, and infrastructure best practices.
The Senior SRE provides technical leadership across the SRE function, owns reliability strategy, release management, major incident leadership, and security posture initiatives.
All SRE roles include participation in a 24×7 on-call rotation.
Please note: We are currently seeking candidates who are available to join immediately or within a short notice period. Applications from candidates with extended notice periods will not be considered at this time.
How You Will Make an Impact:
Lead major (P0/P1) production incidents, including cross-team coordination
Define and evolve reliability standards, SLIs, SLOs, and alerting practices
Architect and review highly available, fault-tolerant systems on GCP
Drive large-scale automation initiatives to reduce operational risk and on-call load
Lead release management strategy for production deployments
Own complex Terraform designs and infrastructure reviews
Partner with Security and Platform Engineering to define and strengthen security posture
Mentor Junior SREs and SREs and review operational documentation
Communicate clearly with senior leadership during critical incidents
Skills you will have fine-tuned:
Bachelor’s degree in Computer Science, Information Technology, Engineering, or a related field
6+ years of experience in SRE, DevOps, or platform reliability roles
Deep expertise in GCP, including Kubernetes (GKE), CloudSQL, Spanner, networking, and IAM
Experience supporting high-volume, mission-critical payment or financial systems
Advanced hands-on experience with Terraform and infrastructure automation
Proven ability to lead major incidents and influence cross-functional teams
Excellent communication, documentation, and stakeholder management skills
What's in it for you:
We offer you the opportunity to be an innovator, challenge the status quo, and redefine the payments category
Competitive salary in a fast-growing start-up
Medical insurance with coverage for dependents (parents, spouse, children)
Opportunity for personal and professional growth in a dynamic industry
Please note, we retain applicant data for 12 months for recruitment purposes and future opportunities, after which it is securely deleted. At CellPoint Digital we are committed to be globally compliant with robust data protection.