SVP, Site Reliability Engineer
SGX · Singapore, SG · On-site
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
BUILD MARKETS. SHAPE ECONOMIES.
At SGX Group, we create markets. We turn ideas into products and expand access to new opportunities. We are building the future of the exchange in-house: from architecture and platforms to the critical systems that power markets. The biggest decisions are still open to people who join us now.
The opportunity
The SVP, Site Reliability Engineering is a strategic executive leadership role responsible for shaping SGX’s enterprise reliability, observability, automation, and operational resilience agenda across critical platforms and services. This leader will define the long-term SRE vision and operating model, elevate engineering and resilience standards, and ensure reliability is embedded as a strategic differentiator that supports business growth, trusted operations, regulatory confidence, and customer outcomes. Working closely with senior stakeholders across business, technology, infrastructure, security, risk, compliance, and audit, the role leads senior SRE leadership teams and drives alignment between reliability investments, critical service commitments, and enterprise transformation priorities.
SGX Ambition & Career Path:
This is a rare opportunity to shape enterprise reliability at one of Asia’s leading market infrastructures, where technology resilience, market integrity, and customer trust are inseparable. The role offers meaningful executive visibility, direct impact on business-critical outcomes, and the platform to build a world-class SRE capability in a highly regulated, innovation-driven environment.
The Site Reliability Engineering team
Site Reliability Engineering keeps the platforms behind SGX Group's business and market infrastructure available, observable, and recoverable.
The team sets the standards other engineering teams work to across service levels, error budgets, observability, incident response and automation. It works across engineering, infrastructure, security, and product, and it owns the practice as well as seen as the SME for ensuring resilience and proactive maintenance and continuous improvement of the estate .
The next phase is about establishing SRE as a discipline rather than a function, moving reliability decisions upstream into design, and reducing the manual work that currently sits behind keeping services up.
What you will do
Reliability Engineering: Define enterprise reliability strategy and resilience objectives.
Observability: Establish enterprise observability strategy and investment roadmap.
Incident Management: Define the enterprise incident management framework and resilience standards.
SLO / SLA Management: Define enterprise service performance strategy.
Automation: Define the function automation strategy and lead the build-out of agentic AI adoption and improvements.
Performance Engineering: Define the organisational scalability and performance roadmap.
Capacity & Resilience Planning: Own enterprise…