SVP, Site Reliability Engineer
Singapore, SG
Job Summary
The SVP, Site Reliability Engineering is a strategic executive leadership role responsible for shaping SGX’s enterprise reliability, observability, automation, and operational resilience agenda across critical platforms and services. This leader will define the long-term SRE vision and operating model, elevate engineering and resilience standards, and ensure reliability is embedded as a strategic differentiator that supports business growth, trusted operations, regulatory confidence, and customer outcomes. Working closely with senior stakeholders across business, technology, infrastructure, security, risk, compliance, and audit, the role leads senior SRE leadership teams and drives alignment between reliability investments, critical service commitments, and enterprise transformation priorities.
SGX Ambition & Career Path:
This is a rare opportunity to shape enterprise reliability at one of Asia’s leading market infrastructures, where technology resilience, market integrity, and customer trust are inseparable. The role offers meaningful executive visibility, direct impact on business-critical outcomes, and the platform to build a world-class SRE capability in a highly regulated, innovation-driven environment.
Job Responsibilities
- Reliability Engineering: Define enterprise reliability strategy and resilience objectives.
- Observability: Establish enterprise observability strategy and investment roadmap.
- Incident Management: Define the enterprise incident management framework and resilience standards.
- SLO / SLA Management: Define enterprise service performance strategy.
- Automation: Define the function automation strategy and lead the build-out of agentic AI adoption and improvements.
- Performance Engineering: Define the organisational scalability and performance roadmap.
- Capacity & Resilience Planning: Own enterprise resilience and continuity planning.
- Cloud & Platform Operations: Define cloud operations and platform reliability strategy.
- Operational Risk Management: Help build the function operational risk strategy.
- Engineering Leadership: Build enterprise SRE capability and workforce development strategy.
- Agentic AI for Reliability & Operations: Help define the function strategy for AI-enabled reliability engineering, including how AI agents support resilience, incident management, observability, operational risk reduction and measurable improvements in MTTD, MTTR and service stability.
- Partner with senior stakeholders across business, engineering, infrastructure, operations, security, risk, compliance, and audit to align reliability priorities with business strategy, regulatory obligations, and critical service commitments.
- Provide executive leadership during major incidents and crisis scenarios, ensuring decisive cross-functional action, strong stakeholder communication, and continuous improvement.
Job Requirements
- Experience: Distinguished executive-level experience leading Site Reliability Engineering, platform engineering, production engineering or large-scale technology operations within complex, always-on, mission-critical environments.
- Track record: Proven track record of defining enterprise SRE strategies, scaling leadership teams and embedding reliability standards, observability practices and engineering disciplines across critical platforms, with strong fluency in DORA, SLIs, SLOs and error budgets.
- Technical stack: Strong experience with observability platforms such as Datadog, New Relic or Grafana Cloud, and cloud-native infrastructure including AWS multi-account, Kubernetes/EKS and Terraform.
- Standards & environment: Experience operating in regulated, high-availability environments, with strong understanding of operational resilience, technology risk, governance, audit and compliance expectations.
- Communication & leadership: Exceptional executive presence, stakeholder management and communication skills, with strong commercial judgement, vendor management capability and the ability to attract, inspire and retain top-tier engineering talent.
- Education & certifications: Bachelor’s degree in Computer Science, Engineering, Information Systems, or a related discipline; advanced qualifications are advantageous.
Nice to Have
- Exposure to financial market infrastructure, including securities and derivatives trading, clearing, settlement, market operations, or exchange-related systems.
Job Segment:
Compliance, Computer Science, Risk Management, Information Systems, Legal, Technology, Finance