S

Senior Vice President - Site Reliability Engineering

SGX Group

Singapore · Full Time

Be the first to apply

Experience
Any
Salary
Openings
1
Posted
41 minutes ago
Work mode
In office
Education
Bachelor's degree in Computer Science, Engineering, Information Systems, or related field
Resume
Required to apply

Where you'll work

Sign in to tell us what does and doesn't work for you here — it sharpens every match we show you.

Job description

About the Role

At SGX Group, renowned for creating markets and shaping economic futures, the Senior Vice President (SVP) of Site Reliability Engineering plays a pivotal leadership role. This strategic executive is accountable for formulating and guiding SGX's enterprise-wide agenda on reliability, observability, automation, and operational resilience across essential platforms and services. The SVP will craft a long-term vision and operational model for Site Reliability Engineering (SRE), raise engineering and resilience benchmarks, and ensure that reliability becomes a strategic advantage that fuels business expansion, trusted operations, regulatory compliance, and superior customer outcomes.

Collaborating extensively with senior leaders in business, technology, infrastructure, security, risk, compliance, and audit, the role leads senior SRE teams and aligns reliability investments with critical service goals and organizational transformation priorities.

Site Reliability Engineering Team

The SRE team at SGX Group maintains the availability, observability, and recoverability of the platforms underpinning the company's business and market infrastructure. The team establishes standards across service levels, error budgets, observability, incident response, and automation. Working cross-functionally with engineering, infrastructure, security, and product teams, the SRE group serves as subject matter experts and stewards of resilience, proactive maintenance, and continuous improvement of the platform estate. The next evolution focuses on embedding SRE as a discipline by integrating reliability considerations into design and minimizing manual interventions to keep services operational.

Key Responsibilities

  • Define and drive the enterprise-wide reliability engineering strategy and resilience goals.
  • Develop the organizational observability strategy along with a corresponding investment roadmap.
  • Establish an enterprise incident management framework coupled with resilience standards.
  • Set performance goals through defining service level objectives (SLOs) and agreements (SLAs) at the enterprise scale.
  • Lead the automation strategy for the function including spearheading agentic AI adoption to enhance operational capabilities.
  • Plan for organizational scalability and performance optimization through performance engineering roadmaps.
  • Own enterprise capacity planning, resilience, and continuity strategy to ensure operational robustness.
  • Define the strategy for cloud operations and platform reliability across multi-account AWS, Kubernetes/EKS, and Terraform environments.
  • Contribute to operational risk management by shaping the function’s risk strategy.
  • Provide engineering leadership to build enterprise-wide SRE capabilities and design workforce development programs.
  • Champion agentic AI applications within reliability engineering for improving incident management, observability, risk reduction, and enhancing metrics such as mean time to detection (MTTD), mean time to recovery (MTTR), and overall service stability.
  • Partner with executives and stakeholders across business units, engineering, infrastructure, security, risk, compliance, and audit functions to synchronize reliability priorities with business strategy, regulatory mandates, and critical service commitments.
  • Lead executive management efforts during major incidents and crisis events to ensure prompt cross-functional response, clear communications, and continuous improvement post-events.

Candidate Profile

Required Experience and Skills

  • Extensive executive-level leadership experience in Site Reliability Engineering, platform or production engineering, or large-scale technology operations supporting complex and mission-critical environments.
  • Proven expertise in defining enterprise-wide SRE strategies and scaling leadership while integrating reliability standards, observability practices, and engineering disciplines in critical platforms. Fluency in DORA metrics, SLIs, SLOs, and error budgets is essential.
  • Strong familiarity with observability tools like Datadog, New Relic, or Grafana Cloud and in managing cloud-native infrastructure including AWS multi-account setups, Kubernetes/EKS, and Terraform.
  • Experience working in regulated, high-availability environments with comprehensive knowledge of operational resilience, technology risk management, governance, audit, and compliance requirements.
  • Outstanding executive presence paired with excellent stakeholder management and communications capabilities; commercial acumen and experience in vendor management; demonstrated ability to attract, motivate, and retain superior engineering talent.
  • A bachelor's degree in Computer Science, Engineering, Information Systems, or a related field is mandatory; advanced qualifications will be considered advantageous.

Preferred Background

  • Prior exposure to financial market infrastructure involving securities and derivatives trading, clearing and settlement, market operations, or exchange-related systems is a strong differentiator.

Why This Role Is Essential

Reliability is fundamental to SGX Group’s market infrastructure, where it directly impacts trading activities and capital flow. The SVP role uniquely determines the definition, metrics, and trade-offs related to reliability. The position offers a challenging yet rewarding opportunity to embed and cultivate Site Reliability Engineering as a core discipline rather than inheriting it, driving meaningful impact in a dynamic, regulated environment.

About SGX Group

SGX Group is recognized globally as a trusted international marketplace, valued for its consistent stability and openness. Based in Singapore, the organization facilitates price discovery, capital formation, and risk management across diverse asset classes. Supported by resilient infrastructure and reliable clearing, SGX brings together issuers, investors, and intermediaries to create enduring markets.

Minimum education

Bachelor's Degree

Tools & software

Terraform required

How they work

Communication Problem Solving Leadership Strategic Thinking Relationship Building

Leave it if you'd like a reply — we won't use it for anything else.

Click to browse, drag & drop, or paste a screenshot

PNG, JPG, GIF, MP4, WebM, MOV · Max 20MB each · Up to 5 files

🤖
Online · instant AI help
Broxer