Sign up to access all features of our service
  • Job search
  • Favorites
  • Create a CV
    New
  • Salaries
  • Subscriptions

SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology

17000 - 23000 SGD
Full-time

DBS

Role Summary

The SVP, Site Reliability Engineering (SRE), will lead and oversee the 24/7 infrastructure operations and reliability engineering function across critical platforms including Hypervisors (VPC, EPC, OPC), OpenShift , Windows, Databases, TWS, and Mainframe environments .

This role is responsible for driving resilience, scalability, automation, and operational excellence across hybrid cloud and on-premises environments, while ensuring alignment with business, risk, and regulatory expectations.

Key Responsibilities

Leadership & Governance

  • Lead and manage a distributed 24/7 SRE infrastructure team , including shift-based operations and command center functions

  • Define and execute the SRE strategy aligned to enterprise technology and business priorities

  • Establish strong governance across incident, problem, change, release, and capacity management

  • Drive SLA/SLO/SLI frameworks to ensure service reliability and performance targets

Infrastructure & Platform Ownership

  • Oversee end-to-end reliability of infrastructure platforms:

  • Cloud & Container : VPC, OpenShift, Kubernetes

  • Compute & Virtualization : Hypervisors (VMware/others), private cloud platforms

  • Enterprise Platforms : Windows, Unix/Linux, TWS, Mainframe, Databases

  • Ensure high availability, resilience, and disaster recovery readiness across all critical systems

  • Own infrastructure lifecycle including capacity planning, patching, upgrades, and decommissioning

Reliability Engineering & Automation

  • Champion SRE principles including error budgets, toil reduction, and automation-first mindset

  • Drive end-to-end observability strategy (monitoring, logging, tracing)

  • Lead initiatives to reduce MTTR, incident volume, and manual operational effort

  • Scale automation across deployment, patching, incident resolution, and self-healing capabilities

Operational Excellence

  • Ensure 24/7 monitoring, incident response, and recovery processes are robust and continuously improved

  • Lead major incident management and command bridge coordination for critical outages

  • Conduct RCA, trend analysis, and preventive engineering improvements

  • Embed ITIL best practices across service management processes

Risk, Compliance & Security

  • Identify infrastructure risks and drive proactive mitigation strategies

  • Ensure compliance with regulatory, audit, and internal security requirements

  • Partner with security teams on hardening, vulnerability management, and access controls

Stakeholder & Cross-Functional Collaboration

  • Collaborate with application, DevOps, security, architecture, and business teams to improve system reliability

  • Provide leadership in large-scale transformation programs (cloud adoption, infra modernization, SRE maturity)

  • Act as a key interface with senior management and external stakeholders

People & Talent Development

  • Build and develop a high-performing SRE organization across L1/L2/L3 layers

  • Drive fungibility, cross-skilling, and leadership development within the team

  • Mentor senior leaders and establish clear career progression frameworks

Requirements

Experience

  • 18+ years of experience in IT infrastructure, SRE, or production operations

  • Proven leadership in managing large-scale 24/7 infrastructure teams in banking/financial services

  • Strong experience in hybrid cloud, data center, and enterprise platforms

Technical Expertise

  • Deep expertise in:

  • Cloud platforms (private/public cloud architectures)

  • Container platforms (OpenShift/Kubernetes)

  • Hypervisors & virtualization technologies

  • Operating systems (Windows, Linux/Unix)

  • Databases (MariaDB, Postgres, MSSQL, Redis, DB2)

  • Enterprise scheduling & legacy systems (TWS, Mainframe)

  • Strong understanding of DevOps, CI/CD, and infrastructure as code

Leadership & Functional Skills

  • Strong strategic thinking with ability to translate business goals into technology outcomes

  • Excellent incident leadership and crisis management skills

  • Proven track record of driving automation and operational transformation

  • Strong stakeholder management and executive communication skills

Other Skills

  • Expertise in ITIL / Service Management frameworks

  • Strong analytical, problem-solving, and decision-making capabilities

  • Ability to manage high-pressure situations and multiple priorities

Key Success Metrics (Optional for your slide/JD refinement)

  • Infrastructure availability (SLA/SLO adherence)

  • Reduction in MTTR / incident volume

  • Automation coverage & reduction in manual toil

  • Capacity utilization and cost optimization

  • Audit and compliance adherence

Vacancy posted 11 days ago
Similar jobs that could be interesting for youBased on the SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology in Singapore vacancy
  • 5000 - 7500 SGD

     ...Tower B ~ Medical & Dental Benefits Provided ~ Entitled to Yearly Bonus & Performance Bonus Responsibilities ( Lead & Senior SRE | Site Reliability Engineer ) High Availability and Stability Maintenance of Application Systems: ~ Includes daily monitoring, alert... 

    REOLINK TECHNOLOGY PTE. LTD.

    Singapore
    24 days ago
  • 7000 - 13000 SGD

     ...Responsibilities: Ensure the stability, reliability, and high availability of the...  ...automation platforms and engineering efficiency tools to advance...  ...the application of AI technologies in operations scenarios,...  ...with 3+ years of experience in Site Reliability Engineering (SRE)... 

    XIAOMI TECHNOLOGIES SINGAPORE PTE. LTD.

    Singapore
    19 days ago
  •  ...network of the world's industry leaders. As a leading global long-term investor, we work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    4 days ago
  • 8000 - 10000 SGD

     ...Job Summary We are seeking an experienced Site Reliability Engineer (SRE) / Production Support Engineer with 10+ years of experience in enterprise infrastructure, production support, system administration, and SRE operations. The ideal candidate will have strong hands-on... 

    EXPRESS PTE. LTD.

    Singapore
    11 days ago
  •  ...We are looking for a hands-on SRE Leader with a strong software engineering background to build and lead PatSnap’s Singapore SRE team. You will combine team leadership...  ...technical design and development, improving the reliability of our global SaaS platform while building platforms... 

    Patsnap

    Singapore
    9 days ago
  • 5000 - 7500 SGD

     ...science or a related field is preferred. # Experiences as Senior SRE or leading a small team is preferable. # At least 5 years of...  ...etc.). # Experience with automation operations and container technologies (Docker, Kubernetes). # Familiarity with CI / CD processes... 

    REOLINK TECHNOLOGY PTE. LTD.

    Singapore
    16 days ago
  •  ...We Are Singapore Pools was established by the Singapore government on 23 May 1968 to provide safe and trusted betting to counter...  ..., to conserving the environment. Job Purpose The Site Reliability Engineer (SRE) drives enterprise operational resilience by architecting... 

    Singapore Pools

    Singapore
    18 days ago
  • 11250 - 22500 SGD

     ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...communities we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide... 

    GIC PRIVATE LIMITED

    Singapore
    10 days ago
  • 6000 - 12000 SGD

     ...Play a key role in ensuring system reliability at one of the world’s most iconic and largest...  ...financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the Corporate and Investment Bank, Payments Technology Team, you will use technology to solve... 

    JPMORGAN CHASE BANK, N.A.

    Singapore
    11 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    1 day ago
  •  ...Play a key role in ensuring system reliability at one of the world’s most iconic and...  ...largest financial institutions. As a Site Reliability Engineer II at JPMorgan Chase within the Corporate and Investment Bank, Payments Technology Team, you will use technology to solve... 

    JPMorgan Chase & Co.

    Singapore
    11 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    3 days ago
  • 7000 - 12000 SGD

     ...to ensure the scalability and reliability of microservices. Build Observability...  ...resources, such as node groups, CPU, memory, HPA scheduling,...  ...ISO27001, SOC2, and GDPR. Lead incident response efforts to...  ...solutions. Assist with product/technology selection, including... 

    TP-LINK CORPORATION PTE. LTD.

    Singapore
    20 days ago
  • 6500 - 13000 SGD

     ...About the Role We're hiring a Site Reliability Engineer (SRE) to join a global engineering team supporting mission-critical financial markets platforms undergoing a large-scale cloud, cyber security, and platform modernisation programme. This is a hybrid Production Engineering... 

    ALLEGIS GROUP SINGAPORE PRIVATE LIMITED

    Singapore
    12 days ago
  •  ...We are seeking a skilled and passionate Engineer to join our team to build and operate a Whole-of-Government (WoG) runtime platform. As a Site Reliability Engineer, you will be responsible for designing and operating GitLab, AWS and Kubernetes-based infrastructure and solutions... 

    AvePoint

    Singapore
    more than 2 months ago
  • 7000 - 9500 SGD

     ...Hybrid) Job Type: Full-Time Contract (Subjected to renewal/conversion) Join a high-impact engineering team supporting large-scale digital government platforms. As a Site Reliability Engineer (SRE), you will help ensure platform reliability, scalability, security and... 

    SCIENTEC CONSULTING PTE. LTD.

    Singapore
    4 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    8 days ago
  • 6600 - 13200 SGD

     ...We are looking for seasoned engineers to join our team! You will participate...  ...open-source and proprietary technologies, components, libraries, tools...  ...) to ensure consistency, reliability, and scalability across the...  ...7+ years of experience in an SRE/Platform Engineering/DevOps... 

    F5 NETWORKS SINGAPORE PTE LTD

    Singapore
    20 days ago
  • 15000 - 27000 SGD

     ...Business Function Group Operations enables and empowers the bank...  ..., quality & control, technology, people capability and innovation...  ...role is responsible for shaping governance, resiliency and control frameworks...  ...Risk Management & Controls: Lead the development of risk-based... 

    DBS BANK LTD.

    Singapore
    4 days ago
  • 9000 - 14000 SGD

     ...My Clients: I am currently supporting multiple technology companies in Singapore, including leading global internet platforms, large-scale consumer technology...  ...teams, who are actively expanding their Site Reliability Engineering (SRE) functions. We are hiring across... 

    DADACONSULTANTS PTE. LTD.

    Singapore
    6 days ago
  • 8500 - 9500 SGD

     ...telemetry integration. Work with technologies such as OpenTelemetry, Grafana,...  ...recommend improvements to system reliability and performance. Work with engineering, infrastructure, network, security...  ...years of experience in Observability, SRE, Platform Engineering,... 

    NEXBRIDGE RECRUITMENT PTE. LTD.

    Singapore
    14 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    12 days ago
  •  ...world’s industry leaders. As a leading global long-term investor, we...  .... Enterprise and Data Technology Enterprise & Data Technology...  ...into one coordinated group, we create seamless solutions...  ...the quality bar for data and engineering talent, while attracting, developing... 

    GIC

    Singapore
    3 days ago
  •  ...individual's freedom.   OKX is a leading crypto exchange, and the...  ...markets. We are safe and reliable, backed by our Proof of Reserves...  ...er. OKX is part of OKG, a group that brings the value of Blockchain...  ...data collection, alert governance, and cost data visualization across... 

    OKX

    Singapore
    24 days ago
  • 7500 - 12000 SGD

     ...Job Description We are seeking an experienced Platform Engineer / Site Reliability Engineer (SRE) to support the deployment, operations, and reliability of enterprise platform services. The ideal candidate will have strong hands-on experience with Kubernetes , observability... 

    BOUNTEOUSXACCOLITE SINGAPORE PTE. LTD.

    Singapore
    13 days ago
  • 8000 - 15000 SGD

     ...managing their trading and price risks. Led by a group of visionary leaders, APEX aims to establish a leading commodity and financial derivatives trading...  ...the candidate will be supporting Information Technology (“ IT ”) governance and data function under IT & Operations... 

    ASIA PACIFIC EXCHANGE PTE. LTD.

    Singapore
    6 days ago
  •  ...Binance is a leading global blockchain ecosystem behind the world’s largest cryptocurrency exchange by trading volume and registered users...  ...for our industry-leading security, user fund transparency, trading engine speed, deep liquidity, and an unmatched portfolio of digital-... 

    Binance

    Singapore
    2 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    17 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    16 days ago
  •  ...network of the world’s industry leaders. As a leading global long-term investor, we Work at...  ...we invest in worldwide. Technology Group We experiment, design, and lead a 24×...  ...and risk management. We deliver secure, reliable, and integrated solutions, and provide insights... 

    GIC

    Singapore
    8 days ago

Do you want to receive more vacancies?

Subscribe and receive similar vacancies to SVP, Site Reliability Engineering Lead, SRE & Governance, Group Technology. Be the first to apply!