锦鲤求职

汇丰控股

Head of SSv Reliability Engineering

锦鲤会替你打开官网网申、按简历自动填表提交,你只需要在关键步骤确认。

岗位描述

Organisational Design, Implementation Strategy & Governance Lead a global Reliability Engineering initiative for a 1,200-person engineering organisation, setting the strategy, outcomes and adoption plan to raise reliability maturity across teams and regions. Design and stand up the SRE operating model for SSv Technology — evaluate central, embedded and hybrid structures, and implement the model best suited to custody, fund accounting, transfer agency and post-trade platforms across regions. Establish the SRE function's charter, mandate and governance, and secure alignment from SSv Technology and CIB Technology leadership. Set the multi-year roadmap and priorities for SRE maturity — observability, incident management, automation, resiliency engineering, and toil reduction — with clear measures of progress (e.g., SLO adoption, incident trends, toil reduction, and reliability improvements). Partner with regional technology leads, principal engineers and application engineering leads to align global practice while adapting to local regulatory and operational realities in China and across the globe. Establish capacity and performance engineering practices (capacity planning, load testing, latency/throughput SLIs, saturation monitoring) for critical services. Establish SRE culture and practice from zero: error budgets, blameless postmortems, toil reduction, reliability-as-code, and a shared operational language across application teams. Create an education and enablement programme to upskill ~1,200 engineers on Reliability Engineering concepts and processes (e.g., SLI/SLO thinking, error budgets, incident learning, automation/toil reduction), supported by playbooks, training materials, communities of practice and targeted coaching. Embed Reliability Engineering into day-to-day delivery by introducing pragmatic design reviews and reliability guardrails for critical changes, ensuring teams build resilient services without creating unnecessary bureaucracy. Define career paths, skills frameworks and training for SRE talent, and act as the internal sponsor for the discipline. Hire, grow and mentor a reliability engineering team based in China; build a durable talent pipeline in a competitive local market. Formal training or certification in SRE concepts, extensive applied experience, including prior experience building or scaling a reliability engineering/SRE function from the ground up, ideally in a regulated financial-services environment. Demonstrated Reliability Engineering mindset with evidence of designing self-healing systems and driving automation-first reliability improvements proactively (not only via process). Proven experience leading new initiatives from scratch at scale (strategy, operating model, mobilisation, adoption and delivery across large engineering populations). Demonstrated experience designing organisational models (central vs embedded vs hybrid) for reliability, platform, or infrastructure engineering functions. Proven track record hiring, growing and retaining technical teams, ideally within the China/APAC technology talent market. Strong systems thinking and problem-solving capability to assess complex, cross-domain, cross-regional production issues and drive end-to-end resolution. Demonstrated major incident leadership, coordinating effectively across global teams to restore service rapidly and drive root-cause remediation. Hands-on technology experience across cloud (AWS/Azure/GCP), automation and scripting (Python, Shell, PowerShell, Ansible, Terraform), and modern distributed platforms (containers, microservices, Kubernetes/OpenShift). Strong observability and service management expertise (Dynatrace, Splunk, Grafana; ITIL Incident/Problem/Change/Availability), with excellent verbal and written communication in English Working knowledge of custody, fund services, and/or post-trade workflows, or strong transferable capital markets technology experience. ITIL v4 certification (or equivalent service management experience) preferred. Experience integrating support models and standardising operating practice across multiple regions or legal entities. Background in organisational or cultural change leadership, particularly introducing a new engineering discipline into an established technology organisation. Knowledge of Securities Services domains: custody, fund accounting, transfer agency, middle office, or post-trade processing would be an advantage. Experience partnering with Front Office or client-facing teams to support business-critical, high-availability platforms.

运维/技术支持方向的申请准备

先核对岗位要求与自己的经历,再准备档案、投递和面试。

阅读求职步骤与示例 →