Client introduction
Our client is an established financial institution licensed and regulated by the Monetary Authority of Singapore, serving business customers across the region. Backed by a well-capitalised international technology group, it operates a cloud-first estate with a small on-premise footprint retained for disaster recovery.
The bank is now hiring a Senior IT Infrastructure Engineer (Cloud Operations & DR) to take broad ownership of its cloud infrastructure, operational resilience and day-to-day infrastructure services. This is a senior individual contributor role within a deliberately lean team reporting into senior technology leadership. It will suit someone who remains genuinely hands-on while being equally comfortable with vendors, incidents, regulatory expectations and stakeholders across Singapore and China.
Job responsibilities
- Own and operate the bank's production cloud infrastructure, covering deployment, configuration, optimisation, monitoring, capacity planning and cost management
- Lead disaster recovery and business continuity planning, testing and remediation, ensuring critical services can be restored within agreed recovery objectives and preparing the supporting evidence for audit and management review
- Manage the limited on-premise and disaster recovery environment, including hardware, networking equipment, maintenance and technology refresh
- Manage patching and vulnerability remediation in line with internal policy and agreed service levels, including pre and post deployment validation
- Lead infrastructure incident and problem management, coordinating internal teams and vendors through resolution, root cause analysis and post-incident reporting
- Maintain monitoring and alerting across the estate, proactively identifying performance, availability and security risks
- Own the bank's local infrastructure relationship with the wider group network function, specifying requirements, reviewing designs and holding delivery to account
- Manage cloud and hybrid network connectivity, including routing, firewalls, segmentation, DNS and connectivity between production and disaster recovery environments
- Oversee cloud providers, data centre partners and technology vendors, covering due diligence, service quality review, escalation and performance management
- Partner with cybersecurity stakeholders on access controls, vulnerability management, infrastructure hardening and audit readiness
- Contribute to infrastructure design and continuous improvement, including automation and infrastructure-as-code initiatives, periodic architecture reviews, and opportunities to improve operational efficiency through AIOps
Job requirements
Essential