Company: 1011 United Overseas Bank Ltd
About UOB
United Overseas Bank Limited (UOB) is a leading bank in ASEAN with a global network in Southeast Asia, Asia Pacific, Europe and North America. Operating through our head office in Singapore and banking subsidiaries in China, Indonesia, Malaysia, Thailand and Vietnam, we have a global network of about 430 branches and offices in 19 markets. At the heart of UOB is our culture, shaped by the UOB Way and anchored on our four values – Honourable, Enterprising, United and Committed. For more than 90 years, these values have guided how we do right by our customers, collaborate with one another and create long-term value for the communities we operate in. As One Bank, we are committed to helping our colleagues build sustainable careers grounded in purpose, supported by strong values, and enriched with meaningful opportunities to grow.
Job Description
The incumbent is responsible for the engineering, architecture, lifecycle management, reliability, security, governance, risk management, and operational excellence of the organization's DevOps infrastructure platforms.
This role serves as the Technical Lead and Subject Matter Expert for all DevOps farm of compute infrastructure. It consists of VM’s Linux servers, enterprise operating systems, OpenShift Container Platform (OCP), cloud infrastructure, middleware platforms, DevOps tooling, and supporting application infrastructure.
The incumbent will design, build, operate, maintain, modernize, secure, and continuously improve platform services while ensuring availability, scalability, compliance, performance, and operational resilience.
The role also provides leadership in infrastructure standards, obsolescence management, audit readiness, capacity planning, disaster recovery, risk mitigation, and technical governance across on-premises, cloud, and hybrid environments.
You will be responsible for ensuring the reliability and availability of DevOps CI/CD server systems. The primary responsibility is to provide full support for DevOps tools issues and service availability. In the event of DevOps tool service failures, you will perform root cause analysis and restore systems to a steady state as quickly as possible.
The role includes providing support for all DevOps tools version upgrades, Vulnerability management, Obsolescence remediation, OS/VM patching, SSL renewals on weekends schedule.
This role also focuses on continuous improvement and efficiency through automation initiatives, script development, and the transformation of DevOps tools and operations
Responsibilities
Platform Engineering & Architecture
- Design, build, implement, and maintain enterprise-grade infrastructure platforms supporting DevOps and application delivery.
- Define infrastructure standards, reference architectures, engineering patterns, and best practices.
- Lead infrastructure modernization initiatives across servers, cloud platforms, middleware, and container platforms.
- Design highly available, scalable, resilient, and secure platform architectures.
- Drive infrastructure automation and Infrastructure-as-Code (IaC) adoption.
Linux & Server Engineering
- Manage enterprise Linux environments including RHEL, Rocky Linux, Ubuntu, or equivalent.
- Maintain operating system hardening, patching, upgrades, and lifecycle management.
- Perform server build automation, configuration management, troubleshooting, and performance optimization.
- Lead root cause analysis for critical production incidents.
- Manage enterprise server capacity planning and resource optimization.
OpenShift & Container Platform Management
- Design, deploy, manage, and optimize OpenShift Container Platform (OCP) clusters.
- Ensure cluster reliability, scalability, security, and operational health.
- Manage cluster upgrades, patching, node lifecycle, and capacity expansion.
- Implement platform observability, logging, monitoring, and alerting.
- Collaborate with application teams to improve platform adoption and containerization practices.
Cloud Infrastructure Engineering
- Build and operate cloud infrastructure across public, private, and hybrid cloud environments.
- Design cloud landing zones, networking, security controls, and governance frameworks.
- Manage cloud infrastructure provisioning using automation and Infrastructure-as-Code.
- Optimize cloud service consumption, performance, resiliency, and operational costs.
- Ensure cloud platforms comply with enterprise architecture and security standards.
DevOps Platform Engineering
- Manage and maintain DevOps toolchains and supporting infrastructure.
- Support CI/CD platforms and enable self-service capabilities for development teams.
- Drive automation initiatives across provisioning, deployment, monitoring, patching, and operational workflows.
- Ensure platform reliability and service availability for engineering teams.
Middleware & Web Infrastructure
- Build, maintain, and troubleshoot Apache HTTP Server and related web infrastructure components.
- Manage reverse proxies, load balancers, ingress controllers, TLS certificates, and web security controls.
- Ensure high availability and performance optimization of middleware services.
Operational Excellence
- Establish and maintain operational procedures, runbooks, standards, and technical documentation.
- Define and monitor SLAs, SLOs, and platform KPIs.
- Lead incident response, problem management, and post-incident reviews.
- Maintain proactive monitoring and observability capabilities.
Risk Management & Compliance
- Identify infrastructure risks and develop mitigation strategies.
- Drive vulnerability remediation and security hardening activities.
- Ensure compliance with security, regulatory, audit, and corporate governance requirements.
- Maintain infrastructure risk registers and technical debt inventories.
- Support disaster recovery, business continuity, and resilience testing exercises.
Audit & Governance
- Lead infrastructure audit preparedness and evidence collection.
- Ensure operational controls are implemented, documented, and periodically reviewed.
- Support internal, external, customer, and regulatory audits.
- Maintain asset inventories, lifecycle status, and configuration baselines.
Obsolescence & Lifecycle Management
- Own infrastructure lifecycle and technology refresh planning.
- Develop roadmaps for hardware, OS, middleware, container platform, and cloud service upgrades.
- Monitor vendor support status and proactively manage end-of-life/end-of-support risks.
- Drive platform modernization and technology standardization initiatives.
Technical Leadership
- Act as the highest-level technical escalation point for infrastructure-related issues.
- Mentor engineers and provide guidance on platform engineering best practices.
- Contribute to engineering standards, architecture reviews, and strategic technology planning.
- Collaborate with security, architecture, application, and operations teams.
Requirements
- At least 8 years of experience in Linux system administration, infrastructure engineering, or platform engineering.
- At least 5 years of experience managing enterprise container platforms such as Red Hat OpenShift.
- At least 5 years of experience managing cloud infrastructure in AWS, Azure, GCP, or equivalent environments.
- Strong experience with server operating systems including:
- Red Hat Enterprise Linux (RHEL)
- Rocky Linux
- Ubuntu Linux
- Strong experience with:
- Apache HTTP Server
- Nginx
- Load Balancers
- Reverse Proxies
- TLS/SSL certificate management
- Experience managing enterprise-scale virtualized environments.
- Experience implementing Infrastructure-as-Code practices.
- Experience supporting internal and external audits.
DevOps & Automation
- Strong experience with:
- Git
- Jenkins
- GitLab CI/CD
- Azure DevOps
- Ansible
- Terraform
- ArgoCD
- Experience building automated provisioning and deployment solutions.
- Strong scripting and automation skills using:
- Bash
- Python
- Shell scripting
Container & Kubernetes
- Strong understanding of:
- Kubernetes
- OpenShift
- Container runtimes
- Service Mesh
- Ingress Controllers
- Container Networking
- Persistent Storage
Infrastructure Operations
- Experience in:
- Platform monitoring
- Logging
- Observability
- Capacity planning
- Incident management
- Root cause analysis
- Disaster recovery
Security, Risk & Governance
- Experience implementing:
- OS hardening
- Security baselines
- Vulnerability remediation
- Risk management controls
- Compliance frameworks
Soft Skills
- Strong analytical and problem-solving skills.
- Ability to lead technical initiatives independently.
- Excellent stakeholder management and communication skills.
- Ability to work across infrastructure, security, architecture, and application teams.
Additional Requirements
Be a Part of the UOB Family
UOB is an equal opportunity employer. UOB does not discriminate on the basis of a candidate's age, race, gender, color, religion, sexual orientation, physical or mental disability, or other non-merit factors. All employment decisions at UOB are based on business needs, job requirements and qualifications. If you require any assistance or accommodations to be made for the recruitment process, please inform us when you submit your online application.
Apply now and make a Difference