As the Data Center Site Lead, you will serve as the main accountable owner for all matters related to data center critical operations, working collaboratively with internal and external stakeholders (mainly colocation provider) to ensure operational excellence. You will work hand-in-hand with access management and security systems to adhere to Oracle's stringent data center security standards/policies. Additionally, you will actively involve in incident, change, and vendor management for data center space, cooling, and power systems. Your proactive approach to identifying and mitigating risks associated with data center operations will be vital to the success of this role.
1. Process discipline — Must have
Candidate should be strong in data center operational processes, including SOP, MOP, EOP, SOO, change control, incident review, problem management, risk register management, corrective/preventive actions, audit readiness, and compliance tracking.
2. Technical expertise — Must have
Candidate should have strong working knowledge in at least one of the following areas: Mechanical systems, Electrical systems, TCS loop/liquid cooling, Project Management, Server & Network infrastructure, or Commissioning.
You do not need to be the deepest SME in every area, but should have enough technical understanding to challenge, validate, coordinate, and drive closure with Colo, DCFE, Design, NW, Security, Access Management, and other stakeholders.
3. Data center security practices — Must have
Candidate should understand data center security standards and be able to coordinate closely with Access Management, Security Systems, Colo, and internal stakeholders to ensure compliance with Oracle’s security policies, access control requirements, audit expectations, and site security procedures.
4. Business continuity and shift coverage — Must have
Candidate should understand 24x7 data center operational expectations, business continuity requirements, backup coverage, cross-training, shift planning, weekend/public holiday support, and escalation management. The role may require support outside normal business hours based on business needs.
5. Colo and stakeholder governance — Must have
Candidate should be able to collaborate effectively with colocation providers and internal stakeholders to maintain 100% data center availability. This includes leading weekly, monthly, and quarterly operations reviews, tracking action items, reviewing performance, and ensuring site-specific processes are established and followed.
6. Incident, change, problem, and risk management — Must have
Candidate should be able to lead and manage incidents, changes, and problems related to data center space, cooling, power, security, and operational readiness. The candidate should proactively identify risks, initiate corrective/preventive actions, and implement global or industry best practices where applicable.
7. Build-to-operations handover — Should have
Candidate should have knowledge of new data center build/deployment support, project-to-operations handover, operational readiness review, site acceptance, as-built documentation, open defect closure, handover checklist, and lessons learned from build-to-operations transition activities.
8. Team leadership and mentoring — Should have
Candidate should be able to oversee, guide, and mentor data center technicians and engineers while supporting team development, operational discipline, and ownership at site level.
9. Reporting, audit, and compliance — Should have
Candidate should be able to provide regular reports and audits on data center performance, security, risks, compliance status, open actions, and operational readiness.
10. Liquid cooling / TCS loop exposure — Good to have
Exposure to liquid cooling, TCS loop, GPU/high-density deployments, or related infrastructure will be an added advantage as Malaysia continues to scale with high-density workloads.
Career Level - IC4
Copyright © 2026 Grabjobs Pte.Ltd. All Rights Reserved.