Deskripsi Pekerjaan
Vantage Data Centers is a leading global provider of wholesale data centers, powering and connecting the world's largest hyperscalers, cloud providers, and enterprises. We are committed to delivering the highest standards of safety, security, and reliability across our rapidly growing portfolio.
We are seeking a world-class Senior Reliability Engineer, Mechanical Systems, to join our APAC Operations team. Based at our state-of-the-art campus in Johor, Malaysia, you will serve as the key technical authority for the reliability, maintainability, and performance of all critical mechanical infrastructure across the region. This is a high-visibility role that directly impacts the uptime and efficiency of some of the largest data centers in Asia Pacific.
In this role, you will bridge the gap between design, construction, and operations. Your deep expertise in mechanical systemsâincluding chilled water plants, cooling towers, pumping arrays, CRAH/CRAC units, and Building Management Systems (BMS/EPMS)âwill be essential in optimizing performance, driving energy efficiency, and mitigating operational risks. You will lead Root Cause Analysis (RCA) investigations, develop reliability-centered maintenance strategies, and mentor site engineering teams.
If you are a results-driven mechanical engineer with a passion for solving complex challenges in a hyper-growth environment, join Vantage and help us define the future of critical data center infrastructure.
Tanggung Jawab
- Own the mechanical reliability strategy for the APAC region, ensuring 100% uptime for all critical cooling and mechanical systems.
- Lead systematic Root Cause Analysis (RCA) for significant mechanical failures and develop comprehensive corrective action plans (CAPs).
- Partner with Design, Construction, and Operations teams to influence new building standards and retrofit existing systems for improved reliability.
- Analyze asset performance data (MTBF, MTTR) to drive lifecycle management, spare parts strategy, and capital replacement planning.
- Conduct risk assessments (FMEA) and implement reliability-centered maintenance (RCM) processes to eliminate single points of failure.
- Provide expert technical support for critical incident response, emergency repairs, and complex troubleshooting across all APAC sites.
- Champion energy efficiency and water conservation initiatives, optimizing mechanical plant operations without compromising resilience.
- Develop and deliver technical training programs to elevate the capabilities of local site operations engineers.
Kualifikasi
- Bachelorâs Degree in Mechanical Engineering or a closely related engineering field (Masterâs degree preferred).
- 7+ years of experience in mission-critical facility operations, with a specific focus on data center mechanical systems.
- Deep expertise in large-scale chilled water systems, evaporative cooling, adiabatic systems, and precision air conditioning (CRAH/CRAC).
- Strong analytical skills with proven experience in Root Cause Analysis (RCA), Failure Mode and Effects Analysis (FMEA), and statistical reliability modeling.
- Experience with Building Management Systems (BMS/SCADA/EPMS) and thermal dynamic modeling.
- Excellent project management skills with the ability to manage complex cross-functional initiatives.
- Superior communication and influencing skills, comfortable presenting to senior leadership and global stakeholders.
- Certified Maintenance & Reliability Professional (CMRP), Certified Reliability Leader (CRL), or Professional Engineer (PE/PEng) status is highly desired.