Assistant Manager – Critical Facilities
- Mandaluyong City, Metro Manila, Philippines
- Full-Time
- On-Site
Job Description:
The Assistant Manager is responsible for supporting the safe, reliable, and continuous operation of mission-critical data center facilities and infrastructure.
The role works closely with the Senior Manager – Data Center Operations in managing day-to-day operations, maintenance activities, technical teams, vendors, incidents, and operational performance.
The position oversees the operational readiness of critical infrastructure such as power, cooling, HVAC, UPS, generators, electrical distribution, and related facility systems, ensuring that the data center maintains high availability and operates in accordance with safety, quality, and performance standards.
This role also serves as an operational escalation point, coordinating technical teams and service providers during incidents, maintenance activities, and infrastructure-related concerns.
Key Responsibilities:
1. Data Center Operations
- Support the day-to-day operation of mission-critical data center facilities and infrastructure.
- Ensure continuous availability and reliable operation of critical power, cooling, HVAC, electrical, and mechanical systems.
- Supervise daily facilities operations, equipment monitoring, maintenance activities, and operational support.
- Ensure that standard operating procedures, maintenance procedures, change management processes, and operational protocols are consistently followed.
-
Coordinate preventive and corrective maintenance activities for critical systems, including:
- UPS systems
- Generators
- CRAC/CRAH units
- HVAC and cooling systems
- Electrical distribution systems
- Cooling infrastructure
- Other critical facility systems
- Ensure equipment and facilities remain operationally ready and within required performance parameters.
2. Team Supervision
- Supervise data center engineers, technicians, operators, and other operational personnel.
- Assign daily tasks and ensure appropriate staffing and shift coverage for 24/7 data center operations.
- Support workforce scheduling, shift planning, and operational handovers.
- Provide technical guidance and support to team members in handling operational issues.
- Monitor team performance, readiness, compliance, and adherence to safety and operational procedures.
- Support training and skills development programs for operations personnel.
- Serve as an escalation point for operational issues encountered during assigned shifts or maintenance windows.
3. Infrastructure Monitoring & Performance
- Monitor the performance and operational condition of critical data center infrastructure.
-
Track and analyze key operational metrics, including:
- PUE (Power Usage Effectiveness)
- MTBF (Mean Time Between Failures)
- MTTR (Mean Time to Repair)
- Uptime / Availability
- RCI (Rack Cooling Index)
- Ensure operational KPIs and service-level commitments are consistently achieved.
- Identify operational inefficiencies, recurring issues, and performance gaps.
- Recommend and implement initiatives to improve reliability, efficiency, and infrastructure performance.
- Maintain visibility of critical infrastructure through monitoring and operational dashboards.
4. Incident & Change Management
- Coordinate the response to equipment failures, facility incidents, alarms, and other operational issues.
- Ensure incidents are properly escalated and resolved within established SLAs/OLAs.
- Coordinate technical teams and vendors during critical incidents.
- Lead or support Root Cause Analysis (RCA) following significant operational incidents.
- Ensure corrective and preventive actions are properly documented and implemented.
- Maintain accurate incident records, operational logs, and reports.
- Ensure proper change management procedures are followed for maintenance, upgrades, and infrastructure modifications.
- Coordinate with Engineering and IT teams during infrastructure upgrades, deployments, and planned maintenance windows.
5. Vendor & Service Provider Management
- Coordinate with third-party contractors, maintenance providers, and service vendors supporting data center operations.
- Monitor vendor performance and ensure compliance with agreed service levels and operational requirements.
- Coordinate preventive maintenance schedules, service activities, and work orders.
- Review maintenance activities and ensure work is completed according to required standards.
- Support contract and service management activities.
- Escalate recurring vendor performance issues and operational concerns to management.
6. Safety, Compliance & Operational Standards
- Ensure data center operations comply with company policies, safety requirements, industry standards, and established operational procedures.
- Promote safe working practices across the data center operations team and contractors.
- Support internal and external audits and ensure required operational documentation is maintained.
- Ensure maintenance, inspection, testing, and compliance records are complete and up to date.
- Support the implementation and continuous improvement of operational and safety procedures.
7. Reporting & Cross-Functional Coordination
- Prepare regular reports covering operational performance, maintenance activities, incidents, equipment status, and infrastructure KPIs.
- Maintain operational dashboards and performance tracking reports.
- Provide timely updates and escalation to the Senior Manager – Data Center Operations.
- Coordinate closely with Engineering, IT, Network Operations, Security, Facilities, and other support functions.
- Support management in identifying operational risks and developing mitigation plans.
- Participate in operational reviews, planning meetings, and continuous improvement initiatives.
Qualifications:
- Bachelor's degree in Mechanical Engineering, Electrical Engineering, Facilities Management, Computer Science, or a related field.
- 6–8 years of experience in data center operations, critical facilities, mission-critical infrastructure, or a similar environment.
- Experience supervising engineers, technicians, operators, or technical teams.
-
Strong working knowledge of critical data center infrastructure, particularly:
- HVAC / cooling systems
- CRAC/CRAH systems
- UPS
- Generators
- Electrical distribution
- Mechanical systems
- Critical power and cooling infrastructure
- Experience in 24/7 mission-critical operations is highly preferred.
- Familiarity with data center monitoring systems and DCIM (Data Center Infrastructure Management) platforms.
- Working knowledge of incident management, change management, preventive/corrective maintenance, and operational risk management.
- Familiarity with frameworks such as ITIL is an advantage.
- Strong incident response, problem-solving, coordination, and people management skills.
- Strong communication and reporting skills, particularly when dealing with technical teams and external vendors.
- Certifications such as CDCP, CDCS, Uptime Institute ATS/ATS+, or equivalent are an advantage.