jobs in RecruitFirst

全职 Data Centre Operations Lead (Shift Working Hours) 工作, 薪水, RecruitFirst Selangor 公司招聘中 - Ricebowl

Data Centre Operations Lead (Shift Working Hours)

Undisclosed
分享
保存

工作地点

  • Sepang, Selangor Sepang Selangor Malaysia

职位描述

岗位职责

Company: A growing Data Center consultancy firm

Position: Data Center Operations Lead (Shift)

Location: Cyberjaya, Selangor

Working Days & Hours: Rotational basis (9am - 9pm AND 9pm - 9am)

Responsibilities:

  • Oversee and monitor the day-to-day operations of the data center during assigned shifts.

  • Ensure the proper functioning of critical systems such as power supply, cooling, network infrastructure, and security systems.

  • Conduct regular inspections to ensure compliance with operational standards, safety protocols, and regulatory requirements.

  • Monitor the performance of servers, storage devices, and network components to identify and resolve any issues or potential failures.

  • Respond to and escalate any operational issues or alarms promptly, ensuring minimal downtime and quick resolution.

  • Supervise and coordinate the activities of the engineering team, ensuring they adhere to safety guidelines and operational procedures.

  • Assign tasks to team members and ensure adequate coverage for all critical systems during the shift.

  • Provide training and guidance to junior staff on operational procedures, troubleshooting, and safety protocols.

  • Serve as the first point of contact for engineers and technicians during emergencies or technical issues.

  • Schedule and oversee preventive maintenance on critical infrastructure and equipment (e.g., HVAC systems, power generators, UPS systems, etc.).

  • Troubleshoot and resolve hardware, software, and networking issues, either directly or by coordinating with other engineers or vendors.

  • Ensure that backup systems and disaster recovery processes are in place and functioning as expected.

  • Document any incidents, troubleshooting steps, and resolutions for future reference and improvement.

  • Act as the point of escalation for any significant incidents or emergencies that occur during the shift.

  • Lead the team in emergency response procedures, including power outages, cooling failures, network disruptions, and security breaches.

  • Ensure a thorough investigation of incidents to identify root causes and prevent recurrence.

  • Ensure that the data center operates in compliance with industry standards, regulations, and best practices (e.g., ISO, SOC 2).

  • Enforce strict physical security measures and access controls to protect the data center from unauthorized access or threats.

  • Participate in audits, inspections, and reviews to maintain certifications and operational standards.

  • Prepare detailed shift reports, incident reports, and performance metrics for review by senior management.

  • Maintain logs of maintenance activities, system failures, and other critical operations for record-keeping and continuous improvement.

  • Monitor and report on the performance of key systems and suggest improvements to ensure reliability and efficiency.

  • Collaborate with IT, facilities management, and other relevant departments to ensure alignment between operational needs and business goals.

  • Coordinate with third-party service providers or vendors for equipment servicing, maintenance, and emergency response.

  • Participate in ongoing training and development to stay up-to-date with new technologies, tools, and best practices in data center management.

  • Recommend and implement improvements to operational processes and workflows to enhance efficiency, reduce costs, and improve system uptime.

Requirements:

  • Diploma/Degree in Electrical Engineering, Mechanical Engineering, Facilities Management, Data Centre Engineering or a related field.

  • Minimum 5–7 years of relevant experience in data center operations, critical facilities, facilities management or engineering operations.

  • At least 2–3 years of experience in a supervisory or team lead role.

  • Strong hands-on knowledge of critical data center infrastructure, including UPS, generators, electrical distribution, HVAC/CRAC/CRAH, cooling systems and BMS.

  • Experience in preventive maintenance, troubleshooting, incident management and emergency response.

  • Experience managing contractors, vendors and maintenance activities.

  • Familiarity with CMMS, DCIM or other facilities/data center monitoring systems.

  • Good knowledge of safety procedures, operational standards and regulatory requirements.

  • Strong leadership, problem-solving, communication and decision-making skills.

  • Able to work on a rotating shift basis, including weekends and public holidays.

  • Experience in a 24/7 mission-critical environment will be an added advantage.

重要安全守则

申请工作时,切勿提供您的银行或信用卡详细资料。不要转账或完成无关的在线调查问卷。如果您发现可疑内容,请举报此招聘广告。

了解更多