- Kuala Lumpur, Kuala Lumpur Kuala Lumpur WP Kuala Lumpur Malaysia
工作地点
职位描述
岗位职责
About the role
This role involves monitoring network devices, servers, applications and IT infrastructure to ensure service availability, performance and operational stability. You will detect and investigate incidents, perform initial troubleshooting and escalate unresolved issues to appropriate support teams. You will log, update, track and close incidents and service requests in accordance with agreed SLAs and operational procedures, review and respond to system alerts and events, and execute, monitor and support scheduled batch jobs and processes.
Key responsibilities
Monitor network devices, servers, applications and IT infrastructure to ensure service availability, performance and operational stability
Detect and investigate incidents, perform initial troubleshooting and escalate unresolved issues to the appropriate support teams
Log, update, track and close incidents and service requests in accordance with agreed SLAs and operational procedures
Review and respond to system alerts and events, assess impact and initiate corrective actions where required
Execute, monitor and support scheduled batch jobs and processes, ensuring successful completion and timely resolution of failures or exceptions
Support scheduled maintenance, system upgrades, change implementations and service restoration activities
Maintain accurate operational documentation, incident reports, shift handover records and SOP updates
Coordinate with internal teams, outsourced personnel, vendors and stakeholders to ensure timely communication and issue resolution
Ensure incidents and service requests are handled within the agreed response and resolution timelines
Participate in 24x7 shift rotation, including weekends and public holidays, to support continuous IT operations
About you
Diploma or higher in IT, Computer Science, Network Engineering or related discipline
Minimum 2 years relevant experience in NOC, IT Operations, Service Desk or IT Support
Hands-on experience in monitoring network, server, application or infrastructure environments
Experience handling incidents, service requests, tickets and SLA-based operations
Able to perform L1 troubleshooting and escalate appropriately
Able to interpret alerts/events and take appropriate initial action
Experience with monitoring or IT management tools; Dynatrace / ManageEngine is an advantage
Familiar with batch-job/process monitoring and basic operational procedures
Good communication, analytical and problem-solving skills
Willing and able to work 24x7 shifts, weekends and public holidays
重要安全守则
申请工作时,切勿提供您的银行或信用卡详细资料。不要转账或完成无关的在线调查问卷。如果您发现可疑内容,请举报此招聘广告。