- Kuala Lumpur Federal Territory Malaysia
工作地点
职位描述
岗位职责
Role Summary
We are looking for a highly skilled Infrastructure Engineer with deep expertise in VMware virtualization and strong L2 capabilities across cloud, storage, containers, OS, and network domains. The role involves managing enterprise scale environments, ensuring high availability, performance optimization, and supporting incident management across multiple infrastructure layers.
Key Responsibilities
Act as single point of contact for the Site/Infra
Lead VMware platform operations (vSphere, ESXi, vCenter) including provisioning, performance tuning, patching, and troubleshooting
Handle complex incidents (P1/P2) and drive RCA and permanent fixes across infrastructure layers
Perform L2-level support for:
o NetApp Storage (volumes, snapshots, performance checks)
o OpenShift (cluster health, pods, nodes, deployments)
o Linux (RHEL/CentOS) and Windows Server environments
o Network issues (basic routing, firewall, connectivity debugging)
Ensure infrastructure availability, capacity planning, and compliance adherence
Collaborate with L3 teams/vendors for critical issues and escalations
Automate repetitive tasks using scripts/tools (PowerCLI, Shell, Ansible)
Support DR activities, failover testing, and backup validation
Contribute to continuous improvement initiatives (automation, monitoring optimization, alert reduction)
Maintain technical documentation, SOPs, and runbooks
Technical Skills Required
Primary (Must Have)
Expert-level VMware
o vSphere, ESXi, vCenter administration
o HA, DRS, vMotion, snapshots, datastore management
o Performance tuning and troubleshooting
Secondary (L2 Level – Must Have)
NetApp Storage
o Basic administration, volume management, snapshots, replication awareness
OpenShift / Containers
o Cluster monitoring, pod troubleshooting, basic deployments
Linux
o System administration, logs analysis, patching, troubleshooting
Windows Server
o AD basics, services troubleshooting, patching
Networking
o TCP/IP, DNS, routing fundamentals, firewall basics
---
Good to Have
Exposure to cloud platforms (Azure/AWS/GCP)
Experience in monitoring tools (Grafana, Dynatrace, Prometheus, etc.)
Automation using PowerCLI, Ansible, Python
Knowledge of ITIL processes (Incident, Change Management)
Certification in any of the above technologies
Knowledge of AI /LLM/ML
重要安全守则
申请工作时,切勿提供您的银行或信用卡详细资料。不要转账或完成无关的在线调查问卷。如果您发现可疑内容,请举报此招聘广告。