- Kuala Lumpur Federal Territory Malaysia
Working Location
Job Description
Responsibilities
Role Summary
We are looking for a highly skilled Infrastructure Engineer with deep expertise in VMware virtualization and strong L2 capabilities across cloud, storage, containers, OS, and network domains. The role involves managing enterprise scale environments, ensuring high availability, performance optimization, and supporting incident management across multiple infrastructure layers.
Key Responsibilities
Act as single point of contact for the Site/Infra
Lead VMware platform operations (vSphere, ESXi, vCenter) including provisioning, performance tuning, patching, and troubleshooting
Handle complex incidents (P1/P2) and drive RCA and permanent fixes across infrastructure layers
Perform L2-level support for:
o NetApp Storage (volumes, snapshots, performance checks)
o OpenShift (cluster health, pods, nodes, deployments)
o Linux (RHEL/CentOS) and Windows Server environments
o Network issues (basic routing, firewall, connectivity debugging)
Ensure infrastructure availability, capacity planning, and compliance adherence
Collaborate with L3 teams/vendors for critical issues and escalations
Automate repetitive tasks using scripts/tools (PowerCLI, Shell, Ansible)
Support DR activities, failover testing, and backup validation
Contribute to continuous improvement initiatives (automation, monitoring optimization, alert reduction)
Maintain technical documentation, SOPs, and runbooks
Technical Skills Required
Primary (Must Have)
Expert-level VMware
o vSphere, ESXi, vCenter administration
o HA, DRS, vMotion, snapshots, datastore management
o Performance tuning and troubleshooting
Secondary (L2 Level – Must Have)
NetApp Storage
o Basic administration, volume management, snapshots, replication awareness
OpenShift / Containers
o Cluster monitoring, pod troubleshooting, basic deployments
Linux
o System administration, logs analysis, patching, troubleshooting
Windows Server
o AD basics, services troubleshooting, patching
Networking
o TCP/IP, DNS, routing fundamentals, firewall basics
---
Good to Have
Exposure to cloud platforms (Azure/AWS/GCP)
Experience in monitoring tools (Grafana, Dynatrace, Prometheus, etc.)
Automation using PowerCLI, Ansible, Python
Knowledge of ITIL processes (Incident, Change Management)
Certification in any of the above technologies
Knowledge of AI /LLM/ML
Important Information
Never provide your bank or credit card details when applying for jobs. Do not transfer any money or complete unrelated online surveys. If you see something suspicious, Report this Job ad.