Key Responsibilities
- Build, deploy and maintain the cloud infrastructure that runs our own platforms and our clients’ hosted environments, working to designs set by senior engineers and architects.
- Run day-to-day operations across cloud and on-premises environments — provisioning, patching, monitoring, backup and incident response.
- Take part in migration and deployment projects, from environment build through UAT to go-live support.
- Document what you build: configurations, runbooks, network diagrams and as-built records.
- Learn the platform and the business behind it and grow into owning environments of your own.
Cloud Infrastructure
- Provision and configure compute, storage, networking and identity services on at least one major cloud platform (AWS, Azure or Google Cloud).
- Apply and maintain infrastructure baselines — instance sizing, storage tiers, security groups, load balancers, auto-scaling and backup policies.
- Support hybrid setups that link cloud environments to on-premises data centres over VPN or dedicated circuits.
- Help build and test disaster recovery environments and take part in DR drills against agreed RTO and RPO targets.
- Track cost and utilisation and raise it when resources are over-provisioned or idle.
Containers and Orchestration
- Build and maintain container images and run workloads under Docker or a comparable runtime.
- Deploy and support applications on Kubernetes or a managed equivalent (EKS, AKS, GKE) — deployments, services, ingress, config maps, secrets and persistent volumes.
- Work with CI/CD pipelines to move builds through development, staging and production.
- Troubleshoot container and cluster problems: failing pods, resource limits, image pulls and service-to-service networking.
Systems and Networking
- Administer Linux servers (RHEL, Rocky, Ubuntu or equivalent) — users and permissions, packages, services, storage, OS and kernel patching, and shell scripting.
- Configure and troubleshoot core network services: DNS, DHCP, NTP, routing, VLANs, NAT, VPN and firewall rules.
- Diagnose connectivity and performance problems with standard tools — ping, traceroute, dig, tcpdump, ss, curl — and read the output correctly.
- Maintain monitoring and alerting and respond to alerts within agreed service targets.
Databases and Application Support
- Carry out routine MySQL tasks: user and privilege management, schema deployment, backup and restore, log checks and basic query troubleshooting.
- Deploy and support application services on web and application servers (Apache, Nginx, Tomcat or PHP-based stacks).
- Support release deployments and post-release verification alongside the development and QA teams.
Security and Compliance
- Work within our ISO/IEC 27001:2022 and PCI DSS v4.0 control environment — access control, change management, logging, patching and vulnerability remediation.
- Harden systems to the agreed baselines and close vulnerability scan and penetration test findings within target timeframes.
- Produce evidence for internal and external audits when asked.
Requirements
- 1–3 years’ experience in cloud, infrastructure or systems engineering. Fresh graduates with solid hands-on project or lab experience are welcome to apply.
- Hands-on experience implementing or supporting workloads on at least one major cloud platform — AWS, Azure or Google Cloud. Certification (AWS Solutions Architect or SysOps Associate, AZ-104, Google ACE) is an advantage, not a substitute for practical experience.
- Proficient in Linux administration and comfortable working entirely from the command line.
- Working knowledge of TCP/IP, subnetting, DNS, DHCP, VLANs, routing, NAT and firewall concepts.
- Basic MySQL — able to read and write SQL, run a backup and a restore, and explain what an index does.
- Scripting in Bash, and ideally Python, to automate routine work.
- Familiar with Git and with working from tickets and change records.
- Clear written English for documentation, and the confidence to speak to internal stakeholders and clients.
- Strong analytical and troubleshooting skills, and the discipline to test before applying a change to production.
- Degree or diploma in Computer Science, Information Technology, Engineering or a related field — or equivalent practical experience.
Advantageous
- Docker and Kubernetes in a production or near-production setting.
- Infrastructure as code — Terraform, CloudFormation or Ansible.
- Virtualisation platforms — VMware, Hyper-V or Proxmox.
- Monitoring and observability tooling — Zabbix, Prometheus, Grafana, CloudWatch or Azure Monitor.
- Exposure to regulated environments such as financial services or payments, and to audit or compliance work.
- Experience in a managed services or client-facing technical role.
Working Arrangements
- Based at our Kuala Lumpur office, Vertical Bangsar South, on standard office hours.
- Participation in an on-call rotation covering production incidents and scheduled maintenance windows, which fall outside office hours.
- Occasional onsite work at data centres and client premises, including out-of-hours cutovers and DR drills.
Pay: RM3,200.00 - RM3,600.00 per month
Benefits:
- Professional development
- Work from home
Work Location: In person