Position: Site Reliability Engineer
Employment Type: Permanent
Location: Bangsar South, KL
Working Arrangement: Hybrid: 1 day onsite and 4 days remote per week, Monday to Friday.
Working Hours in MY time: 9AM - 6PM
About the Role
Support the client’s digital transformation by ensuring the reliability, resilience, performance, and availability of digital systems across Azure Cloud and supporting platforms. Collaborate with agile teams to drive automation, DevSecOps, and continuous improvement in application and infrastructure reliability.
Key Responsibilities
- Monitor and maintain application availability, reliability, performance, security, and operational SLAs/SLOs.
- Support incident management, root cause analysis, permanent remediation, and application health monitoring.
- Develop and enhance automation using Infrastructure as Code, Configuration as Code, and DevOps practices.
- Maintain and improve DevSecOps workflows, infrastructure resilience, and cost optimization.
- Collaborate with development and operations teams to embed site reliability and operational excellence.
- Document technical processes, automation solutions, and administration activities.
Qualifications
- Strong understanding of Docker, Kubernetes, Azure Cloud, and high-availability systems.
- Experience in SRE, monitoring, automation, scalability, and customer-facing production support.
- Hands-on exposure to Dynatrace, Azure Monitor, Splunk, or similar monitoring platforms.
- Basic coding or scripting skills for troubleshooting and permanent issue resolution.
- Experience with incident management, service management frameworks, and agile methodologies.
- Strong analytical, problem-solving, collaboration, and continuous improvement skills.
Pay: RM7,000.00 - RM10,000.00 per month
Benefits:
Application Question(s):
- how much are you minimum to maximum salary expectations in RM?
Work Location: Hybrid remote in Kuala Lumpur (Kuala Lumpur)