Junior DevOps Operation Engineer – CI/CD Monitoring
Location: Petaling Jaya
Employment type : Contract
Contract Duration : 6 months, extendable up to 1 year.
Working arrangement : Onsite
Salary Range : Max RM7.5k
Open to : Malaysian and PR
Job Summary
We are looking for a motivated DevOps / Operations Engineer to join our team in managing, maintaining, and optimizing our cloud-native infrastructure, AI platforms, and database systems. In this role, you will support Kubernetes cluster operations, infrastructure automation, system monitoring, and troubleshooting while working closely with cross-functional teams to ensure the reliability, stability, and performance of our production services.
Key Responsibilities
- Deploy, operate, monitor, and troubleshoot Kubernetes (K8s) clusters and containerized applications to maintain stable production environments.
- Develop automation scripts and internal tools using Python, Java, or Go to improve operational efficiency and reduce manual tasks.
- Manage and maintain database systems, including routine maintenance, performance tuning, backup and recovery, and issue resolution.
- Build and maintain observability platforms, including log collection, metrics monitoring, alerting, and performance tracking.
- Collaborate with software development teams to optimize CI/CD pipelines and improve deployment efficiency.
- Perform daily system health checks, respond to incidents, conduct root cause analysis (RCA), and implement continuous improvements.
Required Qualifications & Core Skills
- 3 years of relevant experience in DevOps, system operations, cloud-native technologies, or a related field is an advantage.
- Basic knowledge of Kubernetes (K8s), container technologies, or cloud-native environments.
- Proficiency in at least one programming language such as Python, Java, or Go.
- Exposure to AI/ML platforms or related infrastructure is a plus.
- Fluent in Mandarin for daily communication and proficient in English for technical documentation.
- Good understanding of Linux operating systems, networking fundamentals, and cloud-native concepts.
- Strong analytical, problem-solving, and troubleshooting skills with a willingness to learn.
Preferred Skills
- Familiarity with OpenSearch, Grafana, ELK Stack, or other monitoring and observability tools.
- Knowledge of the Argo ecosystem (Argo CD, Argo Workflows) for GitOps and workflow orchestration.
- Exposure to CI/CD pipelines, infrastructure automation, or observability platforms through academic projects, internships, or professional experience.