Key Responsibilities
- Provide Level 2 (Day 2) operational support for mission-critical applications and platform services.
- Investigate, troubleshoot, and resolve production incidents, including performing root cause analysis (RCA).
- Support application deployments, rollback activities, and post-deployment verification.
- Build, maintain, and optimize CI/CD pipelines for automated build, testing, security scanning, and deployments.
- Deploy, manage, and troubleshoot containerized applications using Docker, Podman, and Kubernetes.
- Administer Linux and Windows environments, including patching, configuration, and preventive maintenance.
- Monitor system health, application performance, logs, and alerts to ensure service availability and reliability.
- Collaborate with application, infrastructure, cybersecurity, and project teams to support operational excellence.
- Maintain technical documentation, deployment procedures, operational runbooks, and support compliance requirements.
- Participate in a 24/7 support roster, including standby and after-hours support for production incidents and deployments, as required.
Key Requirements
- Diploma, Bachelor's Degree, or equivalent qualifications or relevant professional experience in Computer Science, Information Technology, Software Engineering, Computer Engineering, or a related field.
- 3+ years of relevant experience in DevOps, Application Support, Production Support, System Engineering, Platform Administration, or a similar role.
- Strong knowledge of Linux administration, application deployment, and production troubleshooting.
- Hands-on experience with GitLab CI/CD, Jenkins, Git, Docker, Podman, and Kubernetes.
- Familiarity with React, Node.js, and Spring Boot applications.
- Experience with scripting languages such as Bash, Python, PowerShell, or Shell scripting.
- Knowledge of monitoring and observability tools such as Prometheus, Grafana, Nagios, Zabbix, or eG Enterprise.
- Understanding of SQL/NoSQL databases, secure CI/CD practices, vulnerability management, and certificate/secret management.
- Experience with SonarQube, VMware, Kafka, API integrations, High Availability (HA), Disaster Recovery (DR), or ITIL processes will be an advantage.
- Strong analytical, troubleshooting, communication, documentation, and stakeholder management skills.
- Ability to participate in a 24/7 support environment, including standby and after-hours support, as required.
apply, please kindly email your updated resume to *************.
Only shortlisted applicants will be notified.
APBA TG Human Resource Pte Ltd (14C7275) || Akshya R (R24122440)