Location: Punggol (Hybrid)
Job Type: Full-Time Contract (Subjected to renewal/conversion)
Join a high-impact engineering team supporting large-scale digital government platforms. As a Site Reliability Engineer (SRE), you will help ensure platform reliability, scalability, security and operational excellence while working across cloud, DevSecOps and modern infrastructure technologies.
Key Responsibilities
- Define, monitor and improve Service Level Objectives (SLOs) and Service Level Indicators (SLIs) to maintain platform reliability and availability.
- Design and implement observability solutions covering logging, monitoring, metrics and distributed tracing.
- Lead incident response, troubleshooting and post-incident reviews, driving improvements to prevent recurring issues.
- Automate infrastructure provisioning, configuration and deployment using Infrastructure-as-Code (IaC) tools.
- Manage and optimise Kubernetes and containerised environments across cloud and on-premises infrastructure.
- Support capacity planning, performance tuning and platform scalability.
- Improve CI/CD and DevSecOps practices while reducing manual operational work and system toil.
- Collaborate closely with software engineering and product teams to embed reliability and operability into the development lifecycle.
- Ensure platforms meet relevant security, hardening and regulatory requirements.
Requirements
- Minimum 4 years of relevant experience in Cloud Engineering, Site Reliability Engineering (SRE), or related technical roles.
- Working knowledge of cloud platforms such as AWS, Azure or GCP.
- Hands-on experience with Kubernetes, containers and container orchestration.
- Experience with CI/CD and DevSecOps tools such as GitLab, Fortify, Jira, Confluence or equivalent.
- Proficiency in at least one scripting or programming language such as Python, Go or Bash.
- Experience with Terraform, Ansible or other Infrastructure-as-Code tools.
- Experience implementing or supporting observability platforms such as ELK Stack, Prometheus, Grafana or equivalent.
- Good understanding of networking, system security, troubleshooting and infrastructure operations.
- Strong communication, problem-solving and ownership mindset.
- Experience with GitOps practices and workflows would be advantageous
By submitting your resume, you consent to the collection, use, and disclosure of your personal information per ScienTec’s Privacy Policy (*************).
This authorizes us to:
• Contact you about potential opportunities.
• Delete personal data as it is not required at this application stage.
All applications will be processed with strict confidence. Only shortlisted candidates will be contacted.
Elaine Wong (Zoe) - R26162114
ScienTec Consulting Pte Ltd - 11C5781