ASTEK has been providing IT and Engineering solutions for some of the world’s largest industrial and services groups for more than 35 years with 10,000 passionate experts in 22 countries throughout Europe, APAC, Middle East and the Americas
Currently, we are looking for Senior Platform Infrastructure Engineer which would be based in Singapore
Requirements and Responsibilities:
- Bachelor’s degree in computer science, Engineering, Information Systems, or related field
- Experience in cloud infrastructure engineering and operations
- Strong hands-on experience with at least one cloud platform such as AWS, Azure, or GCP
- Experience with Terraform, Ansible, GitHub Actions, or similar automation tooling
- Experience operating Kubernetes or container platforms
- Familiarity with observability platforms such as Dynatrace, Prometheus, Grafana, or CloudWatch
- Good understanding of infrastructure security
- Experience supporting production environments and operational incident response
- Scripting or programming knowledge in Python, Bash, or similar languages
- Strong troubleshooting and stakeholder communication skills
- Experience leveraging AI/ML driven tools for operational insights, automation, observability or infrastructure optimisation
- Design, build, automate, and operate modern infrastructure platforms across cloud and hybrid environments. The role focuses on improving reliability, observability, automation, security posture, and operational resilience for critical systems supporting healthcare and government digital services
- Work closely with application, security, cloud, and operations teams to enable scalable and resilient platform capabilities while reducing operational toil and technical debt
- Build and maintain cloud infrastructure platforms across AWS, Azure, Kubernetes, and hybrid environments
- Develop and maintain Infrastructure-as-Code using Terraform, Ansible, and CI/CD pipelines
- Support platform observability, monitoring, logging, alerting, and operational dashboards
- Improve infrastructure reliability, resiliency, failover readiness, and operational standards
- Support container platforms including Kubernetes and related ecosystem tooling
- Automate operational processes, patching, provisioning, compliance checks, and reporting
- Support vulnerability remediation and infrastructure security hardening activities
- Troubleshoot infrastructure and platform incidents across compute, network, storage, and cloud services
- Partner with application teams to improve cloud-native adoption and operational readiness