Principal Linux Engineer
Our client is a leading global financial markets and technology organisation operating highly available, business-critical infrastructure globally.
We are looking for a Principal Linux Engineer to take senior technical ownership of large-scale Linux infrastructure, driving platform reliability, resilience, performance and operational excellence.
Key Responsibilities
- Own the reliability, availability and performance of large-scale Linux environments.
- Act as the senior technical authority for complex Linux issues and major production incidents.
- Lead infrastructure upgrades, patching, failover, recovery and lifecycle management.
- Drive automation and operational improvements using Python, Bash/Shell and Ansible/Puppet/Salt.
- Support enterprise bare-metal servers, storage and high-availability infrastructure.
- Drive monitoring, observability, capacity management and infrastructure resilience.
- Partner with Engineering, SRE, Security and Infrastructure teams on platform improvements.
- Provide technical guidance and mentorship to engineering teams.
Requirements
- 8+ years of Linux infrastructure / systems engineering experience in large-scale 24x7 environments.
- Deep expertise in Linux, performance tuning, troubleshooting and system fundamentals.
- Strong experience with Ansible, Puppet or Salt.
- Strong Python and Shell/Bash scripting skills.
- Experience with bare-metal servers, SAN/NAS, RAID and NVMe.
- Exposure to Grafana, Prometheus, AWS/GCP, Docker and Kubernetes.
- Experience with Terraform, Git and CI/CD.
- Strong experience managing critical incidents, production environments and high-availability infrastructure.
- Ability to operate as a technical leader and subject matter expert in a complex global environment.
Work Schedule: Wednesday–Sunday, 7:00 AM–4:00 PM SGT, including weekend operational support.