jobs in Autonomai Recruitment

全职 Site Reliability Engineer 工作, 薪水, Autonomai Recruitment 公司招聘中 - Ricebowl

Site Reliability Engineer

Autonomai Recruitment

Undisclosed

Singapore

分享
保存

工作地点

  • Singapore

职位描述

岗位职责

About the Firm


We're working with a leading proprietary trading firm committed to world-class research, bringing together exceptional talent in Mathematics, Physics, and Computer Science to push scientific and technological boundaries and apply cutting-edge research to global financial markets. The culture is built on innovation, intellectual honesty, and a relentless competitive edge — with collaboration and mutual respect at its core. Beyond trading, the firm designs and deploys technologies that extend well beyond the trading floor, funds start-ups across industries, and partners with leading global research organisations and universities.


The Trading Infrastructure team is a global organisation of engineers who architect, build, and maintain world-class infrastructure — from colo design and implementation, to optimising exchange connectivity, to building low-latency Wide Area Networks. The team leverages research and automation to continuously adapt and scale infrastructure in line with the evolving trading business.


We're looking for exceptional talent who can collaborate effectively across global teams and help take this infrastructure to the next level.


What You'll Do

  • Develop deep technical expertise in your assigned product area and tech stack
  • Own production deployment, configuration, and release processes
  • Drive performance, reliability, and operability through continuous improvement
  • Build and maintain production tooling that supports deployment, orchestration, monitoring, and system diagnostics
  • Define and maintain observability, SLI/SLOs, and performance metrics in partnership with product owners
  • Leverage metrics and capacity planning to ensure scalability and uptime
  • Collaborate across engineering teams to troubleshoot and resolve complex production incidents
  • Lead and coordinate incident response, root cause analysis, and post-mortems
  • Influence architecture and promote best practices by aligning with global SRE teams
  • Document processes and procedures; provide mentorship and cross-training to peers
  • Actively manage operational risk for production changes


What You'll Need

  • Degree in Computer Science, a related field, or equivalent professional experience
  • 5+ years of relevant work experience in an IT ops role, such as DevOps, SRE, Linux Systems Engineering, or Network Engineering
  • Expert-level proficiency in C++
  • A rigorous, detail-oriented approach to operations
  • Strong understanding of the Linux operating system, including network and system configuration, kernel internals, scheduling, and performance tuning
  • Strong understanding of networking concepts such as routing, multicast, LLDP, VLANs, and Ethernet
  • A deep sense of ownership and desire to meet business priorities with urgency
  • Ability to handle shared operational and periodic on-call duties
  • Reliable and predictable availability

重要安全守则

申请工作时,切勿提供您的银行或信用卡详细资料。不要转账或完成无关的在线调查问卷。如果您发现可疑内容,请举报此招聘广告。

了解更多