jobs in Autonomai Recruitment

Kerja Sepenuh Masa, Site Reliability Engineer di Autonomai Recruitment - Maukerja

Site Reliability Engineer

Autonomai Recruitment

Undisclosed

Singapore

Kongsi
Simpan

Lokasi Kerja

  • Singapore

Penerangan Kerja

Tanggungjawab

About the Firm


We're working with a leading proprietary trading firm committed to world-class research, bringing together exceptional talent in Mathematics, Physics, and Computer Science to push scientific and technological boundaries and apply cutting-edge research to global financial markets. The culture is built on innovation, intellectual honesty, and a relentless competitive edge — with collaboration and mutual respect at its core. Beyond trading, the firm designs and deploys technologies that extend well beyond the trading floor, funds start-ups across industries, and partners with leading global research organisations and universities.


The Trading Infrastructure team is a global organisation of engineers who architect, build, and maintain world-class infrastructure — from colo design and implementation, to optimising exchange connectivity, to building low-latency Wide Area Networks. The team leverages research and automation to continuously adapt and scale infrastructure in line with the evolving trading business.


We're looking for exceptional talent who can collaborate effectively across global teams and help take this infrastructure to the next level.


What You'll Do

  • Develop deep technical expertise in your assigned product area and tech stack
  • Own production deployment, configuration, and release processes
  • Drive performance, reliability, and operability through continuous improvement
  • Build and maintain production tooling that supports deployment, orchestration, monitoring, and system diagnostics
  • Define and maintain observability, SLI/SLOs, and performance metrics in partnership with product owners
  • Leverage metrics and capacity planning to ensure scalability and uptime
  • Collaborate across engineering teams to troubleshoot and resolve complex production incidents
  • Lead and coordinate incident response, root cause analysis, and post-mortems
  • Influence architecture and promote best practices by aligning with global SRE teams
  • Document processes and procedures; provide mentorship and cross-training to peers
  • Actively manage operational risk for production changes


What You'll Need

  • Degree in Computer Science, a related field, or equivalent professional experience
  • 5+ years of relevant work experience in an IT ops role, such as DevOps, SRE, Linux Systems Engineering, or Network Engineering
  • Expert-level proficiency in C++
  • A rigorous, detail-oriented approach to operations
  • Strong understanding of the Linux operating system, including network and system configuration, kernel internals, scheduling, and performance tuning
  • Strong understanding of networking concepts such as routing, multicast, LLDP, VLANs, and Ethernet
  • A deep sense of ownership and desire to meet business priorities with urgency
  • Ability to handle shared operational and periodic on-call duties
  • Reliable and predictable availability

Peringatan Penting

Jangan pernah kongsikan maklumat bank atau kad kredit anda semasa memohon pekerjaan. Elakkan membuat sebarang pembayaran atau mengisi survey yang tidak berkaitan. Jika ada yang mencurigakan, sila laporkan iklan pekerjaan ini segera.

Lebih Lanjut