Search by job, company or skills

Site Reliability Engineer

5-7 Years
  • Posted 2 days ago
  • Be among the first 10 applicants

Job Description

About the Firm

We're working with a leading proprietary trading firm committed to world-class research, bringing together exceptional talent in Mathematics, Physics, and Computer Science to push scientific and technological boundaries and apply cutting-edge research to global financial markets. The culture is built on innovation, intellectual honesty, and a relentless competitive edge — with collaboration and mutual respect at its core. Beyond trading, the firm designs and deploys technologies that extend well beyond the trading floor, funds start-ups across industries, and partners with leading global research organisations and universities.

The Trading Infrastructure team is a global organisation of engineers who architect, build, and maintain world-class infrastructure — from colo design and implementation, to optimising exchange connectivity, to building low-latency Wide Area Networks. The team leverages research and automation to continuously adapt and scale infrastructure in line with the evolving trading business.

We're looking for exceptional talent who can collaborate effectively across global teams and help take this infrastructure to the next level.

What You'll Do

  • Develop deep technical expertise in your assigned product area and tech stack
  • Own production deployment, configuration, and release processes
  • Drive performance, reliability, and operability through continuous improvement
  • Build and maintain production tooling that supports deployment, orchestration, monitoring, and system diagnostics
  • Define and maintain observability, SLI/SLOs, and performance metrics in partnership with product owners
  • Leverage metrics and capacity planning to ensure scalability and uptime
  • Collaborate across engineering teams to troubleshoot and resolve complex production incidents
  • Lead and coordinate incident response, root cause analysis, and post-mortems
  • Influence architecture and promote best practices by aligning with global SRE teams
  • Document processes and procedures; provide mentorship and cross-training to peers
  • Actively manage operational risk for production changes

What You'll Need

  • Degree in Computer Science, a related field, or equivalent professional experience
  • 5+ years of relevant work experience in an IT ops role, such as DevOps, SRE, Linux Systems Engineering, or Network Engineering
  • Expert-level proficiency in C++
  • A rigorous, detail-oriented approach to operations
  • Strong understanding of the Linux operating system, including network and system configuration, kernel internals, scheduling, and performance tuning
  • Strong understanding of networking concepts such as routing, multicast, LLDP, VLANs, and Ethernet
  • A deep sense of ownership and desire to meet business priorities with urgency
  • Ability to handle shared operational and periodic on-call duties
  • Reliable and predictable availability

More Info

Job Type:
Industry:
Function:
Employment Type:

Job ID: 151599239

Similar Jobs

Singapore

Skills:

JavaNginxUnixDnsLdapMicroservicesShellLinuxAnsibleKubernetesPythonAWSaptGoPackerFaaSoci

Singapore

Skills:

ShellPerlLinuxPythonDevOps solutionssoftware testing and validationGolarge-scale system operation

Singapore, Kallang

Skills:

UnixElkRpaLinuxPower AutomateSeleniumPythonOAPMPlaywright

Singapore

Skills:

LinuxautomationKubernetesScriptingmicroservices architectureobservability monitoring toolscloud-hosted applicationscontainer technologies

Singapore

Skills:

GithubSentryPostgreSQLPrometheusArtifactoryGrafanaJiraDatadogNew RelicJenkinsDockerLinuxAnsibleMySQLDynatraceMongoDBSplunkPythonKubernetes