Site Reliability Engineer - Selby Jennings - #2120199

eFinancialCareers


Date: 5 days ago
City: London
Contract type: Full time
Work schedule: Full day
eFinancialCareers

Our client, a world renowned hedge fund, is seeking a Site Reliability Engineer to join its world-class engineering organization in London. This role sits at the intersection of software engineering and infrastructure, focusing on the reliability, scalability, and performance of the technology platforms that power global trading and investment operations.

Working closely with software engineers, quantitative researchers, traders, and infrastructure teams, you will be responsible for building automation, improving observability, and ensuring critical production systems operate at the highest levels of availability and efficiency.

Key Responsibilities

  • Design, build, and maintain highly reliable, scalable, and automated infrastructure platforms.
  • Drive improvements in system performance, monitoring, observability, and operational efficiency.
  • Troubleshoot and resolve complex production incidents across distributed systems.
  • Develop tools and automation to reduce operational overhead and improve platform resilience.
  • Partner with engineering teams to improve system design, deployment processes, and operational readiness.
  • Participate in incident management and root cause analysis, ensuring lessons learned are incorporated into future improvements.
  • Support mission-critical trading and research environments in a fast-paced, high-performance setting.

Requirements

  • Strong software engineering skills in Python, Go, C++, Java, or a similar language.
  • Deep Linux systems knowledge and experience operating large-scale production environments.
  • Experience with Kubernetes, containerisation technologies, and cloud infrastructure.
  • Strong understanding of networking, distributed systems, and infrastructure automation.
  • Experience with monitoring and observability tools such as Prometheus, Grafana, Splunk, or similar.
  • Proven track record of solving complex reliability, scalability, or performance challenges.
  • Excellent problem-solving skills and ability to operate effectively in high-pressure environments.

Preferred Backgrounds

  • Technology companies operating large-scale distributed systems.
  • High-frequency trading firms, hedge funds, or electronic trading environments.
  • Cloud infrastructure, platform engineering, or production engineering teams.
  • Software engineers with a strong interest in reliability and infrastructure.

How to apply

To apply for this job you need to authorize on our website. If you don't have an account yet, please register.

Post a resume

Similar jobs

Data Governance Foundation Senior Lead Analyst SVP - Citi

eFinancialCareers,
12 hours ago
Discover your future at Citi Working at Citi is far more than just a job. A career with us means joining a team of approximately 219,000 dedicated people from around the globe. At Citi, you'll have the opportunity to grow...
eFinancialCareers

Chief Executive Officer

Clover HR Services Limited t/a Clover HR,
£90,000 / year
12 hours ago
Imroc is seeking a Chief Executive Officer to provide strategic, values-led leadership as the organisation enters its next phase of development. This is a significant leadership opportunity for someone who can combine a strong commitment to recovery and lived experience...

Interim HR Operations Manager

Morgan Law,
£50,000 - £53,000 / year
14 hours ago
We are seeking an experienced and proactive interim People Operations Manager to lead the delivery of high-quality HR operational services across the employee lifecycle. This is a fantastic opportunity for a skilled HR professional who thrives in a fast-paced environment...
Morgan Law