Site Reliability Engineer

Company: Tempest Vane Partners
Apply for the Site Reliability Engineer
Location: London
Job Description:

London | Hedge Fund | Permanent

The Client

Tempest Vane Partners has partnered with a globally recognised investment management business seeking an experienced Infrastructure Engineer to join a high-calibre platform engineering team in London.

The Role

This opportunity sits within a technology-driven organisation where engineering is fundamental to business performance. The successful candidate will work on highly scalable infrastructure supporting large-scale research, analytics, and trading environments across an international platform.

The position offers significant exposure to modern cloud-native engineering, distributed systems, automation, and platform reliability initiatives within a fast-moving and technically sophisticated environment.

The Opportunity

The incoming engineer will play a key role in the design, improvement, and automation of critical infrastructure systems across both cloud and on-premise environments.

Working closely with infrastructure, development, and platform teams, the role will focus heavily on scalability, resilience, observability, and operational efficiency.

Core responsibilities

  • Engineering and maintaining enterprise-scale Kubernetes environments
  • Supporting the evolution toward cloud-native and distributed architectures
  • Driving automation initiatives across infrastructure and operational workflows
  • Developing and promoting Infrastructure-as-Code best practices
  • Enhancing platform stability, availability, and system performance
  • Contributing to monitoring, observability, and incident response capabilities
  • Partnering with global engineering teams on platform improvements and technical delivery

Candidate Requirements

Successful candidates are likely to demonstrate experience across several of the following areas:

  • Strong scripting or software engineering capability using Python, Golang, Bash, or similar
  • Deep understanding of Kubernetes and container-based infrastructure
  • Experience with Terraform, Ansible, Puppet, or equivalent automation tooling
  • Knowledge of hybrid infrastructure and distributed systems environments
  • Familiarity with observability and monitoring technologies such as Prometheus, Grafana, ELK, or Jaeger
  • Exposure to data streaming or workflow technologies including Kafka or Airflow
  • Strong troubleshooting capability with a focus on automation and reliability engineering

For further information or a confidential discussion, please apply directly via Tempest Vane Partners.

#J-18808-Ljbffr…

Posted: July 19th, 2026