Lead Site Reliability Engineer

ZA | FNZ GLOBAL MANAGEMENT LIMITED (INCORPORATED IN ENGLAND AND WALES)

Apply to this job
Edinburgh - UK London - UK Until 8/21/2026 5+ years exp First posted March 3, 2026 Last posted March 3, 2026
Job description

Role Purpose

The Site Reliability Engineer will work closely with Application, Infrastructure, and Network Engineering teams to ensure the reliability, scalability, and performance of FNZ platforms. This role focuses on deploying, integrating, and providing ongoing operational support for mission-critical systems, leveraging modern automation and cloud-native practices.

Key Responsibilities

·       Maintain high availability and performance of FNZ platforms.

·       Implement monitoring, alerting, and observability solutions to proactively detect and resolve issues.

·       Collaborate with engineering teams to design and implement robust deployment pipelines.

·       Ensure smooth integration of applications with infrastructure and network components.

·       Use Terraform for provisioning and managing infrastructure across environments.

·       Operate and optimize workloads onprem and public cloud.

·       Manage and troubleshoot application delivery networks, load balancing, and traffic routing.

·       Configure and support F5 Distributed Cloud or similar CDN/ADC technologies.

·       Participate in on-call rotations, perform root cause analysis, and implement preventive measures.

·       Work cross-functionally with Application, Infrastructure, and Network Engineering teams to deliver reliable services.

Required Skills & Experience

·       Kubernetes (K8s): Deep understanding of container orchestration and cluster management.

·       Terraform: Strong experience in Infrastructure as Code for cloud and on-prem environments.

·       Public Cloud: Hands-on experience with AWS, Azure, or GCP.

·       F5 Distributed Cloud or Similar: Knowledge of CDN/ADC platforms and their integration.

·       Networking Fundamentals: Expertise in application delivery networks, load balancing, traffic routing, and troubleshooting.

·       Observability Tools: Familiarity with Splunk, NewRelic, or similar.

·       Scripting & Automation: Proficiency in Terraform, Bash, or similar languages.

Desirable Skills

·       Experience with CI/CD pipelines and GitOps workflows.

·       Knowledge of SRE principles.

·       Familiarity with security best practices.

Key Attributes

·       Strong problem-solving and troubleshooting skills.

·       Ability to work collaboratively across multiple teams.

·       Passion for automation and reducing operational toil.

Reporting Line

Reports to: Head of Platform Operations/Application Engineering.

Works closely with Application Engineering, Infrastructure Engineering, Network Engineering teams.

#LI-CM1

About FNZ

FNZ is committed to opening up wealth so that everyone, everywhere can invest in their future on their terms. We know the foundation to do that already exists in the wealth management industry, but complexity holds firms back. 

We created wealth’s growth platform to help. We provide a global, end-to-end wealth management platform that integrates modern technology with business and investment operations. All in a regulated financial institution. 

We partner with the world’s leading financial institutions, with over US$2.2 trillion in assets on platform (AoP).

Together with our clients, we empower nearly 30 million people across all wealth segments to invest in their future.

About this role

Summary

Ensure reliable, scalable, and high-performance platforms through automation, monitoring, and cloud practices.

Job title

Lead Site Reliability Engineer

Experience level

senior level

Minimum experience

5+ years exp

Industry

finance

Location requirements

Edinburgh or London, UK; remote flexible

Salary

Not specified

Management role

No

Skills & keywords

Required skills

kubernetesterraformawsazuregcpf5 distributed cloudnetworkingobservability toolsbash

Preferred skills

ci/cd pipelinesgitopssre principlessecurity best practices

Specializations

kubernetesterraformcloudnetworkingautomation
Locations

Structured locations inferred from the posting.

Edinburgh, UK

On-site City

London, UK

On-site City