Site Reliability Engineer, iCloud

London Until 9/22/2026 5+ years exp First posted July 24, 2026 Last posted July 24, 2026
Job description

People at Apple don’t just build products — they craft experiences our customers love and depend on. Apple Services Engineering (ASE) builds and supports the systems that make many of these daily experiences possible. If you’ve used Apple products, you’ve likely interacted with us. Apple Services Site Reliability Engineering (SRE) teams are responsible for the systems and services that directly support those customers and their experiences. We are looking for an SRE with experience in building and supporting highly available customer-facing services.

Description

Apple Services’ scale is BIG. Operating at our scale, across multiple geographies and servicing hundreds of millions of users presents unique challenges. As a Software Developer in SRE at Apple, you'll need to solve these problems using data, teamwork, and your own expertise. ASE Products Site Reliability teams are responsible for the reliability and performance of the server software stack that powers products like iCloud Photos, Mail, Drive, Backup and many more. We do that by focusing on reliability best practices from service inception to production, collaborating deeply with product development teams to deliver a superlative product and shared vision while leveraging data and automation as first principles. We run a mix of open source, vendor licensed, and internally developed tools to manage the end to end SDLC of our products. You'll learn these tools and have opportunities to improve them.

Minimum Qualifications

Strong sense of ownership, customer service, and integrity proven through clear communication.
BS in Computer Science or related field, or equivalent employment
5 + years experience in managing and scaling distributed systems in a public, private, or hybrid cloud environment
Strong experience with deploying, supporting and supervising new and existing services, platforms, and application stacks
Experience with scale testing, disaster recovery, and capacity planning
Experience with observability platforms with Splunk, Grafana, Prometheus.
Demonstrable fluency in at least one of the following languages: Java, Python, or Go.
Experience with Kubernetes, Nginx, Envoy, Prometheus, and/or Docker.

Preferred Qualifications

Understanding of standard networking protocols and components such as: HTTP, DNS, ECMP, TCP/IP, ICMP, the OSI Model, Subnetting and Load Balancing strategies.
Understanding of the Linux Operating System, including Kernel, Memory, Process, Threads, Static / Shared Libraries, IPC, Signals.
Experience in developing iOS apps using Xcode and Swift.
Experience in OpenTelemetry Standards / distributed tracing like jaeger

About this role

Summary

Manage and support highly available customer-facing services for Apple iCloud.

Job title

Site Reliability Engineer, iCloud

Experience level

5+ years

Minimum experience

5+ years exp

Industry

technology

Location requirements

London, on-site required

Salary

Not specified

Management role

No

Skills & keywords

Required skills

distributed systemscloud environmentsservice deploymentmonitoring platformsJavaPythonGoKubernetesNginxEnvoyDocker

Preferred skills

networking protocolsLinux OSiOS appsXcodeSwiftdistributed tracing

Specializations

distributed systemscloud computingservice reliabilitymonitoringKubernetes
Locations

Structured locations inferred from the posting.

London, UK

On-site City