Site Reliability Engineering Manager
Apple
Apply to this jobDo you want to help build some of the largest and most consequential enterprise and customer technology systems in the world? Join Apple’s Information Systems and Technology (IS&T) organization. IS&T is the engine behind everything Apple does for customers and for the people who build for them. It’s Apple’s central nervous system. Supporting 2.5 billion active Apple devices, processing billions of secure transactions, and keeping the technology that defines modern life running flawlessly, IS&T makes the impossible feel effortless.Do you love building solutions to handle global complexity and immense scale? Imagine what you could do here. AI & Data Platforms (AiDP) is IS&T's engine for AI-powered innovation. The team brings together data, application development, and machine learning — including generative AI — along with data services and customer success functions, to help IS&T build solutions more efficiently and streamline the adoption and embedding of generative AI across Apple.
Description
We are seeking an experienced Site Reliability Engineering (SRE) Manager to support scalable and resilient distributed systems that power Apple's data pipelines and analytics platforms. Our Enterprise Data Warehouse landscape caters to a wide variety of real-time, near real-time and batch analytical solutions. These solutions are an integral part of business functions like Sales, Operations, Finance, AppleCare, Marketing and Internet Services, enabling business drivers to make critical decisions. We utilizes proprietary and open source technologies such as Kafka, Spark, Iceberg, Airflow, and others to build these solutions. If you are passionate about addressing infrastructure challenges at scale, both on-premises and in the cloud, and focused on optimizing scalable solutions by prioritizing ease of use and maintenance, you will discover exciting opportunities in AiDP.
As a hands-on SRE Manager, you’ll lead by example—actively driving operational excellence, contributing to code, and ensuring system reliability. You will be deeply involved in incident response across complex, distributed data platforms designed to support data exploration, analytics, and reporting solutions. These platforms operate at the unique intersection of high data volume and hybrid infrastructure, spanning both cloud and on-premise environments.
Minimum Qualifications
10+ years of experience in Site Reliability Engineering (SRE) or a related domain.
2+ years of direct people-management experience, including leading, hiring, developing, and building engineering teams.
Hands-on experience supporting and maintaining applications in cloud or hybrid environments.
Expertise in cloud-native services, including ETL frameworks (e.g., Apache Spark, Flink) and messaging systems (e.g., Kafka).
Strong knowledge of cloud infrastructure & services (e.g., AWS, GCP, Kubernetes).
Experience with Observability tools (e.g., Prometheus, Grafana, CloudWatch).
Programming experience in Python, Java, or Scala.
Proven ability to lead incident response, perform root cause analysis, and drive system reliability improvements.
Bachelor’s degree or equivalent.
Preferred Qualifications
Hands-on experience supporting enterprise data systems on distributed architectures.
Exposure to data visualization tools such as Tableau, Business Objects, or ThoughtSpot, with experience supporting and troubleshooting related issues.
Experience with modern & distributed databases such as Snowflake, Cassandra, SingleStore, or SAP HANA.
Experience using Generative AI or automation tools for issue detection, alerting, or remediation.
Solid understanding of system design, data structures, and incident management best practices.
Summary
Lead and support scalable, resilient distributed systems for Apple’s data platforms.
Job title
Site Reliability Engineering Manager
Experience level
10+ years
Minimum experience
10+ years exp
Industry
technology
Location requirements
Bengaluru or Hyderabad, remote work allowed
Salary
Not specified
Management role
Yes
Required skills
Preferred skills
Specializations
Structured locations inferred from the posting.
Bengaluru, Karnataka, India
Hyderabad, Telangana, India