Source description
About the role
Partner with US SRE counterparts and product engineering to define and execute the reliability engineering vision, strategy, and roadmap for Apple Data Platform managed services. Lead day-to-day execution of reliability initiatives, including sprint planning, prioritization, and retrospectives, focusing on operational outcomes. Establish and own SLIs, SLOs, and error budgets, using metrics and observability to guide engineering trade-offs and reliability investments. Promote automation and operational efficiency through tooling, testing, deployment pipelines, and self-healing systems. Mentor and develop engineers through regular one-on-ones, career planning, and performance feedback, fostering a culture of ownership and continuous improvement. Collaborate with recruiting to attract and hire top reliability engineering talent. Advocate for reliability, resilience, and operational excellence across multiple product and platform teams. Handle production on-call and incident management responsibilities. Job Requirements Minimum 5 years of experience in site reliability engineering or a related field. Strong knowledge of SRE principles and technical experience. Experience with Jupyter Notebooks, Spark, Flink, and other open-source products used by our platform teams. Ability to build, scale, and mentor a high-performing SRE team. Strong leadership skills and ability to establish strong cross-functional partnerships. Passionate about operating mission-critical, globally distributed systems and driving long-term reliability improvements.
More at Apple
Related open roles
Senior Software Engineering Manager – Manufacturing Intelligence, Agentic Systems & Physical AI
Bangalore
Senior Quality Engineering Manager
Hyderabad
Senior Quality Engineering Manager
Hyderabad
Senior Quality Engineering Manager
Hyderabad
Site Reliability Engineering Manager, Apple Data Platform (India)
India
Manager - Business Intelligence Data Engineering for Retail Customer Care
Hyderabad