Senior Site Reliability Engineer
Summary
The Media Platforms SRE team under the Apple Service Engineering division is one of the most exciting examples of Apple’s long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, and Apple Books. And they do it on a massive scale, meeting Apple’s high expectations with high performance to deliver a huge variety of entertainment in over 35 languages to more than 150 countries.
These engineers build secure, end-to-end solutions. They develop the custom software used to process all the creative work, the tools that providers use to deliver that media, all the server-side systems, and the APIs for many Apple services.
Thanks to Apple’s unique integration of hardware, software, and services, engineers here partner to get behind a single unified vision. That vision always includes a deep commitment to strengthening Apple’s privacy policy, one of Apple’s core values. Although services are a bigger part of Apple’s business than ever before, these teams remain small, forward-thinking, and cross-functional, offering greater exposure to the array of opportunities here.
Description
Media Platforms SRE is responsible for designing, building and running a diverse set of services that ingest, transform and manage all media content across products like the App Store, Apple Music, Apple Fitness+ and, Apple TV+. Whether it’s the latest episode of the Morning Show, a new album from your favorite artist or a brand new iOS app destined for the App Store. You will make sure our systems are highly available, fast and efficient. You’ll interact with a diverse range of system types -- public facing web & API developer services like App Store Connect & TestFlight, Media Processing pipelines that enables audio & video transcoding at scale as well as internal business tools/systems and machine learning models. Culturally we believe in a close partnership with our development teams and aim to design & build new services together. We're passionate about software and automation in SRE and develop a variety of tooling and infrastructure. Our services run on mixed platforms - many run on our bare metal fleet which you will help transition to modern Cloud infrastructure and Kubernetes micro-services and work with state of the art observability tools and monitoring platforms.
Minimum Qualifications
- BS degree in computer science or equivalent field with 6+ years experience or MS degree in computer science or equivalent field with 4+ years experience
- At least 6 years in a Reliability Engineering, DevOps or infrastructure focused role
- Advanced experience with programming languages (Golang, Python, Java, C++)
- Deep systems and infrastructure knowledge
- Excellent troubleshooting and problem solving skills
Preferred Qualifications
- Passion for designing and building reliable systems
- Familiarity with microservices architecture and container orchestration with Kubernetes
- Demonstrated ability to deliver results on time with high quality
- Automation advocate - you truly believe in removing operation load with software
- Experience with deploying, supporting and monitoring new and existing services, platforms, and application stacks