SlipstreamJobs tracks this role from the company's public career site. Apply directly on the employer's site.
Apple Services Engineering (ASE) is seeking a Site Reliability Engineer to join the Private Cloud Compute (PCC) team in London. PCC represents a groundbreaking approach to cloud intelligence, extending the security and privacy of Apple devices into the cloud while unlocking advanced intelligence for users without compromising privacy.
You will be responsible for the availability, reliability, and automation of critical systems and services that enable PCC to deliver cloud intelligence at scale. This role offers the opportunity to directly shape how Apple builds and operates services globally while maintaining the highest standards of operational excellence.
Key Responsibilities:
- Deploy, support, and monitor new and existing services, platforms, and application stacks
- Conduct scale testing to measure, tune, and optimize system performance
- Enhance, architect, and deliver software to improve availability, scalability, and security of Apple's internet services
- Build and run systems, infrastructure, and applications through automation
- Participate in periodic on-call duties to ensure service reliability
Required Qualifications:
- BS in Computer Science or related field, or equivalent professional experience
- 4+ years managing and scaling distributed systems in public, private, or hybrid cloud environments
- Strong experience deploying, supporting, and supervising services, platforms, and application stacks
- Excellent troubleshooting and problem-solving skills
- Experience with scale testing, disaster recovery, and capacity planning
- Demonstrated ability to write programs in Java, Go, Python, or Perl
- Experience with Kubernetes, Nginx, Envoy, Prometheus, and/or Docker
- Passion for automation and eliminating repetitive manual processes
Preferred Qualifications:
- Understanding of standard networking protocols (HTTP, DNS, TCP/IP, load balancing)
- Deep knowledge of Linux operating systems (kernel, memory, processes, threads)
- Experience with configuration management systems (Puppet, Chef, Ansible, Salt)
- Familiarity with complex distributed system architectures