People at Apple don't just build products, they craft the kind of experience that has revolutionized entire industries. The diverse collection of our people and their ideas inspire innovation in everything we do. Imagine what you could do here! Join Apple, and help us leave the world better than we found it.
The Apple Services Engineering (ASE) team builds and provides systems and infrastructure that power Apple's services (such as iCloud, Apple Music, Apple Intelligence, and Maps). We are the foundation on which Apple's software developers build the products that our customers love. Our services have to scale globally, stay highly available, and \\"just work.\\" If you love designing, engineering, and running systems and infrastructure that will help millions of customers, then this is the place for you!
Description
The ASE Compute team is looking for a Site Reliability Engineer to deploy and manage a large Kubernetes platform that Apple's services run on, partnering with engineering teams across the company to solve complex problems using both open-source and in-house tooling. You will contribute to the development of our controllers and namespace management infrastructure, working alongside senior engineers to strengthen the reliability of our Kubernetes services. You will learn to write well-tested code, participate in design reviews, and gradually take ownership of features. You'll have the opportunity to engage with the upstream community, gain hands-on experience with production-scale systems, and build the technical foundation to support service teams across Apple. The role also offers room to build AI-assisted tooling that accelerates triage, operational workflows, and infrastructure automation for the whole team.
Minimum Qualifications
Hands-on experience in Linux systems administration and containerization with enterprise distributions such as RHEL, Oracle Linux, or CentOS
Proficiency in Python or Go for scripting and tooling
Solid understanding of Linux fundamentals: file systems, process management, user and group administration, and package managementWorking knowledge of networking concepts including TCP/IP, DNS, DHCP, and basic firewall configuration
Experience with version control systems such as Git and configuration management (Puppet, Ansible, or equivalent)
Strong written and verbal communication skills
Preferred Qualifications
Site Reliability Engineering, DevOps, or Infrastructure focused experience
Experience with third-party cloud platforms (AWS, Google Cloud Platform, or Azure)
Experience with containerization and orchestration technologies such as Docker or Kubernetes
Familiarity with bare-metal provisioning and lifecycle management at datacenter scale
Understanding of cloud-native observability (Prometheus, Thanos, Splunk, or similar)
Familiarity with CI/CD pipelines and DevOps practices
Knowledge of OS security hardening, encryption, and regulatory compliance frameworks
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
- Dice Id: 90733111
- Position Id: ecf3db98b63c67c582ac2d3bbd49890
- Posted 30+ days ago