Service Reliability Engineer (SRE)

Washington, WA, US • Posted 9 days ago • Updated 9 hours ago
Full Time
On-site
Fitment

Dice Job Match Score™

📋 Comparing job requirements...

Job Details

Skills

  • Art
  • Media
  • Computer Hardware
  • Privacy
  • Cloud Computing
  • Data Processing
  • Data Analysis
  • Music
  • Scalability
  • Build Automation
  • Build Tools
  • Network
  • Agile
  • Reliability Engineering
  • DevOps
  • Computer Science
  • Management
  • Apache Hadoop
  • Apache Spark
  • Apache Flink
  • Kubernetes
  • Amazon Web Services
  • Quick Learner
  • Analytical Skill
  • Problem Solving
  • Conflict Resolution
  • Java
  • Big Data
  • Migration
  • Communication

Summary

The Apple Services Engineering team (ASE) is one of the most exciting examples of Apple's long-held passion for combining art and technology. These are the people who power the App Store, Apple TV, Apple Music, Apple Podcasts, and Apple Books. And they do it at an extensive scale, meeting our high expectations with dedication to deliver a huge variety of entertainment in over 35 languages to more than 150 countries.

These engineers build secure, end-to-end solutions. They develop the custom software used to process all the creative work, the tools that providers use to deliver that media, all the server-side systems, and the APIs for many Apple services.

Thanks to Apple's unique integration of hardware, software, and services, engineers here partner to get behind a single unified vision. That vision always includes a deep commitment to strengthening Apple's privacy policy, one of our core values. Although services are a bigger part of Apple's business than ever before, these teams remain small, and multi-functional, offering greater exposure to the array of opportunities here.

Description

The Service Reliability Engineer (SRE) role in Apple Services Engineering requires a mix of strategic engineering and design along with hands-on, technical work. This SRE will configure, tune, and fix multi-tiered systems to achieve optimal application performance, stability and availability.

We manage jobs as well as applications on bare-metal and cloud computing platforms to deliver data processing for many of Apple's global products. Our teams work with exabytes of data, petabytes of memory, and tens of thousands of jobs to enable predicable and performant data analytics enabling features in Apple Music, TV+, Appstore and other world class products.

If you love designing, running systems that will impact millions of users then this is the place for you!

THE MAIN RESPONSIBILITIES FOR THIS POSITION INCLUDE:

- Support java based applications & Spark/Flink jobs on Baremetal, AWS & Kubernetes

- Ability to understand the application requirements (Performance, Security, Scalability etc.) and assess the right services/topology on AWS, Baremetal & Kubernetes

- Build automation to enable self-healing systems

- Build tools to monitor high performance & alert the low latency applications

- Ability to troubleshoot application specific, core network, system & performance issues.

- Involvement in challenging and fast paced projects supporting Apple's business by delivering innovative solutions.

- Monitor production, staging, test and development environments for a myriad of applications in an agile and dynamic organization.

Minimum Qualifications

At least 5 years in a Site Reliability Engineering (SRE), DevOps role

BS degree in computer science or equivalent field with 5+ years or MS degree with 3+ years experience, or equivalent.

5+ years of running services in a large scale *nix environment

Understanding of SRE principles and goals along with prior on-call experience

Extensive experience in managing the applications on AWS & Kubernetes

Deep understanding and experience in one or more of the following - Hadoop, Spark, Flink, Kubernetes, AWS

Preferred Qualifications

Fast learner with excellent analytical problem solving and interpersonal skills

Experience supporting Java applications

Experience on Big Data Technologies

Experience working with geographically distributed teams and implement high level projects and migrations

Strong communication skills and ability deliver results on time with high quality
Employers have access to artificial intelligence language tools (“AI”) that help generate and enhance job descriptions and AI may have been used to create this description. The position description has been reviewed for accuracy and Dice believes it to correctly reflect the job opportunity.
  • Dice Id: 90733111
  • Position Id: 4ebe6d43204e3e026bedb78bd31c0803
  • Posted 9 days ago
Create job alert
Set job alertNever miss an opportunity! Create an alert based on the job you applied for.

Similar Jobs

No location provided

Today

Full-time

USD 81,100.00 - 187,000.00 per year

Redmond, Washington

Today

Full-time

USD 165,000.00 - 270,000.00 per year

Remote

5d ago

Easy Apply

Contract

Depends on Experience

Redmond, Washington

Today

Full-time

USD 125,000.00 - 150,000.00 per year

Search all similar jobs