Apple

Apple

Posted via Apple Careers

Service Reliability Engineer, G&A Solutions Engineering

Always Hiring

Posted Sep 29, 2026

Role at a glance

Job function
Software Engineering & IT DevOps & Site Reliability Engineering
Salary
Not Disclosed
Location
Austin, Texas, United States
Experience
4+ years of experience in a Site Reliability Engineering, production support or related role, supporting large-scale, enterprise-level...
Education
Bachelor's degree in Computer Science or work related equivalent experience

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The Service Reliability Engineer on Apple's General and Administrative Solutions Engineering team supports global, mission-critical production systems. The role focuses on maintaining service health, stability, efficiency, and reliability.

What You'll Do

  • Monitor service performance, identify potential bottlenecks, and implement solutions to improve efficiency and resilience.
  • Lead incident response, resolve incidents, and conduct root cause analysis.
  • Develop automation to streamline operational tasks and reduce manual intervention.
  • Collaborate with development teams on monitoring, alerting, and scalability practices for new services.
  • Maintain run-books and service level objectives, participate in on-call rotations, and define and supervise service level indicators.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Minimum qualifications include 4+ years in Site Reliability Engineering, production support, or a related role supporting large-scale enterprise services; proficiency in at least one programming language and scripting language; experience with cloud platforms and cloud-native technologies; hands-on experience with monitoring and alerting tools; and a bachelor's degree in Computer Science or equivalent work-related experience.

Required

  • Strong proficiency in at least one programming language, such as Python, Java, or Go, and scripting languages, such as Bash or PowerShell.
  • Experience with cloud platforms such as AWS, Azure, or GCP, and cloud-native technologies such as Kubernetes or Docker.
  • Hands-on experience with monitoring and alerting tools such as Prometheus, Grafana, Splunk, or Datadog.

Preferred

  • Familiarity with CI/CD pipelines and DevOps practices.
  • Experience with database technologies such as MySQL, PostgreSQL, or NoSQL databases.
  • Knowledge of ITIL frameworks and incident management processes.
  • Experience with vibe coding.
  • Understanding of Linux/Unix system administration.
  • Experience with configuration management tools such as Ansible, Chef, or Puppet.

Original job description

Content provided by the employer

Summary

Do you have a passion for ensuring the reliability, scalability, and performance of critical services? Are you a highly motivated and expert engineer with a strong understanding of Site Reliability Engineering (SRE) principles and a desire to automate and improve processes? Join Apple's General and Administrative (G&A) Solutions Engineering team as a Service Reliability Engineer and play a vital role in supporting our global, mission-critical production systems.

Description

You'll be at the forefront of maintaining the health, stability, and efficiency of our services, working with a diverse range of technologies and platforms. You will collaborate with Engineers, Data Engineers, DBAs, and network specialists to proactively identify and resolve potential issues, automate repetitive tasks, and drive continuous improvement initiatives. Your expertise will directly impact the reliability of our systems, enabling Apple to deliver innovative products and services to our customers.

Responsibilities

Proactively monitor service performance, identify potential bottlenecks, and implement solutions to optimize efficiency and resilience

Lead incident response efforts, driving rapid resolution and conducting thorough root cause analysis (RCA)

Develop and implement automation strategies to streamline operational tasks, improve service resilience, and reduce manual intervention

Apply SRE principles to maintain highly reliable and scalable service infrastructure

Collaborate closely with development teams to ensure that new services are designed for operational perfection, incorporating best practices for monitoring, alerting, and scalability

Contribute to the creation and maintenance of comprehensive documentation, including run-books, service level objectives (SLOs)

Participate in on-call rotations, providing 24/7 support for critical services and responding to incidents with a sense of urgency

Find opportunities for process improvement and drive initiatives to enhance the efficiency and effectiveness of the service reliability team

Champion a culture of continuous learning and knowledge sharing within the team

Define and supervise key service level indicators (SLIs) to measure and improve service reliability

Minimum Qualifications

4+ years of experience in a Site Reliability Engineering, production support or related role, supporting large-scale, enterprise-level services

Strong proficiency in at least one programming language (e.g., Python, Java, Go) and scripting languages (e.g., Bash, PowerShell)

Experience with cloud platforms (e.g., AWS, Azure, GCP) and cloud-native technologies (e.g., Kubernetes, Docker)

Hands-on experience with monitoring and alerting tools (e.g., Prometheus, Grafana, Splunk, Datadog)

Bachelor's degree in Computer Science or work related equivalent experience

Preferred Qualifications

Familiarity with CI/CD pipelines and DevOps practices

Experience with database technologies (e.g., MySQL, PostgreSQL, NoSQL databases)

Knowledge of ITIL frameworks and incident management processes

Experience with vibe coding

Understanding of Linux/Unix system administration

Experience with configuration management tools (Ansible, Chef, Puppet)

Application Deadline

Apple accepts applications to this posting on an ongoing basis.

Apple

About the company

Apple

Large Enterprise

Apple Inc. is a global technology company known for its innovative products and services, including the iPhone, iPad, Mac computers, and Apple Watch. Founded in 1976, Apple has continuously pushed the boundaries of design and functionality, earning a reputation for high-quality consumer electronics and software solutions like iOS and macOS. With a strong commitment to user experience and privacy, Apple also leads in digital services, offering platforms such as the App Store, Apple Music, and iCloud. The company's focus on sustainability and corporate responsibility further enhances its standing as a leader in the technology sector.