Apple

Apple

Posted via Apple Careers

Technical Operations & Site Reliability Engineer, Customer Systems

Always Hiring

Posted Aug 4, 2026

Role at a glance

Job function
Software Engineering & IT DevOps & Site Reliability Engineering
Salary
Not Disclosed
Location
Sunnyvale, California, United States
Work arrangement
On-site
Employment
Full-time
Education
B.S in Computer Science, Computer Engineering or similar or equivalent related work experience.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

The TechOps Engineer supports the reliability, availability, and performance of business-critical, globally distributed systems. The role focuses on production operations, automation, incident response, and collaboration with support, engineering, and business operations teams.

What You'll Do

  • Manage large-scale production outages, lead incident response, and improve response efficiency.
  • Design, build, and maintain automation and software tools for monitoring and managing distributed systems and reducing repetitive...
  • Plan system health monitoring and incident communications for critical global applications, and drive operational metrics and KPI alignment.
  • Partner with cross-functional teams and engineers to improve reliability, stability, efficiency, and operational processes.
  • Maintain architecture, infrastructure, and procedure documentation; write status and incident reports, and create training materials and...

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Minimum qualifications include interpreting operational data and hands-on experience with production monitoring, log analysis, troubleshooting, and support dashboards; a B.S. in Computer Science, Computer Engineering, or similar, or equivalent related work experience; understanding of networking protocols and components including HTTP, DNS, TCP/IP, ICMP, the OSI Model, subnetting, and load balancing; experience using AI and LLMs for operational efficiency, including model training, optimization, and model utilities; and experience with scripting and automation tools such as Java, JEE, REST, Swift/Objective-C, database schema design, and data access technologies.

Required

  • Experience interpreting operational data from Hubble, ExtraHop, Splunk, or other monitoring tools, plus hands-on experience with...
  • B.S. in Computer Science, Computer Engineering, or similar, or equivalent related work experience.
  • Understanding of HTTP, DNS, TCP/IP, ICMP, the OSI Model, subnetting, and load balancing.
  • Experience using AI and LLMs to enhance operational efficiency through model training, optimization (including Model Context Protocol or...
  • Experience with scripting languages and automation tools such as Java, JEE, REST, Swift/Objective-C, database schema design, and data...

Preferred

  • Experience strategizing and achieving operational excellence in global distributed systems.
  • Fundamental understanding of distributed systems, including microservices, messaging brokers, and versioning.
  • Experience driving operations teams for large-scale mission-critical applications in a 24x7 environment across multiple locations.
  • Understanding of Linux, including kernel, memory, process, threads, static/shared libraries, IPC, and signals.
  • Excellent organizational and documentation skills.
  • Excellent interpersonal skills; proactive, with a strong sense of personal ownership.

Original job description

Content provided by the employer

Summary

At Apple, Customer Experience is at the forefront of everything we do. The Customer Systems Operations team is looking for a highly skilled and motivated TechOps Engineer (Technical Operations & Site Reliability) to join us. The team is responsible for maintaining the reliability, availability, and performance of business-critical, globally distributed systems.

If you have the desire and motivation to design and develop automation solutions to streamline system sustenance, monitoring, and operational workflows, while collaborating closely with support, engineering and business operations teams, this profile is for you. Ideal candidates will combine a passion for operational excellence with strong software engineering skills, and thrive in a fast-paced, change-driven environment focused on continuous improvement and flawless delivery.

Description

Manage large-scale production outages, leading incident response and improving efficiency.
Design, build, and maintain automation solutions to streamline the monitoring, sustenance, and management of large-scale distributed systems.
Develop tools and software (using Java/JEE, REST, Swift/Objective C, Python, Go, or Bash) to automate repetitive operational tasks, reduce manual intervention, and improve system reliability. Utilize AI & LLM models to achieve Operational Excellence in application support.
Plan and execute actionable system health monitoring, incident response, and communication across critical global applications. Drive operational metrics and KPI identification and alignment.
Partner with multi-functional teams to improve reliability, efficiency, stability, and processes.
Be a self-directed problem-solver exhibiting deftness to handle multiple simultaneous competing priorities and deliver solutions in a timely manner.
Create and maintain accurate, up-to-date documentation reflecting architecture, infra configuration, and procedures. Write status and incident reports. Write training material and train users in complex topics.
Partner with a team of highly skilled engineers across the globe and guide their work towards operational excellence, gaining efficiency.
Build a culture where the regional members are responsible for cultivating strong in-region relationships and getting results for our business partners ensuring they remain informed about significant incidents and problems.

Minimum Qualifications

Experience in interpreting operational data from systems like Hubble, ExtraHop, Splunk or other monitoring tools along with hands-on experience of production monitoring systems, log analysis, troubleshooting, and support dashboards.
B.S in Computer Science, Computer Engineering or similar or equivalent related work experience.
Understanding of standard networking protocols and components such as: HTTP, DNS, TCP/IP, ICMP, the OSI Model, Subnetting and Load Balancing.
Experience in using AI and Large Language Models (LLMs) to enhance operational efficiency through tasks such as model training, optimization (including areas like Model Context Protocol or similar methods), and designing effective model utilities.
Experience in scripting languages and automation tools such as Java, JEE, REST, Swift/Objective C, database schema design and data access technologies.

Preferred Qualifications

Experience in strategizing and achieving operational excellence in global distributed systems.
Fundamental understanding of distributed systems including: Micro services, Messaging Brokers, and Versioning.
Experience in driving operations teams for large-scale mission-critical applications working in a 24x7 environment across multiple locations.
Understanding of the Linux Operating System, including Kernel, Memory, Process, Threads, Static / Shared Libraries, IPC, and Signals.
Excellent organizational and documentation skills.
Excellent interpersonal skills. Proactive, with a strong sense of personal ownership.

Pay & Benefits — Sunnyvale, California, United States

At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $150,400 and $277,600, and your base pay will depend on your skills, qualifications, experience, and location.

Apple employees also have the opportunity to become an Apple shareholder through participation in Apple’s discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple’s Employee Stock Purchase Plan. You’ll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses — including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits

Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program.

Application Deadline

Apple accepts applications to this posting on an ongoing basis.

Apple

About the company

Apple

Large Enterprise

Apple Inc. is a global technology company known for its innovative products and services, including the iPhone, iPad, Mac computers, and Apple Watch. Founded in 1976, Apple has continuously pushed the boundaries of design and functionality, earning a reputation for high-quality consumer electronics and software solutions like iOS and macOS. With a strong commitment to user experience and privacy, Apple also leads in digital services, offering platforms such as the App Store, Apple Music, and iCloud. The company's focus on sustainability and corporate responsibility further enhances its standing as a leader in the technology sector.