Apple

Apple

Posted via Apple Careers

Software Engineer - Generative Data Platform, Evaluation

Always Hiring

Posted Oct 1, 2026

Role at a glance

Job function
Software Engineering & IT Software Engineering
Salary
Not Disclosed
Location
San Francisco, California, United States
Experience
3+ years of experience building and operating production software in Python.
Education
Bachelor's degree in Computer Science, Computer Engineering, or a related field, or 3 years of equivalent work experience.

Spotted an issue?

We’ll check it against the original posting.

Log in to report

Role Summary

AI-generated

This software engineer builds and operates a platform that helps Apple teams generate synthetic datasets for machine learning training and evaluation. The role combines hands-on platform engineering with collaboration to help partner teams use generated data and assess its suitability.

What You'll Do

  • Partner with ML engineers, data scientists, and designers to translate their needs into platform features and support their dataset...
  • Help design model ablations that measure how synthetic data affects training and evaluation.
  • Assess differences between synthetic and real data, including visual properties, metadata, distributions, and labels, and build...
  • Help teams evaluate generated output and understand its quality, variability, cost, speed, and feasibility.
  • Integrate image and video generation models into a multi-stage production pipeline.
  • Operate and debug GPU workloads to keep generation reliable and cost-effective.

Generated from the employer's posting. Verify important details before applying.

View full posting

Qualifications

Required: A bachelor's degree in Computer Science, Computer Engineering, or a related field, or 3 years of equivalent work experience; 3+ years building and operating production software in Python; experience debugging production distributed systems or data pipelines using logs, metrics, and task state; experience deploying machine learning models on GPU infrastructure, including dependency management and GPU memory sizing; experience designing input validation and configuration contracts; ability to explain machine learning concepts and tradeoffs to non-technical partners; and hands-on daily use of agentic AI coding tools, including judging their effectiveness and verifying their output.

Preferred

  • Experience working directly with partner teams or internal customers to define and ship features.
  • Experience measuring the domain gap between synthetic and real data, including visual statistics and non-visual properties such as...
  • Experience with generative image, video, or large language model APIs, including structured output.
  • Familiarity with image and video file formats and metadata standards.
  • Experience with batch compute platforms or job schedulers.

Original job description

Content provided by the employer

Summary

Apple's machine learning features are only as good as the data behind them, and our team builds the platform that lets teams create data that fills the gaps real-world collection can't, provides privacy-preserving ways to train and evaluate our models safely, and helps us ensure our products have seen diverse inputs to generalize properly across the many parts of the world where we operate.

In the AI & Machine Learning (AIML) organization, our team builds the self-service tools and platform that turn ideas, text, and images into large-scale generated datasets. We don't produce the datasets ourselves: ML engineers, data scientists, designers, feature teams, ML data operations teams, and Safety and Responsible AI teams across Apple use our platform to generate their own. That data is only valuable if it matches the real world closely enough to train on and to evaluate against, proving our models and features meet Apple's quality bar before they reach customers. As a software engineer on the platform, you'll sit between that technology and the people who depend on it: understanding what they want to create, building the features that make it possible, and helping them get the most from generative AI. It's a hands-on engineering role for someone who enjoys the people side as much as the code.

Description

You'll spend your days moving between engineering and partnership. Synthetic data here does far more than fill gaps: it lets teams build privacy-preserving digital humans, represent locales and domains that real-world collection underserves, and experiment in days instead of waiting on slow, costly data collection. It also gives Safety and Responsible AI red teams the data they need to stress-test models against misuse and edge cases. One day you might be integrating a new image or video generation model and tuning how it runs on GPUs; the next, you might be helping a design team get their first dataset running on the platform, or working with a feature team on model ablations that measure what synthetic data adds to training and evaluation.

You'll help teams understand where synthetic data differs from the real data their models see, whether that's how images look or subtler differences in metadata, distributions, and labels, and then build what closes those gaps. Your core partners are ML engineers and data scientists, but you'll also work with designers, artists, and others who are newer to machine learning, and help make these systems understandable and approachable for them. You'll own features end to end, from shaping the request with partner teams to validating the result on production infrastructure.

Our team builds with AI coding agents every day, and you'll help shape how we use them well. You don't need a research background in generative models; you'll learn the models on the job. What matters most is strong engineering judgment and the ability to bring people along.

Responsibilities

Partner with ML engineers, data scientists, and designers to turn their needs into platform features.

Support teams as they build their own datasets on the platform, and help design model ablations that measure how synthetic data affects training and evaluation.

Explain how ML, vision, and generative models behave to people new to them, including cost, quality, speed, and feasibility.

Assess the domain gap between synthetic and real data, covering visual properties and non-visual ones such as metadata, value distributions, and labels.

Work with feature teams to find where synthetic data falls short for their use case, and build platform capabilities to close those gaps.

Help teams judge generated output, understand why results vary, and get closer to what they want.

Integrate new image and video generation models into a multi-stage production pipeline.

Operate and debug GPU workloads to keep generation reliable and cost-effective.

Minimum Qualifications

Bachelor's degree in Computer Science, Computer Engineering, or a related field, or 3 years of equivalent work experience.

3+ years of experience building and operating production software in Python.

Experience debugging distributed systems or data pipelines in production, using logs, metrics, and task state to find root causes.

Experience deploying machine learning models on GPU infrastructure, including dependency management and GPU memory sizing.

Experience designing input validation and configuration contracts for systems where late failures are costly.

Demonstrated ability to explain machine learning concepts and tradeoffs to non-technical partners such as designers or product managers.

Proven, hands-on use of agentic AI coding tools in day-to-day software development, including judging where they are effective and where their output needs human verification.

Preferred Qualifications

Experience working directly with partner teams or internal customers to define and ship features.

Experience measuring the domain gap between synthetic and real data, including visual statistics and non-visual properties such as metadata and label distributions.

Experience with generative image, video, or large language model APIs, including structured output.

Familiarity with image and video file formats and metadata standards.

Experience with batch compute platforms or job schedulers.

Pay & Benefits — San Francisco, California, United States

At Apple, base pay is one part of our total compensation package and is determined within a range. This provides the opportunity to progress as you grow and develop within a role. The base pay range for this role is between $150,400 and $225,300, and your base pay will depend on your skills, qualifications, experience, and location.

Apple employees also have the opportunity to become an Apple shareholder through participation in Apple’s discretionary employee stock programs. Apple employees are eligible for discretionary restricted stock unit awards, and can purchase Apple stock at a discount if voluntarily participating in Apple’s Employee Stock Purchase Plan. You’ll also receive benefits including: Comprehensive medical and dental coverage, retirement benefits, a range of discounted products and free services, and for formal education related to advancing your career at Apple, reimbursement for certain educational expenses — including tuition. Additionally, this role might be eligible for discretionary bonuses or commission payments as well as relocation. Learn more about Apple Benefits

Note: Apple benefit, compensation and employee stock programs are subject to eligibility requirements and other terms of the applicable plan or program.

Application Deadline

Apple accepts applications to this posting on an ongoing basis.

Apple

About the company

Apple

Large Enterprise

Apple Inc. is a global technology company known for its innovative products and services, including the iPhone, iPad, Mac computers, and Apple Watch. Founded in 1976, Apple has continuously pushed the boundaries of design and functionality, earning a reputation for high-quality consumer electronics and software solutions like iOS and macOS. With a strong commitment to user experience and privacy, Apple also leads in digital services, offering platforms such as the App Store, Apple Music, and iCloud. The company's focus on sustainability and corporate responsibility further enhances its standing as a leader in the technology sector.