Site Reliability Engineer - ML, Apple Ads
Apple · New York City ·
- Category
- SRE
- Experience
- 3+ years
- Company size
- 1000+
Apple · New York City ·
Cisco · Bangalore
draftkings · Boston, MA
PayPal · Chennai, Tamil Nadu, India
The Federal Reserve System · San Francisco, CA
Site Reliability Engineer on Apple Ads' ML platform team, owning the health, performance, and scalability of large-scale AWS infrastructure for ML training, inference, and serving workloads. Day-to-day centers on building automation and platform tooling (Terraform, Kubernetes, Linux) rather than CI/CD pipeline configuration.
At Apple, we focus deeply on our customers’ experience. Apple Ads brings this same approach to advertising, helping people find exactly what they’re looking for and helping advertisers grow their businesses.
Our technology powers ads and sponsorships across Apple Services, including the App Store, Apple News, and MLS Season Pass. Everything we do is designed for trust, connection, and impact: We respect user privacy, integrate advertising thoughtfully into the experience, and deliver value for advertisers of all sizes—from small app developers to big, global brands. Because when advertising is done right, it benefits everyone.
The Site Reliability Engineering team within Apple Ads ensures the reliability, performance, and availability of ML Platform and Services at scale. The team partners closely with Ads engineering, data science and ML platform teams to enable product delivery through design, configuration, and automation of machine learning infrastructure powering Apple Ads applications.
We are looking for a ML Platform Infrastructure Engineer to help build and evolve the next generation of Apple Ads machine learning platform — enabling fast, reliable, and scalable operations across AWS-based environments supporting transactional and analytical workloads.
As a site reliability engineer in Apple Ads focused on machine learning, you will own the health, performance, and scalability of large scale infrastructure powering ML training, inference, serving workloads and associated platform tooling. Your focus will be on building automation that eliminates manual processes, improves platform resilience, and enables teams to move faster with confidence.
This is not a DevOps-only or CI/CD-focused role. We are looking for engineers who build platform solutions, not just configure pipelines.
1global · Berlin, Berlin, Germany