Enter a job title or keyword

Site Reliability Engineer, Apple Ads

Apple


Job Location:

Cupertino, CA - USA

Monthly Salary: Not provided by the employer
Posted: 29 September 2026 (Yesterday)
Application Deadline: 27 December 2026
Vacancies: 1 Vacancy

Job Summary

At Apple we focus deeply on our customers experience. Apple Ads brings this same approach to advertising helping people find exactly what theyre looking for and helping advertisers grow their businesses. nnOur technology powers ads and sponsorships across Apple Services including the App Store Maps Apple News and MLS Season Pass. Everything we do is designed for trust connection and impact: We respect user privacy integrate advertising thoughtfully into the experience and deliver value for advertisers of all sizesfrom small app developers to big global brands. Because when advertising is done right it benefits everyone.

As a Site Reliability Engineer you will be responsible for providing the platform for mission-critical ad-tech systems to maintain constant uptime scale seamlessly and allow for new applications and services to successful candidate will be highly self-motivated and passionate about excellence quality and detail. The SRE will not only support operations but also work closely with the developers and architects within the team to aid in the design and assist with the implementation to improve stability security and scalability.

Build and operate distributed systems using AWS managed services such as EKS MSK and ElastiCache. nDevelop internal tooling and automation frameworks to improve infrastructure reliability cost-efficiency and operational visibility. nCollaborate with engineering teams to define infrastructure architecture troubleshoot complex issues and drive production provision and maintain resilient infrastructure by combining Terraform for core foundational resources with GitOps workflows and Kubernetes-native control planes (e.g. Argo CD Flux Helm kro ACK Crossplane etc.) to enable declarative automated and drift-free continuous or participate in incident response postmortems and continuous improvement cycles to reduce future risk.

3 years of experience supporting internet-facing production systems and distributed cloud programming skills in at least one of: Python Go or expertise with AWS-managed infrastructurenHands-on experience with Linux systems and deep knowledge of its experience with Infrastructure as Code especially foundation in SRE concepts: Monitoring alerting and observability incident response and root cause analysis error budgets SLAs/SLOs and system reliability

Demonstrated experience designing building or integrating AI/LLM-powered tooling and automationsnPassion for customer privacynBuilt tools or services that automate platform operations reduce toil or improve cost managing Kubernetes clusters at scale in production -on experience troubleshooting distributed systems under real-world building and operating infrastructure at building solutions that reduces friction in software communication skills and comfort collaborating across engineering infrastructure and product certifications or broad experience across multiple AWS services is a plus.

Required Experience:

IC


About Company

Company Logo

Ask Siri to name the most successful company in the world and it might respond: Apple. And it's not just out of familial pride. Apple consistently ranks highly in profit, revenue, market capitalization, and consumer cachet. In 2018, the company became the first reach a trillion dollar ... View more

View Profile View Profile