Skills
Site Reliability Engineering
Cloud Infrastructure
Datacenter Operations
Infrastructure Scaling
Private Cloud Compute
AI Services
Automation
Monitoring & Incident Management
Linux/Unix Systems
Scripting (Python, Go, etc.)
Infrastructure as Code (IaC)
Cloud infrastructure
Apple Silicon
Datacenter hardware/software
Kubernetes
Linux
Automation tools (e.g., Ansible, Terraform)
Summary
We are seeking a Site Reliability Engineer to design, build, and operate the cloud infrastructure that powers a global services platform. You will work on scaling Apple Silicon systems in the datacenter, supporting privacy‑focused AI workloads, and advancing the reliability and performance of services that serve billions of users. The role involves high responsibility, influencing the core platform direction, and collaborating with a team that values innovation and excellence.
Highlights
Opportunity to shape the core of Apple's global cloud platform, work on cutting‑edge AI and privacy technologies, and have significant impact on services used by billions of users while collaborating with a high‑performing engineering team.
Description
Summary
Become a Site Reliability Engineer in Apple’s Cloud Service Infrastructure team, part of Apple’s Services Engineering organisation, and help scale the cloud that underpins services for billions of Apple users.
We are building and supporting new and existing infrastructure to support the hyperscaling of Apple Silicon systems in the datacenter.
This allows Apple to provide best in class privacy and power efficiency for AI users (as part of our Private Cloud Compute service) and for all Apple Services.
We are at the cutting edge of Apple’s cloud hardware and software infrastructure, moving the dial so that Apple can provide new and exciting services to our end users.
We help Apple surprise and delight our users.
Description
The Apple Services Engineering Cloud Services SRE organisation is looking for a strong, enthusiastic SRE to join our team in Dublin.
This person will have a tremendous amount of individual responsibility and influence over the direction the core platform of many critical Apple internet services takes for years to come.
You are someone with ideas and real passion for software delivered as a service to improve reuse, efficiency, and simplicity.
This engineer’s work will impact billions of users and be essential to the success of some of the most visible current and future Apple features.
We are domain experts in fleet management, systems, and software engineering.
We build automation and reliability tools to scale the systems reliably.
We respond to alerts and incidents which may pose a risk to the reliability of the platform and we learn from them to improve the future performance of the services.
The team’s focus is on infrastructure capabilities and processes, improving the reliability and efficiency of the systems, at scale.
We have a range of expertise in the team across hardware, networking, distributed systems, reliability, processes, operating systems, software development.
We need people who can bring their own expertise to bear and are happy to teach and learn as we grow the service.
Minimum Qualifications
Strong emphasis on SRE as an engineering subject area, with proficiency in at least one of the following languages (Go, Rust, Python, Swift).A successful track record and proven experience as a backend internet services software developer.Knowledge of the software development life-cycle; including continuous integration, testing methodologies, TDD, and agile development methodologies.Understanding of foundational internet infrastructure services including DNS, DHCP, virtualisation, and monitoring.Experience operating critical, large-scale distributed systems spanning hardware, operating systems, and software.Understanding of SRE principles, including observability, alerting, error budgets, fault analysis, and other common reliability engineering concepts, with a keen eye for opportunities to eliminate toil by code and process improvements.Bachelor’s or Master’s in Computer Science, Computer Engineering, or equivalent experience.
Preferred Qualifications
Experience with large scale server provisioning and maintenance (OpenStack Ironic, Metal3, MAAS, xCat, Netbox, Tinkerbell).Experience with development within the Kubernetes ecosystem, including operator framework, controllers, and CRDs.Experience with UI frameworks such as React or Angular.Hardware bootstrap and associated security (PXE, BIOS, TPM, secure boot, trusted computing).Structured or unstructured storage and caching.Automating operations processes via services and tools.Configuration management and fleet orchestration via Puppet, Chef, Ansible, or othersCloud Services (AWS S3/EC2/CloudFront or equivalent).
At Apple, we believe accessibility is a fundamental human right.
You’ll find that idea reflected in everything here — in our culture, our benefits and our digital tools.
By welcoming as many perspectives as possible, we help you build a career where you feel like you belong.
Learn about accessibility in Apple’s workplace
Role Number: 200674537-0562