Senior Site Reliability Engineer

Apple

Singapore, SGonsitePosted Jul 27, 2026
Posting intelligenceActively listed

Skills

kubernetesbootstrapansibleangularpythonreactswiftrustawsgo

About the role

Become a Site Reliability Engineer in Apple’s Cloud Service Infrastructure team, part of Apple’s Services Engineering organization, and help scale the cloud that underpins services for billions of Apple users.

We are building and supporting new and existing infrastructure to support the hyperscaling of Apple Silicon systems in the datacenter. This allows Apple to provide best in class privacy and power efficiency for AI users (as part of our Private Cloud Compute service) and for all Apple Services. We are at the cutting edge of Apple’s cloud hardware and software infrastructure, moving the dial so that Apple can provide new and exciting services to our end users. We help Apple surprise and delight our users.

Description

The Apple Services Engineering Cloud Services SRE organization is looking for a strong, enthusiastic SRE to join our team in Singapore. This person will have a tremendous amount of individual responsibility and influence over the direction the core platform of many critical Apple internet services takes for years to come. You are someone with ideas and real passion for software delivered as a service to improve reuse, efficiency, and simplicity. This engineer’s work will impact billions of users and be essential to the success of some of the most visible current and future Apple features.

We are domain experts in fleet management, systems, and software engineering. We build automation, instrumentation and tools to scale the systems reliably. We respond to alerts and incidents which may pose a risk to the reliability of the platform and we learn from them to improve the future performance of the services. The team’s focus is on infrastructure capabilities and processes, improving the reliability and efficiency of the systems, at scale.

We have a range of expertise in the team across hardware, networking, distributed systems, reliability, processes, operating systems, software development. We need people who can bring their own expertise to bear and are happy to teach and learn as we grow the service.

Desired Skills:

Experience with large scale server provisioning and maintenance (OpenStack Ironic, Metal3, MAAS, xCat, Netbox, Tinkerbell)

Experience with development within Kubernetes ecosystem, including operator framework, controllers and CRDs

Experience with UI frameworks such as React or Angular

Some exposure to the following:

Hardware bootstrap and associated security (PXE, BIOS, TPM, secure boot, trusted computing)

Structured or unstructured storage and caching

Automating operations processes via services and tools

Configuration management and fleet orchestration via Puppet, Chef, Ansible, or others

Cloud Services (AWS S3/EC2/CloudFront or equivalent)

Preferred Qualifications

Hardware bootstrap and associated security (PXE, BIOS, TPM, secure boot, trusted computing)

Structured or unstructured storage and caching

Automating operations processes via services and tools

Configuration management and fleet orchestration via Puppet, Chef, Ansible, or others

Cloud Services (AWS S3/EC2/CloudFront or equivalent)

Minimum Qualifications

Bachelor’s or Master’s in Computer Science, Computer Engineering, or equivalent experience.

Strong emphasis on SRE as an engineering subject area, with proficiency in at least one of the following languages (Go, Rust, Python, Swift)

A successful track record and proven experience as a backend internet services software developer

Knowledge of the software development lifecycle, including continuous integration, testing methodologies, TDD and agile development methodologies

Understanding of foundational internet infrastructure services including DNS, DHCP, virtualization and monitoring

Experience operating critical, large-scale distributed systems spanning hardware, operating systems, and software

Understanding of SRE principles, including observability, alerting, error budgets, fault analysis, and other common reliability engineering concepts, with a keen eye for opportunities to eliminate toil by code and process improvements

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in Singapore varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for Singapore medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.