Senior Software Engineer - SRE

AirAsia

unknownPosted Jul 14, 2026
Posting intelligenceActively listed

Skills

kubernetesterraformcicdgooglecloud

About the role

Job Description

Job Title: Senior Site Reliability Engineer Location: Kuala Lumpur

About AirAsia MOVE AirAsia MOVE is a leading ASEAN-focused budget travel OTA, part of the Capital A Group. We deliver customer-centric travel solutions by combining innovation with operational excellence. Our goal is to create seamless, reliable, and delightful journeys for travelers across the region.

About the Role

We’re looking for a Senior Site Reliability Engineer to help scale and stabilize our cloud infrastructure and reliability practices as we grow across multiple lines of business.

You’ll lead key initiatives around:

Cloud architecture modernization

Multi-region reliability

Observability and incident response

Reducing toil through automation and self-service

This is a hands-on technical role, where you’ll work across platform, SRE, and application teams to build scalable systems that are resilient, cost-aware, and developer-friendly.

What You’ll Do

Design and implement secure, scalable infrastructure on Google Cloud Platform (GCP)

Lead efforts to build and evolve MOVE’s GCP Landing Zone, including Shared VPC, org structure, IAM, and policy guardrails

Build and improve multi-region architectures for high availability and disaster recovery

Drive infrastructure automation using Terraform, CI/CD, and GitOps practices

Improve observability across teams by standardizing monitoring, tracing, and alerting

Collaborate on incident response and postmortems to reduce MTTR and build resilience

Enforce tagging, FinOps controls, and security policies across GCP projects

Contribute to platform engineering initiatives and developer self-service tools

What We’re Looking For

5+ years in SRE, DevOps, or cloud infrastructure roles

Solid experience with GCP, Terraform, Kubernetes (GKE), or similar cloud providers

Strong hands-on experience in automation and multi-region architecture design

Experience in networking (VPCs, NAT, PSC), IAM, and cloud-native security

Proven ability to debug and support production systems under pressure

Familiarity with monitoring and tracing tools like Cloud Monitoring, OpenTelemetry, Signoz

Exposure to using AI/anomaly detection for alert tuning or reliability insights

Clear communicator who works well with developers, product, and other infra teams

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation varies by seniority, employer size, and location. When this listing publishes a salary band you'll see it in the badge row above the description.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.