DevOps Engineer – StackOps (Incident Management)

Rapsys Technologies Pte. Ltd

Singapore, SGonsite$52k-$140k/yrPosted Jul 6, 2026
Posting intelligenceActively listedReposted 21×, possible evergreen/ghost posting

Skills

elasticsearchterraformpagerdutyansibledatadognodepulumipythonazurejiracicdawsjavascriptgo

About the role

We are seeking a highly technical DevOps Engineer to join our Central Enablement

Team. In this role, you will be the core builder of our enterprise "Resiliency-as-a-

Service" platform.

Operating within the Resiliency Programme, your mission is to design, code, and deploy

the automated infrastructure, toolchains, and workflows that Whole-of-Government

(WOG) engineering teams rely on to manage incidents and minimize downtime. You will

write the integrations that connect disparate observability tools into incident

management product & centralized intelligence hub and build the automation that

accelerates recovery across the organization.

Key Responsibilities

1. Engineering "Resiliency-as-Code"

Build, maintain, and scale using Infrastructure-as-Code (Pulumi/Terraform) to

enable teams to deploy standardized monitoring & incident workflows with a

single click.

Develop self-service onboarding portals and APIs that allow distributed

engineering teams to easily hook their applications into the central resiliency

framework.

2. Building Incident Management Integrations

Develop and maintain robust API integrations and webhooks between

specialized observability platforms (Elasticsearch, AWS/Azure Cloud-native

tools, Dynatrace) and our central IT Service Management system (Jira Service

Management - JSM).

Code the automation that routes alerts, enriches payloads with dependency

metadata, and triggers specific JSM workflows without manual human

intervention.

Develop and manage automations tools – enabling systems integrations into the

central Elastic Intelligence hub.

3. AI Enablement & Auto-Remediation

Implement and configure AIOps capabilities within the observability pipeline to

assist with Root Cause Analysis (RCA) and anomaly detection.

Write complex automation scripts (Python, Go, or Node.js) to execute automated

runbooks and self-healing tasks (auto-remediation) triggered by specific alerts or

AI outputs.

Integrate AI-driven "Incident Scribe" features into JSM to automatically

summarize incident timelines and metrics for Post-Incident Reviews (PIRs).

4. Platform Reliability & CI/CD

Ensure the high availability and performance of the central observability and

incident management pipelines.

Build and maintain CI/CD pipelines to test and deploy configuration changes to

the StackOps architecture seamlessly.

Work closely with the Product Lead to translate resiliency strategies into

scalable, technical deliverables.

Provide consultation and support to platform users, solution and enhance

platform offerings based on user challenges.

Qualifications & Requirements

Experience: 3-5+ years in DevOps, Software Engineering, or Site Reliability

Engineering (SRE), ideally within a central platform team or Internal Developer

Platform (IDP) environment.

Programming & Automation: Strong coding skills in Python, Go, or Node.js,

specifically for writing automation scripts, interacting with REST APIs, and

building custom integrations.

Infrastructure-as-Code: Deep, hands-on experience with Terraform, Pulumi,

Ansible, or similar IaC tools within large-scale AWS or Azure environments.

Observability Tooling: Practical experience configuring and managing telemetry

pipelines using OpenTelemetry, Elastic Stack, Dynatrace, or Datadog.

ITSM Integration: Experience working with the APIs of Jira Service Management

(JSM), ServiceNow, or PagerDuty to build automated alerting and ticketing

workflows.

Pay: $4,328.88 - $11,682.80 per month

Work Location: In person

Compensation

This DevOps / SRE role pays $52k-$140k/yr. Within typical range for devops / sre roles in Singapore.

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in Singapore varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for Singapore medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.