Senior Site Reliability Engineer (SRE)

Vanguard Software

Singapore, SGonsitePosted Jul 1, 2026
Posting intelligenceActively listedReposted 25×, possible evergreen/ghost posting

Skills

kubernetesprometheusterraformjenkinsansiblegrafanadatadogdockergithubgitlabpythonazurehelmcicdgooglecloudawsgo

About the role

Job Summary

We are seeking a Senior Site Reliability Engineer (SRE) to join our growing engineering team. In this role, you will work independently to design, build, and optimize infrastructure and deployment pipelines that ensure the stability, scalability, and security of our systems. You will take full responsibility for automating workflows, improving observability, and enabling development teams to ship code faster and safer. This is an excellent opportunity for an experienced engineer with at least 5 years of work experience who thrives on ownership, reliability, and technical leadership.

Key Responsibilities

Infrastructure & Automation: Design, implement, and maintain scalable cloud infrastructure using Infrastructure as Code (IaC) tools.

CI/CD Pipelines: Build and optimize automated pipelines for testing, deployment, and release management.

Monitoring & Reliability: Establish observability standards, implement monitoring, logging, and alerting systems to ensure system health.

Security & Compliance: Enforce best practices for cloud security, access control, and compliance across environments.

Collaboration: Partner with backend, frontend, and product teams to ensure smooth deployments and reliable system operations.

Process & Mentorship: Improve DevOps processes, share best practices, and mentor junior engineers.

Job Requirements

Bachelor's Degree of Computing, Software Engineering, IT or related field.

Experience: Minimum 5 years of DevOps, Site Reliability Engineering (SRE), or related experience.

Tech Stack: Proficient with cloud platforms (AWS, GCP, or Azure), containerization (Docker, Kubernetes), IaC (Terraform, Ansible, Helm), and CI/CD tools (Jenkins, GitHub Actions, GitLab CI/CD, ArgoCD, etc.).

Systems Knowledge: Strong background in Linux administration, networking, and distributed systems.

Monitoring & Observability: Hands-on experience with tools like Prometheus, Grafana, ELK/EFK, or Datadog.

Scripting & Automation: Proficient in one or more languages (Python, Go, Bash, etc.).

Problem Solving: Skilled at diagnosing complex issues, ensuring high availability, and improving system performance.

System Design: Capable of designing fault-tolerant, secure, and scalable infrastructure with disaster recovery in mind.

Soft Skills

Team Mindset: Collaborate effectively across teams, proactively contributing to company goals.

Ownership: Take responsibility for infrastructure health and ensure continuous improvements.

Adaptability: Open to new technologies, evolving processes, and changing business needs.

Communication: Clearly explain technical topics to both engineers and non-technical stakeholders.

What We Offer

Technical Leadership Opportunities: Lead infrastructure design for high-impact projects and guide DevOps best practices.

Continuous Growth: Access to mentorship, certifications, and a clear career progression path.

High-Performance Collaboration: Work with a talented team in a modern DevOps environment (Agile/CI-CD, GitOps).

Flexibility and Trust: An open culture that values innovation, autonomy, and results-driven decision-making.

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in Singapore varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for Singapore medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.