Site Reliability Engineer, VP
Skills
About the role
Closing date for applications: 15/07/2026
Location Gurugram, India
Job typePermanent | Contract typeFull Time
#R-00281615
Join our digital revolution in NatWest Digital X
In everything we do, we work to one aim. To make digital experiences which are effortless and secure.
So we organise ourselves around three principles: engineer, protect, and operate. We engineer simple solutions, we protect our customers, and we operate smarter.
Job description
This role is based in India and as such all normal working days must be carried out in India.
Join us as a Site Reliability Engineer
In this key role, you’ll improve, drive, and embed non-functional and operational characteristics such as availability, performance, efficiency, change management, monitoring, security, incident response, and capacity planning of our products and services
You’ll enjoy significant stakeholder interaction, working in collaboration with engineers to ensure a principled approach to deliver change in a safe and secure way
This is a chance to join an inclusive team with a collaborative ethos and a commitment to innovation and professional development
We're offering this role at vice president level
What you'll do
As our Site Reliability Engineer, you’ll work closely with our feature team and other colleagues to meet defined service level objectives and continually improve systems and environments. You’ll define error budgets that support finding the right balance between risk and reliability.
You’ll also provide structure and help to our release process, suggesting and making improvements where possible. You’ll scale systems sustainably through mechanisms like automation, evolving them by pushing for changes that improve reliability and velocity. We’ll also look to you to coach and provide guidance to colleagues and the wider team, leading where required.
In addition to this, you’ll:
Proactively contribute new ideas and innovations to meet short term and longer-term goals
Continually balance and manage any potential risks
Be accountable for the day-to-day health of both production and non-production environments and respond to any incidents as required
Provide technical expertise and input to establish the risk tolerance of products and services
Communicate incident status updates clearly and frequently to other teams, customers and stakeholders
The skills you'll need
We’re looking for someone with at least 12 years of experience of reliability systems thinking and experience of software engineering. You’ll need experience of using a data driven and scientific approach to fact finding. We’ll also look for financial services knowledge, and the ability to identify wider business impact, risk and opportunity, and make connections across key outputs and processes
We’re also looking for:
Good knowledge and experience of programming languages
Strong knowledge of deploy and release services, automation, and troubleshooting
Experience of utilising tools and technology across the software development lifecycle
Experience using mathematical and statistical models to assess trends
Strong communication skills with the ability to proactively engage with a wide range of stakeholders
Welcome to our Gurugram hub
Spanning 437,000 sq. ft., our campus in Gurugram features two state-of-the-art towers – 1A and 2A at the Candor TechSpace in Sector 21.
Key facts:
Surrounded by 28 acres
Space for 4,100 colleagues
Opened in 2010
Our tech stack
Here’s just some of the technologies we use.
Front end
JavaScript
ReactJS
AngularJS
Back end
Python
Java
Microsoft Dynamics
DevOps
AWS
GitLab
Google Cloud Platform
Data
Kafka
Hadoop
PostgreSQL
Snowflake
Questions about this role
Want AI Applyd to auto-apply to roles like this?
We tailor your resume per posting, fill the forms, and track replies for you.