Site Reliability Engineer

Darktrace

Cambridge, UKonsitePosted Jun 16, 2026
Posting intelligenceActively listedReposted 14×, possible evergreen/ghost posting

Skills

kubernetespythonazuregooglecloudawsgo

About the role

Darktrace is a global leader in AI for cybersecurity that keeps organizations ahead of the changing threat landscape every day. Founded in 2013, Darktrace provides the essential cybersecurity platform protecting nearly 10,000 organizations from unknown threats using its proprietary AI.

The Darktrace Active AI Security Platform™ delivers a proactive approach to cyber resilience to secure the business across the entire digital estate – from network to cloud to email. Breakthrough innovations from our R&D teams have resulted in over 200 patent applications filed. Darktrace’s platform and services are supported by over 2,400 employees around the world. To learn more, visit http://www.darktrace.com.

Job Description:

About The Role

We’re looking for a Site Reliability Engineer (SRE) to bring deep expertise in a key reliability domain and help shape the future of our platform reliability strategy.

SRE sits at the heart of our operational trifecta alongside Platform Engineering and DevSecOps. In this role, you’ll act as the go-to authority in your area of specialism, working across teams to embed best practices, solve complex reliability challenges, and improve system resilience at scale.

Unlike a generalist SRE, this role focuses on a core domain of expertise—such as observability, performance engineering, data infrastructure reliability, security-focused SRE, or network reliability—while influencing reliability standards across the wider engineering organisation.

Key Responsibilities

Domain Expertise & Strategy

Act as the subject matter expert in your chosen reliability domain

Define and implement standards, frameworks, and best practices across SRE, Platform Engineering, and DevSecOps

Stay current with industry trends and bring innovative ideas into the organisation

Engineering & Delivery

Design and implement solutions to complex, cross-cutting reliability challenges

Build tooling, automation, and frameworks to improve system resilience and scalability

Lead deep-dive investigations into systemic issues and drive long-term fixes

Collaboration & Platform Integration

Partner with Platform Engineering to ensure your domain is embedded within the internal developer platform

Collaborate with DevSecOps to integrate security, compliance, and resilience practices

Contribute to cross-team initiatives that improve reliability across the stack

Incident & Operational Excellence

Play a key role in incident response, particularly within your specialism

Contribute to on-call rotations and continuous improvement of operational processes

Develop runbooks, documentation, and training materials to support teams

Essential

What You’ll Bring

Proven experience in Site Reliability Engineering, DevOps, or infrastructure engineering

Deep expertise in at least one of the following areas:

Observability & monitoring (metrics, logging, distributed tracing)

Performance engineering & capacity planning

Data infrastructure reliability (databases, streaming, pipelines)

Security-focused SRE (hardening, compliance automation, secrets management)

Network reliability & traffic management

Strong programming skills (e.g. Go, Python, or similar)

Experience with cloud platforms (AWS, GCP, Azure) and Kubernetes

Strong communication skills, with the ability to explain complex technical concepts clearly

Self-driven with the ability to identify and prioritise high-impact work independently

Desirable

Experience building internal developer platforms or tooling

Contributions to open-source, technical blogs, or public speaking

Experience working in regulated environments

Familiarity with SLO frameworks and error budget management

Relevant certifications in your specialist domain

Success Measures

Improved reliability and performance within your domain of specialism

Adoption of best practices across SRE, Platform Engineering, and DevSecOps

Reduction in incidents and faster resolution times

Scalable, well-integrated solutions within the internal platform

Strong collaboration across teams and measurable improvements in operational maturity

Why Join Us?

Shape reliability strategy in a modern, cloud-native engineering environment

Work on complex, high-impact systems at scale

Collaborate with expert teams across Platform Engineering and DevSecOps

Take ownership of a domain and drive meaningful, organisation-wide impact

Benefits:

23 days’ holiday + all public holidays, rising to 25 days after 2 years of service,

Additional day off for your birthday,

Private medical insurance which covers you, your cohabiting partner and children,

Life insurance of 4 times your base salary,

Salary sacrifice pension scheme,

Enhanced family leave,

Confidential Employee Assistance Program,

Cycle to work scheme.

Darktrace is committed to providing reasonable accommodations to qualified individuals with disabilities in accordance with applicable laws. If you require a reasonable accommodation to participate in the application or interview process, please contact your Talent Partner.

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in United Kingdom varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for United Kingdom medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.