Site Reliability Engineer

Reapit ANZ

unknownPosted Jun 25, 2026
Posting intelligenceActively listedReposted 17×, possible evergreen/ghost posting

Skills

cloudformationterraformawsgo

About the role

Reapit – Who are we?

Reapit is the original, end-to-end business technology provider for estate agencies of all sizes. We’ve been helping sales and lettings agents to build relationships and grow their businesses for more than 25 years. Our technology connects property professionals in Europe, the Middle East, Australia, and New Zealand with buyers, sellers, tenants and landlords to power the relationships that change lives.

In Australia, Reapit stands as the preferred technology choice among the nation's leading estate agents and agencies. Tailored to the unique demands of the Australian property market, Reapit provides successful leaders with unparalleled tools across sales, property management, client relations, and data analytics, reinforcing their position at the pinnacle of real estate excellence.

What you’ll be doing

Reporting to the ANZ DevOps Lead you’ll be involved in:

Deploy, and maintain robust, scalable AWS infrastructure utilizing Infrastructure as Code (IaC) principles (e.g., CloudFormation, Terraform).

Implement, and maintain comprehensive monitoring, logging, and alerting solutions to ensure system health and performance.

Respond promptly and effectively to critical system alerts and incidents, performing root cause analysis (RCA) and implementing preventative measures.

Manage and execute scheduled maintenance windows, coordinating necessary system updates, patching, and upgrades with minimal downtime.

Provide out-of-hours on-call support for major incidents and escalations to restore critical service functionality quickly and efficiently.

Automate repetitive operational tasks ("toil") to increase system efficiency, reduce manual effort, and free up engineering time.

Drive continuous improvement in system reliability, performance, and recoverability (Disaster Recovery/Business Continuity Planning).

Collaborate closely with development teams (DevOps) to improve the entire software lifecycle, focusing on service stability and release engineering.

Establish and refine Service Level Indicators (SLIs), Service Level Objectives (SLOs), and Service Level Agreements (SLAs) for critical services.

Conduct capacity planning and performance testing to ensure the AWS environment can handle current and future load requirements.

Who we're looking for

At Reapit, we prioritise hiring individuals who share our values and possess the right attitudes and behaviours for success. Whilst some of the listed requirements may be important, don’t worry if you don’t meet all of them, we’d still like to hear from you.

Technically excellent engineer - You have deep hands-on expertise with cloud infrastructure, automation, and the entire DevOps toolchain, and you're not afraid to get into the weeds to solve complex technical challenges

Genuine team player - You believe the team's success is your success, actively collaborate across disciplines, and contribute to a positive team culture where everyone can do their best work

Exceptional communicator - You can articulate technical concepts clearly to both engineers and non-technical stakeholders, actively listen to understand problems, and document your work thoroughly

Passionate about the craft - You genuinely love building reliable systems, take pride in clean infrastructure code, and stay energized by the challenge of improving operational excellence

Ownership-driven professional - You take full accountability for your systems, follow through on commitments, and don't need to be asked twice to see things through to completion

Collaborative problem-solver - You seek input from others, build consensus around solutions, and approach technical disagreements with curiosity rather than ego

Continuous learner who shares knowledge - You actively develop your skills, stay current with industry trends, and generously share what you learn through documentation, mentoring, and pairing sessions

Calm and decisive under pressure - You maintain composure during production incidents, make sound decisions with incomplete information, and help keep the team focused on resolution rather than blame

What your impact and success looks like

We expect your success and impact in the early stages of your career with us to look something like this:

Within 1 Month

Complete onboarding with demonstrated understanding of our key systems, infrastructure architecture, deployment processes, and on-call procedures

Successfully respond to and resolve production incidents independently for at least three core services, creating or updating runbooks based on your experience

Become the go to person for our main Australian system

Within 3 Months

Own incident response end-to-end for your assigned systems, including leading post-mortems and driving remediation work to completion

Deliver at least two substantial automation or tooling improvements that measurably reduce operational overhead, improve deployment speed, or enhance system reliability

Implement monitoring and alerting enhancements for critical services that improve observability, reduce alert fatigue, or decrease mean time to detection

Within 6 Months

Deliver a major operations project to production (such as: delivery of a new product/system, infrastructure migration, disaster recovery implementation)

Demonstrate quantifiable improvements in reliability metrics for systems under your ownership

Establish yourself as an expert in your assigned systems, helping team members, influencing operational decisions, and contributing to technical improvements

We operate a Flexible Working Policy and we would like for you to work from our Brisbane or Sydney office as required.

Don't tick all the boxes? Neither do we

“We are a 2025 Circle Back Initiative Employer – we commit to respond to every applicant.”

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation varies by seniority, employer size, and location. When this listing publishes a salary band you'll see it in the badge row above the description.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.