Site Reliability Engineer
Skills
About the role
Job Description:
Key Responsibilities
Monitor application health and performance using tools such as Splunk, Dynatrace, Grafana, or Datadog
Perform end-to-end application troubleshooting, root cause analysis, and resolution
Analyze logs, metrics, and traces to proactively identify and resolve issues
Develop and maintain scripts (Python/Shell/Bash) for automation and operational efficiency
Work extensively in Unix/Linux environments for application and infrastructure support
Write and optimize complex SQL queries for data analysis and issue resolution
Support and maintain databases such as Oracle, PostgreSQL, MySQL, or Cassandra
Manage and support Apache Tomcat Server deployments and configurations
Participate in DevOps pipeline creation, maintenance, and optimization
Collaborate with cross-functional teams for deployment, upgrades, and incident resolution
Ensure system stability, performance, and high availability
✅ Must-Have Skills
Splunk and Dynatrace
Strong Application Troubleshooting skills
Scripting: Python, Shell, Bash
Unix/Linux operating systems
Strong SQL writing experience
Database concepts with any of:
Oracle
PostgreSQL
MySQL
Cassandra
Shell Scripting
Apache Tomcat Server
Grafana or Datadog
Basics of DevOps pipeline creation and maintenance
Good-to-Have Skills
Knowledge of ITIL / TISM processes
Change and Deployment Management
Incident Management
Experience in production support or SRE environments
Behavioral Competencies
Strong analytical and problem-solvi ng skills
Excellent communication and collaboration abilities
Ability to work in high-pressure, production environments
Continuous learning mindset and ownership attitude
Location
Sydney
Job Function
TECHNOLOGY
Role
Engineer
Job Id
423558
Desired Skills
Splunk
Questions about this role
Want AI Applyd to auto-apply to roles like this?
We tailor your resume per posting, fill the forms, and track replies for you.