Staff Site Reliability Engineer - 26248
Skills
About the role
Senior Site Reliability Engineer
Why YOU want this position:
At Enverus, we’re committed to empowering the global quality of life by helping our customers make energy affordable and accessible to the world.
We are the most trusted energy-dedicated SaaS company, with a platform built to maximize value from generative AI, and our innovative solutions are reshaping the way energy is consumed and managed. By offering anytime, anywhere access to analytics and insights, we’re helping our customers make better decisions that help provide communities around the world with clean, affordable energy.
The energy industry is changing fast. But we’ve continued to lead the way in energy technology, creating intelligent connections across the entire energy ecosystem, from renewables, power and utilities, to oil and gas and financial institutions. Our solutions create more efficient production and distribution, capital allocation, renewable energy development, investment and sourcing, and help reduce costs by automating crucial business operations. Of course, this wouldn’t be possible without our people, which is why we have built a team of individuals from a diverse range of backgrounds.
Are you ready to help power the global quality of life? Join Enverus, and be a part of creating a brighter, more sustainable tomorrow.
We are currently seeking a Staff Site Reliability Engineer to join our Cloud Engineering team that manages our entire AWS presence. This role will be based in Canada.
This role offers the opportunity to join a rapidly growing company delivering industry-leading solutions to customers in the world’s most dynamic and fastest growing sector.
Performance Objectives
Work on a team that manages our entire global AWS presence
The team you will be working with is responsible for keeping our infrastructure humming as new releases and maintenance updates are rolled out
You will help organize, secure, and automate existing infrastructure and deployments
You will work closely with developers to provide feedback and drive operational improvements within our products and operations infrastructure
You will be responsible for ensuring that our platform is stable and balanced
Maintain high site up time, while embracing rapid change and growth
Scale infrastructure to meet increasing demand and evolving technology
Help the dev teams working on our code bases realize zero down-time deployments
Develop and improve operational practices and procedures
Implement, monitor, and maintain CI/CD frameworks
You will coordinate and participate in on-call rotations (day shifts)
Automate, automate, automate…
Competitive Candidate Profile
5+ years of professional Windows and Linux server administration
3+ years of Amazon Web Services (AWS) administration
3+ years of experience within a high-performance, 24x7, DevOps, SysOps, or Operations team
You have excellent communication and collaboration skills
You demonstrate the ability to succeed in a high-pressure environment with rapidly changing priorities
You are an excellent problem solver, and willing to roll up your sleeves to take on any issue thrown your way
You have a desire not just to resolve problems, but to fully understand them and prevent them in the future
You seek out opportunities to improve, fix bugs, and challenge assumptions
You have experience working with global teams (North America, Europe, Asia)
You have experience with the following technologies:
Kubernetes (Container Orchestration)
Infrastructure as Code (Terraform, Cloudformation)
AWS (or other major cloud providers)
Python, C#, or Golang programming experience is a plus
Azure and Azure local experience is a plus
You prefer to lead the charge, not just keep up with it
Questions about this role
Want AI Applyd to auto-apply to roles like this?
We tailor your resume per posting, fill the forms, and track replies for you.