
SITE RELIABILITY ENGINEER
Skills
About the role
SITE RELIABILITY ENGINEER
COUNTRY
United States
FORMAT
On-site
Svitla Systems Inc. is looking for a Site Reliability Engineer for a full-time position (40 hours per week) in Seattle, USA. Our client is a provider of a cloud data platform that manages exabytes of the world's most demanding data, unifying files, objects, and every workload across edge, core, and cloud.
You'll be one of the first hires on a team with a single mandate: find out how the company breaks before the customers do. The platform manages exabytes of data for more than 1,100 customers across on-prem and every major cloud, and these are mission-critical workloads where a missed edge case becomes a customer's bad day. It is a test-centric SRE role for an engineer who thinks like a breaker. You'll put on the customer's hat, work out how a feature will actually be used, and design tests that push it to its limits across hardware and cloud environments. You'll automate the testing that the principal engineers currently run by hand, and decide what gets tested, how often, and why. You'll help build this function from the ground up, including setting the quality bar and deciding which builds are good enough to ship.
REQUIREMENTS:
3+ years of experience in building and operating automated testing, validation, and/or certification for complex software systems.
Strong programming ability in C.
Experience with distributed file systems or parallel file systems would be a major plus.
A real breaker's instinct. You look for edge cases and ask, "What happens if I do this?" before anyone asks you to.
A track record of building tests yourself, not just running test plans handed to you.
Hands-on experience across both on-premises infrastructure and cloud (AWS, GCP, or Azure), with a real grasp of where each one's limits are.
Strong knowledge of Linux (the team runs Ubuntu)
Knowledge of Python.
Understanding of a data-driven approach to deciding what to test and how often.
Knowledge of orchestration tools (Ansible, Terraform), containers, and Kubernetes.
NICE TO HAVE:
Solid understanding of networks (routing, firewalls, security inspection devices, switch configuration).
Experience with storage (IOPS, Latency, read/write patterns) or protocol (NFS, SMB, S3).
RESPONSIBILITIES:
Design and operationalize testing for new features: work out how customers will actually use them, how to scale-test them, and how to break them.
Automate the manual, repetitive testing our principal engineers run by hand today, using Python and our in-house frameworks on Jenkins and Argo.
Build a data-driven plan for which tests run, how often, and why, plus the framework to schedule and rerun them.
Troubleshoot build and test failures across VM instances and hardware, from compile-time errors to integration failures.
Read cluster output and C error logs to tell a test problem from an infrastructure problem from a real bug.
Set up monitoring and alerting so problems surface early (the team uses OpenMetrics, Grafana, InfluxDB, and Prometheus alongside home-grown tooling).
Help set the quality bar for releases, including a real say in what ships.
Take part in an on-call rotation for the systems your team owns.
WE OFFER
US and EU projects based on advanced technologies.
Competitive compensation based on skills and experience.
Flexibility in workspace, either remote or our welcoming office.
Bonuses for article writing, public talks, and other activities.
Free tech webinars and meetups organized by Svitla.
Regular corporate online activities.
Awesome team, friendly and supportive community!
ABOUT SVITLA
Svitla Systems is a global digital solutions company headquartered in the U.S. and operating across the Americas, Europe, Asia, and APAC. Since 2003, we have served a wide range of clients - from innovative start-ups to Fortune 500 companies.
Our success is built on partnership. By integrating seamlessly with clients’ teams, we create lasting collaborations that drive real results.
We are strong advocates of workplace flexibility, remote culture, individual approach to professional and personal growth.
Our global mission is to build a business that contributes to wellbeing of our partners, personnel, and their families, improves our communities, and makes a lasting difference in the world.
Together, we are coding a brighter tomorrow - and living it.
Join us!
LET'S MEET IN PERSON
Olha Romanenko
RECRUITER
Email:
O.romanenko@svitla.com
LinkedIn:
Olha Romanenko
Questions about this role
Want AI Applyd to auto-apply to roles like this?
We tailor your resume per posting, fill the forms, and track replies for you.