Sr. DevOps Engineer

Kaseya

Bengaluru, INonsitePosted Jul 24, 2026
Posting intelligenceActively listedReposted 2×, possible evergreen/ghost posting

Skills

kubernetesprometheusterraformjenkinsgrafanadatadoggithubpulumipythonazureswiftcicdgooglecloudawsgo

About the role

About Kaseya

Kaseya is the leading provider of AI-powered IT management and cybersecurity software, serving Managed Service Providers (MSPs) and internal IT organizations worldwide. Our comprehensive platform helps organizations efficiently manage, secure, and automate their IT environments, driving operational efficiency and long-term business success.

Backed by Insight Partners, a leading global software investor, Kaseya has experienced sustained double-digit growth and continues to its global footprint. Today, Kaseya supports customers in more than 20 countries and manages over 15 million endpoints worldwide.

Founded in 2000, Kaseya has built a culture centered around innovation, accountability, and results. We are a high-growth, high-performance organization that values individuals who are driven, adaptable, and committed to delivering exceptional outcomes for our customers and teammates alike.

At Kaseya, success comes from embracing challenges, moving with urgency, and continuously raising the bar.

Senior DevOps Engineer

Experience: 8–12 Years

Location: Pune (Hybrid/Onsite)

About the Role

We are looking for an experienced Senior DevOps Engineer with deep expertise in Linux Administration to join our Backup Platform Engineering team. In this role, you will own the reliability, scalability, and performance of large-scale Linux infrastructure that powers our next-generation backup and disaster recovery platform. You will work closely with software engineering, SRE, and platform teams to automate infrastructure, improve operational excellence, and ensure highly available production environments.

Key Responsibilities

Design, build, and manage highly available Linux-based production infrastructure supporting mission-critical backup services.

Administer and optimize large-scale Linux environments, including performance tuning, kernel configuration, storage, networking, and system troubleshooting.

Manage Kubernetes clusters, ensuring reliability, scalability, security, and efficient resource utilization.

Build and maintain Infrastructure as Code (Terraform/Pulumi) following reusable and modular design principles.

Design and enhance CI/CD pipelines using GitHub Actions, Jenkins, ArgoCD, or similar tools.

Develop automation using Shell scripting, Python, or Go to eliminate manual operational tasks.

Implement monitoring, logging, and observability using Prometheus, Grafana, Datadog, or similar platforms.

Drive incident response, root cause analysis, postmortems, and continuous operational improvements.

Collaborate with development teams to improve deployment processes, platform reliability, and production readiness.

Implement infrastructure security best practices including RBAC, secrets management, vulnerability scanning, and audit logging.

Troubleshoot complex infrastructure, networking, storage, and Linux operating system issues in production environments.

Required Skills

8–12 years of experience in DevOps, Platform Engineering, Linux Administration, or Site Reliability Engineering (SRE).

Strong expertise in Linux system administration, including:

Performance tuning

Kernel parameters

Storage & filesystem management

Process management

System troubleshooting

Networking fundamentals

Hands-on experience managing Kubernetes clusters in production.

Strong knowledge of Infrastructure as Code using Terraform or Pulumi.

Experience designing and maintaining CI/CD pipelines using GitHub Actions, Jenkins, ArgoCD, or equivalent.

Experience with monitoring and observability tools such as Prometheus, Grafana, Datadog, ELK, etc.

Strong scripting skills in Shell, Python, or Go for infrastructure automation.

Experience managing production incidents, on-call support, alert tuning, and operational excellence.

Strong understanding of:

DNS

Load Balancers

Firewalls

VPCs

Hybrid networking

Experience implementing security best practices including Vault, RBAC, secrets management, and audit logging.

Preferred Skills

Experience with OpenStack (Nova, Swift, Neutron, Cinder).

Experience managing workloads across AWS, Azure, GCP, and private cloud environments.

Exposure to large-scale distributed infrastructure (5,000+ nodes).

Experience with storage platforms, backup infrastructure, and disaster recovery concepts (RPO/RTO).

Knowledge of cost optimization (FinOps), storage tiering, and infrastructure capacity planning.

Experience with Chaos Engineering and resilience testing.

Familiarity with bare-metal provisioning technologies such as PXE, MaaS, or Ironic.

Ability to read and troubleshoot Go-based services and contribute to automation tooling.

Additional information

Questions about this role

Click "Apply with AI Applyd" above. We auto-fill the application from your resume and answer screening questions in seconds. No copy and paste, no juggling tabs.

Compensation for DevOps / SRE roles in India varies widely by seniority, employer size, and remote vs onsite arrangement. Check the salary range on this listing when published, or browse our DevOps / SRE hub for India medians across recent openings.

Most applications complete in under 90 seconds. You can track the status in your dashboard and watch the screenshot proof land the moment the application submits.

AI Applyd supports Greenhouse, Lever, Ashby, Workday, iCIMS, SmartRecruiters, Personio, Teamtailor and other major ATS platforms. If we can submit through the platform, we do.

Want AI Applyd to auto-apply to roles like this?

We tailor your resume per posting, fill the forms, and track replies for you.