Senior / Principal DevOps / Platform Engineer
Skills
About the role
About the Role
We are looking for a highly experienced Senior / Principal DevOps / Platform Engineer to own and operate the complete infrastructure lifecycle of a secure, cloud-native production platform. This is a hands-on engineering role focused on designing, automating, securing, and maintaining scalable cloud infrastructure while ensuring high availability, reliability, and compliance.
You will work closely with architects and senior engineers in a collaborative environment with direct ownership of production systems and minimal operational overhead.
Key ResponsibilitiesPlatform & Infrastructure
Design, provision, and manage cloud infrastructure across production and non-production environments.
Operate and maintain Kubernetes clusters, including upgrades, autoscaling, node pools, maintenance windows, and Pod Disruption Budgets.
Manage Kubernetes configurations using Kustomize.
Implement workload identity and secure cloud networking (VNets, subnets, DNS, private endpoints, security groups, and peering).
CI/CD & Automation
Build, maintain, and optimize GitHub Actions CI/CD pipelines.
Manage release workflows, environment promotions, artifact versioning, and container image lifecycle.
Enforce security scanning, quality gates, and secure deployment practices.
Maintain local development environments aligned with production.
Secrets & Configuration
Manage centralized secrets and key management.
Maintain Kubernetes ConfigMaps and Secrets using Git-based workflows.
Implement identity-based authentication and role-based access controls.
Container & Image Management
Build and maintain secure, optimized container images.
Manage enterprise container registries, image retention, and lifecycle policies.
Implement immutable production image practices.
Database & Data Platform
Operate managed PostgreSQL and Redis services.
Support backups, disaster recovery, schema migrations, and performance optimization.
Manage object storage and support data orchestration platforms.
Observability & Operations
Implement monitoring, logging, tracing, and OpenTelemetry.
Build dashboards and alerting for platform health and performance.
Manage production incidents, on-call support, and operational runbooks.
Security & Compliance
Configure WAF, API Gateway, Kubernetes network policies, and infrastructure security controls.
Perform vulnerability assessments and infrastructure security reviews.
Ensure encryption, audit logging, and regulatory compliance.
Required Qualifications
5+ years of experience in DevOps, Platform Engineering, or Infrastructure Engineering.
Strong hands-on experience with Kubernetes in production environments.
Expertise in Azure (preferred) and AWS.
Experience with Infrastructure as Code (Terraform, ARM, or Bicep).
Strong knowledge of GitHub Actions and CI/CD automation.
Experience managing PostgreSQL and Redis in production.
Proficiency with Docker and container registry management.
Experience implementing observability using OpenTelemetry.
Strong understanding of cloud networking and infrastructure security.
Proficiency in Python, Bash, YAML, and JSON.
Preferred Qualifications
Experience with regulated or highly secure environments.
GitOps experience using ArgoCD or Flux.
Helm chart development.
Experience with authentication and identity platforms.
Knowledge of analytics or data platform integrations.
Advanced Kubernetes debugging and RBAC management.
Key Skills
Kubernetes
Azure & AWS
Terraform / ARM / Bicep
GitHub Actions
Docker
PostgreSQL
Redis
OpenTelemetry
API Gateway
WAF/CDN
Python
Bash
GitOps
Helm
Cloud Security
Infrastructure as Code
Work Location: Hybrid remote in Noida, Uttar Pradesh (Noida)
Questions about this role
Want AI Applyd to auto-apply to roles like this?
We tailor your resume per posting, fill the forms, and track replies for you.