hirq
← All jobs

ClanX

Senior DevOps Engineer (AWS | Kubernetes | Terraform)

Pune, Mahārāshtra, India · On-site · fulltime_permanent · Engineering

Apply well, not just fast

Create a free account and upload your resume to get a match score, keyword gaps, a tailored resume, a cover letter and interview prep for this job.

About the role

AWSTerraformKubernetesCI/CDAzureDockerGitHub ActionsPrometheusGrafanaObservabilityPythonBashGitServerlessSystem DesignIAM
Lead AWS cloud infrastructure, Kubernetes, Terraform, CI/CD, and observability for Musafir’s cloud-native travel platform, with ownership of reliability, security, automation, and operational excellence. Company Details Musafir is a technology-driven travel platform building scalable digital travel experiences. Website: Musafir Requirements - 8–12 years of experience in Cloud/DevOps engineering - Strong hands-on AWS experience across core cloud services - Strong Kubernetes experience with EKS, ECS, Docker, Helm, ingress, and service mesh - Strong Terraform experience; CloudFormation or CDK experience - Experience with CI/CD using Azure Pipelines or GitHub Actions - Strong scripting skills in Python, Bash, or PowerShell - Experience with OpenTelemetry, Prometheus, Grafana, CloudWatch, and ELK - Strong understanding of AWS security, IAM, KMS, secrets management, and security scanning - Experience with Git and Agile/Scrum - Strong incident management, troubleshooting, RCA, and post-mortem skills - Understanding of high availability, disaster recovery, RTO/RPO, and cost optimization Responsibilities - Design and manage scalable, secure, and cost-optimized AWS infrastructure - Build and operate Kubernetes workloads using EKS, ECS, Docker, Helm, and related tooling - Develop and maintain Infrastructure as Code using Terraform - Build and manage CI/CD pipelines using Azure Pipelines and GitHub Actions - Manage AWS serverless and event-driven services including Lambda, API Gateway, SQS/SNS, and EventBridge - Implement observability using OpenTelemetry, Prometheus, Grafana, CloudWatch, and ELK - Define and monitor SLIs, SLOs, and SLAs - Drive cloud security, compliance, secrets management, and vulnerability scanning - Design backup and disaster recovery strategies aligned with RTO/RPO requirements - Lead production incident response, root-cause analysis, and post-mortems - Improve infrastructure reliability, automation, performance, and cloud costs - Use AI/Copilot to generate and review Terraform, scripts, and CI/CD pipelines - Apply AIOps for anomaly detection, log summarization, and AI-assisted incident RCA - Support infrastructure for AI workloads including GPU/inference nodes, vector databases, and model-serving platforms Job Details Pune — Work from Office Interview Process - Technical Screening - Technical Interview - System Design Interview - Hiring Manager Round - HR Round Important Note ClanX is a recruitment partner, helping Musafir hire Senior DevOps Engineer.