Senior DevOps Engineer
Infra360 · Head Office · India · On-site
Pay: INR 1,500,000 – 2,200,000 a year
Posted Jul 27, 2026
Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.
AWS / Azure / GCP | Multi-Client Ownership | Technical Leadership
About the Role
We are looking for a Senior DevOps Engineer with 5–8 years of strong hands-on production experience to independently own multiple client environments and provide technical leadership to a team of engineers.
In this role, you will typically manage 2–3 client environments , drive cloud and infrastructure architecture decisions, troubleshoot complex production challenges, and lead initiatives across reliability, security, automation, performance, and cost optimization.
You will also provide technical direction and mentorship to 3–5 engineers , ensuring high-quality delivery, strong engineering practices, and effective execution.
The role requires strong architectural thinking, engineering judgment, production troubleshooting skills, and the ability to communicate effectively with both technical teams and client stakeholders.
Key Responsibilities
1. Cloud Infrastructure & Architecture
Design, deploy, manage, and optimize production cloud environments across AWS, Azure, and/or GCP .
Design and review highly available, scalable, secure, and cost-efficient cloud architectures.
Work extensively with core cloud services including compute, storage, databases, networking, IAM, and load balancers .
Own production workloads on EKS, AKS, GKE , including cluster design, upgrades, scaling, troubleshooting, and maintenance.
Evaluate architectural options and recommend solutions based on client requirements, technical constraints, and business objectives.
Identify technical debt, architectural risks, and opportunities to improve reliability, scalability, security, and operational efficiency.
Implement and improve backup, disaster recovery, high availability, and business continuity practices.
2. Kubernetes, Infrastructure as Code & Automation
Design and operate production Kubernetes environments, addressing challenges related to networking, scheduling, scaling, resource utilization, availability, and application behavior .
Establish reusable Kubernetes and infrastructure patterns across client environments.
Develop and maintain infrastructure using Terraform and Infrastructure as Code best practices.
Build reusable IaC modules and standards to improve consistency, scalability, and operational reliability.
Design and implement scalable CI/CD and GitOps workflows using tools such as ArgoCD, Flux, Spinnaker , or similar platforms.
Automate operational processes using Bash, Python , and other appropriate scripting or automation tools.
Identify and eliminate repetitive operational work through automation and engineering improvements.
3. Reliability, Monitoring & Production Operations
Own production reliability and operational excellence across assigned client environments.
Lead troubleshooting of complex infrastructure, Kubernetes, networking, and application performance issues.
Configure and improve monitoring, logging, alerting, and observability using tools…