JobRaahGet matched free

Jobs

Site Reliability Engineer

qualys · Pune · India · On-site

Posted Sep 29, 2026

Apply with JobRaah

Sign up free: we match you to jobs like this, tailor your application and fill the form. 2 free applications every day.

Come work at a place where innovation and teamwork come together to support the most exciting missions in the world! THE OPPORTUNITY The Site Reliability Engineer (SRE) plays a critical role in ensuring the reliability, scalability, performance, and operational excellence of Qualys platforms and services. This role operates at the intersection of software engineering and operations, applying automation, observability, and reliability engineering practices to maintain highly available systems and improve customer experience. WHERE THIS ROLE SITS The SRE role partners closely with Engineering, DevOps, Infrastructure, and Security teams to ensure the engineering systems remain resilient, scalable, and operationally efficient. WHAT YOU WILL DO Reliability and Operations Maintain highly available and scalable applications and services. Monitor application and infrastructure health using observability tools. Respond to incidents, troubleshoot deployment issues, and perform root cause analysis. Participate in on-call rotations and incident response processes. Track and improve service reliability, latency, performance, and efficiency. Automation and Platform Engineering Create automation scripts and tooling to reduce manual operational effort. Support deployments, configuration management, and operational workflows. Leverage Infrastructure as Code (IaC) practices for repeatable infrastructure management. Continuously improve operational processes through automation and self-service capabilities. Collaboration and Continuous Improvement Collaborate with development teams to improve resilience and operational readiness. Document system architectures, operational procedures, and best practices. Drive continuous improvement initiatives focused on reliability and operational excellence. WHAT GOOD LOOKS LIKE Uses automation to eliminate repetitive operational tasks. Responds effectively to incidents and drives meaningful root cause analysis. Collaborates effectively across engineering and infrastructure teams. Continuously improves platform reliability, scalability, and observability. DISTINGUISHING EXPECTATION The SRE is measured by improvements in reliability, automation, operational efficiency, and service availability. Success is demonstrated through proactive problem prevention, reduced operational burden, and continuous improvement of production systems. REQUIRED QUALIFICATIONS Bachelor’s degree in computer science, Engineering, or related field (or equivalent experience). Experience in Site Reliability Engineering, DevOps, Systems Engineering, or Cloud Infrastructure roles. Knowledge of Linux/Unix systems administration. Experience with cloud platforms such as AWS, OCI, or Google Cloud Platform. Proficiency in Python, Go, Bash, Java, or similar scripting/programming languages. Experience with Docker and Kubernetes. Familiarity with Terraform, CloudFormation, or other IaC tools. Experience with CI/CD tools such as…