SeniorAdministrator - AWS IAC, Terraform,Python

Apply Now ↗
📍 Bengaluru, India

About this role

Job Summary

Job Summary : We are looking for a passionate and experienced Site Reliability Engineer (SRE) to join our Cloud Platform team. The ideal candidate will have hands-on experience managing large-scale Kubernetes clusters on public cloud environments (AKS, EKS, or GKE) and a strong understanding of modern SRE and DevOps practices. You will be responsible for ensuring high availability, reliability, scalability, and performance of our cloud-native infrastructure and CI/CD systems. Key Responsibilities: • Manage, monitor, and optimize large-scale Kubernetes clusters hosted on public cloud platforms (Azure AKS, AWS EKS, or Google GKE). • Implement and maintain infrastructure as code using tools such as Terraform. • Collaborate with development and operations teams to improve system reliability and deployment automation. • Build and maintain CI/CD pipelines using Jenkins or similar tools. • Troubleshoot production issues, conduct root cause analysis, and implement preventive measures. • Automate operational tasks using Python or other scripting languages. • Contribute to observability and monitoring improvements using modern tools and best practices. • Participate in on-call rotations and incident response processes. Required Skills and Experience: • 4–8 years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles. • Strong hands-on experience managing Kubernetes clusters in production (AKS/EKS/GKE). • Proficiency with Terraform and cloud infrastructure automation. • Practical experience with Jenkins and CI/CD pipeline management. • Sound understanding of SRE principles (incident management, blameless postmortems, capacity planning, error budgets, etc.). • Good programming or scripting skills in Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux). Job Description : We are looking for a passionate and experienced Site Reliability Engineer (SRE) to join our Cloud Platform team. The ideal candidate will have hands-on experience managing large-scale Kubernetes clusters on public cloud environments (AKS, EKS, or GKE) and a strong understanding of modern SRE and DevOps practices. You will be responsible for ensuring high availability, reliability, scalability, and performance of our cloud-native infrastructure and CI/CD systems.\\\\r\\\\n\\\\r\\\\nKey Responsibilities:\\\\r\\\\n• Manage, monitor, and optimize large-scale Kubernetes clusters hosted on public cloud platforms (Azure AKS, AWS EKS, or Google GKE).\\\\r\\\\n• Implement and maintain infrastructure as code using tools such as Terraform.\\\\r\\\\n• Collaborate with development and operations teams to improve system reliability and deployment automation.\\\\r\\\\n• Build and maintain CI/CD pipelines using Jenkins or similar tools.\\\\r\\\\n• Troubleshoot production issues, conduct root cause analysis, and implement preventive measures.\\\\r\\\\n• Automate operational tasks using Python or other scripting languages.\\\\r\\\\n• Contribute to observability and monitoring improvements using modern tools and best practices.\\\\r\\\\n• Participate in on-call rotations and incident response processes.\\\\r\\\\n\\\\r\\\\n\\\\r\\\\nRequired Skills and Experience:\\\\r\\\\n• 4–8 years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles.\\\\r\\\\n• Strong hands-on experience managing Kubernetes clusters in production (AKS/EKS/GKE).\\\\r\\\\n• Proficiency with Terraform and cloud infrastructure automation.\\\\r\\\\n• Practical experience with Jenkins and CI/CD pipeline management.\\\\r\\\\n• Sound understanding of SRE principles (incident management, blameless postmortems, capacity planning, error budgets, etc.).\\\\r\\\\n• Good programming or scripting skills in Python (preferred) or similar langua

Key Responsibilities

 We are looking for a passionate and experienced Site Reliability Engineer (SRE) to join our Cloud Platform team. The ideal candidate will have hands-on experience managing large-scale Kubernetes clusters on public cloud environments (AKS, EKS, or GKE) and a strong understanding of modern SRE and DevOps practices. You will be responsible for ensuring high availability, reliability, scalability, and performance of our cloud-native infrastructure and CI/CD systems. Key Responsibilities: • Manage, monitor, and optimize large-scale Kubernetes clusters hosted on public cloud platforms (Azure AKS, AWS EKS, or Google GKE). • Implement and maintain infrastructure as code using tools such as Terraform. • Collaborate with development and operations teams to improve system reliability and deployment automation. • Build and maintain CI/CD pipelines using Jenkins or similar tools. • Troubleshoot production issues, conduct root cause analysis, and implement preventive measures. • Automate operational tasks using Python or other scripting languages. • Contribute to observability and monitoring improvements using modern tools and best practices. • Participate in on-call rotations and incident response processes. Required Skills and Experience: • 4–8 years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles. • Strong hands-on experience managing Kubernetes clusters in production (AKS/EKS/GKE). • Proficiency with Terraform and cloud infrastructure automation. • Practical experience with Jenkins and CI/CD pipeline management. • Sound understanding of SRE principles (incident management, blameless postmortems, capacity planning, error budgets, etc.). • Good programming or scripting skills in Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux).

 

Skill Requirements

 We are looking for a passionate and experienced Site Reliability Engineer (SRE) to join our Cloud Platform team. The ideal candidate will have hands-on experience managing large-scale Kubernetes clusters on public cloud environments (AKS, EKS, or GKE) and a strong understanding of modern SRE and DevOps practices. You will be responsible for ensuring high availability, reliability, scalability, and performance of our cloud-native infrastructure and CI/CD systems. Key Responsibilities: • Manage, monitor, and optimize large-scale Kubernetes clusters hosted on public cloud platforms (Azure AKS, AWS EKS, or Google GKE). • Implement and maintain infrastructure as code using tools such as Terraform. • Collaborate with development and operations teams to improve system reliability and deployment automation. • Build and maintain CI/CD pipelines using Jenkins or similar tools. • Troubleshoot production issues, conduct root cause analysis, and implement preventive measures. • Automate operational tasks using Python or other scripting languages. • Contribute to observability and monitoring improvements using modern tools and best practices. • Participate in on-call rotations and incident response processes. Required Skills and Experience: • 4–8 years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles. • Strong hands-on experience managing Kubernetes clusters in production (AKS/EKS/GKE). • Proficiency with Terraform and cloud infrastructure automation. • Practical experience with Jenkins and CI/CD pipeline management. • Sound understanding of SRE principles (incident management, blameless postmortems, capacity planning, error budgets, etc.). • Good programming or scripting skills in Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux).

 

Other Requirements

 We are looking for a passionate and experienced Site Reliability Engineer (SRE) to join our Cloud Platform team. The ideal candidate will have hands-on experience managing large-scale Kubernetes clusters on public cloud environments (AKS, EKS, or GKE) and a strong understanding of modern SRE and DevOps practices. You will be responsible for ensuring high availability, reliability, scalability, and performance of our cloud-native infrastructure and CI/CD systems. Key Responsibilities: • Manage, monitor, and optimize large-scale Kubernetes clusters hosted on public cloud platforms (Azure AKS, AWS EKS, or Google GKE). • Implement and maintain infrastructure as code using tools such as Terraform. • Collaborate with development and operations teams to improve system reliability and deployment automation. • Build and maintain CI/CD pipelines using Jenkins or similar tools. • Troubleshoot production issues, conduct root cause analysis, and implement preventive measures. • Automate operational tasks using Python or other scripting languages. • Contribute to observability and monitoring improvements using modern tools and best practices. • Participate in on-call rotations and incident response processes. Required Skills and Experience: • 4–8 years of experience in Site Reliability Engineering, DevOps, or Cloud Infrastructure roles. • Strong hands-on experience managing Kubernetes clusters in production (AKS/EKS/GKE). • Proficiency with Terraform and cloud infrastructure automation. • Practical experience with Jenkins and CI/CD pipeline management. • Sound understanding of SRE principles (incident management, blameless postmortems, capacity planning, error budgets, etc.). • Good programming or scripting skills in Python (preferred) or similar languages. • Strong analytical, troubleshooting, and problem-solving abilities. • Excellent written and verbal communication skills. • Experience with Prometheus, Grafana, or OpenTelemetry for observability. • Exposure to GitOps practices and tools (e.g. Flux).

 

Frequently Asked Questions

Is the salary disclosed for the SeniorAdministrator - AWS IAC, Terraform,Python position at HCLTech?
The salary for this SeniorAdministrator - AWS IAC, Terraform,Python role at HCLTech is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the SeniorAdministrator - AWS IAC, Terraform,Python position at HCLTech located?
This SeniorAdministrator - AWS IAC, Terraform,Python role at HCLTech is based in Bengaluru, India. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
How do I apply for the SeniorAdministrator - AWS IAC, Terraform,Python position at HCLTech?
Click the "Apply Now" button on this page. You will be redirected to HCLTech's official application portal hosted on successfactors where you can submit your application directly.
When was the SeniorAdministrator - AWS IAC, Terraform,Python job at HCLTech posted?
This SeniorAdministrator - AWS IAC, Terraform,Python position at HCLTech was posted on Sep 16, 2026. Apply as soon as possible — early applications are often reviewed first.
SeniorAdministrator - AWS IAC, Terraform,Python
HCLTech
Apply for this role ↗

You'll be redirected to HCLTech's official application page on successfactors.