Site Reliability Technical Operations Manager

jciยท JIN Johnson Controls (India) Private Limited
Apply Now โ†—
๐Ÿ“ Pune-Maharashtra-IndiaFull time

About this role

Summary:

The Site Reliability Engineering team at Johnson Controls is seeking a Technical Reliability & Support Manager to lead cloud product reliability, operational support, and production stability across global cloud applications and platforms. This role will be responsible for managing day-to-day technical operations, L2/L3 support coordination, incident response, service reliability governance, and continuous improvement for cloud products hosted across platforms such as Azure (primarily), Google Cloud, Ali Cloud, and other cloud environments.

The role will partner closely with Engineering, SRE, Security, Observability, and external support partners to ensure production issues are resolved quickly, recurring problems are eliminated, support processes are standardized, and operational risks are proactively identified and addressed.


Primary Duties:

  • Lead technical operations and support management for cloud products across global environments
  • Manage day-to-day L2/L3 support activities, incident response, escalations, and production issue resolution
  • Own service reliability governance for assigned cloud products, including availability, incident trends, MTTR, recurring issues, and operational risks
  • Partner with Engineering, SRE, Security, Observability and Platform teams to improve service stability and operational readiness
  • Drive incident management, problem management, RCA/PCA reviews, and corrective action tracking for production issues
  • Ensure timely communication during incidents, including stakeholder updates, executive summaries, customer-impact statements, and resolution updates
  • Establish and manage support processes aligned with ITIL practices, including incident, problem, change, service request, and escalation management
  • Define, track, and report operational KPIs such as availability, MTTR, incident volume, severity trends, backlog, service requests, alert noise, and SLA performance
  • Identify recurring operational pain points and drive permanent fixes through engineering backlog, automation, monitoring improvements, and process enhancements
  • Collaborate with Observability teams to ensure critical application, infrastructure, database, and integration components are properly monitored and alerted
  • Support implementation and adoption of SLIs, SLOs, SLAs, error budgets, and reliability / supportability scorecards for cloud products
  • Ensure production readiness for new product releases, migrations, infrastructure changes, and platform transformations
  • Lead operational reviews with internal teams, vendors, and external support partners to ensure accountability and continuous improvement
  • Drive automation of manual support tasks, ticket workflows, reporting, and operational runbooks to improve efficiency and consistency
  • Ensure support activities, incidents, changes, and action items are properly documented and tracked in tools such as Jira, ServiceNow, Confluence, or equivalent platforms
  • Manage support handoffs across global teams and ensure clear ownership, escalation paths, and communication protocols
  • Provide technical leadership during high-severity incidents, major outages, migrations, and critical customer-impacting events
  • Ensure compliance with security, audit, operational, and IT governance standards for cloud product support
  • Maintain operational documentation, SOPs, runbooks, escalation matrices, support models, and knowledge base articles
  • Mentor support engineers and technical teams on reliability practices, incident handling, RCA quality, and operational excellence

Qualifications:

  • 10+ years of experience in technical operations, production support, SRE, cloud support, or reliability engineering.
  • Strong experience managing cloud-hosted applications and support operations.
  • Good knowledge of Azure (preferred), with exposure to AWS, GCP, or Ali Cloud.
  • Experience with incident management, problem management, change management, and operational governance.
  • Familiarity with cloud-native technologies, microservices, APIs, databases, containers, and Kubernetes.
  • Understanding of observability tools such as Grafana, Datadog, ELK, Logz.io, Azure Monitor, or similar platforms.
  • Knowledge of reliability practices including SLIs, SLOs, SLAs, monitoring, alerting, and automation.
  • Strong troubleshooting skills across applications, infrastructure, networking, databases, and cloud services.
  • Experience with Jira, Confluence, or similar platforms.
  • Excellent leadership, communication, stakeholder management, and vendor coordination skills.
  • Ability to lead high-severity incidents and drive operational excellence in a fast-paced environment.

ย Mandatory Skills:

  • Cloud Operations / SRE leadership
  • L2 production support (good to have L2 Production Support)
  • Incident & escalation management
  • RCA/PCA and problem management
  • Azure cloud and cloud-native platforms
  • Observability, monitoring & KPI reporting
  • ITIL-based support processes
  • Automation and runbook development
  • Executive communication & stakeholder management
  • Leadership in high-pressure production environments

Frequently Asked Questions

Is the salary disclosed for the Site Reliability Technical Operations Manager position at jci?
The salary for this Site Reliability Technical Operations Manager role at jci is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the Site Reliability Technical Operations Manager position at jci located?
This Site Reliability Technical Operations Manager role at jci is based in Pune-Maharashtra-India. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the Site Reliability Technical Operations Manager role at jci full-time or part-time?
This is listed as a Full time position. It is posted as a Site Reliability Technical Operations Manager role in the JIN Johnson Controls (India) Private Limited department at jci.
Which team or department does the Site Reliability Technical Operations Manager at jci belong to?
This Site Reliability Technical Operations Manager position is part of the JIN Johnson Controls (India) Private Limited department at jci. See the full job description for more information about the team structure and responsibilities.
How do I apply for the Site Reliability Technical Operations Manager position at jci?
Click the "Apply Now" button on this page. You will be redirected to jci's official application portal hosted on workday where you can submit your application directly.
When was the Site Reliability Technical Operations Manager job at jci posted?
This Site Reliability Technical Operations Manager position at jci was posted on Sep 30, 2026. Apply as soon as possible โ€” early applications are often reviewed first.
Site Reliability Technical Operations Manager
jci
Apply for this role โ†—

You'll be redirected to jci's official application page on Workday.