📍 Noida, India

About this role

Job Summary

Define and maintain production readiness standards across platform, data, model, application, and securitylayers.
Establish SLO/SLI frameworks for latency, availability, quality, safety, and drift implement error budgetpolicies.
Publish reference architectures for LLM apps, RAG, vector stores, agent frameworks, and batch/streaminference.
Curate deployment blueprints (canary/shadow, blue–green, A/B) for models and prompts with rollbackguidance.
Standardize observability patterns for prompts, embeddings, latency, cost, quality, and safety telemetry.
Own capacity engineering (token/concurrency budgets, GPU/CPU sizing, vector scaling, cache hierarchies).
Define resilience patterns (timeouts, circuit breakers, fallbacks, idempotent retries, semantic/prompt caching).
Set AI security baselines (secrets, private networking, egress controls) and mandate red‑team & safetyevaluations.
Maintain compliance mappings (e.g., ISO 27001, SOC 2, GDPR/DPDP, HIPAA where applicable).
Provide CI/CD pipelines, SDKs, Helm/Terraform templates, and policy‑as‑code for consistent delivery.
Author PRR checklists, runbooks/playbooks, and DR/BCP blueprints (RTO/RPO, multi‑region/site failover).Drive enablement (trainings, brown-bags) and maintain knowledge repositories and decision records.
Partner with solution teams to validate architecture and non‑functional requirements (scale, latency, cost,safety).
Conduct Production Readiness Reviews (PRRs) and certify releases across performance, security, privacy,and compliance.
Implement observability (tracing, metrics, logs),
Key

Key Responsibilities

1. Provide architectural leadership in infrastructure management by designing and implementing scalable solutionsusing cloud platforms (AWS, Azure, GCP) and automation tools (Terraform, Ansible).
2. Define and govern infrastructure strategies by leveraging virtualization technologies (VMware, Hyper-V) tooptimize resource utilization and operational efficiency.
3. Lead technical assessments and solutioning for infrastructure modernization, utilizing monitoring andmanagement platforms (Nagios, SolarWinds) to ensure reliability and performance.
4. Drive cross-domain technical alignment by integrating security frameworks and compliance standards (ISO27001, SOC 2) into infrastructure consulting deliverables.
5. Mentor consulting teams in adopting enterprise-scale design patterns and best practices for hybrid and multi-cloud environments.
6. Collaborate with stakeholders to translate business requirements into technical roadmaps, utilizing advancedinfrastructure analytics and automation.

Skill Requirements

Other Requirements

Frequently Asked Questions

Is the salary disclosed for the COE Lead position at HCLTech?
The salary for this COE Lead role at HCLTech is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the COE Lead position at HCLTech located?
This COE Lead role at HCLTech is based in Noida, India. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
How do I apply for the COE Lead position at HCLTech?
Click the "Apply Now" button on this page. You will be redirected to HCLTech's official application portal hosted on successfactors where you can submit your application directly.
When was the COE Lead job at HCLTech posted?
This COE Lead position at HCLTech was posted on Sep 17, 2026. Apply as soon as possible — early applications are often reviewed first.

You'll be redirected to HCLTech's official application page on successfactors.