IND Staff Engineer, Reliability

thehartford· Hartford Global Services Private Limited
Apply Now ↗
📍 India GCC-Puppalaguda VillageFull time

About this role

IND Staff Engineer, Reliability - GCC070

We’re determined to make a difference and are proud to be an insurance company that goes well beyond coverages and policies. Working here means having every opportunity to achieve your goals – and to help others accomplish theirs, too. Join our team as we help shape the future.

Key Responsibilities

·       Partner with Enterprise governors to ascertain key reliability, security, and resilience requirements set by The Hartford and bring those requirements into the Platform team for implementation

·       Patternize resilience capabilities into useful tools, services, and products to be used by customers to ensure users of the Platform build fault-tolerant systems

·       Develop governing frameworks to ensure each release is compliant with the standards we expect

·       Liaise with key business and technical customers to understand predictive applications and their infrastructure. Through this consultation, you would be working with them to build resilience and reliability capabilities into their application.

·       Drive IM, cloud ops, and RE efforts across the platform by applying industry best practices and maturing existing the problem management lifecycle by building standards and contributing to runbooks, standard operating procedures, and incident management lifecycle

·       Performance engineering of deployed analytics and AI solutions across the portfolio to ascertain enhancement opportunities

Required Skills & Experience:

  • Experience programming in Python to build automation tools, operational scripts, and platform support capabilities, including infrastructure and reliability automation.
  • Experience using Infrastructure as Code to provision and manage cloud environments, including Terraform and/or CloudFormation, with a focus on repeatability, security, and scalability.
  • Experience deploying and operating systems on public cloud platforms such as AWS and/or Google Cloud Platform, including familiarity with serverless architectures, multi-region deployments, and recoverability strategies.
  • Experience designing and operationalizing resilience, reliability, and disaster recovery capabilities for distributed systems and ML/AI platforms, including performance engineering and fault-tolerant system design.
  • Experience building and maintaining CI/CD pipelines using tools such as GitHub and Jenkins, including embedding security checks, compliance gates, and automated validation into deployment workflows.
  • Experience applying core reliability engineering concepts, including authoring runbooks, operational guides, and automation to support resilient platform operations.
  • Experience designing observability solutions, including logging, monitoring, and alerting using tools such as Splunk, with dashboards and metrics that surface service health, SLO/SLA adherence, and early-warning signals for ML and data workloads.
  • Hands-on experience with incident and problem management practices, including ITIL-based processes, postmortems, and blameless root cause analysis, as well as disaster recovery planning, failover testing, and resilience frameworks such as FMEA.
  • Foundational knowledge of networking fundamentals and operations architecture to support IT service management (ITSM) automation and distributed system reliability.
  • Familiarity with relational databases such as Snowflake or other RDBMS platforms, with an understanding of data reliability, availability, and consistency requirements in analytics and ML environments.

 

Nice to Have

  • Knowledge of NIST 800-171
  • Working in Agile / a consultative mindset – we function as an agile team, which places a responsibility and accountability on developers to take ownership of their work.
    • We agree on acceptance criteria as a goal, and work to find the best ways to implement a solution which meets those criteria. Seldom do our stories define precisely how to build a solution giving space for engineers to design together and implement the best possible solutions
  • A mind for resilience – not necessarily a security engineer, but someone keen on establishing best practices, following those practices, and fostering adoption of those practices at all levels
  • Exceptional Communication and collaboration skills – we never work alone, and often work with customers across DS&A and the Enterprise.

About Us | Our Culture | What It’s Like to Work Here

Frequently Asked Questions

Is the salary disclosed for the IND Staff Engineer, Reliability position at thehartford?
The salary for this IND Staff Engineer, Reliability role at thehartford is not publicly listed. Click "Apply Now" to learn more about the compensation package on their official careers page.
Where is the IND Staff Engineer, Reliability position at thehartford located?
This IND Staff Engineer, Reliability role at thehartford is based in India GCC-Puppalaguda Village. The position is listed as on-site or hybrid. Check the full job description or apply directly to confirm the work arrangement.
Is the IND Staff Engineer, Reliability role at thehartford full-time or part-time?
This is listed as a Full time position. It is posted as a IND Staff Engineer, Reliability role in the Hartford Global Services Private Limited department at thehartford.
Which team or department does the IND Staff Engineer, Reliability at thehartford belong to?
This IND Staff Engineer, Reliability position is part of the Hartford Global Services Private Limited department at thehartford. See the full job description for more information about the team structure and responsibilities.
How do I apply for the IND Staff Engineer, Reliability position at thehartford?
Click the "Apply Now" button on this page. You will be redirected to thehartford's official application portal hosted on workday where you can submit your application directly.
When was the IND Staff Engineer, Reliability job at thehartford posted?
This IND Staff Engineer, Reliability position at thehartford was posted on Oct 1, 2026. Apply as soon as possible — early applications are often reviewed first.
IND Staff Engineer, Reliability
thehartford
Apply for this role ↗

You'll be redirected to thehartford's official application page on Workday.