Department:Cloud Operations / AWS Operations
Location:Pune / Mumbai / Bangalore
Experience:2–5 Years
Employment Type:Full-Time
Shift:Rotational / 24×7 Managed Services
Reporting To:Lead – Cloud Support / Cloud Operations Manager
Role Overview:
- As a Cloud Engineer – L2 at Cloud.in, you will be responsible for managing, troubleshooting, optimizing and securing AWS cloud infrastructure for customer environments.
- The role requires hands-on experience across AWS infrastructure, networking, compute, storage, databases, security, monitoring, automation, backup/DR and cloud operations.
- The L2 Engineer will act as a technical escalation point for L1, independently troubleshoot complex incidents, perform root-cause analysis, implement approved changes and contribute to automation, migrations and infrastructure improvements.
Key ResponsibilitiesAWS Infrastructure Management:
- Manage the complete lifecycle of AWS infrastructure including provisioning, configuration, monitoring, optimization and decommissioning.
- Design and manage multi-tier AWS environments following AWS Well-Architected principles.
- Manage compute infrastructure using EC2, Auto Scaling and Load Balancers.
- Configure and manage VPCs, subnets, route tables, security groups, NACLs, VPN and VPC peering.
- Manage AWS storage services including S3 and EFS.
- Manage RDS databases including: Read replicas/ Automated backups/ Point-in-time recovery
- Security and access controls.
- Configure and manage CloudFront and Route 53.
- Manage IAM users, groups, roles and policies.
- Manage ACM certificates and certificate lifecycle.
- Manage encryption keys and key-rotation requirements.
- Configure SNS and integrations with other AWS services.
Monitoring & Incident Management:
- Monitor infrastructure availability, performance, logs, alarms and application health.
- Handle L2 escalations from L1 and independently troubleshoot complex issues.
- Perform root-cause analysis for recurring incidents and infrastructure failures.
- Troubleshoot Linux and Windows server issues.
- Troubleshoot network latency, connectivity, server crashes and application availability issues.
- Analyze CloudWatch metrics and logs to identify performance and availability issues.
- Ensure incidents are resolved within agreed SLA timelines.
- Participate in major incident management and provide technical updates to stakeholders.
- Prepare RCA reports and recommend preventive actions.
Networking & Security:
- Configure and troubleshoot VPC networking, routing, VPN, peering and connectivity.
- Manage security groups and NACLs.Troubleshoot TCP/IP, DNS, HTTP/HTTPS and network connectivity issues.
- Configure secure cloud environments based on customer/project requirements.
- Implement security best practices for AWS resources.
- Manage IAM permissions and cross-account access.Support security, compliance and audit requirements.
Backup & Disaster Recovery:
- Configure and manage AWS backup solutions.
- Define and maintain backup lifecycle policies.
- Support Disaster Recovery implementation and testing.
- Validate backup integrity and restoration procedures.
- Assist in designing resilient and fault-tolerant infrastructure.
Automation & Optimization:
- Identify repetitive operational activities and automate them.
- Use scripting such as Shell, Python or PowerShell for administration and automation.
- Work with Infrastructure as Code / CloudFormation where applicable.
- Optimize AWS infrastructure for performance, availability and cost.
- Monitor cloud consumption and identify cost-optimization opportunities.
- Recommend appropriate AWS resources based on business and technical requirements.
Migration & Projects:
- Support and execute on-premises to AWS cloud migrations.
- Participate in migration planning and execution with minimal downtime.
- Support hybrid-cloud environments where required.
- Assist with infrastructure design and implementation.
- Support cross-account resource sharing and AWS Organizations.
- Participate in new customer onboarding and cloud implementation projects.
Documentation & Process:
- Create and maintain SOPs, technical documentation, runbooks and knowledge-base articles.
- Maintain internal technical wiki and operational documentation.
- Follow ITIL-based incident, problem and change-management processes.
- Participate in CAB/change-management activities.
- Maintain code repositories and infrastructure documentation.
- Communicate technical issues, resolutions and recommendations to customers and internal stakeholders.
Must-Have Technical Skills:
- Strong hands-on AWS experience.Strong knowledge of: EC2/ VPC/ IAM/ S3/ EFS /RDS/ ALB / ELB/ Auto Scaling/ CloudWatch/ CloudFront / Route 53/ ACM/ SSM/ SNS.
- Strong understanding of AWS networking.
- Strong Linux administration skills.
- Working knowledge of Windows Server.TCP/IP, DNS, routing, firewalls, VPN and load-balancing concepts.
- Experience troubleshooting infrastructure and application availability issues.
- Experience with monitoring, logging and alerting.
- Understanding of backup and DR.Scripting / automation experience.
- Experience with incident, problem and change management.Ability to perform RCA and resolve complex technical issues.
Good-to-Have Skills:
- AWS Certified Solutions Architect – Associate/Professional.
- AWS Certified SysOps Administrator. (Mandatory )
- AWS Certified DevOps Engineer.
- AWS Certified Security Specialty.
- RHCE / equivalent Linux certification.
- Infrastructure as Code – CloudFormation / Terraform.CI/CD and DevOps tools.Docker / Kubernetes exposure.
- Experience with AWS Organizations and multi-account architecture.Hybrid-cloud experience across AWS / Azure / GCP.
- Cloud migration experience.Cloud cost optimization / FinOps exposure.Security and compliance frameworks.
Soft Skills:
- Strong analytical and problem-solving ability.
- Excellent verbal and written communication.
- Strong customer-facing skills.
- Ability to independently handle technical escalations.
- Ability to prioritize incidents based on business impact.
- Strong ownership and accountability.
- Good documentation and knowledge-sharing practices.
- Ability to mentor and support L1 engineers.
- Ability to work in a fast-paced 24×7 managed-services environment.
- Willingness to learn new AWS services and technologies.
- Strong team player with professional work ethics.
Education & Certification Preferred:
- B.E./B.Tech/B.Sc./BCA/MCA or equivalent technical qualification.
- AWS certification strongly preferred.
- Linux/Windows administration certification preferred.
Experience:
- 2–5 years of hands-on experience in AWS Cloud Operations, Cloud Support, System Administration, Infrastructure Management or Managed Services.
- Candidates should have practical experience managing production or mission-critical infrastructure.Key Performance Expectations.
- SLA adherence for incidents, service requests and changes.
- Effective resolution of L2 escalations.
- Reduction in repeat incidents through RCA and preventive actions.
- Infrastructure availability and performance.
- Successful implementation of approved changes.
- Backup and DR compliance.
- Cloud security and operational compliance.
- Cost optimization initiatives.
- Automation of repetitive operational activities.
- Quality and completeness of technical documentation.
- Customer satisfaction and communication.
- Mentoring and technical support to L1 engineers.
Work Environment:
- This is a customer-facing AWS Managed Services role supporting production and mission-critical environments.
- The position may require rotational shifts, 24×7 support, weekend/holiday coverage and on-call support for critical incidents and planned maintenance activities.