Skip to main content
H

hostinservices

Cloud Engineer - L2

Pune2-5 yrsPosted 3 days ago

Skills

AWSLinuxPythonDevOpsTerraformCI/CDDockerKubernetesAzureGoogle CloudAnalytical SkillsProblem Solving

Job description

Department: Cloud Operations / AWS Operations

Location: Pune / Mumbai / Bangalore

Experience: 2–5 Years

Employment Type: Full-Time

Shift: Rotational / 24×7 Managed Services

Reporting To: Lead – Cloud Support / Cloud Operations Manager

Role Overview:

As a Cloud Engineer – L2 at Cloud.in, you will be responsible for managing, troubleshooting, optimizing and securing AWS cloud infrastructure for customer environments.

The role requires hands-on experience across AWS infrastructure, networking, compute, storage, databases, security, monitoring, automation, backup/DR and cloud operations.

The L2 Engineer will act as a technical escalation point for L1, independently troubleshoot complex incidents, perform root-cause analysis, implement approved changes and contribute to automation, migrations and infrastructure improvements.

Key ResponsibilitiesAWS Infrastructure Management:

Manage the complete lifecycle of AWS infrastructure including provisioning, configuration, monitoring, optimization and decommissioning.

Design and manage multi-tier AWS environments following AWS Well-Architected principles.

Manage compute infrastructure using EC2, Auto Scaling and Load Balancers.

Configure and manage VPCs, subnets, route tables, security groups, NACLs, VPN and VPC peering.

Manage AWS storage services including S3 and EFS.

Manage RDS databases including: Read replicas/ Automated backups/ Point-in-time recovery

Security and access controls.

Configure and manage CloudFront and Route 53.

Manage IAM users, groups, roles and policies.

Manage ACM certificates and certificate lifecycle.

Manage encryption keys and key-rotation requirements.

Configure SNS and integrations with other AWS services.

Monitoring & Incident Management:

Monitor infrastructure availability, performance, logs, alarms and application health.

Handle L2 escalations from L1 and independently troubleshoot complex issues.

Perform root-cause analysis for recurring incidents and infrastructure failures.

Troubleshoot Linux and Windows server issues.

Troubleshoot network latency, connectivity, server crashes and application availability issues.

Analyze CloudWatch metrics and logs to identify performance and availability issues.

Ensure incidents are resolved within agreed SLA timelines.

Participate in major incident management and provide technical updates to stakeholders.

Prepare RCA reports and recommend preventive actions.

Networking & Security:

Configure and troubleshoot VPC networking, routing, VPN, peering and connectivity.

Manage security groups and NACLs.Troubleshoot TCP/IP, DNS, HTTP/HTTPS and network connectivity issues.

Configure secure cloud environments based on customer/project requirements.

Implement security best practices for AWS resources.

Manage IAM permissions and cross-account access.Support security, compliance and audit requirements.

Backup & Disaster Recovery:

Configure and manage AWS backup solutions.

Define and maintain backup lifecycle policies.

Support Disaster Recovery implementation and testing.

Validate backup integrity and restoration procedures.

Assist in designing resilient and fault-tolerant infrastructure.

Automation & Optimization:

Identify repetitive operational activities and automate them.

Use scripting such as Shell, Python or PowerShell for administration and automation.

Work with Infrastructure as Code / CloudFormation where applicable.

Optimize AWS infrastructure for performance, availability and cost.

Monitor cloud consumption and identify cost-optimization opportunities.

Recommend appropriate AWS resources based on business and technical requirements.

Migration & Projects:

Support and execute on-premises to AWS cloud migrations.

Participate in migration planning and execution with minimal downtime.

Support hybrid-cloud environments where required.

Assist with infrastructure design and implementation.

Support cross-account resource sharing and AWS Organizations.

Participate in new customer onboarding and cloud implementation projects.

Documentation & Process:

Create and maintain SOPs, technical documentation, runbooks and knowledge-base articles.

Maintain internal technical wiki and operational documentation.

Follow ITIL-based incident, problem and change-management processes.

Participate in CAB/change-management activities.

Maintain code repositories and infrastructure documentation.

Communicate technical issues, resolutions and recommendations to customers and internal stakeholders.

Must-Have Technical Skills:

Strong hands-on AWS experience.Strong knowledge of: EC2/ VPC/ IAM/ S3/ EFS /RDS/ ALB / ELB/ Auto Scaling/ CloudWatch/ CloudFront / Route 53/ ACM/ SSM/ SNS.

Strong understanding of AWS networking.

Strong Linux administration skills.

Working knowledge of Windows Server.TCP/IP, DNS, routing, firewalls, VPN and load-balancing concepts.

Experience troubleshooting infrastructure and application availability issues.

Experience with monitoring, logging and alerting.

Understanding of backup and DR.Scripting / automation experience.

Experience with incident, problem and change management.Ability to perform RCA and resolve complex technical issues.

Good-to-Have Skills:

AWS Certified Solutions Architect – Associate/Professional.

AWS Certified SysOps Administrator. (Mandatory )

AWS Certified DevOps Engineer.

AWS Certified Security Specialty.

RHCE / equivalent Linux certification.

Infrastructure as Code – CloudFormation / Terraform.CI/CD and DevOps tools.Docker / Kubernetes exposure.

Experience with AWS Organizations and multi-account architecture.Hybrid-cloud experience across AWS / Azure / GCP.

Cloud migration experience.Cloud cost optimization / FinOps exposure.Security and compliance frameworks.

Soft Skills:

Strong analytical and problem-solving ability.

Excellent verbal and written communication.

Strong customer-facing skills.

Ability to independently handle technical escalations.

Ability to prioritize incidents based on business impact.

Strong ownership and accountability.

Good documentation and knowledge-sharing practices.

Ability to mentor and support L1 engineers.

Ability to work in a fast-paced 24×7 managed-services environment.

Willingness to learn new AWS services and technologies.

Strong team player with professional work ethics.

Education & Certification Preferred:

B.E./B.Tech/B.Sc./BCA/MCA or equivalent technical qualification.

AWS certification strongly preferred.

Linux/Windows administration certification preferred.

Experience

2–5 years of hands-on experience in AWS Cloud Operations, Cloud Support, System Administration, Infrastructure Management or Managed Services.

Candidates should have practical experience managing production or mission-critical infrastructure.Key Performance Expectations.

SLA adherence for incidents, service requests and changes.

Effective resolution of L2 escalations.

Reduction in repeat incidents through RCA and preventive actions.

Infrastructure availability and performance.

Successful implementation of approved changes.

Backup and DR compliance.

Cloud security and operational compliance.

Cost optimization initiatives.

Automation of repetitive operational activities.

Quality and completeness of technical documentation.

Customer satisfaction and communication.

Mentoring and technical support to L1 engineers.

Work Environment:

This is a customer-facing AWS Managed Services role supporting production and mission-critical environments.

The position may require rotational shifts, 24×7 support, weekend/holiday coverage and on-call support for critical incidents and planned maintenance activities.

Apply on hostinservices